Commits · 28fa014a1fd214ffdbac3fb76ae87a6c71f3f99d · chenpangpang / transformers

07 Dec, 2020 1 commit
- Use word_ids to get labels in run_ner (#8962) · 7f9ccffc
  Sylvain Gugger authored Dec 07, 2020
```
* Use word_ids to get labels in run_ner

* Add sanity check
```
  7f9ccffc
05 Dec, 2020 1 commit

Don't pass in token_type_ids to BART for GLUE (#8929) · 8dfc8c72

Ethan Perez authored Dec 05, 2020

Without this fix, training a `BARTForSequenceClassification` model with `run_pl_glue.py` gives `TypeError: forward() got an unexpected keyword argument 'token_type_ids'`, because BART does not have token_type_ids. I've solved this issue in the same way as it's solved for the "distilbert" model, and I can train BART models on SNLI without errors now.

8dfc8c72

04 Dec, 2020 2 commits

[seq2seq] document the caveat of leaky native amp (#8930) · df311a5c

Stas Bekman authored Dec 04, 2020



* document the caveat of leaky native amp

* Update examples/seq2seq/README.md
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

df311a5c

[s2s finetune_trainer] add instructions for distributed training (#8884) · 4c3d98dd
Stas Bekman authored Dec 03, 2020

4c3d98dd

01 Dec, 2020 1 commit
- start using training_args.parallel_mode (#8882) · 379005c9
  Stas Bekman authored Dec 01, 2020
  
  379005c9
30 Nov, 2020 3 commits

[s2s trainer] fix DP mode (#8823) · 7f34d757

Stas Bekman authored Nov 30, 2020

* fix DP case on multi-gpu

* make executable

* test all 3 modes

* use the correct check for distributed

* dp doesn't need a special case

* restore original name

* cleanup

7f34d757

Remove deprecated `evalutate_during_training` (#8852) · 55302990

Sylvain Gugger authored Nov 30, 2020



* Remove deprecated `evalutate_during_training`

* Update src/transformers/training_args_tf.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

55302990

token-classification: use is_world_process_zero instead of deprecated is_world_master() (#8828) · 19fa01ce
Stefan Schweter authored Nov 30, 2020

19fa01ce

26 Nov, 2020 4 commits
- potpurri of small fixes (#8807) · ddf3c646
  Stas Bekman authored Nov 26, 2020
  
  ddf3c646
- Fix PPLM (#8779) · 52708d26
  chutaklee authored Nov 27, 2020
```
* Fix pplm

* fix style

* make style
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
```
  52708d26
- Revert "finetune.py: specifying generation min_length (#8478)" (#8805) · 8f07f5c4
  Patrick von Platen authored Nov 26, 2020
```
This reverts commit 5aa361f3.
```
  8f07f5c4
- finetune.py: specifying generation min_length (#8478) · 5aa361f3
  Daniel Khashabi authored Nov 25, 2020
  
  5aa361f3
24 Nov, 2020 3 commits

[core] implement support for run-time dependency version checking (#8645) · 82d443a7

Stas Bekman authored Nov 24, 2020



* implement support for run-time dependency version checking

* try not escaping !

* use findall that works on py36

* small tweaks

* autoformatter worship

* simplify

* shorter names

* add support for non-versioned checks

* add deps

* revert

* tokenizers not required, check version only if installed

* make a proper distutils cmd and add make target

* tqdm must be checked before tokenizers

* workaround the DistributionNotFound peculiar setup

* handle the rest of packages in setup.py

* fully sync setup.py's install_requires - to check them all

* nit

* make install_requires more readable

* typo

* Update setup.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* restyle

* add types

* simplify

* simplify2
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

82d443a7

fix rag index names in eval_rag.py example (#8730) · a7d73cfd
Quentin Lhoest authored Nov 24, 2020

a7d73cfd

Support various BERT relative position embeddings (2nd) (#8276) · 2c83b3c3

zhiheng-huang authored Nov 24, 2020



* Support BERT relative position embeddings

* Fix typo in README.md

* Address review comment

* Fix failing tests

* [tiny] Fix style_doc.py check by adding an empty line to configuration_bert.py

* make fix copies

* fix configs of electra and albert and fix longformer

* remove copy statement from longformer

* fix albert

* fix electra

* Add bert variants forward tests for various position embeddings

* [tiny] Fix style for test_modeling_bert.py

* improve docstring

* [tiny] improve docstring and remove unnecessary dependency

* [tiny] Remove unused import

* re-add to ALBERT

* make embeddings work for ALBERT

* add test for albert
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

2c83b3c3

23 Nov, 2020 2 commits
- Fix max length in run_plm script (#8738) · 367f497d
  Sylvain Gugger authored Nov 23, 2020
  
  367f497d
- [trainer] make generate work with multigpu (#8716) · 1e45bef0
  Stas Bekman authored Nov 23, 2020
```
* make generate work with multigpu

* better fix - thanks @sgugger
```
  1e45bef0
22 Nov, 2020 1 commit
- Fix many typos (#8708) · e1f3156b
  Santiago Castro authored Nov 22, 2020
  
  e1f3156b
20 Nov, 2020 1 commit

Fix rag finetuning + add finetuning test (#8585) · 8062fa63

Quentin Lhoest authored Nov 20, 2020

* replace init_ddp_connection for index init

* style

* add finetune test

* add test data

* move generate tensors to device

* add test on EM metric

* style

* allow multi process test

* keep gloo process group for retrieval

* add multi-gpu test

* use custom accelerator

* clean test finetune

* minor

* style

* style

* typo

* use python call instead of imported main fumction

* return_dict fix in modeling_rag

* use float32 in retrieval

* store as float32 as well in the custom knowledge dataset example

* style

* rename to finetune_rag

* style

* update readme

* rename utils and callbacks to utils_rag and callbacks_rag

* fix test

* patrick's comments

* generate dummy data in the finetue test script

* remove dummy data files

* style

8062fa63

19 Nov, 2020 6 commits
- [examples/seq2seq] fix PL deprecation warning (#8577) · 0ad45e10
  Stas Bekman authored Nov 19, 2020
```
* fix deprecation warning

* fix
```
  0ad45e10
- Fix run_ner script (#8664) · 20b65860
  Sylvain Gugger authored Nov 19, 2020
```
* Fix run_ner script

* Pin datasets
```
  20b65860
- Fix a few last paths for the new repo org (#8666) · cb3e5c33
  Sylvain Gugger authored Nov 19, 2020
  
  cb3e5c33
- fix small typo (#8644) · a79a96dd
  Matthias authored Nov 19, 2020
```
Fixed a small typo on the XLNet and permutation language modelling section
```
  a79a96dd
- Better filtering of the model outputs in Trainer (#8633) · 4208f496
  Sylvain Gugger authored Nov 19, 2020
```
* Better filtering of the model outputs in Trainer

* Fix examples tests

* Add test for Lysandre
```
  4208f496
- fix missing return dict (#8653) · 62cd9ce9
  Quentin Lhoest authored Nov 19, 2020
  
  62cd9ce9
18 Nov, 2020 5 commits
- Update README.md (#8635) · 28d16e7a
  Tim Isbister authored Nov 19, 2020
  
  28d16e7a
- [s2s] distillation apex breaks return_dict obj (#8631) · d86d57fa
  Stas Bekman authored Nov 18, 2020
```
* apex breaks return_dict obj

* style
```
  d86d57fa
- Fix training from scratch in new scripts (#8623) · a0c62d24
  Sylvain Gugger authored Nov 18, 2020
  
  a0c62d24
- fix to adjust for #8530 changes (#8612) · cdf1b7ae
  Stas Bekman authored Nov 18, 2020
  
  cdf1b7ae
- [s2s] broken test (#8613) · 2819da02
  Stas Bekman authored Nov 18, 2020
  
  2819da02
17 Nov, 2020 4 commits

Remove deprecated (#8604) · dd52804f

Sylvain Gugger authored Nov 17, 2020



* Remove old deprecated arguments
Co-authored-by: LysandreJik <lysandre.debut@reseau.eseo.fr>

* Remove needless imports

* Fix tests
Co-authored-by: LysandreJik <lysandre.debut@reseau.eseo.fr>

dd52804f

these should run fine on multi-gpu (#8582) · f0435f5a
Stas Bekman authored Nov 17, 2020

f0435f5a

Tokenizers: ability to load from model subfolder (#8586) · 042a6aa7

Julien Chaumond authored Nov 17, 2020



* <small>tiny typo</small>

* Tokenizers: ability to load from model subfolder

* use subfolder for local files as well

* Uniformize model shortcut name => model id

* from s3 => from huggingface.co
Co-authored-by: Quentin Lhoest <lhoest.q@gmail.com>

042a6aa7

Reorganize repo (#8580) · c89bdfbe

Sylvain Gugger authored Nov 16, 2020

* Put models in subfolders

* Styling

* Fix imports in tests

* More fixes in test imports

* Sneaky hidden imports

* Fix imports in doc files

* More sneaky imports

* Finish fixing tests

* Fix examples

* Fix path for copies

* More fixes for examples

* Fix dummy files

* More fixes for example

* More model import fixes

* Is this why you're unhappy GitHub?

* Fix imports in conver command

c89bdfbe

16 Nov, 2020 1 commit

Switch `return_dict` to `True` by default. (#8530) · 1073a2bd

Sylvain Gugger authored Nov 16, 2020

* Use the CI to identify failing tests

* Remove from all examples and tests

* More default switch

* Fixes

* More test fixes

* More fixes

* Last fixes hopefully

* Use the CI to identify failing tests

* Remove from all examples and tests

* More default switch

* Fixes

* More test fixes

* More fixes

* Last fixes hopefully

* Run on the real suite

* Fix slow tests

1073a2bd

15 Nov, 2020 1 commit

[breaking|pipelines|tokenizers] Adding slow-fast tokenizers equivalence tests... · f4e04cd2

Thomas Wolf authored Nov 15, 2020


[breaking|pipelines|tokenizers] Adding slow-fast tokenizers equivalence tests pipelines - Removing sentencepiece as a required dependency (#8073)

* Fixing roberta for slow-fast tests

* WIP getting equivalence on pipelines

* slow-to-fast equivalence - working on question-answering pipeline

* optional FAISS tests

* Pipeline Q&A

* Move pipeline tests to their own test job again

* update tokenizer to add sequence id methods

* update to tokenizers 0.9.4

* set sentencepiecce as optional

* clean up squad

* clean up pipelines to use sequence_ids

* style/quality

* wording

* Switch to use_fast = True by default

* update tests for use_fast at True by default

* fix rag tokenizer test

* removing protobuf from required dependencies

* fix NER test for use_fast = True by default

* fixing example tests (Q&A examples use slow tokenizers for now)

* protobuf in main deps extras["sentencepiece"] and example deps

* fix protobug install test

* try to fix seq2seq by switching to slow tokenizers for now

* Update src/transformers/tokenization_utils_base.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/tokenization_utils_base.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

f4e04cd2

12 Nov, 2020 2 commits
- Try to understand and apply Sylvain's comments (#8458) · 27b3ff31
  Julien Plu authored Nov 12, 2020
  
  27b3ff31
- quick fix on concatenating text to support more datasets (#8474) · 924c624a
  zeyuyun1 authored Nov 12, 2020
  
  924c624a
11 Nov, 2020 2 commits

[s2s] distill t5-large -> t5-small (#8376) · 81ebd706
Sumithra Bhakthavatsalam authored Nov 11, 2020
```
Co-authored-by: Sam Shleifer <sshleifer@gmail.com>
```
81ebd706

Example NER script predicts on tokenized dataset (#8468) · a38d1c7c

sarnoult authored Nov 11, 2020

The new run_ner.py script tries to run prediction on the input
test set `datasets["test"]`, but it should be the tokenized set
`tokenized_datasets["test"]`

a38d1c7c