Commits · d160782a53a226b8f12815cdfa7a3cd29544fd38 · chenpangpang / transformers

01 Sep, 2021 1 commit

Add template for adding flax models (#12441) · d160782a

Jonathan Chang authored Sep 01, 2021



* Add option to add flax

* Add flax template for __init__.py

* Add flax template for .rst

* Copy TF modeling template

* Add a missing line in modeling_tf_... template

* Update first half of modeling_flax_..

* Update encoder flax template

* Copy test_modeling_tf... as test_modeling_flax...

* Replace some TF to Flax in test_modeling_flax_...

* Replace tf to np

some function might not work, like _assert_tensors_equal

* Replace remaining tf to np (might not work)

* Fix cookiecutter

* Add Flax in to_replace_... template

* Update transformers-cli add-new-model

* Save generate_flax in configuration.json

This will be read by transformers-cli

* Fix to_replace_... and cli

* Fix replace cli

* Fix cookiecutter name

* Move docstring earlier to avoid not defined error

* Fix a missing Module

* Add encoder-decoder flax template from bart

* Fix flax test

* Make style

* Fix endif

* Fix replace all "utf-8 -> unp-8"

* Update comment

* Fix flax template (add missing ..._DOCSTRING)

* Use flax_bart imports in template (was t5)

* Fix unp

* Update templates/adding_a_new_model/tests

* Revert "Fix unp"

This reverts commit dc9002a41d902c4f9b07343eab1cb350c8b7fd57.

* Remove one line of copied from to suppress CI error

* Use generate_tensorflow_pytorch_and_flax

* Add a missing part

* fix typo

* fix flax config

* add examples for flax

* small rename

* correct modeling imports

* correct auto loading

* corrects some flax tests

* correct small typo

* correct as type

* finish modif

* correct more templates

* final fixes

* add file testers

* up

* make sure tests match template regex

* correct pytorch

* correct tf

* correct more tf

* correct imports

* minor error

* minor error

* correct init

* more fixes

* correct more flax tests

* correct flax test

* more fixes

* correct docs

* update

* fix
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

d160782a

31 Aug, 2021 1 commit

Set missing seq_length variable when using inputs_embeds with ALBERT & Remove... · ef8d6f2b

Jongheon Kim authored Aug 31, 2021

Set missing seq_length variable when using inputs_embeds with ALBERT & Remove code duplication (#13152)

* Set seq_length variable when using inputs_embeds

* remove code duplication

ef8d6f2b

19 Aug, 2021 1 commit
- Update namespaces inside torch.utils.data to the latest. (#13167) · 91ff480e
  Allan Lin authored Aug 19, 2021
```
* Update torch.utils.data namespaces to the latest.

* Format

* Update Dataloader.

* Style
```
  91ff480e
06 Aug, 2021 1 commit

[WIP] Disentangle auto modules from other modeling files (#13023) · 9870093f

Sylvain Gugger authored Aug 06, 2021

* Initial work

* All auto models

* All tf auto models

* All flax auto models

* Tokenizers

* Add feature extractors

* Fix typos

* Fix other typo

* Use the right config

* Remove old mapping names and update logic in AutoTokenizer

* Update check_table

* Fix copies and check_repo script

* Fix last test

* Add back name

* clean up

* Update template

* Update template

* Forgot a )

* Use alternative to fixup

* Fix TF model template

* Address review comments

* Address review comments

* Style

9870093f

03 Aug, 2021 1 commit
- Fix template for inputs docstrings (#12976) · 790f1c95
  Sylvain Gugger authored Aug 03, 2021
  
  790f1c95
21 Jul, 2021 1 commit
- Expose get_config() on ModelTesters (#12812) · c3d9ac76
  Lysandre Debut authored Jul 21, 2021
```
* Expose get_config() on ModelTesters

* Typo
```
  c3d9ac76
08 Jul, 2021 1 commit

Init pickle (#12567) · 0a6b9048

Sylvain Gugger authored Jul 08, 2021

* Try to pickle transformers

* Deal with special objs better

* Make picklable

0a6b9048

07 Jul, 2021 1 commit

Remove tf.roll wherever not needed (#12512) · 0d2bffad

Michal Szutenberg authored Jul 07, 2021

It was used in shift_right.
After this change TF code is more similar to Pytorch implementations
Also, TF graphs are optimized (one node less)

0d2bffad

28 Jun, 2021 1 commit

[Examples] Added context manager to datasets map (#12367) · 04dbea31

Bhadresh Savani authored Jun 28, 2021

* added cotext manager to datasets map

* fixed style and spaces

* fixed warning of deprecation

* changed desc

04dbea31

26 Jun, 2021 1 commit
- updated example template (#12365) · 9a754594
  Bhadresh Savani authored Jun 26, 2021
  
  9a754594
25 Jun, 2021 1 commit
- remove extra white space from log format (#12360) · 4a872cae
  Stas Bekman authored Jun 25, 2021
  
  4a872cae
23 Jun, 2021 1 commit
- Add all XxxPreTrainedModel to the main init (#12314) · 9eda6b52
  Sylvain Gugger authored Jun 23, 2021
```
* Add all XxxPreTrainedModel to the main init

* Add to template

* Add to template bis

* Add FlaxT5
```
  9eda6b52
22 Jun, 2021 1 commit

Fix for the issue of device-id getting hardcoded for token_type_ids during Tracing [WIP] (#11252) · af6e01c5

Hamid Shojanazeri authored Jun 22, 2021



* registering a buffer for token_type_ids, to pass the error of device-id getting hardcoded when tracing

* sytle format

* adding persistent flag to the resgitered buffers that prevent from adding them to the state_dict and addresses the Backward compatibility issue

* adding the try catch to the fix as persistent flag is only available from PT >1.6

* adding version check

* added the condition to only use the token_type_ids buffer when its autogenerated not passed by user

* adding comments and making the conidtion where token_type_ids are None to use the registered buffer

* taking out position-embeddding from the if block

* adding comments

* handling the case if buffer for position_ids was not registered

* reverted the changes on position_ids, fix the issue with size of token_type_ids buffer, moved the modification for generated token_type_ids to Bertmodel, instead of Embeddings

* reverting the token_type_ids in case of None to the previous version

* reverting changes on position_ids adding back the if block

* changes added by running make fix-copies

* changes added by running make fix-copies and added the import version as it was getting used

* changes added by running make fix-copies

* changes added by running make fix-copies

* fixing the import format

* fixing the import format

* modified to use temp tensor for trimed and expanded token_type_ids buffer

* changes made by fix-copies after temp tensor modifications

* changes made by fix-copies after temp tensor modifications

* changes made by fix-copies after temp tensor modifications

* clean up

* clean up

* clean up

* clean up

* Nit

* Nit

* Nit

* modified according to support device conversion on traced models

* modified according to support device conversion on traced models

* modified according to support device conversion on traced models

* modified according to support device conversion on traced models

* changes based on latest in master

* Adapt templates

* Add version import
Co-authored-by: Ubuntu <ubuntu@ip-172-31-32-81.us-west-2.compute.internal>
Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>

af6e01c5

18 Jun, 2021 1 commit
- Depreciate pythonic Mish and support PyTorch 1.9 version of Mish (#12240) · f3558bbc
  Xa9aX ツ authored Jun 18, 2021
```
* Moved Mish to Torch 1.9 version

* Run black formatting
```
  f3558bbc
14 Jun, 2021 1 commit
- consistent nn. and nn.functional: p2 templates (#12153) · a156da9a
  Stas Bekman authored Jun 14, 2021
  
  a156da9a
07 Jun, 2021 2 commits

Fixes bug that appears when using QA bert and distilation. (#12026) · f8bd8c6c

François Lagunas authored Jun 07, 2021

* Fixing bug that appears when using distilation (and potentially other uses).
During backward pass Pytorch complains with:
RuntimeError: one of the variables needed for gradient computation has been modified by an inplace operation
This happens because the QA model code modifies the start_positions and end_positions input tensors, using clamp_ function: as a consequence the teacher and the student both modifies the inputs, and backward pass fails.

* Fixing all models QA clamp_ bug.

f8bd8c6c

fix docs of past_key_values (#12049) · 185122ef
Suraj Patil authored Jun 07, 2021

185122ef

25 May, 2021 1 commit
- Add option to log only once in multinode training (#11819) · f086652b
  Sylvain Gugger authored May 25, 2021
```
* Add option to long only once in multinode training

* Use an alternate property
```
  f086652b
06 May, 2021 1 commit
- Re-styling in seq2seq attention (#11613) · 7eee950a
  Sylvain Gugger authored May 06, 2021
  
  7eee950a
26 Apr, 2021 2 commits

[Examples] Fixes inconsistency around eval vs val and predict vs test (#11380) · 1d30ec95

Bhadresh Savani authored Apr 26, 2021

* added changes for uniformity

* modified files

* corrected typo

* fixed qa scripts

* fix typos

* fixed predict typo in qa no trainer

* fixed test file

* reverted trainer changes

* reverted trainer changes in custom exmaples

* updated readme

* added changes in deepspeed test

* added changes for predict and eval

1d30ec95

TF BART models - Add `cross_attentions` to model output and fix... · 38a716cd

Daniel Stancl authored Apr 26, 2021

TF BART models - Add `cross_attentions` to model output and fix cross-attention head masking (#10699)

* Add cross_attn_head_mask to BART

* Fix cross_attentions in TFBart-like models

* This commit enables returning of `cross_attentions`
for TFBart-like models

* It also fixes attention head masking in cross-attenion module

* Update TF model templates

* Fix missing , in TF model templates

* Fix typo: congig -> config

38a716cd

23 Apr, 2021 1 commit

Fix cross-attention head mask for Torch encoder-decoder models (#10605) · e3ff165a

Daniel Stancl authored Apr 23, 2021

* Fix cross-attention head mask for Torch BART models

* Fix head masking for cross-attention module for the following
models: BART, Blenderbot, Blenderbot_small, M2M_100, Marian, MBart,
Pegasus

* Enable test_headmasking for M2M_100 model

* Fix cross_head_mask for FSMT, LED and T5

* This commit fixes `head_mask` for cross-attention modules
in the following models: FSMT, LED, T5

* It also contains some smaller changes in doc so that
it is be perfectly clear the shape of `cross_head_mask`
is the same as of `decoder_head_mask`

* Update template

* Fix template for BartForCausalLM

* Fix cross_head_mask for Speech2Text models

* Fix cross_head_mask in templates

* Fix args order in BartForCausalLM template

* Fix doc in BART templates

* Make more explicit naming

* `cross_head_mask` -> `cross_attn_head_mask`

* `cross_layer_head_mask` -> `cross_attn_layer_head_mask`

* Fix doc

* make style quality

* Fix speech2text docstring

e3ff165a

21 Apr, 2021 1 commit
- Honor contributors to models (#11329) · 74712e22
  Sylvain Gugger authored Apr 21, 2021
```
* Honor contributors to models

* Fix typo

* Address review comments

* Add more authors
```
  74712e22
09 Apr, 2021 1 commit
- Make `get_special_tokens_mask` consider all tokens (#11163) · 45fc8c79
  Sylvain Gugger authored Apr 09, 2021
  
  45fc8c79
07 Apr, 2021 1 commit
- fix: The 'warn' method is deprecated (#11105) · c9035e45
  Stas Bekman authored Apr 07, 2021
```
* The 'warn' method is deprecated

* fix test
```
  c9035e45
31 Mar, 2021 1 commit

Enforce string-formatting with f-strings (#10980) · acc3bd9d

Sylvain Gugger authored Mar 31, 2021



* First third

* Styling and fix mistake

* Quality

* All the rest

* Treat %s and %d

* typo

* Missing )

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

acc3bd9d

29 Mar, 2021 1 commit

Fixes in the templates (#10951) · 700229f8

Sylvain Gugger authored Mar 29, 2021



* Fixes in the templates

* Define in all cases

* Dimensionality -> Dimension
Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>

700229f8

23 Mar, 2021 2 commits

[Examples] Added predict stage and Updated Example Template (#10868) · 7ef40120

Bhadresh Savani authored Mar 23, 2021



* added predict stage

* added test keyword in exception message

* removed example specific saving predictions

* fixed f-string error

* removed extra line
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

7ef40120

Update the example template for a no Trainer option (#10865) · bf1f43fb
Sylvain Gugger authored Mar 23, 2021

bf1f43fb

10 Mar, 2021 1 commit

Copy tokenizer files in each of their repo (#10624) · 2295d783

Sylvain Gugger authored Mar 10, 2021

* Move tokenizer files in each repo

* Fix mBART50 tests

* Fix mBART tests

* Fix Marian tests

* Update templates

2295d783

09 Mar, 2021 1 commit
- added max_sample args and metrics changes (#10602) · ac17f711
  Bhadresh Savani authored Mar 09, 2021
  
  ac17f711
05 Mar, 2021 1 commit

Fix embeddings for PyTorch 1.8 (#10549) · 7da995c0

Sylvain Gugger authored Mar 05, 2021

* Fix embeddings for PyTorch 1.8

* Try with PyTorch 1.8.0

* Fix embeddings init

* Fix copies

* Typo

* More typos

7da995c0

03 Mar, 2021 1 commit
- [T5] Fix speed degradation bug t5 (#10496) · 2d2ed2cc
  Patrick von Platen authored Mar 03, 2021
```
* fix speed degradation bug t5

* fix for all models

* fix code quality
```
  2d2ed2cc
25 Feb, 2021 1 commit

Bugfix: Removal of padding_idx in BartLearnedPositionalEmbedding (#10200) · 894db670

mingruimingrui authored Feb 25, 2021



* Assumption of padding_idx <2 might not stand

* Use offset instead of 2

* Fix with black

* Change behavior to warning instead for backward compatibility.

* Fix with black

* Remove warning

* Make padding_idx non-required

* padding_idx fix for blenderbot

* padding_idx fix for blenderbot_small

* padding_idx fix for led

* padding_idx fix for mbart

* Remove extra whitespaces

* padding_idx fix for template

* Fix padding_idx passed to nn.Embedding mistake

* Fixed padding_idx passed to positional embedding in template

* Remove padding_idx from pytorch learned positional embeddings

* Remove accidentally added quotes

* Remove padding_idx from tf learned positional embeddings

* Remove zeroing of weights in __init__
Co-authored-by: Wang Ming Rui <mingrui.wang@C02CJTUYMD6M.local>

894db670

17 Feb, 2021 1 commit

Making TF BART-like models XLA and AMP compliant (#10191) · 83d803ba

Julien Plu authored Feb 17, 2021

* Update BART

* Update Blenderbot

* Update BlenderbotSmall

* Update Marian

* Update MBart

* Update MBart

* Update Pegasus

* Update template

* Fix Marian and Pegasus

* Apply style

* Default initializer

* Default initializer

* Default initializer

* Remove int32 casts

* Fix template

* Remove more cast

83d803ba

15 Feb, 2021 2 commits
- Add AMP for Albert (#10141) · 31b0560a
  Julien Plu authored Feb 15, 2021
  
  31b0560a
- Fix TF template (#10189) · 57021887
  Julien Plu authored Feb 15, 2021
```
* Fix template

* Update Seq2Seq tests
```
  57021887
11 Feb, 2021 2 commits
- Update README.md · 8e13b735
  Patrick von Platen authored Feb 11, 2021
  
  8e13b735
- Update ADD_BIG_BIRD.md · d6b4f48e
  Patrick von Platen authored Feb 11, 2021
  
  d6b4f48e
09 Feb, 2021 1 commit
- Update ADD_BIG_BIRD.md · 4cda2d73
  Patrick von Platen authored Feb 09, 2021
  
  4cda2d73