Commits · e3990d137aee4728a1127f55cbcf76970df571ce · chenpangpang / transformers

04 Sep, 2020 2 commits

fix (#6946) · e3990d13
Patrick von Platen authored Sep 04, 2020

e3990d13

Fix mixed precision issue in TF DistilBert (#6915) · a75e3198

Yih-Dar authored Sep 04, 2020

* Remove hard-coded uses of float32 to fix mixed precision use in TF Distilbert

* fix style

* fix gelu dtype issue in TF Distilbert

* fix numeric overflow while using half precision

a75e3198

03 Sep, 2020 12 commits

[s2s] support early stopping based on loss, rather than rouge (#6927) · e95d262f
Sam Shleifer authored Sep 03, 2020

e95d262f
[s2s] use --eval_beams command line arg (#6926) · 207ed8cb
Sam Shleifer authored Sep 03, 2020

207ed8cb
move wandb/comet logger init to train() to allow parallel logging (#6850) · 0f360d3d
krfricke authored Sep 03, 2020
```
* move wandb/comet logger init to train() to allow parallel logging

* Setup wandb/comet loggers on first call to log()
```
0f360d3d
[s2s] allow task_specific_params=summarization_xsum (#6923) · 39ed68d5
Sam Shleifer authored Sep 03, 2020

39ed68d5
[s2s]: script to convert pl checkpoints to hf checkpoints (#6911) · 5a318f07
Sam Shleifer authored Sep 03, 2020
```
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
```
5a318f07
tweak tar command in readme (#6919) · b8e4906c
brett koonce authored Sep 03, 2020

b8e4906c
Corrected link to paper (#6905) · a66db7d8
Stefan Engl authored Sep 03, 2020

a66db7d8
Added a link to the thesis. (#6906) · 55d61ce8
David Mark Nemeskey authored Sep 03, 2020

55d61ce8

Loodos model cards had errors on "Usage" section. It is fixed. Also... · 653a79cc

abdullaholuk-loodos authored Sep 03, 2020


Loodos model cards had errors on "Usage" section. It is fixed. Also "electra-base-turkish-uncased" model removed from s3 and re-uploaded as "electra-base-turkish-uncased-discriminator". Its README added. (#6921)
Co-authored-by: Abdullah Oluk <abdullaholuk123@gmail.com>

653a79cc

[model_card] link to correctly cased piaf dataset · 5a3aec90
Julien Chaumond authored Sep 03, 2020
```
cc @psorianom @rachelker
```
5a3aec90
Template updates (#6914) · 722b5807
Sylvain Gugger authored Sep 03, 2020

722b5807

Adding the LXMERT pretraining model (MultiModal languageXvision) to... · ea2c6f1a

Antonio V Mendoza authored Sep 03, 2020


Adding the LXMERT pretraining model (MultiModal  languageXvision)  to HuggingFace's suite of models (#5793)

* added template files for LXMERT and competed the configuration_lxmert.py

* added modeling, tokization, testing, and finishing touched for lxmert [yet to be tested]

* added model card for lxmert

* cleaning up lxmert code

* Update src/transformers/modeling_lxmert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/modeling_tf_lxmert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/modeling_tf_lxmert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/modeling_lxmert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* tested torch lxmert, changed documtention, updated outputs, and other small fixes

* Update src/transformers/convert_pytorch_checkpoint_to_tf2.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/convert_pytorch_checkpoint_to_tf2.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/convert_pytorch_checkpoint_to_tf2.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* renaming, other small issues, did not change TF code in this commit

* added lxmert question answering model in pytorch

* added capability to edit number of qa labels for lxmert

* made answer optional for lxmert question answering

* add option to return hidden_states for lxmert

* changed default qa labels for lxmert

* changed config archive path

* squshing 3 commits: merged UI + testing improvments + more UI and testing

* changed some variable names for lxmert

* TF LXMERT

* Various fixes to LXMERT

* Final touches to LXMERT

* AutoTokenizer order

* Add LXMERT to index.rst and README.md

* Merge commit test fixes + Style update

* TensorFlow 2.3.0 sequential model changes variable names

Remove inherited test

* Update src/transformers/modeling_tf_pytorch_utils.py

* Update docs/source/model_doc/lxmert.rst
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update docs/source/model_doc/lxmert.rst
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/modeling_tf_lxmert.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* added suggestions

* Fixes

* Final fixes for TF model

* Fix docs
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

ea2c6f1a

02 Sep, 2020 12 commits
- test_tf_common: remove un_used mixin class parameters (#6866) · 4ebb52af
  Puneetha Pai authored Sep 02, 2020
  
  4ebb52af
- [testing] fix ambiguous test (#6898) · e71f32c0
  Stas Bekman authored Sep 02, 2020
```
Since `generate()` does:
```
          num_beams = num_beams if num_beams is not None else self.config.num_beams
```
This test fails if `model.config.num_beams > 1` (which is the case in the model I'm porting).

This fix makes the test setup unambiguous by passing an explicit `num_beams=1` to `generate()`.

Thanks.
```
  e71f32c0
- Output attention takes an s (#6903) · 8f2723ca
  Sylvain Gugger authored Sep 02, 2020
```
* Fix output_attention -> output_attentions

* Formatting

* One unsaved file
```
  8f2723ca
- fix error class instantiation (#6634) · 485da722
  Yohei Tamura authored Sep 02, 2020
  
  485da722
- [pipelines] Text2TextGenerationPipeline (#6744) · 4230d30f
  Suraj Patil authored Sep 02, 2020
```
* add Text2TextGenerationPipeline

* remove max length warning

* remove comments

* remove input_length

* fix typo

* add tests

* use TFAutoModelForSeq2SeqLM

* doc

* typo

* add the doc below TextGenerationPipeline

* doc nit

* style

* delete comment
```
  4230d30f
- fix typo in comments (#6838) · 6b242812
  Prajjwal Bhargava authored Sep 02, 2020
  
  6b242812
- [doc] typos (#6867) · 7351ef83
  Stas Bekman authored Sep 02, 2020
```
* [doc] typos

fixed typos

* Update README.md
```
  7351ef83
- minor docs grammar fixes (#6889) · ee1bff06
  Harry Wang authored Sep 02, 2020
  
  ee1bff06
- fix warning for position ids (#6884) · 8abd7f69
  Patrick von Platen authored Sep 02, 2020
  
  8abd7f69
- Update modeling_bert.py (#6897) · 7cb0572c
  Parthe Pandit authored Sep 02, 2020
```
outptus -> outputs in example of BertForPreTraining
```
  7cb0572c
- Model card for huBERT (#6893) · e3c55ceb
  David Mark Nemeskey authored Sep 02, 2020
```
* Create README.md

Model card for huBERT.

* Update README.md

lowercase h

* Update model_cards/SZTAKI-HLT/hubert-base-cc/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com>
```
  e3c55ceb
- fix QA example for PT (#6890) · 1889e96c
  Patrick von Platen authored Sep 02, 2020
  
  1889e96c
01 Sep, 2020 14 commits
- [model_cards] Fix file path for flexudy/t5-base-multi-sentence-doctor · d822ab63
  Julien Chaumond authored Sep 02, 2020
  
  d822ab63
- Create README.md (#6598) · ad5fb33c
  Rohan Rajpal authored Sep 02, 2020
  
  ad5fb33c
- Create README.md (#6602) · f9dadcd8
  Rohan Rajpal authored Sep 02, 2020
  
  f9dadcd8
- Update multilingual passage rereanking model card (#6788) · f5d69c75
  Igli Manaj authored Sep 01, 2020
```
Fix range of possible score, add inference .
```
  f5d69c75
- Model card for primer/BART-Squad2 (#6801) · 5d820f3c
  Tom Grek authored Sep 01, 2020
  
  5d820f3c
- added model card for flexudys t5 model (#6759) · 8b884dad
  zolekode authored Sep 01, 2020
```
Co-authored-by: zolekode <pascal.zoleko@fau.de>
```
  8b884dad
- loodos turkish model cards added (#6840) · bff6d517
  hakan authored Sep 02, 2020
  
  bff6d517
- Create README.md (#6887) · 502d194b
  Manuel Romero authored Sep 01, 2020
```
Add language meta attribute
```
  502d194b
- Create README.md (#6888) · d082edf2
  Manuel Romero authored Sep 01, 2020
```
Add language meta attribute
```
  d082edf2
- Create README.md (#6886) · dacbee9a
  Abed khooli authored Sep 02, 2020
```
* Create README.md

model card for  akhooli/xlm-r-large-arabic-sent

* Update model_cards/akhooli/xlm-r-large-arabic-sent/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com>
```
  dacbee9a
- Create README.md (#6885) · e2971e61
  Abed khooli authored Sep 01, 2020
  
  e2971e61
- [EncoderDecoder] Add xlm-roberta to encoder decoder (#6878) · 4d1a3ffd
  Patrick von Platen authored Sep 01, 2020
```
* finish xlm-roberta

* finish docs

* expose XLMRobertaForCausalLM
```
  4d1a3ffd
- Create README.md (#6883) · 31199263
  Patrick von Platen authored Sep 01, 2020
```
* Create README.md

* Update README.md
```
  31199263
- Add cache_dir to save features TextDataset (#6879) · 21d71923
  Jin Young (Daniel) Sohn authored Sep 01, 2020
```
* Add cache_dir to save features TextDataset

This is in case the dataset is in a RO filesystem, for which is the case
in tests (GKE TPU tests).

* style
```
  21d71923