Commits · 2c335037bd33b75e8c7bdc4a82e3526abb2aa95c · chenpangpang / transformers

18 Jan, 2022 9 commits

[Fix doc example] Wrong checkpoint name (#15079) · 979ca24e

Yih-Dar authored Jan 18, 2022



* fix doc example - MarianForCausalLM example

* try to keep copies

* fix copies

* fix more similar doc examples

* fix more

* fix style
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>

979ca24e

fix: #14486 do not use BertPooler in DPR (#15068) · 7b3d4df4

PaulLerner authored Jan 18, 2022



* fix: #14486 do not use BertPooler in DPR

* fix tf dpr as well

* finish
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

7b3d4df4

Add MAE (#15120) · 74bec986

NielsRogge authored Jan 18, 2022

* First draft

* More improvements

* More improvements

* More improvements

* Fix embeddings

* Add conversion script

* Finish conversion script

* More improvements

* Fix forward pass

* Remove print statements

* Add weights initialization

* Add initialization of decoder weights

* Add support for other models in the conversion script

* Fix patch_size for huge model

* Fix most of the tests

* Fix integration test

* Fix docs

* Fix archive_list

* Apply suggestions from code review

* Improve documentation

* Apply more suggestions

* Skip some tests due to non-deterministic behaviour

* Fix test_initialization

* Remove unneccessary initialization of nn.Embedding

* Improve docs

* Fix dummies

* Remove ViTMAEFeatureExtractor from docs

* Add model to README and table of contents

* Delete inference file

74bec986

[MBartTokenizer] remove dep on xlm-roberta tokenizer (#15201) · 2ae3be54
Suraj Patil authored Jan 18, 2022

2ae3be54
[ASR pipeline] correct with lm pipeline (#15200) · 497346d0
Patrick von Platen authored Jan 18, 2022
```
* [ASR pipeline] correct with lm pipeline

* improve error
```
497346d0

Fix deprecation warnings for int div (#15180) · 531336bb

Sylvain Gugger authored Jan 18, 2022



* Fix deprecation warnings for int div
Co-authored-by: mgoldey <matthew.goldey@gmail.com>

* Fix import

* ensure that tensor output is python scalar

* make backward compatible

* make code more readable

* adapt test functions
Co-authored-by: mgoldey <matthew.goldey@gmail.com>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

531336bb

Add REALM (#13292) · 22454ae4

Li-Huai (Allan) Lin authored Jan 18, 2022



* REALM initial commit

* Retriever OK (Update new_gelu).

* Encoder prediction score OK

* Encoder pretrained model OK

* Update retriever comments

* Update docs, tests, and imports

* Prune unused models

* Make embedder as a module `RealmEmbedder`

* Add RealmRetrieverOutput

* Update tokenization

* Pass all tests in test_modeling_realm.py

* Prune RealmModel

* Update docs

* Add training test.

* Remove completed TODO

* Style & Quality

* Prune `RealmModel`

* Fixup

* Changes:
1. Remove RealmTokenizerFast
2. Update docstrings
3. Add a method to RealmTokenizer to handle candidates tokenization.

* Fix up

* Style

* Add tokenization tests

* Update `from_pretrained` tests

* Apply suggestions

* Style & Quality

* Copy BERT model

* Fix comment to avoid docstring copying

* Make RealmBertModel private

* Fix bug

* Style

* Basic QA

* Save

* Complete reader logits

* Add searcher

* Complete searcher & reader

* Move block records init to constructor

* Fix training bug

* Add some outputs to RealmReader

* Add finetuned checkpoint variable names parsing

* Fix bug

* Update REALM config

* Add RealmForOpenQA

* Update convert_tfrecord logits

* Fix bugs

* Complete imports

* Update docs

* Update naming

* Add brute-force searcher

* Pass realm model tests

* Style

* Exclude RealmReader from common tests

* Fix

* Fix

* convert docs

* up

* up

* more make style

* up

* upload

* up

* Fix

* Update src/transformers/__init__.py

* adapt testing

* change modeling code

* fix test

* up

* up

* up

* correct more

* make retriever work

* update

* make style

* finish main structure

* Resolve merge conflict

* Make everything work

* Style

* Fixup

* Fixup

* Update training test

* fix retriever

* remove hardcoded path

* Fix

* Fix modeling test

* Update model links

* Initial retrieval test

* Fix modeling test

* Complete retrieval tests

* Fix

* style

* Fix tests

* Fix docstring example

* Minor fix of retrieval test

* Update license headers and docs

* Apply suggestions from code review

* Style

* Apply suggestions from code review

* Add an example to RealmEmbedder

* Fix
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

22454ae4

[Fix doc example] TFRagModel (#15187) · b25067d8

Yih-Dar authored Jan 18, 2022



* fix doc example - NameError: name 'PATH' is not defined

* fix name 'TFRagModel' is not defined

* correct TFRagRagSequenceForGeneration

* fix name 'tf' is not defined

* fix style
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>

b25067d8

`is_ctc` needs to be updated to `self.type == "ctc". (#15194) · dea563c9
Nicolas Patry authored Jan 18, 2022
```
* `is_ctc` needs to be updated to `self.type == "ctc".

* Adding fast test for this functionality.
```
dea563c9

17 Jan, 2022 3 commits
- [Fix doc example] UniSpeechSatForPreTraining (#15152) · 32090c72
  Yih-Dar authored Jan 18, 2022
```
* fix doc example - cannot import name 'UniSpeechSatFeatureEncoder'

* fix ckpt name
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
  32090c72
- Mark bad tokenizers version (#15188) · 6f8e644f
  Sylvain Gugger authored Jan 17, 2022
  
  6f8e644f
- Fix dtype issue in TF BART (#15178) · 9a2dabae
  Matt authored Jan 17, 2022
  
  9a2dabae
14 Jan, 2022 8 commits
- update from keras2onnx to tf2onnx (#15162) · ebc4edfe
  Joao Gante authored Jan 14, 2022
  
  ebc4edfe
- Better dummies (#15148) · 1b730c3d
  Sylvain Gugger authored Jan 14, 2022
```
* Better dummies

* See if this fixes the issue

* Fix quality

* Style

* Add doc for DummyObject
```
  1b730c3d
- Fixing flaky test (hopefully). (#15154) · b212ff9f
  Nicolas Patry authored Jan 14, 2022
```
* Fixing flaky test (hopefully).

* tf compliant.
```
  b212ff9f
- TF Bert inference - support `np.ndarray` optional arguments (#15074) · 7d9a33fb
  Joao Gante authored Jan 14, 2022
```
* TF Bert inference - support np.ndarray optional arguments

* apply np input tests to all TF architectures
```
  7d9a33fb
- fix BertTokenizerFast `tokenize_chinese_chars` arg (#15158) · 51d7ebf2
  SaulLu authored Jan 14, 2022
```
* add new test

* fix in init

* more relevant test
```
  51d7ebf2
- fix doc example - object has no attribute 'lm_logits' (#15143) · 4aa16fce
  Yih-Dar authored Jan 14, 2022
```
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
  4aa16fce
- Make sure all submodules are properly registered (#15144) · 7cbf8429
  Sylvain Gugger authored Jan 14, 2022
```
* Make sure all submodules are properly registered

* Try to fix tests

* Fix tests
```
  7cbf8429
- add TF glu activation function (#15146) · c4f7eb12
  Joao Gante authored Jan 14, 2022
  
  c4f7eb12
13 Jan, 2022 3 commits

Enable AMP for xla:gpu device in trainer class (#15022) · 6e058e84

Yanming Wang authored Jan 13, 2022

* Multiple fixes of trainer class with XLA GPU

* Make fp16 valid for xla:gpu

* Add mark_step in should_log to reduce compilation overhead

6e058e84

Deprecates AdamW and adds `--optim` (#14744) · 7b83feb5

Manuel R. Ciosici authored Jan 13, 2022



* Add AdamW deprecation warning

* Add --optim to Trainer

* Update src/transformers/optimization.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/optimization.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/optimization.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/optimization.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/training_args.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/training_args.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/training_args.py

* fix style

* fix

* Regroup adamws together
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Change --adafactor to --optim adafactor

* Use Enum for optimizer values

* fixup! Change --adafactor to --optim adafactor

* fixup! Change --adafactor to --optim adafactor

* fixup! Change --adafactor to --optim adafactor

* fixup! Use Enum for optimizer values

* Improved documentation for --adafactor
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Add mention of no_deprecation_warning
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Rename OptimizerOptions to OptimizerNames

* Use choices for --optim

* Move optimizer selection code to a function and add a unit test

* Change optimizer names

* Rename method
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Rename method
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Remove TODO comment
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Rename variable
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Rename variable
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Rename function

* Rename variable

* Parameterize the tests for supported optimizers

* Refactor

* Attempt to make tests pass on CircleCI

* Add a test with apex

* rework to add apex to parameterized; add actual train test

* fix import when torch is not available

* fix optim_test_params when torch is not available

* fix optim_test_params when torch is not available

* re-org

* small re-org

* fix test_fused_adam_no_apex

* Update src/transformers/training_args.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/training_args.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/training_args.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Remove .value from OptimizerNames

* Rename optimizer strings s|--adam_|--adamw_|

* Also rename Enum options

* small fix

* Fix instantiation of OptimizerNames. Remove redundant test

* Use ExplicitEnum instead of Enum

* Add unit test with string optimizer

* Change optimizer default to string value
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>
Co-authored-by: Stas Bekman <stas@stason.org>

7b83feb5

fix doc example - AssertionError: has to be configured as a decoder. (#15124) · 74837171
Yih-Dar authored Jan 13, 2022
```
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
74837171

12 Jan, 2022 3 commits

Add ONNX configuration classes to docs (#15121) · 021f2ea9

lewtun authored Jan 12, 2022

* Add ONNX classes to main package

* Remove permalinks from ONNX guide

* Fix ToC entry

* Revert "Add ONNX classes to main package"

This reverts commit eb794a5b00d66b0b4eab234987301676d8357630.

* Add ONNX classes to main doc

* Fix syntax highlighting in doc

* Fix text

* Add FeaturesManager to doc

* Use paths to reference ONNX classes

* Add FeaturesManager to init

* Add missing ONNX paths

021f2ea9

Fix #14357 (#15001) · 68209044
Yih-Dar authored Jan 12, 2022
```
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
68209044

Pipeline ASR with LM. (#15071) · 68cc4ccd

Nicolas Patry authored Jan 12, 2022



* Pipeline ASR with LM.

* Revamped into `self.decoder`.

* Fixing.

* 2nd fix.

* Update src/transformers/pipelines/__init__.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Fixing.
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

68cc4ccd

11 Jan, 2022 9 commits

Update TF test_step to match train_step (#15111) · 44eaa2b3

Matt authored Jan 11, 2022

* Update TF test_step to match train_step

* Update compile() warning to be clearer about what to pass

44eaa2b3

Fix saving FlaubertTokenizer configs (#14991) · 57b980a6

Vladimir Maryasin authored Jan 11, 2022

All specific tokenizer config properties must be passed to its base
class (XLMTokenizer) in order to be saved. This was not the case for
do_lowercase config. Thus it was not saved by save_pretrained() method
and saving and reloading the tokenizer changed its behaviour.

This commit fixes it.

57b980a6

Update ONNX docs (#14904) · 16f0b7d7

lewtun authored Jan 11, 2022



* Remove docs for deprecated ONNX export

* Tidy up the CLI help messages

* Revamp ONNX docs

* Update auto-config table

* Use DistilBERT as example for consistency

* Wrap up first pass at ONNX docs

* Fix table check

* Add tweaks and introduction

* Add cross-ref

* Fix missing import

* Fix style

* Add permalinks to ONNX configs

* Clarify role of OrderedDict

* Update docs/source/serialization.mdx
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Add doctest syntax to code blocks

* Remove permalinks

* Revert "Remove permalinks"

This reverts commit 099701daf0db27823457867938efdb2d4f22a7c1.
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

16f0b7d7

Add Nystromformer (#14659) · 28e09143

novice authored Jan 11, 2022



* Initial commit

* Config and modelling changes

Added Nystromformer-specific attributes to config and removed all decoder functionality from modelling.

* Modelling and test changes

Added Nystrom approximation and removed decoder tests.

* Code quality fixes

* Modeling changes and conversion script

Initial commits to conversion script, modeling changes.

* Minor modeling changes and conversion script

* Modeling changes

* Correct modeling, add tests and documentation

* Code refactor

* Remove tokenizers

* Code refactor

* Update __init__.py

* Fix bugs

* Update src/transformers/__init__.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/__init__.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/__init__.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update docs/source/model_doc/nystromformer.mdx
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/configuration_nystromformer.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/configuration_nystromformer.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/configuration_nystromformer.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/configuration_nystromformer.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/convert_nystromformer_original_pytorch_checkpoint_to_pytorch.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/nystromformer/configuration_nystromformer.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update modeling and test_modeling

* Code refactor

* .rst to .mdx

* doc changes

* Doc changes

* Update modeling_nystromformer.py

* Doc changes

* Fix copies

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update configuration_nystromformer.py

* Fix copies

* Update tests/test_modeling_nystromformer.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update test_modeling_nystromformer.py

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Fix code style

* Update modeling_nystromformer.py

* Update modeling_nystromformer.py

* Fix code style

* Reformat modeling file

* Update modeling_nystromformer.py

* Modify NystromformerForMultipleChoice

* Fix code quality

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Code style changes and torch.no_grad()

* make style

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

28e09143

change metric_key_prefix in seq2seq_trainer.py (#15099) · 285131bf
JejuWayfarer authored Jan 11, 2022
```
It solves the problem that metric_key_prefix is different from trainer.
```
285131bf

Adds IBERT to models exportable with ONNX (#14868) · c4fa908f

Virus authored Jan 11, 2022

* Add IBertOnnxConfig and tests

* add all the supported features for IBERT and remove outputs in IbertOnnxConfig

* use OnnxConfig

* fix codestyle

* remove serialization.rst

* codestyle

c4fa908f

[Wav2Vec2ProcessorWithLM] improve decoder downlaod (#15040) · efb35a41
Patrick von Platen authored Jan 11, 2022

efb35a41
fix doc example - TypeError: forward() got an unexpected keyword argument 'input_ids' (#15092) · 68810aa2
Yih-Dar authored Jan 11, 2022
```
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
68810aa2
Take gradient accumulation into account when defining samplers (#15095) · ca76618d
Sylvain Gugger authored Jan 11, 2022
```
* Take gradient accumulation into account when defining samplers

* style
```
ca76618d

10 Jan, 2022 5 commits

Add TFVisionEncoderDecoderModel (#14148) · b67fd797

Yih-Dar authored Jan 10, 2022



* Start the work on TFVisionEncoderDecoderModel

* Expose TFVisionEncoderDecoderModel

* fix import

* Add modeling_tf_vision_encoder_decoder to _ignore_modules in get_model_modules()

* reorder

* Apply the fix for checkpoint loading as in #14016

* remove attention_mask + fix VISION_DUMMY_INPUTS

* A minimal change to make TF generate() work for vision models as encoder in encoder-decoder setting

* fix wrong condition: shape_list(input_ids) == 2

* add tests

* use personal TFViTModel checkpoint (for now)

* Add equivalence tests + projection layer

* style

* make sure projection layer can run

* Add examples

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Clean comments (need to work on TODOs for PyTorch models)

* Remove TF -> PT in check_pt_tf_equivalence for TFVisionEncoderDecoderModel

* fixes

* Revert changes in PT code.

* Update tests/test_modeling_tf_vision_encoder_decoder.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Add test_inference_coco_en for TF test

* fix quality

* fix name

* build doc

* add main_input_name

* Fix ckpt name in test

* fix diff between master and this PR

* fix doc

* fix style and quality

* fix missing doc

* fix labels handling

* Delete auto.rst

* Add the changes done in #14016

* fix prefix

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* make style
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

b67fd797

[DOC] fix doc examples for bart-like models (#15093) · 3e9fdcf0
Suraj Patil authored Jan 10, 2022
```
* fix doc examples

* remove double colons
```
3e9fdcf0
Fix style · af9cb949
Sylvain Gugger authored Jan 10, 2022

af9cb949

fix doc example - AttributeError: type object 'RagModel' has no attribute... · 533624c5

Yih-Dar authored Jan 10, 2022


fix doc example - AttributeError: type object 'RagModel' has no attribute 'from_question_encoder_generator_pretrained' (#15076)
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>

533624c5

support the trocr small models (#14893) · b2c477fc

Minghao Li authored Jan 10, 2022



* support the trocr small models

* resolve conflict

* Update docs/source/model_doc/trocr.mdx
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update docs/source/model_doc/trocr.mdx
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update docs/source/model_doc/trocr.mdx
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/trocr/processing_trocr.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/trocr/processing_trocr.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/trocr/processing_trocr.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Update src/transformers/models/trocr/processing_trocr.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* fix unexpected indent in processing_trocr.py

* Update src/transformers/models/trocr/processing_trocr.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* update the docstring of processing_trocr

* remove extra space
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

b2c477fc