Commits · 0ecdf6de036f5c06de8dc31e6508358e71482cdd · chenpangpang / transformers

"vscode:/vscode.git/clone" did not exist on "5e620a92cf7e6c312435db86ec55e13b75dece75"

22 Sep, 2021 6 commits

Patch training arguments issue (#13700) · 0ecdf6de

Lysandre Debut authored Sep 22, 2021



* Patch training arguments issue

* Update src/transformers/training_args.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

0ecdf6de

Allow only textual inputs to VisualBert (#13687) · 50c746ee
Gunjan Chhablani authored Sep 22, 2021

50c746ee
Fix non-negligible difference between GPT2 and TFGP2 (#13679) · 93624bfe
Yih-Dar authored Sep 22, 2021
```
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
93624bfe

Assertions to exceptions (#13692) · a0c08aa3

MocktaiLEngineer authored Sep 22, 2021



* Raise exceptions instead of using assertions for control flow #12789

* # coding=utf-8

* Raise exceptions instead of using assertions for control flow

* Raise exceptions instead of using assertions for control flow

* Update src/transformers/tokenization_utils.py

Raise exceptions instead of using assertions for control flow
Co-authored-by: Suraj Patil <surajp815@gmail.com>

* Update src/transformers/tokenization_utils.py

Raise exceptions instead of using assertions for control flow
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Raise exceptions instead of using assertions for control flow

* test

* Raise exceptions instead of using assertions for control flow
Co-authored-by: MocktaiLEngineer <kavinarasu22@gmail.com>
Co-authored-by: Suraj Patil <surajp815@gmail.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

a0c08aa3

Make gradient_checkpointing a training argument (#13657) · 27d46397

Sylvain Gugger authored Sep 22, 2021



* Make gradient_checkpointing a training argument

* Update src/transformers/modeling_utils.py
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Update src/transformers/configuration_utils.py
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>

* Fix tests

* Style

* document Gradient Checkpointing as a performance feature

* Small rename

* PoC for not using the config

* Adapt BC to new PoC

* Forgot to save

* Rollout changes to all other models

* Fix typo
Co-authored-by: Stas Bekman <stas00@users.noreply.github.com>
Co-authored-by: Stas Bekman <stas@stason.org>

27d46397

[Wav2Vec2FeatureExtractor] Fix `extractor.pad()` dtype backwards compatibility (#13693) · 75f6641e
Anton Lozhkov authored Sep 22, 2021
```
* Force dtype, add tests

* Local torch imports

* Remove unused logic (always ndarray)
```
75f6641e

21 Sep, 2021 12 commits

[AutoTokenizer] Allow creation of tokenizers by tokenizer type (#13668) · 8e908c8c
Patrick von Platen authored Sep 22, 2021
```
* up

* up
```
8e908c8c
up (#13688) · 2608944d
Patrick von Platen authored Sep 22, 2021

2608944d

Update modeling_flax_wav2vec2.py (#13680) · 8565d38f

Kamal Raj authored Sep 22, 2021

conv kernel_size to Tuple,
Flax Version 0.3.5 breaking change, https://github.com/google/flax/releases/tag/v0.3.5

8565d38f

Skip FlaxWav2Vec2 test until fixed · d16bec95
Sylvain Gugger authored Sep 21, 2021

d16bec95

Layoutlm onnx support (Issue #13300) (#13562) · ddd4d02f

Nishant Prabhu authored Sep 22, 2021



* Add support for exporting PyTorch LayoutLM to ONNX

* Added tests for converting LayoutLM to ONNX

* Add support for exporting PyTorch LayoutLM to ONNX

* Added tests for converting LayoutLM to ONNX

* cleanup

* Removed regression/ folder

* Add support for exporting PyTorch LayoutLM to ONNX

* Added tests for converting LayoutLM to ONNX

* cleanup

* Fixed import error

* Remove unnecessary import statements

* Changed max_2d_positions from class variable to instance variable of the config class

* Add support for exporting PyTorch LayoutLM to ONNX

* Added tests for converting LayoutLM to ONNX

* cleanup

* Add support for exporting PyTorch LayoutLM to ONNX

* cleanup

* Fixed import error

* Changed max_2d_positions from class variable to instance variable of the config class

* Use super class generate_dummy_inputs method
Co-authored-by: Michael Benayoun <mickbenayoun@gmail.com>

* Add support for Masked LM, sequence classification and token classification
Co-authored-by: Michael Benayoun <mickbenayoun@gmail.com>

* Removed uncessary import and method

* Fixed code styling

* Raise error if PyTorch is not installed

* Remove unnecessary import statement
Co-authored-by: Michael Benayoun <mickbenayoun@gmail.com>

ddd4d02f

Add push_to_hub to no_trainer examples (#13659) · b7d264be

Sylvain Gugger authored Sep 21, 2021

* Add push_to_hub to no_trainer examples

* Quality

* Document integration

* Roll out to other examples

b7d264be

[SinusoidalPositionalEmbedding] incorrect dtype when make_weights in forward (#13665) · a722c301
Stas Bekman authored Sep 21, 2021

a722c301

[SequenceFeatureExtractor] Rewrite padding logic from pure python to numpy (#13650) · 1417978c

Anton Lozhkov authored Sep 21, 2021

* Test np padding

* Pass feature extraction tests

* Update type hints

* Fix flaky integration tests

* Try a more stable waveform

* Add to_numpy jax support

* int32 attention masks

* Refactor normalization tests

1417978c

Typo "UNKWOWN" -> "UNKNOWN" (#13675) · 8d533e6a
Kamal Raj authored Sep 21, 2021

8d533e6a

[FLAX] Question Answering Example (#13649) · 78807d86

Kamal Raj authored Sep 21, 2021

* flax qa example

* Updated README:  Added Large model

* added utils_qa.py FULL_COPIES

* Updates:
1. Copyright Year updated
2. added dtype arg
3. passing seed and dtype to load model
4. Check eval flag before running eval

* updated README

* updated code comment

78807d86

beit-flax (#13515) · a2dec768

Kamal Raj authored Sep 21, 2021

* beit-flax

* updated FLAX_BEIT_MLM_DOCSTRING

* removed bool_masked_pos from classification

* updated Copyright

* code refactoring: x -> embeddings

* updated test: rm from_pt

* Update docs/source/model_doc/beit.rst

* model code dtype updates and
other changes according to review

* relative_position_bias
revert back to pytorch design

a2dec768

Add Speech AutoModels (#13655) · 48fa42e5
Patrick von Platen authored Sep 21, 2021
```
* upload

* correct

* correct

* correct

* finish

* up

* up

* up again
```
48fa42e5

20 Sep, 2021 10 commits

Fix typo distilbert doc (#13643) · ea921365
flozi00 authored Sep 20, 2021

ea921365
fix research_projects/mlm_wwm readme.md examples (#13646) · 28d5700a
Lowin authored Sep 21, 2021
```
the variables of run example is not correct
```
28d5700a

Dynamically load model code from the Hub (#13467) · 002a078a

Sylvain Gugger authored Sep 20, 2021



* Dynamic model

* Use defensive flag

* Style

* Doc and arg rename

* Arg rename

* Add tests

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Address review comments

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

002a078a

Change https:/ to https:// (#13644) · aeb2dac0
flozi00 authored Sep 20, 2021

aeb2dac0

[megatron_gpt2] checkpoint v3 (#13508) · 0af901e8

Stas Bekman authored Sep 20, 2021

* [megatron_gpt2] checkpoint v3

* bug fix

* fixes

* switch to default  from  - which is what the current megatron-lm uses

* cleanup

* back compat

0af901e8

Update modeling_tf_deberta.py (#13654) · 936b3fde
Kamal Raj authored Sep 20, 2021
```
Fixed expand_dims axis
```
936b3fde
Fix mT5 documentation (#13639) · 04976a32
Ayaka Mikazuki authored Sep 20, 2021
```
* Fix MT5 documentation

The abstract is incomplete

* MT5 -> mT5
```
04976a32
[Fix]Make sure the args tb_writer passed to the TensorBoardCallback works (#13636) · fe379f85
Chengjiang Li authored Sep 20, 2021

fe379f85

Add FNet (#13045) · d8049331

Gunjan Chhablani authored Sep 20, 2021



* Init FNet

* Update config

* Fix config

* Update model classes

* Update tokenizers to use sentencepiece

* Fix errors in model

* Fix defaults in config

* Remove position embedding type completely

* Fix typo and take only real numbers

* Fix type vocab size in configuration

* Add projection layer to embeddings

* Fix position ids bug in embeddings

* Add minor changes

* Add conversion script and remove CausalLM vestiges

* Fix conversion script

* Fix conversion script

* Remove CausalLM Test

* Update checkpoint names to dummy checkpoints

* Add tokenizer mapping

* Fix modeling file and corresponding tests

* Add tokenization test file

* Add PreTraining model test

* Make style and quality

* Make tokenization base tests work

* Update docs

* Add FastTokenizer tests

* Fix fast tokenizer special tokens

* Fix style and quality

* Remove load_tf_weights vestiges

* Add FNet to  main README

* Fix configuration example indentation

* Comment tokenization slow test

* Fix style

* Add changes from review

* Fix style

* Remove bos and eos tokens from tokenizers

* Add tokenizer slow test, TPU transforms, NSP

* Add scipy check

* Add scipy availabilty check to test

* Fix tokenizer and use correct inputs

* Remove remaining TODOs

* Fix tests

* Fix tests

* Comment Fourier Test

* Uncomment Fourier Test

* Change to google checkpoint

* Add changes from review

* Fix activation function

* Fix model integration test

* Add more integration tests

* Add comparison steps to MLM integration test

* Fix style

* Add masked tokenization fix

* Improve mask tokenization fix

* Fix index docs

* Add changes from review

* Fix issue

* Fix failing import in test

* some more fixes

* correct fast tokenizer

* finalize

* make style

* Remove additional tokenization logic

* Set do_lower_case to False

* Allow keeping accents

* Fix tokenization test

* Fix FNet Tokenizer Fast

* fix tests

* make style

* Add tips to FNet docs
Co-authored-by: patrickvonplaten <patrick.v.platen@gmail.com>

d8049331

fix typo (#13647) · 87d5057d
Suraj Patil authored Sep 20, 2021

87d5057d

17 Sep, 2021 9 commits

Fix GPT2Config parameters in GPT2ModelTester (#13630) · b518aaf1
calpt authored Sep 17, 2021

b518aaf1
Updated tiny distilbert models (#13631) · 300ee0c7
Lysandre Debut authored Sep 17, 2021

300ee0c7
fix some docstring in encoder-decoder models (#13611) · afb07a79
Yih-Dar authored Sep 17, 2021
```
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
afb07a79
Cloned tensors after indexing in _compute_attn_output_with_global_indices (#13613) · 19b7acdd
Alessandro Suglia authored Sep 17, 2021
```
Co-authored-by: Alessandro Suglia <asuglia@fb.com>
```
19b7acdd
Use `config_dict_or_path` for deepspeed.zero.Init (#13614) · ce32c69c
Alex Hedges authored Sep 17, 2021

ce32c69c

Removed console spam from misfiring warnings (#13625) · 0eb02871

Matt authored Sep 17, 2021

* Removed misfiring warnings

* Revert "Removed misfiring warnings"

This reverts commit cea90de325056b9c1cbcda2bd2613a785c1639ce.

* Retain the warning, but only when the user actually overrides things

* Fix accidentally breaking just about every model on the hub simultaneously

* Style pass

0eb02871

Fix special tokens not correctly tokenized (#13489) · da8beaaf

Li-Huai (Allan) Lin authored Sep 17, 2021

* Fix special tokens not correctly tokenized

* Add testing

* Fix

* Fix

* Use user workflows instead of directly assigning variables

* Enable test of fast tokenizers

* Update test of canine tokenizer

da8beaaf

[Trainer] Add nan/inf logging filter (#13619) · 1f9dcfc1

Patrick von Platen authored Sep 17, 2021

* finish

* add test

* push

* remove unnecessary code

* up

* correct test

* Update src/transformers/training_args.py

1f9dcfc1

Optimize Token Classification models for TPU (#13096) · eae7a96b

Ibraheem Moosa authored Sep 17, 2021

* Optimize Token Classification models for TPU

As per the XLA document XLA cannot handle masked indexing well. So token classification
models for BERT and others use an implementation based on `torch.where`. This implementation
works well on TPU. 

ALBERT token classification model uses the masked indexing which causes performance issues
on TPU. This PR fixes this issue by following the BERT implementation.

* Same fix for ELECTRA

* Same fix for LayoutLM

eae7a96b

16 Sep, 2021 3 commits

XLMR tokenizer is fully picklable (#13577) · e02ed0ee
Benjamin Davidson authored Sep 16, 2021
```
* made tokenizer fully picklable

* remove whitespace

* added testcase
```
e02ed0ee

Properly use test_fetcher for examples (#13604) · af5c6ae5

Sylvain Gugger authored Sep 16, 2021

* Properly use test_fetcher for examples

* Fake example modification

* Fake modeling file modification

* Clean fake modifications

* Run example tests for any modification.

af5c6ae5

[deepspeed] replaced deprecated init arg (#13587) · bec2e3f5
Stas Bekman authored Sep 16, 2021
```
* [deepspeed] replaced deprecated init arg

* Trigger CI
```
bec2e3f5