Commits · 4c41c6622c556d2b5bf2df1f7adad0f3a6a6e445 · chenpangpang / transformers

15 Mar, 2021 3 commits
- Wrong link to super class (#10709) · 4c41c662
  cronoik authored Mar 15, 2021
```
Documentation was referring to slow tokenizer class while it should be the fast tokenizer.
```
  4c41c662
- enable loading Mbart50Tokenizer with AutoTokenizer (#10690) · fcf10214
  Suraj Patil authored Mar 15, 2021
```
* enable auto tokenizer for mbart50 tokenizers

* fix imports
```
  fcf10214
- make rag tests smaller (#10679) · bd8f6caf
  Patrick von Platen authored Mar 15, 2021
  
  bd8f6caf
12 Mar, 2021 7 commits

AdamW is now supported by default (#9624) · 4c32f9f2
Stas Bekman authored Mar 12, 2021

4c32f9f2

Pass encoder outputs into GenerationMixin (#10599) · fa35cda9

ymfa authored Mar 12, 2021

* Pass encoder_outputs into generate()

* Remove an if-statement

* Reformat

* Minimize changes to generate()

* Comment on input_ids

fa35cda9

fix: #10628 expanduser path in TrainingArguments (#10660) · 00cad2e5

PaulLerner authored Mar 12, 2021



* fix: #10628 expanduser path in TrainingArguments

* docs: explain why we expand paths in TrainingArguments

* Style
Co-authored-by: Sylvain Gugger <sylvain.gugger@gmail.com>

00cad2e5

Add auto_wrap option in fairscale integration (#10673) · e8246f78
Sylvain Gugger authored Mar 12, 2021
```
* Add auto_wrap option in fairscale integration

* Style
```
e8246f78
TensorFlow tests: having from_pt set to True requires torch to be installed. (#10664) · 184ef8ec
Lysandre Debut authored Mar 12, 2021
```
* TF model exists for Blenderbot 400M

* Marian

* RAG
```
184ef8ec

Adding new parameter to `generate`: `max_time`. (#9846) · 543d0549

Nicolas Patry authored Mar 12, 2021

* [WIP] Adding new parameter to `generate`:  `max_time`.

Generation by tokens number is sometimes a bit clunky because we don't
know how many tokens are good enough or even how many tokens are in
the payload (for pipelines users for instance). This leads to hard
to understand behavior.

This PR proposes a new argument `max_time` which is a float of seconds
for the allowed time for `generate` to run on.
Ideally combinations of `max_tokens=None`, `max_time=2` could be used to
generate as many tokens as possible within time budget.

NB: Another possible approach consists of passing a callback to `generate`
  putting the caller in charge of the actual decision of when to stop
  generating tokens. It opens the door to 'which args should we pass'
  to this callback. It's hard to imagine other use-cases for this
  early stopping behavior than time (that are not already covered by
  parameters of generate)

* Revamp with StoppingCriteria

* Removing deprecated mentions.

* Forgot arguments to stopping criteria.

* Readding max_length it's not just used as a stopping criteria.

* Default value for `stopping_criteria`.

* Address @patrickvonplaten comments.

- More docstrings
- Actual doc
- Include in global namespace
- Remove TF work.

* Put back `max_length` (deprecation different PR).

* Doc quality.

* Fixing old behavior without `stopping_criteria` but with `max_length`.

Making sure we don't break that in the future.

* Adding more tests for possible inconsistencies between

`max_length` and `stopping_criteria`.

* Fixing the torch imports.

543d0549

Adjust loss difference (#10669) · ea46e3fa
Lysandre Debut authored Mar 12, 2021

ea46e3fa

11 Mar, 2021 17 commits
- fix typing error for HfArgumentParser for Optional[bool] (#10672) · c526bde3
  Benjamin Fineran authored Mar 11, 2021
```
* fix typing error for TrainingArguments Optional[bool]

* updating equality check for Optional[bool]
```
  c526bde3
- Tentative fix for HFArgumentParser in Python 3.8 · fa1a8d10
  Sylvain Gugger authored Mar 11, 2021
  
  fa1a8d10
- Fix broken link (#10656) · 2f848519
  WybeKoper authored Mar 11, 2021
```
* Fixed broken link

* fixed max length violation
Co-authored-by: WybeKoper <WybeKoper@users.noreply.github.com>
```
  2f848519
- Add DeBERTa to MODEL_FOR_PRETRAINING_MAPPING (#10668) · a01ea31b
  jeswan authored Mar 11, 2021
```
* add deberta to pretraining mapping

* add deberta_v2 to PRETRAINING_MAPPING
```
  a01ea31b
- Specify minimum version for sacrebleu (#10662) · 9fbb4cdc
  Lysandre Debut authored Mar 11, 2021
  
  9fbb4cdc
- Fix integration slow tests (#10670) · fda703a5
  Sylvain Gugger authored Mar 11, 2021
```
* PoC

* Fix slow tests for the PT1.8 Embedding problem
```
  fda703a5
- Onnx fix test (#10663) · 3ab68203
  Funtowicz Morgan authored Mar 11, 2021
```
* Allow to pass kwargs to model's from_pretrained when using pipeline.

* Disable the use of past_keys_values for GPT2 when exporting to ONNX.

* style

* Remove comment.

* Appease the documentation gods

* Fix style
Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>
```
  3ab68203
- Fixes Pegasus tokenization tests (#10671) · a637ae00
  Lysandre Debut authored Mar 11, 2021
  
  a637ae00
- Conversion to tensors requires padding (#10661) · 7e442874
  Lysandre Debut authored Mar 11, 2021
  
  7e442874
- W2v2 test require torch (#10665) · 2adc8c92
  Lysandre Debut authored Mar 11, 2021
```
* Adds a @require_torch to a test that requires it

* Tokenizer too

* Style
```
  2adc8c92
- [S2T] fix example in docs (#10667) · 055ed78f
  Suraj Patil authored Mar 11, 2021
  
  055ed78f
- Remove special treatment for custom vocab files (#10637) · 89693e17
  Sylvain Gugger authored Mar 11, 2021
```
* Remove special path for custom vocab files

* Update src/transformers/tokenization_utils_base.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Expand error message
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
```
  89693e17
- S2S + M2M100 should be available in tokenization_auto (#10657) · 6d9e11a1
  Lysandre Debut authored Mar 11, 2021
```
* S2S + M2M100 should be available in tokenization_auto

* Requires sentencepiece

* SentencePiece for S2T as well :)
```
  6d9e11a1
- [XLSR-Wav2Vec2] Add multi-lingual Wav2Vec2 models (#10648) · 602d63f0
  Patrick von Platen authored Mar 11, 2021
```
* add conversion script

* add wav2vec2 xslr models

* finish

* Update docs/source/model_doc/xlsr_wav2vec2.rst
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
```
  602d63f0
- Ensure metric results are JSON-serializable (#10632) · 63c295ac
  Sylvain Gugger authored Mar 11, 2021
  
  63c295ac
- Update README.md (#10647) · 27d9e05c
  ArvidYin authored Mar 11, 2021
```
correct spell error: 'nether'
```
  27d9e05c
- merge_file -> merges_file (#10653) · 053f0197
  Lysandre Debut authored Mar 11, 2021
  
  053f0197
10 Mar, 2021 8 commits

Document Trainer limitation on custom models (#10635) · 26a33cfd
Sylvain Gugger authored Mar 10, 2021

26a33cfd

Extend trainer logging for sm (#10633) · 49c61a4a

Philipp Schmid authored Mar 10, 2021

* renamed logging to hf_logging

* changed logging from hf_logging to logging and loggin to native_logging

* removed everything trying to fix import Trainer error

* adding imports again

* added custom add_handler function to logging.py

* make style

* added remove_handler

* added another conditional to assert

49c61a4a

Fix GPU tests with speech · 1aa9c13f
Sylvain Gugger authored Mar 10, 2021

1aa9c13f

Copy tokenizer files in each of their repo (#10624) · 2295d783

Sylvain Gugger authored Mar 10, 2021

* Move tokenizer files in each repo

* Fix mBART50 tests

* Fix mBART tests

* Fix Marian tests

* Update templates

2295d783

Speech2TextTransformer (#10175) · d26b37e7

Suraj Patil authored Mar 10, 2021



* s2t

* fix config

* conversion script

* fix import

* add tokenizer

* fix tok init

* fix tokenizer

* first version working

* fix embeds

* fix lm head

* remove extra heads

* fix convert script

* handle encoder attn mask

* style

* better enc attn mask

* override _prepare_attention_mask_for_generation

* handle attn_maks in encoder and decoder

* input_ids => input_features

* enable use_cache

* remove old code

* expand embeddings if needed

* remove logits bias

* masked_lm_loss => loss

* hack tokenizer to support feature processing

* fix model_input_names

* style

* fix error message

* doc

* remove inputs_embeds

* remove input_embeds

* remove unnecessary docstring

* quality

* SpeechToText => Speech2Text

* style

* remove shared_embeds

* subsample => conv

* remove Speech2TextTransformerDecoderWrapper

* update output_lengths formula

* fix table

* remove max_position_embeddings

* update conversion scripts

* add possibility to do upper case for now

* add FeatureExtractor and Processor

* add tests for extractor

* require_torch_audio => require_torchaudio

* add processor test

* update import

* remove classification head

* attention mask is now 1D

* update docstrings

* attention mask should be of type long

* handle attention mask from generate

* alwyas return attention_mask

* fix test

* style

* doc

* Speech2TextTransformer => Speech2Text

* Speech2TextTransformerConfig => Speech2TextConfig

* remove dummy_inputs

* nit

* style

* multilinguial tok

* fix tokenizer

* add tgt_lang setter

* save lang_codes

* fix tokenizer

* add forced_bos_token_id to tokenizer

* apply review suggestions

* add torchaudio to extra deps

* add speech deps to CI

* fix dep

* add libsndfile to ci

* libsndfile1

* add speech to extras all

* libsndfile1 -> libsndfile1

* libsndfile

* libsndfile1-dev

* apt update

* add sudo to install

* update deps table

* install libsndfile1-dev on CI

* tuple to list

* init conv layer

* add model tests

* quality

* add integration tests

* skip_special_tokens

* add speech_to_text_transformer in toctree

* fix tokenizer

* fix fp16 tests

* add tokenizer tests

* fix copyright

* input_values => input_features

* doc

* add model in readme

* doc

* change checkpoint names

* fix copyright

* fix code example

* add max_model_input_sizes in tokenizer

* fix integration tests

* add do_lower_case to tokenizer

* remove clamp trick

* fix "Add modeling imports here"

* fix copyrights

* fix tests

* SpeechToTextTransformer => SpeechToText

* fix naming

* fix table formatting

* fix typo

* style

* fix typos

* remove speech dep from extras[testing]

* fix copies

* rename doc file,

* put imports under is_torch_available

* run feat extract tests when torch is available

* dummy objects for processor and extractor

* fix imports in tests

* fix import in modeling test

* fxi imports

* fix torch import

* fix imports again

* fix positional embeddings

* fix typo in import

* adapt new extractor refactor

* style

* fix torchscript test

* doc

* doc

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* fix docs, copied from, style

* fix docstring

* handle imports

* remove speech from all extra deps

* remove s2t from seq2seq lm mapping

* better names

* skip training tests

* add install instructions

* List => Tuple

* doc

* fix conversion script

* fix urls

* add instruction for libsndfile

* fix fp16 test
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

d26b37e7

Add new GLUE example with no Trainer. (#10555) · efb5c0a4
Sylvain Gugger authored Mar 10, 2021
```
* Add new GLUE example with no Trainer.

* Style

* Address review comments
```
efb5c0a4
remove final_logits_bias (#10606) · 44f64132
Suraj Patil authored Mar 10, 2021

44f64132

Fixes an issue in `text-classification` where MNLI eval/test datasets are not... · 6f52fce6

Allen Wang authored Mar 09, 2021

Fixes an issue in `text-classification` where MNLI eval/test datasets are not being preprocessed. (#10621)

* Fix MNLI tests

* Linter fix

6f52fce6

09 Mar, 2021 5 commits

Fix tests of TrainerCallback (#10615) · 72d9e039

Sylvain Gugger authored Mar 09, 2021



* Fix tests of TrainerCallback

* Update tests/test_trainer_callback.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

72d9e039

Fairscale FSDP fix model save (#10596) · 0d909f6b
Sylvain Gugger authored Mar 09, 2021
```
* Hotfix fairscale FSDP

* Evaluation works

* Save on process zero
```
0d909f6b
added max_sample args and metrics changes (#10602) · ac17f711
Bhadresh Savani authored Mar 09, 2021

ac17f711

Trigger add sm information (#10610) · c19c811a

Philipp Schmid authored Mar 09, 2021

* added sm to ua

* update id

* removed id

* removed comments

* added env variable

* changed variable name

* make quality happy

* added sguggers feedback

* make styling happy and remove brackets

* added sm to ua

* update id

* removed id

* removed comments

* added env variable

* changed variable name

* make quality happy

* added sguggers feedback

* make styling happy and remove brackets

c19c811a

layerdrop 0 (#10604) · 20c10258
Suraj Patil authored Mar 09, 2021

20c10258