Commits · eca77f4719531ecaabe9ec6b2dee6075a391d98a · chenpangpang / transformers

23 Mar, 2022 1 commit

Updates the default branch from master to main (#16326) · eca77f47

Lysandre Debut authored Mar 23, 2022



* Updates the default branch from master to main

* Links from `master` to `main`

* Typo

* Update examples/flax/README.md
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

eca77f47

22 Mar, 2022 4 commits

Adopt framework-specific blocks for content (#16342) · 77321481

Steven Liu authored Mar 22, 2022

* ✨ refactor code samples with framework-specific blocks

* ✨ update training.mdx

* 🖍 apply feedback

77321481

Fix code repetition in serialization guide (#16346) · 62cbd842
Omar Sanseviero authored Mar 22, 2022

62cbd842

[GLPN] Improve docs (#16331) · a2379b92

NielsRogge authored Mar 22, 2022



* Add link to notebook

* Add link

* Fix bug
Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>

a2379b92

Add GLPN (#16199) · 0c55d47c

NielsRogge authored Mar 22, 2022



* First draft

* Fix logits calculation

* Improve tests

* Add copied from statements

* Fix base_model_prefix

* Improve implementation, upload new models

* Update design

* Fix integration test

* Add model to README and toctree

* Add document image

* Apply suggestions from code review

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Add decoder_hidden_size attribute

* Update design of decoder

* Add DepthEstimatorOutput class

* Rename in_index to head_in_index and add feature extractor tests

* Apply suggestions from code review

* Apply suggestions from code review

* Update pretrained model name and add to doc tests

* Remove test.py script

* Update copied from statements and clean up
Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

0c55d47c

21 Mar, 2022 9 commits

Add Flaubert OnnxConfig to Transformers (#16279) · 0aac9ba2

Thomas Chaigneau authored Mar 21, 2022



* Add Flaubert to ONNX to make it available for conversion.

* Fixed features for FlauBERT. fixup command remove flaubert to docs list.
Co-authored-by: ChainYo <t.chaigneau.tc@gmail.com>

0aac9ba2

Update troubleshoot with more content (#16243) · 5a42bb43
Steven Liu authored Mar 21, 2022
```
* 📝 first draft

* 🖍 apply feedback
```
5a42bb43

[SegFormer] Remove unused attributes (#16285) · fbb45430

NielsRogge authored Mar 21, 2022



* Remove unused attributes

* Add link to blog and add clarification about input size

* Improve readability of the code
Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>

fbb45430

Remove disclaimer from Longformer docs (#16296) · 3f0f75e4
Gunjan Chhablani authored Mar 21, 2022

3f0f75e4
Fix a typo (add a coma) (#16291) · abf3cc70
PolarisRisingWar authored Mar 21, 2022
```
As mentioned: https://github.com/huggingface/transformers/issues/16277
```
abf3cc70

Fixed Error Raised Due to Wrongly Accessing Training Sample (#16115) · f3938680

Aflah authored Mar 21, 2022



* Update training.mdx

Fixed Error Raised Due to Wrongly Accessing Training Sample

* Ran make style

* Revert to Old Commit

* Apply suggestions from code review
Co-authored-by: Suraj Patil <surajp815@gmail.com>

f3938680

Draft a guide with our code quirks for new models (#16237) · 4ecb022e

Sylvain Gugger authored Mar 21, 2022



* Draft a guide with our code quirks for new models

* Apply suggestions from code review
Co-authored-by: Suraj Patil <surajp815@gmail.com>
Co-authored-by: Joao Gante <joao@huggingface.co>

* Apply suggestions from code review
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Suraj Patil <surajp815@gmail.com>
Co-authored-by: Joao Gante <joao@huggingface.co>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

4ecb022e

Framework split for Spanish version of doc quicktour.mdx (#16215) · c36b8565

Omar U. Espejel authored Mar 21, 2022



* Apply framework changes

* Fix italics

* Fix nits

* correct syntax
Co-authored-by: Omar Espejel <espejelomar@Omars-MacBook-Air.local>

c36b8565

Add Slack notification support for doc tests (#16253) · c1af180d

Patrick von Platen authored Mar 21, 2022

* up

* up

* up

* fix

* yeh

* ups

* Empty test commit

* correct quicktour

* correct

* correct

* up

* up

* uP

* uP

* up

* up

* uP

* up

* up

* up

* up

* up

* up

* up

* up

* up

* up

* Update src/transformers/models/van/modeling_van.py

* finish

* apply suggestions

* remove folder

* revert to daily testing

c1af180d

18 Mar, 2022 1 commit
- Fix links in guides (#16182) · ffc319e7
  Steven Liu authored Mar 18, 2022
```
* 🖍 fix links in guides

* 🖍 apply feedback
```
  ffc319e7
17 Mar, 2022 2 commits

[Deepspeed] non-HF Trainer doc update (#16238) · 47cccb53
Stas Bekman authored Mar 17, 2022

47cccb53

[Generate Docs] Correct docs (#16133) · 8a96b0f1

Patrick von Platen authored Mar 17, 2022



* [Generate Docs] Correct docs

* Apply suggestions from code review
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>
Co-authored-by: Steven Liu <59462357+stevhliu@users.noreply.github.com>

8a96b0f1

16 Mar, 2022 1 commit

Swin support for any input size (#15986) · 667b823b

Francesco Saverio Zuppichini authored Mar 16, 2022



* padding done

* correctly return one attention per layer

* almost correct, attentions are not flatten one tuple per stage

* tests green

* doc

* conversations

* reshaping hidden_states

* view in the test

* reshape_hidden_states in Encoder and Model

* new outputs with reshaped_hidden_states

* conversations

* doc

* Update docs/source/model_doc/swin.mdx
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* conversations

* fix tests

* minor changes

* resolved conversations

* attentions one per stage

* typo

* typos

* typos

* function signature

* CI

* clean up tests
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

667b823b

15 Mar, 2022 4 commits

Framework split (#16030) · 4f4e5ddb
Sylvain Gugger authored Mar 15, 2022
```
* First files

* More files

* Last files

* Style
```
4f4e5ddb

[Fix doc example] Fix first example for the custom_datasets tutorial (#16087) · bcaf5660

Markus Sagen authored Mar 15, 2022

* Fix inconsistent example variable naming

- Example code for a sequence classification in Tensorflow had spelling mistakes and incorrect and inconsistent naming
- Changed variable naming to be consistent with the two other TF examples

* Fix incorrect incorrect training examples

bcaf5660

Added spanish translation of quicktour.mdx (#16158) · daa49447

Daniel Espejel authored Mar 15, 2022



* Added spanish translation of quicktour.mdx

* Suggestions applied in the revision of the translation
Co-authored-by: Omar U. Espejel <espejelomar@gmail.com>
Co-authored-by: Omar U. Espejel <espejelomar@gmail.com>

daa49447

Visual Attention Network (VAN) (#16027) · 0a057201

Francesco Saverio Zuppichini authored Mar 15, 2022



* encoder works

* addded files

* norm in stage

* convertion script

* tests

* fix copies

* make fix-copies

* fixed __init__

* make fix-copies

* fix

* shapiro test needed

* make fix-copie

* minor changes

* make style + quality

* minor refactor conversion script

* rebase + tests

* removed unused variables

* updated doc

* toctree

* CI

* doc

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* resolved conversations

* make fixup

* config passed to modules

* config passed to modules

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* conversations

* conversations

* copyrights

* normal test

* tests
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

0a057201

14 Mar, 2022 4 commits

[WIP] Resnet (#15770) · e3008c67

Francesco Saverio Zuppichini authored Mar 14, 2022



* first commit

* ResNet model correctly implemented.

basic modeling + weights conversion is done

removed unused doc

mdx file

doc and conversion script

added feature_extractor to auto

test

minor changes + style + quality

doc

test

Delete process.yml

A left over from my attempt of running circleci locally

* minor changes

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* new test format

* minor changes from conversations

* minor changes from conversations

* make style + quality

* readded the tests

* test + README

* minor changes from conversations

* error in README

* make fix-copies

* removed regression for classification head

* make quality

* fixed loss control flow

* fixed loss control flow

* resolved conversations

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* READMEs

* index.mdx

* minor changes

* updated tests and models

* unused import

* outputs

* Update docs/source/model_doc/resnet.mdx
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* added embeddings_size

* Apply suggestions from code review
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* conversation

* added push to hub

* test

* embedding_size

* make fix-copies

* resolved conversations

* CI

* changed organization

* minor changes

* CI

* minor changes

* conversations

* conversation

* doc

* tests

* removed unused docstring

* conversation

* removed unused outputs

* CI
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

e3008c67

Spanish translation of the file training.mdx (#16047) · efd6e9a8

Yhary Arias authored Mar 14, 2022



* Spanish translation of the file training.mdx

* Settings - Spanish translation of the file training.mdx

* Latest changes to the Spanish translation of the training.mdx file

* Delete Hugging.mdx

* Last changes to the training fil Espanish version

* Latest modifications

* Latest changes, document ready for PR

* Nits
Co-authored-by: Yhary Arias <yharystefa@gmail.com>
Co-authored-by: Omar U. Espejel <espejelomar@gmail.com>

efd6e9a8

Fix and document Zero Shot Image Classification (#16079) · 802984ad
Omar Sanseviero authored Mar 14, 2022

802984ad
Add TFCamembertForCausalLM and ONNX integration test (#16073) · 6e1e88fd
lewtun authored Mar 14, 2022
```
* Make Camembert great again!

* Add Camembert to TensorFlow ONNX tests
```
6e1e88fd

12 Mar, 2022 1 commit

[Deepspeed] add support for bf16 mode (#14569) · 580dd87c

Stas Bekman authored Mar 11, 2022



* [WIP] add support for bf16 mode

* prep for bf16

* prep for bf16

* fix; zero2/bf16 is ok

* check bf16 is available

* test fixes

* enable zero3_bf16

* config files

* docs

* split stage_dtype; merge back to non-dtype-specific config file

* fix doc

* cleanup

* cleanup

* bfloat16 => bf16 to match the PR changes

* s/zero_gather_fp16_weights_on_model_save/zero_gather_16bit_weights_on_model_save/; s/save_fp16_model/save_16bit_model/

* test fixes/skipping

* move

* fix

* Update docs/source/main_classes/deepspeed.mdx
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* backticks

* cleanup

* cleanup

* cleanup

* new version

* add note about grad accum in bf16
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

580dd87c

11 Mar, 2022 3 commits

Audio/vision task guides (#15808) · ae2dd42b

Steven Liu authored Mar 11, 2022

* 📝 first draft of audio/vision guides

* ✨ make fixup

* 🖍 fix typo

* 🖍 close parentheses

* 🖍 apply feedback

* 🖍 apply feedback, make fixup

* 🖍 more fixup for perceiver

* 🖍 apply feedback

* ✨ make fixup

* 🖍 fix data collator

ae2dd42b

Update troubleshoot guide (#16001) · 5b4c97d0
Steven Liu authored Mar 11, 2022
```
* 📝 first draft

* 🖍 apply feedback

* 🖍 apply feedback
```
5b4c97d0
Force default brnahc name via the config · f7708e1b
Sylvain Gugger authored Mar 11, 2022

f7708e1b

10 Mar, 2022 3 commits

updating fine-tune classifier documentation (#16063) · 96ac7549
David S. Batista authored Mar 10, 2022

96ac7549

[Docs] Improve PyTorch, Flax generate API (#15988) · 6ce11c2c

Patrick von Platen authored Mar 10, 2022

* Move generate docs

* up

* Update docs/source/_toctree.yml

* correct

* correct some stuff

* correct tests

* more fixes

* finish generate

* add to doc stest

* finish

* finalize

* add warning to generate method

6ce11c2c

Add Document Image Transformer (DiT) (#15984) · 0835119b

NielsRogge authored Mar 10, 2022



* Add conversion script

* Improve script

* Fix bug

* Add option to push to hub

* Add support for classification models

* Update model name

* Upload feature extractor files first

* Remove hash checking

* Fix config

* Add id2label

* Add import

* Fix id2label file name

* Fix expected shape

* Add model to README

* Improve docs

* Add integration test and fix CI

* Fix code style

* Add missing init

* Add model to SPECIAL_MODULE_TO_TEST_MAP
Co-authored-by: Niels Rogge <nielsrogge@Nielss-MacBook-Pro.local>

0835119b

09 Mar, 2022 3 commits

Add FlaxBartForCausalLM (#15995) · b256f351

Sanchit Gandhi authored Mar 09, 2022

* add causal lm

* add CausalLM tests

* Add FlaxBartForCausalLM

* Add EncoderDecoder model tests

* change docstring

* make repo-consistency

* suggested changes

* remove jax ops

* correction

* rename pre-trained decoder model

b256f351

Add ONNX export for ViT (#15658) · 50dd314d

lewtun authored Mar 09, 2022



* Add ONNX support for ViT

* Refactor to use generic preprocessor

* Add vision dep to tests

* Extend ONNX slow tests to ViT

* Add dummy image generator

* Use model_type to determine modality

* Add deprecation warnings for tokenizer argument

* Add warning when overwriting the preprocessor

* Add optional args to docstrings

* Add minimum PyTorch version to OnnxConfig

* Refactor OnnxConfig class variables from CONSTANT_NAME to snake_case

* Add reasonable value for default atol
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

50dd314d

[Doctests] Move doctests to new GPU & Fix bugs (#15969) · c1aaa439

Patrick von Platen authored Mar 09, 2022



* test

* up

* up

* Empty test commit

* up

* update tests

* up

* fix some vision models

* correct

* correct docs

* Trigger notification

* finalize

* check

* correct quicktour

* Apply suggestions from code review

* improve doctests

* Trigger Build

* next try

* next try

* and again

* Output current clone information

* Output current clone information

* Correct path

* add tf round again

* revert to daily job
Co-authored-by: Lysandre <lysandre.debut@reseau.eseo.fr>

c1aaa439

07 Mar, 2022 1 commit

Update training scripts docs (#15931) · 38cc3506

Steven Liu authored Mar 07, 2022

* 📝 first draft

* 🖍 apply feedback

* 🖍 remove examples from toctree

* 🗑 remove examples from docs/source

38cc3506

04 Mar, 2022 3 commits

Constrained Beam Search [*With* Disjunctive Decoding] (#15761) · 5c6f57ee

Chan Woo Kim authored Mar 05, 2022



* added classes to get started with constrained beam search

* in progress, think i can directly force tokens now but not yet with the round robin

* think now i have total control, now need to code the bank selection

* technically works as desired, need to optimize and fix design choices leading to undersirable outputs

* complete PR #1 without disjunctive decoding

* removed incorrect tests

* Delete k.txt

* Delete test.py

* Delete test.sh

* revert changes to test scripts

* genutils

* full implementation with testing, no disjunctive yet

* shifted docs

* passing all tests realistically ran locally

* removing accidentally included print statements

* fixed source of error in initial PR test

* fixing the get_device() vs device trap

* fixed documentation docstrings about constrained_beam_search

* fixed tests having failing for Speech2TextModel's floating point inputs

* fix cuda long tensor

* added examples and testing for them and founx & fixed a bug in beam_search and constrained_beam_search

* deleted accidentally added test halting code with assert False

* code reformat

* Update tests/test_generation_utils.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Update tests/test_generation_utils.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Update tests/test_generation_utils.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Update tests/test_generation_utils.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Update tests/test_generation_utils.py

* fixing based on comments on PR

* took out the testing code that should but work fails without the beam search moditification ; style changes

* fixing comments issues

* docstrings for ConstraintListState

* typo in PhrsalConstraint docstring

* docstrings improvements

* finished adding what is sort of an opinionated implementation of disjunctive generation, but it revealed errors in inner beam search logic during testing.

* fixed bug found in constrained beam search that used beam_idx that were not global across all the batches

* disjunctive constraint working 100% correctly

* passing all tests

* Accidentally included mlruns

* Update src/transformers/generation_beam_constraints.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Update src/transformers/generation_beam_constraints.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* complete overhaul of type complexities and other nits

* strict type checks in generate()

* fixing second round of feedback by narsil

* fixed failing generation test because of type check overhaul

* generation test fail fix

* fixing test fails
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

5c6f57ee

Add missing support for Flax XLM-RoBERTa (#15900) · 01485cee

Javier de la Rosa authored Mar 04, 2022



* Adding Flax XLM-RoBERTa

* Add Flax to __init__

* Adding doc and dummy objects

* Add tests

* Add Flax XLM-R models autodoc

* Fix tests

* Add Flask XLM-RoBERTa to TEST_FILES_WITH_NO_COMMON_TESTS

* Update src/transformers/models/xlm_roberta/modeling_flax_xlm_roberta.py
Co-authored-by: Suraj Patil <surajp815@gmail.com>

* Update tests/xlm_roberta/test_modeling_flax_xlm_roberta.py
Co-authored-by: Suraj Patil <surajp815@gmail.com>

* Update tests/xlm_roberta/test_modeling_flax_xlm_roberta.py
Co-authored-by: Suraj Patil <surajp815@gmail.com>

* Remove test on large Flask XLM-RoBERTa

* Add tokenizer to the test
Co-authored-by: Suraj Patil <surajp815@gmail.com>

01485cee

Making MaskFormerForInstanceSegmentation. (#15934) · 89c7d9cf

Nicolas Patry authored Mar 04, 2022

Small adjustments.

Adding in type hint.

Last fix ?

Only include the default dict thing, not the pipelines.

89c7d9cf