Commits · fc1d97f29d7b98e82ae17fc5ac49229e2859bcca · chenpangpang / transformers

30 Nov, 2021 6 commits

VisionTextDualEncoder (#13511) · fc1d97f2

Suraj Patil authored Nov 30, 2021



* init vision_text_dual_encoder

* fix merge

* remove extra heads

* fix tests

* remove VISION_TEXT_DUAL_ENCODER_PRETRAINED_CONFIG_ARCHIVE_MAP

* remove archive map

* fix imports

* fix more imports

* fix init

* delete tokenizers

* fix imports

* clean

* support clip's vision model

* handle None config

* begin tests

* more test and few fixes

* warn about newly init weights

* more tests

* add loss to model

* remove extra classes from doc

* add processor

* doc and small fixes

* add start docstr

* update flax model

* flax tests

* more flax tests

* doc

* quality

* doc and quality

* fix doc

* doc

* remove comments

* update warning

* quality

* fix docs

* Apply suggestions from code review
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* replace asserts, fix imports

* update imports

* fix import

* address some review comments

* fix check

* reduce tolerance

* fix test

* add flax integration test

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* address Sylvain's comments

* fix style

* add pt_flax_equivalence test in PT tests

* add pt integration test

* update test

* use pre-trained checkpoint in examples
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

fc1d97f2

use functional interface for softmax in attention (#14198) · 6ed9882d

Thomas Viehmann authored Nov 30, 2021

* use functional interface instead of instantiating module and immediately calling it

* fix torch.nn.functional to nn.functional. Thank you Stas!

6ed9882d

Add documentation for multi-label classification (#14168) · 4176bc16
giacomo snidero authored Nov 30, 2021
```
* "update example docstring multilabel example

* update example docstring multilabel example
```
4176bc16

[Flax] Add FlaxBlenderbot (#13633) · faacd747

Daniel Stancl authored Nov 30, 2021



* Init Flax implementation for Blenderbot

* Add a majority of stuff except for tests

* make style quality

* Add tests and fix some bugs

* Add tests

* Clean source code and fix some bugs

* Fix copies and docs

* Fix jax device condition for tests

* Fix layer norm in the encoder

* Fix a few typos in the test file

* make fix-copies

* make fix-copies

* fix layer norm

* Fix Flax params dtype (#13090)

* Fix PR reference (#13098)

* make fix-copies

* Update tests/test_modeling_flax_blenderbot.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Suraj Patil <surajp815@gmail.com>

faacd747

Fix backend regex (#14566) · 254fef67
Sylvain Gugger authored Nov 30, 2021

254fef67

Tapas tf (#13393) · c468a87a

Kamal Raj authored Nov 30, 2021

* TF Tapas first commit

* updated docs

* updated logger message

* updated pytorch weight conversion
script to support scalar array

* added use_cache to tapas model config to
work properly with tf input_processing

* 1. rm embeddings_sum
2. added # Copied
3. + TFTapasMLMHead
4. and lot other small fixes

* updated docs

* + test for tapas

* updated testing_utils to check
is_tensorflow_probability_available

* converted model logits post processing using
numpy to work with both PT and TF models

* + TFAutoModelForTableQuestionAnswering

* added TF support

* added test for
TFAutoModelForTableQuestionAnswering

* added test for
TFAutoModelForTableQuestionAnswering pipeline

* updated auto model docs

* fixed typo in import

* added tensorflow_probability to run tests

* updated MLM head

* updated tapas.rst with TF  model docs

* fixed optimizer import in docs

* updated convert to np
data from pt model is not
`transformers.tokenization_utils_base.BatchEncoding`
after pipeline upgrade

* updated pipeline:
1. with torch.no_gard removed, pipeline forward handles
2. token_type_ids converted to numpy

* updated docs.

* removed `use_cache` from config

* removed floats_tensor

* updated code comment

* updated Copyright Year and
logits_aggregation Optional

* updated docs and comments

* updated docstring

* fixed model weight loading

* make fixup

* fix indentation

* added tf slow pipeline test

* pip upgrade

* upgrade python to 3.7

* removed from_pt from tests

* revert commit f18cfa9

c468a87a

29 Nov, 2021 6 commits

Add model checkpointing to push_to_hub and PushToHubCallback (#14492) · 6fc38adf

Matt authored Nov 29, 2021



* Add checkpointing to push_to_hub and PushToHubCallback

* Add checkpoint loading

* Add missing default value

* Correct method name

* make style

* Moving everything to the right location

* make style

* Revert changes to file_utils.py

* Update src/transformers/keras_callbacks.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/keras_callbacks.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Adding docstrings and comments to clarify code

* make style

* Fix organization positional arg

* Fix load_repo_checkpoint to no longer accidentally create empty repos

* make style

* Remove unnecessary 'organization' argument in load_repo_checkpoint

* Avoid private `_create_or_get_repo` method

* make style

* Update src/transformers/modeling_tf_utils.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

6fc38adf

Fix sentinel token IDs in data collator for Flax T5 pretraining script (#14477) · 8332327d
Rahul Nadkarni authored Nov 29, 2021

8332327d
[Flax] token-classification model steps enumerate start from 1 (#14547) · 2bd950ca
Kamal Raj authored Nov 29, 2021
```
* step start from 1

* Updated cur_step calcualtion
```
2bd950ca
[Generate] Fix generate with inputs_embeds on GPU (#14564) · cea17acd
Patrick von Platen authored Nov 29, 2021

cea17acd
Rename ImageGPT (#14526) · 25156eb2
NielsRogge authored Nov 29, 2021
```
* Rename

* Add MODEL_FOR_CAUSAL_IMAGE_MODELING_MAPPING
```
25156eb2

LayoutLMv2FeatureExtractor now supports non-English languages when applying Tesseract OCR. (#14514) · 4ee0b755

Štěpán Műller authored Nov 29, 2021



* Added the lang argument to apply_tesseract in feature_extraction_layoutlmv2.py, which is used in pytesseract.image_to_data.

* Added ocr_lang argument to LayoutLMv2FeatureExtractor.__init__, which is used when calling apply_tesseract

* Updated the documentation of the LayoutLMv2FeatureExtractor

* Specified in the documentation of the LayoutLMv2FeatureExtractor that the ocr_lang argument should be a language code.

* Update src/transformers/models/layoutlmv2/feature_extraction_layoutlmv2.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* Split comment into two lines to adhere to the max line size limit.

* Update src/transformers/models/layoutlmv2/feature_extraction_layoutlmv2.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

4ee0b755

28 Nov, 2021 1 commit
- Tokenizers docs: Specify which class contains `__call__` method (#14379) · ebbe8cc3
  Xing Han Lu authored Nov 28, 2021
```
* Update tokenizer.rst

* Apply `make fixup`
```
  ebbe8cc3
26 Nov, 2021 4 commits

unfreeze initial cache in gpt models (#14535) · 69511cdc
Suraj Patil authored Nov 26, 2021

69511cdc
Fixes (#14534) · 2318bf77
Lysandre Debut authored Nov 26, 2021

2318bf77
Quicktour updates (#14533) · c15f4f20
Lysandre Debut authored Nov 26, 2021

c15f4f20

added save_directories for _psave_pretrained_pt and _tf, changed model to... · 1bbd6fcd

Chris Fregly authored Nov 26, 2021

     added save_directories for _psave_pretrained_pt and _tf, changed model to tf_model and pt_model, enable the notebook to run cleanly from top to bottom without error (#14529)

* added save_directories for _psave_pretrained_pt and _tf, changed model to tf_model and pt_model, enable the notebook to run cleanly from top to bottom without error

* Update quicktour.rst

* added >>>

* dependencies

* added space

1bbd6fcd

25 Nov, 2021 2 commits
- Fix a slow test. (#14527) · 04683c06
  Nicolas Patry authored Nov 25, 2021
  
  04683c06
- clear ~/.cache/torch_extensions between builds (#14520) · d1fd64e7
  Stas Bekman authored Nov 25, 2021
  
  d1fd64e7
24 Nov, 2021 3 commits

[Tests] Improve vision tests (#14458) · 3772af49
NielsRogge authored Nov 24, 2021
```
* Improve tests

* Install vision for tf tests
```
3772af49
Fix feature extraction utils import (#14515) · f2e90bcb
Lysandre Debut authored Nov 24, 2021

f2e90bcb

add cache_dir for tokenizer verification loading (#14508) · 6c4d688f

Vladimir Maryasin authored Nov 24, 2021

When loading a pretrained tokenizer, a verification is done to ensure
that the actual tokenizer class matches the class it was called from.
If the tokenizer is absent, its config file is loaded from the repo.

However, the cache_dir for downloading is not provided, which leads to
ignoring of the user-specified cache_dir, storing files in several
places and and may result in incorrect warnings when the default
cache_dir is unreachsble.

This commit fixes that.

6c4d688f

23 Nov, 2021 1 commit

[deepspeed] zero inference (#14253) · 956a4831

Stas Bekman authored Nov 23, 2021



* [deepspeed] zero inference

* only z3 makes sense for inference

* fix and style

* docs

* rework

* fix test

* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* responding to suggestions
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

956a4831

22 Nov, 2021 6 commits

Switch from using sum for flattening lists of lists in group_texts (#14472) · 69e16abf

Nicholas Broad authored Nov 22, 2021



* remove sum for list flattening

* change to chain(*)

* make chain object a list

* delete empty lines

per sgugger's suggestions
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Nicholas Broad <nicholas@nmbroad.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

69e16abf

fixes some key names for in LayoutLMv2 / LayoutXLM tokenizers (#14493) · 0b7d053c
Valentin authored Nov 22, 2021
```
in case of left padding_side there was a copy/paste error
assigning the bbox data to the labels
```
0b7d053c

Auto processor (#14465) · 204d2513

Sylvain Gugger authored Nov 22, 2021



* Add AutoProcessor class

* Init and tests

* Add doc

* Fix init

* Update src/transformers/models/auto/processing_auto.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Reverts to tokenizer or feature extractor when available

* Adapt test
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

204d2513

[test] add test for --config_overrides (#14466) · 11f65d41
Stas Bekman authored Nov 22, 2021
```
* add test for --config_overrides

* remove unneeded parts of the test
```
11f65d41
Improve a add-new-pipeline docs a bit (#14485) · e0e2da11
Daniel Stancl authored Nov 22, 2021

e0e2da11
Moving pipeline tests from `Narsil` to `hf-internal-testing`. (#14463) · a4553e6c
Nicolas Patry authored Nov 22, 2021
```
* Moving everything to `hf-internal-testing`.

* Fixing test values.

* Moving to other repo.

* Last touch?
```
a4553e6c

21 Nov, 2021 2 commits
- Fix dummy objects for quantization (#14478) · 1a92bc57
  Sylvain Gugger authored Nov 21, 2021
```
* Fix dummy objects for quantization

* Add more models
```
  1a92bc57
- add Tuple as possible type hint for EvalPredictions label_ids (#14473) · c9d2cf85
  Alexander Measure authored Nov 21, 2021
```
* Update trainer_utils.py

* add Tuple type hints to all label_ids outputs

affects EvalLoopOutput and PredicctionOutput
```
  c9d2cf85
19 Nov, 2021 5 commits

Add QDQBert model and quantization examples of SQUAD task (#14066) · a59e7c1e

Shang Zhang authored Nov 19, 2021



* clean up branch for add-qdqbert-model

* README update for QAT example; update docstrings in modeling_qdqbert.py

* Update qdqbert.rst

* Update README.md

* Update README.md

* calibration data using traning set; QAT example runs in fp32

* re-use BERTtokenizer for qdqbert

* Update docs/source/model_doc/qdqbert.rst
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update docs/source/model_doc/qdqbert.rst
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update docs/source/model_doc/qdqbert.rst
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* remove qdqbert tokenizer

* Update qdqbert.rst

* update evaluate-hf-trt-qa.py

* update configuration_qdqbert.py

* update modeling_qdqbert.py: add copied statement; replace assert with ValueError

* update copied from statement

* add is_quantization_available; run make fix-copies

* unittest add require_quantization

* add backend dependency to qdqbert model

* update README; update evaluate script; make style

* lint

* docs qdqbert update

* circleci build_doc add pytorch-quantization for qdqbert

* update README

* update example readme with instructions to upgrade TensorRT to 8.2

* Update src/transformers/models/qdqbert/configuration_qdqbert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/models/qdqbert/configuration_qdqbert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/models/qdqbert/configuration_qdqbert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Update src/transformers/models/qdqbert/configuration_qdqbert.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* change quantization to pytorch_quantization for backend requirement

* feed_forward_chunking not supported in QDQBert

* make style

* update model docstrings and comments in testing scripts

* rename example to quantization-qdqbert; rename example scripts from qat to quant

* Update src/transformers/models/qdqbert/modeling_qdqbert.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* rm experimental functions in quant_trainer

* qa cleanup

* make fix-copies for docs index.rst

* fix doctree; use post_init() for qdqbert

* fix early device assignment for qdqbert

* fix CI:Model templates runner
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

a59e7c1e

Adding support for `hidden_states` and `attentions` in unbatching (#14420) · 81fe8afa
Nicolas Patry authored Nov 19, 2021
```
support.
```
81fe8afa
[Generation] Allow `inputs_embeds` as an input (#14443) · f25a9332
Patrick von Platen authored Nov 19, 2021
```
* up

* finalize

* finalize

* finish

* Update src/transformers/generation_utils.py

* apply feedback
```
f25a9332
[ImageGPT] Small fixes (#14460) · 0490b988
NielsRogge authored Nov 19, 2021
```
* Add integration test

* Fix typo
```
0490b988
Add GitPython to quality tools (#14459) · 331c3d2a
Lysandre Debut authored Nov 19, 2021
```
* Update setup.py

* Update setup.py

* Update setup.py

* Remove GitPython install
```
331c3d2a

18 Nov, 2021 4 commits

[Speech Recognition] More examples · efea0f86
Patrick von Platen authored Nov 18, 2021
```
Add more XLS-R training runs to the official examples
```
efea0f86
[Bert, et al] fix early device assignment (#14447) · 72a6bf33
Stas Bekman authored Nov 18, 2021
```
* fix early device assignment

* more models
```
72a6bf33
Fix finite IterableDataset test on multiple GPUs (#14445) · 83ef8bca
Sylvain Gugger authored Nov 18, 2021

83ef8bca

Add ImageGPT (#14240) · da36c557

NielsRogge authored Nov 18, 2021

* First draft

* More improvements

* Improve conversion script

* Fix init weights for layer norm

* Fix correct model for conversion script

* Don't tie input and output embeddings

* Add print statements for debugging

* Add print statements for debugging

* Fix vocab size of model

* Improve documentation, remove fast tokenizer

* Add ImageGPTForImageClassification, improve docs

* Fix docs issue

* Set verbosity level back to info

* Improve tests

* Fix tests and add figure

* Delete tokenizer file

* Remove ImageGPTTokenizer from init files

* Remove ImageGPTLayer from init files

* Remove ImageGPT tokenizer from docs

* First draft of ImageGPTFeatureExtractor

* Fix typo

* Fix bug

* More improvements

* Apply suggestions from code review, add tests for feature extractor

* Fix layernorm

* Update save_pretrained method

* Fix issue

* Make all tests of ImageGPTFeatureExtractor pass

* Update code examples

* Rename model inputs to pixel_values

* Improve code examples

* Update init_weights to post_init

* Fix post_init

da36c557