Commits · 3c7b965bcd5745213134d6363e1879e0d70fa13c · chenpangpang / transformers

"examples/research_projects/visual_bert/requirements.txt" did not exist on "e0e0675ac7a55f2bb7863813f3c7fd797de9375f"

21 Sep, 2022 2 commits
- Add some tests for check_dummies (#19146) · 3c7b965b
  Sylvain Gugger authored Sep 21, 2022
  
  3c7b965b
- Fix dummy creation for multi-frameworks objects (#19144) · 451df725
  Sylvain Gugger authored Sep 21, 2022
  
  451df725
27 Jun, 2022 1 commit

Add a TF in-graph tokenizer for BERT (#17701) · ee0d001d

Matt authored Jun 27, 2022

* Add a TF in-graph tokenizer for BERT

* Add from_pretrained

* Add proper truncation, option handling to match other tokenizers

* Add proper imports and guards

* Add test, fix all the bugs exposed by said test

* Fix truncation of paired texts in graph mode, more test updates

* Small fixes, add a (very careful) test for savedmodel

* Add tensorflow-text dependency, make fixup

* Update documentation

* Update documentation

* make fixup

* Slight changes to tests

* Add some docstring examples

* Update tests

* Update tests and add proper lowercasing/normalization

* make fixup

* Add docstring for padding!

* Mark slow tests

* make fixup

* Fall back to BertTokenizerFast if BertTokenizer is unavailable

* Fall back to BertTokenizerFast if BertTokenizer is unavailable

* make fixup

* Properly handle tensorflow-text dummies

ee0d001d

17 May, 2022 1 commit
- Fix dummy creation script (#17304) · 032d63b9
  Sylvain Gugger authored May 17, 2022
  
  032d63b9
23 Mar, 2022 1 commit

Reorganize file utils (#16264) · 4975002d

Sylvain Gugger authored Mar 23, 2022

* Split file_utils in several submodules

* Fixes

* Add back more objects

* More fixes

* Who exactly decided to import that from there?

* Second suggestion to code with code review

* Revert wront move

* Fix imports

* Adapt all imports

* Adapt all imports everywhere

* Revert this import, will fix in a separate commit

4975002d

14 Jan, 2022 1 commit

Better dummies (#15148) · 1b730c3d

Sylvain Gugger authored Jan 14, 2022

* Better dummies

* See if this fixes the issue

* Fix quality

* Style

* Add doc for DummyObject

1b730c3d

21 Nov, 2021 1 commit
- Fix dummy objects for quantization (#14478) · 1a92bc57
  Sylvain Gugger authored Nov 21, 2021
```
* Fix dummy objects for quantization

* Add more models
```
  1a92bc57
16 Nov, 2021 1 commit
- Add forward method to dummy models (#14419) · 3e8d17e6
  Sylvain Gugger authored Nov 16, 2021
```
* Add forward method to dummy models

* Fix quality
```
  3e8d17e6
15 Jun, 2021 1 commit
- Have dummy processors have a `from_pretrained` method (#12145) · d07b540a
  Lysandre Debut authored Jun 15, 2021
  
  d07b540a
11 Jun, 2021 1 commit

Add from_pretrained to dummy timm objects (#12097) · 3b1f5caf

Lysandre Debut authored Jun 11, 2021



* Add from_pretrained to dummy timm

* Fix at the source

* Update utils/check_dummies.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Missing pretrained dummies

* Style
Co-authored-by: Sylvain Gugger <sylvain.gugger@gmail.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

3b1f5caf

26 Apr, 2021 1 commit
- make style (#11442) · 32dbb2d9
  Patrick von Platen authored Apr 26, 2021
  
  32dbb2d9
07 Apr, 2021 1 commit

Dummies multi backend (#11100) · 11505fa1

Sylvain Gugger authored Apr 07, 2021

* Replaces requires_xxx by one generic method

* Quality and update check_dummies

* Fix inits check

* Post-merge cleanup

11505fa1

06 Apr, 2021 1 commit

Auto feature extractor (#11097) · 403d530e

Sylvain Gugger authored Apr 06, 2021

* AutoFeatureExtractor

* Init and first tests

* Tests

* Damn you gitignore

* Quality

* Defensive test for when not all backends are here

* Use pattern for Speech2Text models

403d530e

26 Mar, 2021 1 commit

Add ImageFeatureExtractionMixin (#10905) · b0595d33

Sylvain Gugger authored Mar 26, 2021

* Add ImageFeatureExtractionMixin

* Add dummy vision objects

* Add require_vision

* Add tests

* Fix test

b0595d33

07 Jan, 2021 1 commit

Transformers fast import part 2 (#9446) · 758ed333

Sylvain Gugger authored Jan 07, 2021



* Main init work

* Add version

* Change from absolute to relative imports

* Fix imports

* One more typo

* More typos

* Styling

* Make quality script pass

* Add necessary replace in template

* Fix typos

* Spaces are ignored in replace for some reason

* Forgot one models.

* Fixes for import
Co-authored-by: LysandreJik <lysandre.debut@reseau.eseo.fr>

* Add documentation

* Styling
Co-authored-by: LysandreJik <lysandre.debut@reseau.eseo.fr>

758ed333

12 Nov, 2020 1 commit
- Use LF instead of os.linesep (#8491) · 91a67b75
  Julien Plu authored Nov 12, 2020
  
  91a67b75
10 Nov, 2020 1 commit

Model versioning (#8324) · 70f622fa

Julien Chaumond authored Nov 10, 2020

* fix typo

* rm use_cdn & references, and implement new hf_bucket_url

* I'm pretty sure we don't need to `read` this file

* same here

* [BIG] file_utils.networking: do not gobble up errors anymore

* Fix CI 😇



* Apply suggestions from code review
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Tiny doc tweak

* Add doc + pass kwarg everywhere

* Add more tests and explain

cc @sshleifer let me know if better
Co-Authored-By: Sam Shleifer <sshleifer@gmail.com>

* Also implement revision in pipelines

In the case where we're passing a task name or a string model identifier

* Fix CI 😇



* Fix CI

* [hf_api] new methods + command line implem

* make style

* Final endpoints post-migration

* Fix post-migration

* Py3.6 compat

cc @stefan-it

Thank you @stas00
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sam Shleifer <sshleifer@gmail.com>

70f622fa

20 Oct, 2020 1 commit
- Add Flax dummy objects (#7918) · 6d4f8bd0
  Sylvain Gugger authored Oct 20, 2020
  
  6d4f8bd0
18 Oct, 2020 1 commit

[Dependencies|tokenizers] Make both SentencePiece and Tokenizers optional dependencies (#7659) · ba8c4d0a

Thomas Wolf authored Oct 18, 2020

* splitting fast and slow tokenizers [WIP]

* [WIP] splitting sentencepiece and tokenizers dependencies

* update dummy objects

* add name_or_path to models and tokenizers

* prefix added to file names

* prefix

* styling + quality

* spliting all the tokenizer files - sorting sentencepiece based ones

* update tokenizer version up to 0.9.0

* remove hard dependency on sentencepiece 🎉

* and removed hard dependency on tokenizers 🎉



* update conversion script

* update missing models

* fixing tests

* move test_tokenization_fast to main tokenization tests - fix bugs

* bump up tokenizers

* fix bert_generation

* update ad fix several tokenizers

* keep sentencepiece in deps for now

* fix funnel and deberta tests

* fix fsmt

* fix marian tests

* fix layoutlm

* fix squeezebert and gpt2

* fix T5 tokenization

* fix xlnet tests

* style

* fix mbart

* bump up tokenizers to 0.9.2

* fix model tests

* fix tf models

* fix seq2seq examples

* fix tests without sentencepiece

* fix slow => fast  conversion without sentencepiece

* update auto and bert generation tests

* fix mbart tests

* fix auto and common test without tokenizers

* fix tests without tokenizers

* clean up tests lighten up when tokenizers + sentencepiece are both off

* style quality and tests fixing

* add sentencepiece to doc/examples reqs

* leave sentencepiece on for now

* style quality split hebert and fix pegasus

* WIP Herbert fast

* add sample_text_no_unicode and fix hebert tokenization

* skip FSMT example test for now

* fix style

* fix fsmt in example tests

* update following Lysandre and Sylvain's comments

* Update src/transformers/testing_utils.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/testing_utils.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/tokenization_utils_base.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/tokenization_utils_base.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

ba8c4d0a

05 Oct, 2020 1 commit

Allow soft dependencies in the namespace with ImportErrors at use (#7537) · 28d183c9

Sylvain Gugger authored Oct 05, 2020

* PoC on RAG

* Format class name/obj name

* Better name in message

* PoC on one TF model

* Add PyTorch and TF dummy objects + script

* Treat scikit-learn

* Bad copy pastes

* Typo

28d183c9