Commits · 5936c8c57ccb2bda3b3f28856a7ef992c5c9f451 · chenpangpang / transformers

22 Sep, 2023 9 commits

Fixed unclosed p tags (#26240) · 5936c8c5
HanSeokhyeon authored Sep 23, 2023

5936c8c5

feat: adding num_proc to load_dataset (#26326) · 910faa3e

Phuc Van Phan authored Sep 23, 2023

* feat: adding num_proc to load_dataset

* feat: add add_num_proc for run_mlm_flax

* feat: add num_proc for bart and t5

* chorse: remove

910faa3e

Add image to image pipeline (#25393) · 576cd45a

LeviVasconcelos authored Sep 22, 2023



* Add image to image pipeline

Add image to image pipeline

* remove swin2sr from tf auto

* make ImageToImage importable

* make style

make style

make style

make style

* remove tf support

* remove nonused imports

* fix postprocessing

* add important comments; add unit tests

* add documentation

* remove support for TF

* make fixup

* fix typehint Image.Image

* fix documentation code

* address review request; fix unittest type checking

* address review request; fix unittest type checking

* make fixup

* address reviews

* Update src/transformers/pipelines/image_to_image.py
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>

* enhance docs

* make style

* make style

* improve docetest time

* improve docetest time

* Update tests/pipelines/test_pipelines_image_to_image.py
Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com>

* Update tests/pipelines/test_pipelines_image_to_image.py
Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com>

* make fixup

* undo faulty merge

* undo faulty merge

* add image-to-image to test pipeline mixin

* Update src/transformers/pipelines/image_to_image.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* Update tests/pipelines/test_pipelines_image_to_image.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* improve docs

---------
Co-authored-by: NielsRogge <48327001+NielsRogge@users.noreply.github.com>
Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com>
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

576cd45a

[TTA Pipeline] Fix MusicGen test (#26348) · 914771cb
Sanchit Gandhi authored Sep 22, 2023
```
* fix musicgen pipeline test

* fix wav2vec2 doctest

* revert wav2vec2
```
914771cb

[`core` ] Integrate Flash attention 2 in most used models (#25598) · 368a58e6

Younes Belkada authored Sep 22, 2023



* v1

* oops

* working v1

* fixup

* add some TODOs

* fixup

* padding support + try with module replacement

* nit

* alternative design

* oops

* add `use_cache` support for llama

* v1 falcon

* nit

* a bit of refactor

* nit

* nits nits

* add v1 padding support falcon (even though it seemed to work before)

* nit

* falcon works

* fixup

* v1 tests

* nit

* fix generation llama flash

* update tests

* fix tests + nits

* fix copies

* fix nit

* test- padding mask

* stype

* add more mem efficient support

* Update src/transformers/modeling_utils.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* fixup

* nit

* fixup

* remove it from config when saving

* fixup

* revert docstring

* add more checks

* use values

* oops

* new version

* fixup

* add same trick for falcon

* nit

* add another test

* change tests

* fix issues with GC and also falcon

* fixup

* oops

* Update src/transformers/models/falcon/modeling_falcon.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* add init_rope

* updates

* fix copies

* fixup

* fixup

* more clarification

* fixup

* right padding tests

* add docs

* add FA in docker image

* more clarifications

* add some figures

* add todo

* rectify comment

* Change to FA2

* Update docs/source/en/perf_infer_gpu_one.md
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* split in two lines

* change test name

* add more tests

* some clean up

* remove `rearrange` deps

* add more docs

* revert changes on dockerfile

* Revert "revert changes on dockerfile"

This reverts commit 8d72a66b4b9b771abc3f15a9b9506b4246d62d8e.

* revert changes on dockerfile

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <hi@lysand.re>

* address some comments

* docs

* use inheritance

* Update src/transformers/testing_utils.py
Co-authored-by: Lysandre Debut <hi@lysand.re>

* fixup

* Apply suggestions from code review
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* Update src/transformers/modeling_utils.py

* final comments

* clean up

* style

* add cast + warning for PEFT models

* fixup

---------
Co-authored-by: Felix Marty <9808326+fxmarty@users.noreply.github.com>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>
Co-authored-by: Lysandre Debut <hi@lysand.re>

368a58e6

[doc] fixed indices in obj detection example (#26343) · dcbfd93d
Maria Khalusova authored Sep 22, 2023
```
fixed indexes in obj detection example
```
dcbfd93d

Fix doctest CI (#26324) · c3ecf2d9

Yih-Dar authored Sep 22, 2023



fix doc CI
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>

c3ecf2d9

Use CircleCI `store_test_results` (#26223) · 06ee91ae
Yih-Dar authored Sep 22, 2023
```
store_test_results
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
06ee91ae

[QUICK FIX LINK] Update trainer.py (#26293) · 587b7b16

Gema Parreño authored Sep 22, 2023



* Update trainer.py

Fix link

* Update src/transformers/trainer.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* Update trainer.py

---------
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

587b7b16

21 Sep, 2023 5 commits

More error message fixup, plus some linebreaks! (#26296) · 000e52ae

Matt authored Sep 21, 2023



* More error message fixup, plus some linebreaks!

* Update src/transformers/dynamic_module_utils.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* Update src/transformers/dynamic_module_utils.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* Update src/transformers/dynamic_module_utils.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

---------
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

000e52ae

Porting the torchaudio kaldi fbank implementation to audio_utils (#26182) · 9a307534

Yoach Lacombe authored Sep 21, 2023



* add kaldi fbank

* make style

* add herz_to_mel_kaldi tests

* add mel to hertz kaldi test

* integration tests

* correct test and remove comment

* make style

* Apply suggestions from code review
Co-authored-by: Sanchit Gandhi <93869735+sanchit-gandhi@users.noreply.github.com>

* change parameter name

* Apply suggestions from Arthur review
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* Update remove_dc_offset description

* fix bug  + make style

* fix error in using np.exp instead of np.power

* make style

---------
Co-authored-by: Sanchit Gandhi <93869735+sanchit-gandhi@users.noreply.github.com>
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

9a307534

update hf hub dependency to be compatible with the new tokenizers (#26301) · b132c170
Arthur authored Sep 21, 2023

b132c170
Fix FSMT weight sharing (#26292) · 26ba56cc
Lysandre Debut authored Sep 21, 2023

26ba56cc

Keep relevant weights in fp32 when `model._keep_in_fp32_modules` is set even... · da971b22

fxmarty authored Sep 21, 2023

Keep relevant weights in fp32 when `model._keep_in_fp32_modules` is set even when `accelerate` is not installed (#26225)

* fix bug where weight would not be kept in fp32

* nit

* address review comments

* fix test

da971b22

20 Sep, 2023 10 commits

add custom RMSNorm to `ALL_LAYERNORM_LAYERS` (#26227) · e3a4bd2b

Shijie Wu authored Sep 20, 2023

* add LlamaRMSNorm to ALL_LAYERNORM_LAYERS

* fixup

* add IdeficsRMSNorm to ALL_LAYERNORM_LAYERS and fixup

e3a4bd2b

[`Trainer`] Refactor trainer + bnb logic (#26248) · 0b5024ce
Younes Belkada authored Sep 20, 2023
```
* refactor trainer + bnb logic

* remove logger.info

* oops
```
0b5024ce
include changes from llama (#26260) · f94c9b3d
Arthur authored Sep 20, 2023
```
* include changes from llama

* add a test
```
f94c9b3d
add bbox input validation (#26294) · 00247ea0
Jinho Park authored Sep 20, 2023

00247ea0
fix deepspeed available detection (#26252) · 24553206
fxmarty authored Sep 20, 2023

24553206
Rewrite for custom code warning messages (#26291) · f29fe745
Matt authored Sep 20, 2023
```
Quick britpicking for some warning messages!
```
f29fe745

Integrate AMD GPU in CI/CD environment (#26007) · 2d71307d

Funtowicz Morgan authored Sep 20, 2023

* Add a Dockerfile for PyTorch + ROCm based on official AMD released artifact

* Add a new artifact single-amdgpu testing on main

* Attempt to test the workflow without merging.

* Changed BERT to check if things are triggered

* Meet the dependencies graph on workflow

* Revert BERT changes

* Add check_runners_amdgpu to correctly mount and check availability

* Rename setup to setup_gpu for CUDA and add setup_amdgpu for AMD

* Fix all the needs.setup -> needs.setup_[gpu|amdgpu] dependencies

* Fix setup dependency graph to use check_runner_amdgpu

* Let's do the runner status check only on AMDGPU target

* Update the Dockerfile.amd to put ourselves in / rather than /var/lib

* Restore the whole setup for CUDA too.

* Let's redisable them

* Change BERT to trigger tests

* Restore BERT

* Add torchaudio with rocm 5.6 to AMD Dockerfile (#26050)

fix dockerfile
Co-authored-by: Felix Marty <felix@hf.co>

* Place AMD GPU tests in a separate workflow (correct branch) (#26105)

AMDGPU CI lives in an other workflow

* Fix invalid job name is dependencies.

* Remove tests multi-amdgpu for now.

* Use single-amdgpu

* Use --net=host for now.

* Remote host networking.

* Removed duplicated check_runners_amdgpu step

* Let's tag machine-types with mi210 for now.

* Machine type should be only mi210

* Remove unnecessary push.branches item

* Apply review suggestions moving from `x-amdgpu` to `x-gpu` introducing `amd-gpu` and `miXXX` labels.

* Remove amdgpu from step names.

* finalize

* delete

---------
Co-authored-by: fxmarty <9808326+fxmarty@users.noreply.github.com>
Co-authored-by: Felix Marty <felix@hf.co>
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>

2d71307d

Update bros checkpoint (#26277) · 37c205eb
Jinho Park authored Sep 20, 2023
```
* fix bros integration test

* update bros checkpoint
```
37c205eb
fix name error when accelerate is not available (#26278) · 86ffd5ff
Sourab Mangrulkar authored Sep 20, 2023
```
* fix name error when accelerate is not available

* fix `is_fsdp_available`
```
86ffd5ff

FSDP tests and checkpointing fixes (#26180) · 382ba670

Sourab Mangrulkar authored Sep 20, 2023



* add fsdp tests

* Update test_fsdp.py

* Update test_fsdp.py

* fixes

* checks

* Update trainer.py

* fix

* fixes for saving/resuming checkpoints

* fixes

* add tests and delete debug statements

* fixing tests

* Update test_fsdp.py

* fix tests

* fix tests

* minor nits

* fix code style and quality

* refactor and modularize test code

* reduce the time of tests

* reduce the test time

* fix test

* reduce test time

* reduce test time

* fix failing tests

* fix

* Apply suggestions from code review
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* resolve comments

---------
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

382ba670

19 Sep, 2023 6 commits

[FIX] resize_token_embeddings (#26102) · 8e3980a2

Sam Passaglia authored Sep 20, 2023



* fix roundup command

* add test for resize_token_embeddings

* Update tests/test_modeling_common.py
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

* style

---------
Co-authored-by: Arthur <48595927+ArthurZucker@users.noreply.github.com>

8e3980a2

DeepSpeed ZeRO-3 handling when resizing embedding layers (#26259) · ffbf989f
Sourab Mangrulkar authored Sep 20, 2023
```
* fix failing deepspeed slow tests

* fixes
```
ffbf989f

Fix `Error` not captured in PR doctesting (#26215) · 39df4eca

Yih-Dar authored Sep 19, 2023



* fix

* fix

* fix

* fix

---------
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>

39df4eca

Add ViTMatte (#25843) · 7d6354e0

NielsRogge authored Sep 19, 2023

* First draft

* Simplify image processor

* Fix rebase

* Address comments

* Address more comments

* Address more comments

* Address more comments

* Address more comments

* Improve pad_image

* Add tests

* Update integration test

* Fix image processor tests

* Fix model tests

* Convert checkpoints

* Fix doc tests

* Remove file

* Apply suggestions

* Address comments

* Fix typing hint

* Add batch_norm_eps

* Address comments

* Fix style

7d6354e0

Fix gated repo tests (#26257) · 04191ea1
Lucain authored Sep 19, 2023
```
* Fix gated repo tests

* Apply suggestions from code review
```
04191ea1
Fix some docstring in image processors (#26235) · eb848997
Yih-Dar authored Sep 19, 2023
```
Fix doc
Co-authored-by: ydshieh <ydshieh@users.noreply.github.com>
```
eb848997

18 Sep, 2023 10 commits

Fix the gitlab user mention in issue templates to the correct user (#26237) · e469be34
Ralf Müller-Zimmermann authored Sep 19, 2023

e469be34
[docs] Fix model reference in zero shot image classification example (#26206) · 373d0d99
Aleksandar Ivanovski authored Sep 19, 2023

373d0d99
Update add_new_pipeline.md (#26197) · 500dfb5b
Nino Risteski authored Sep 19, 2023
```
fixed a few typos
```
500dfb5b
Update README.md (#26198) · 7d4e0c23
Nino Risteski authored Sep 19, 2023
```
Fixed a few typos
```
7d4e0c23
[AutoBackbone] Add test (#26094) · de8bec6d
NielsRogge authored Sep 18, 2023
```
* Add test

* Add config_class
```
de8bec6d
Create the return value on device to avoid unnecessary copying from CPU (#26151) · 97f439ae
mksit authored Sep 18, 2023

97f439ae

🌐

[i18n-KO] Translated `whisper.md` to Korean (#26002) · 42791a57

SeongWooChoi authored Sep 19, 2023



* docs: ko-whisper.md

* fix: chatgpt draft

* feat: manual edits

* Feat: manual edits

* fix: resolve suggestions
Co-authored-by: Jungnerd <46880056+jungnerd@users.noreply.github.com>

---------
Co-authored-by: Jungnerd <46880056+jungnerd@users.noreply.github.com>

42791a57

🚨

[`Tokenizer`] attemp to fix add_token issues

🚨

(#23909) · 2da88537

Arthur authored Sep 18, 2023

* fix test for bart. Order is correct now let's skip BPEs

* ouf

* styling

* fix bert....

* slow refactoring

* current updates

* massive refactoring

* update

* NICE!

* update to see where I am at

* updates

* update

* update

* revert

* updates

* updates

* start supporting legacy_save

* styling

* big update

* revert some changes

* nits

* nniiiiiice

* small fixes

* kinda fix t5 with new behaviour

* major update

* fixup

* fix copies

* today's updates

* fix byt5

* upfate

* update

* update

* updates

* update vocab size test

* Barthez does not use not need the fairseq offset ids

* super calll must be after

* calll super

* move all super init

* move other super init

* fixup

* nits

* more fixes

* nits

* more fixes

* nits

* more fix

* remove useless files

* ouch all of them are affected
...

2da88537

[Check] Fix config docstring (#26222) · 835b0a05
Sanchit Gandhi authored Sep 18, 2023

835b0a05
[Permisson] Style fix (#26228) · e5f7e03b
Sanchit Gandhi authored Sep 18, 2023
```
fix copies
```
e5f7e03b