Commits · f0fc0aea6bd885ee90837cd96c3935a5d44e060a · chenpangpang / transformers

08 Sep, 2020 9 commits

pegasus.rst: fix expected output (#7017) · f0fc0aea
Sam Shleifer authored Sep 08, 2020

f0fc0aea
[Longformer] Fix longformer documentation (#7016) · 120176ea
Patrick von Platen authored Sep 08, 2020
```
* fix longformer

* allow position ids to not be initialized
```
120176ea

Fixing FLOPS merge by checking if torch is available (#7013) · 5c4eb4b1

Lysandre Debut authored Sep 08, 2020



* Should check if `torch` is available

* fixed samples_count error, distributed_concat arguments

* style

* Import torch at beginning of file
Co-authored-by: TevenLeScao <teven.lescao@gmail.com>

5c4eb4b1

Floating-point operations logging in trainer (#6768) · 01d340ad

Teven authored Sep 08, 2020



* neFLOs calculation, logging, and reloading (#1)

* testing distributed consecutive batches

* fixed AttributeError from DataParallel

* removed verbosity

* rotate with use_mtime=True

* removed print

* fixed interaction with gradient accumulation

* indent formatting

* distributed neflo counting

* fixed typo

* fixed typo

* mean distributed losses

* exporting log history

* moved a few functions

* floating_point_ops clarification for transformers with parameter-reuse

* code quality

* double import

* made flo estimation more task-agnostic

* only logging flos if computed

* code quality

* unused import

* Update src/transformers/trainer.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Update src/transformers/modeling_utils.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* Sylvain review

* Update src/transformers/modeling_utils.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

* black
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

01d340ad

Funnel transformer (#6908) · d155b38d

Sylvain Gugger authored Sep 08, 2020



* Initial model

* Fix upsampling

* Add special cls token id and test

* Formatting

* Test and fist FunnelTokenizerFast

* Common tests

* Fix the check_repo script and document Funnel

* Doc fixes

* Add all models

* Write doc

* Fix test

* Initial model

* Fix upsampling

* Add special cls token id and test

* Formatting

* Test and fist FunnelTokenizerFast

* Common tests

* Fix the check_repo script and document Funnel

* Doc fixes

* Add all models

* Write doc

* Fix test

* Fix copyright

* Forgot some layers can be repeated

* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Update src/transformers/modeling_funnel.py
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>

* Address review comments

* Update src/transformers/modeling_funnel.py
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>

* Address review comments

* Update src/transformers/modeling_funnel.py
Co-authored-by: Sam Shleifer <sshleifer@gmail.com>

* Slow integration test

* Make small integration test

* Formatting

* Add checkpoint and separate classification head

* Formatting

* Expand list, fix link and add in pretrained models

* Styling

* Add the model in all summaries

* Typo fixes
Co-authored-by: Lysandre Debut <lysandre@huggingface.co>
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com>
Co-authored-by: Sam Shleifer <sshleifer@gmail.com>

d155b38d

fixed trainer tr_loss memory leak (#6999) · 25afb4ea

Stuart Mesham authored Sep 08, 2020

* fixed trainer tr_loss memory leak

* detached returned training loss from computation graph in the Trainer class' training_step() method

* Revert "fixed trainer tr_loss memory leak"

This reverts commit 47226e4e

25afb4ea

Fix typo (#6994) · 1b76936d
Manuel Romero authored Sep 08, 2020

1b76936d
New Community NB "Fine tune GPT-2 with Trainer class" (#7005) · 8235426e
Philipp Schmid authored Sep 08, 2020

8235426e

typo (#7001) · c18f5916

Stas Bekman authored Sep 07, 2020

apologies for the tiny PRs, just sending those as I find them.

c18f5916

07 Sep, 2020 18 commits

README for HooshvareLab/bert-fa-base-uncased (#6990) · 60fc0329

Mehrdad Farahani authored Sep 08, 2020

ParsBERT v2.0 is a fine-tuned and vocab-reconstructed version of ParsBERT, and it's able to be used in other scopes!

It includes these features:
- We added some unused-vocab for use in summarization and other scopes.
- We fine-tuned the model on vast styles of writing in the Persian language.

60fc0329

Add missing arguments for BertWordPieceTokenizer (#5810) · 90ec78b5
Jangwon Park authored Sep 07, 2020

90ec78b5
Conversion scripts shouldn't have relative imports (#6991) · 77cd0e13
Lysandre Debut authored Sep 07, 2020

77cd0e13
Remove misleading docstring · 1650130b
Lysandre authored Sep 07, 2020

1650130b

match CI's version of flake8 (#6941) · 159ef07e

Stas Bekman authored Sep 07, 2020

my flake8 wasn't up-to-date enough `make quality` wasn't reporting the same things CI did - this PR adds the actual required version.

Thinking more about some of these minimal versions - CI will always install afresh and thus will always run the latest version. Is there a way to tell pip to always install the latest versions of certain dependencies on `pip install -i ".[dev]"`, rather than hardcoding the minimals which quickly become outdated?

159ef07e

Create README.md (#6974) · e9d0d4c7
Abed khooli authored Sep 07, 2020

e9d0d4c7

[gen utils] missing else case (#6980) · 848fbe1e

Stas Bekman authored Sep 07, 2020

* [gen utils] missing else case

1. `else` is missing - I hit that case while porting a model. Probably needs to assert there?
2. also the comment on top seems to be outdated (just vocab_size is being set there)

* typo

848fbe1e

Fixed the default number of attention heads in Reformer Configuration (#6973) · f7e80721
tznurmin authored Sep 07, 2020

f7e80721

Create README.md model card (#6964) · e20d8895

Richard Bownes authored Sep 07, 2020



* Create README.md

* Add some custom prompts
Co-authored-by: Julien Chaumond <chaumond@gmail.com>

e20d8895

[testing] add dependency: parametrize (#6958) · b4a9c95f

Stas Bekman authored Sep 07, 2020

unittest doesn't support pytest's super-handy `@pytest.mark.parametrize`, I researched and there are many proposed workarounds, most tedious at best. If we include https://pypi.org/project/parameterized/ in dev dependencies - it will provide a very easy to write parameterization in tests. Same as pytest's fixture, plus quite a few other ways. 

Example:
```
from parameterized import parameterized
@parameterized([
    (2, 2, 4),
    (2, 3, 8),
    (1, 9, 1),
    (0, 9, 0),
])
def test_pow(base, exponent, expected):
   assert_equal(math.pow(base, exponent), expected)
```
(extra `self`var if inside a test class)

To remind the pytest style is slightly different:
```
    @pytest.mark.parametrize("test_input,expected", [("3+5", 8), ("2+4", 6), ("6*9", 42)])
    def test_eval(test_input, expected):
```
More examples here: https://pypi.org/project/parameterized

May I suggest that it will make it much easier to write some types of tests?

b4a9c95f

[docstring] missing arg (#6933) · acfaad74

Stas Bekman authored Sep 07, 2020



* [docstring] missing arg

add the missing `tie_word_embeddings` entry

* cleanup

* Update src/transformers/configuration_reformer.py
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

acfaad74

typo (#6959) · c3317e1f

Stas Bekman authored Sep 07, 2020

there is no var `decoder_input_ids`, but there is `input_ids` for decoder :)

c3317e1f

[model_card] register jplu/tf-xlm-r-ner-40-lang as multilingual · 10c6f94a
Julien Chaumond authored Sep 07, 2020

10c6f94a
Cannot index `None` (#6984) · 9ef9c397
Lysandre Debut authored Sep 07, 2020

9ef9c397
Trainer with grad accum (#6930) · 08de989a
Sylvain Gugger authored Sep 07, 2020
```
* Add warning for gradient accumulation

* Formatting
```
08de989a
[model_card] jplu/tf-xlm-r-ner-40-lang: Fix link · d4aa7284
Julien Chaumond authored Sep 07, 2020
```
cc @jplu
```
d4aa7284

feat: allow prefix for any generative model (#5885) · 995a958d

Boris Dayma authored Sep 07, 2020



* feat: allow padding_text for any generative model

* docs(pipelines.py): correct typo

* Update src/transformers/pipelines.py
Co-authored-by: Sam Shleifer <sshleifer@gmail.com>

* feat: rename padding_text to prefix

* fix: cannot tokenize empty text

* fix: pass prefix arg to pipeline

* test: add prefix to text-generetation pipeline

* style: fix style

* style: clean code and variable name more explicit

* set arg docstring to optional
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>
Co-authored-by: Sam Shleifer <sshleifer@gmail.com>
Co-authored-by: Sylvain Gugger <35901082+sgugger@users.noreply.github.com>

995a958d

[s2s] warn if --fp16 for torch 1.6 (#6977) · ce37be9d
Sam Shleifer authored Sep 06, 2020

ce37be9d

06 Sep, 2020 1 commit
- Correct wrong spacing in README · f72fe1f3
  Patrick von Platen authored Sep 06, 2020
  
  f72fe1f3
05 Sep, 2020 1 commit

create model card for astroGPT (#6960) · d31031f6

Steven Liu authored Sep 05, 2020



* create model card for astroGPT

* Hotlink to actual image file
Co-authored-by: Julien Chaumond <chaumond@gmail.com>

d31031f6

04 Sep, 2020 8 commits
- Create Readme.MD for KanBERTo (#6942) · 56742e9f
  Naveenkhasyap authored Sep 05, 2020
```
* Create Readme.MD for KanBERTo

KanBERTo language model readme for Kannada language.

* Update model_cards/Naveen-k/KanBERTo/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com>
```
  56742e9f
- [doc] remove the implied defaults to :obj:`None`, s/True/ :obj:`True/, etc. (#6956) · 48ff6d51
  Stas Bekman authored Sep 04, 2020
```
* remove the implied defaults to :obj:`None`

* fix bug in the original

* replace to :obj:`True`, :obj:`False`
```
  48ff6d51
- typo (#6952) · eff274d6
  Stas Bekman authored Sep 04, 2020
  
  eff274d6
- [s2s] run_eval.py parses generate_kwargs (#6948) · a4fc0c80
  Sam Shleifer authored Sep 04, 2020
  
  a4fc0c80
- [s2s] distill: --normalize_hidden --supervise_forward (#6834) · 6078b120
  Sam Shleifer authored Sep 04, 2020
  
  6078b120
- [docstring] misc arg doc corrections (#6932) · c5d43a87
  Stas Bekman authored Sep 04, 2020
```
* correct bool types

fix docstring s/int/bool/

* fix description

* fix num_labels to match reality
```
  c5d43a87
- fix (#6946) · e3990d13
  Patrick von Platen authored Sep 04, 2020
  
  e3990d13
- Fix mixed precision issue in TF DistilBert (#6915) · a75e3198
  Yih-Dar authored Sep 04, 2020
```
* Remove hard-coded uses of float32 to fix mixed precision use in TF Distilbert

* fix style

* fix gelu dtype issue in TF Distilbert

* fix numeric overflow while using half precision
```
  a75e3198
03 Sep, 2020 3 commits
- [s2s] support early stopping based on loss, rather than rouge (#6927) · e95d262f
  Sam Shleifer authored Sep 03, 2020
  
  e95d262f
- [s2s] use --eval_beams command line arg (#6926) · 207ed8cb
  Sam Shleifer authored Sep 03, 2020
  
  207ed8cb
- move wandb/comet logger init to train() to allow parallel logging (#6850) · 0f360d3d
  krfricke authored Sep 03, 2020
```
* move wandb/comet logger init to train() to allow parallel logging

* Setup wandb/comet loggers on first call to log()
```
  0f360d3d