Commits · b463651f3eeb313d13d10db28fc08d7a0277cfbf · OpenDAS / ColossalAI

19 Jun, 2023 1 commit
- [nfc] fix dim not defined and fix typo (#3991) · 727c4598
  digger yu authored Jun 19, 2023
  
  727c4598
15 Jun, 2023 1 commit
- fix typo applications/Chat/coati/ (#3947) · d4fb7bfd
  digger yu authored Jun 15, 2023
  
  d4fb7bfd
13 Jun, 2023 2 commits

[evaluate] support gpt evaluation with reference (#3972) · 2925f473
Yuanchen authored Jun 13, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
2925f473

[chat] refactor actor class (#3968) · 9d02590c

Wenhao Chen authored Jun 13, 2023

* refactor: separate log_probs fn from Actor forward fn

* refactor: separate generate fn from Actor class

* feat: update unwrap_model and get_base_model
* unwrap_model returns model not wrapped by Strategy
* get_base_model returns HF model for Actor, Critic and RewardModel

* feat: simplify Strategy.prepare

* style: remove get_base_model method of Actor

* perf: tokenize text in batches

* refactor: move calc_action_log_probs to utils of model

* test: update test with new forward fn

* style: rename forward fn args

* fix: do not unwrap model in save_model fn of naive strategy

* test: add gemini test for train_prompts

* fix: fix _set_default_generate_kwargs

9d02590c

08 Jun, 2023 1 commit
- support UniEval and add CHRF metric (#3924) · 21c4c0b1
  Yuanchen authored Jun 08, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  21c4c0b1
07 Jun, 2023 1 commit

[chat] add distributed PPO trainer (#3740) · b5f05663

Hongxin Liu authored Jun 07, 2023



* Detached ppo (#9)

* run the base

* working on dist ppo

* sync

* detached trainer

* update detached trainer. no maker update function

* facing init problem

* 1 maker 1 trainer detached run. but no model update

* facing cuda problem

* fix save functions

* verified maker update

* nothing

* add ignore

* analyize loss issue

* remove some debug codes

* facing 2m1t stuck issue

* 2m1t verified

* do not use torchrun

* working on 2m2t

* working on 2m2t

* initialize strategy in ray actor env

* facing actor's init order issue

* facing ddp model update issue (need unwarp ddp)

* unwrap ddp actor

* checking 1m2t stuck problem

* nothing

* set timeout for trainer choosing. It solves the stuck problem!

* delete some debug output

* rename to sync with upstream

* rename to sync with upstream

* coati rename

* nothing

* I am going to detach the replaybuffer from trainer and make it a Ray Actor. Two benefits: 1. support TP trainer. 2. asynchronized buffer operations

* experience_maker_holder performs target-revolving _send_experience() instead of length comparison.

* move code to ray subfolder

* working on pipeline inference

* apply comments

* working on pipeline strategy. in progress.

* remove pipeline code. clean this branch

* update remote parameters by state_dict. no test

* nothing

* state_dict sharding transfer

* merge debug branch

* gemini _unwrap_model fix

* simplify code

* simplify code & fix LoRALinear AttributeError

* critic unwrapped state_dict

---------
Co-authored-by: csric <richcsr256@gmail.com>

* [chat] add perfomance evaluator and fix bugs (#10)

* [chat] add performance evaluator for ray

* [chat] refactor debug arg

* [chat] support hf config

* [chat] fix generation

* [chat] add 1mmt dummy example

* [chat] fix gemini ckpt

* split experience to send (#11)
Co-authored-by: csric <richcsr256@gmail.com>

* [chat] refactor trainer and maker (#12)

* [chat] refactor experience maker holder

* [chat] refactor model init

* [chat] refactor trainer args

* [chat] refactor model init

* [chat] refactor trainer

* [chat] refactor experience sending logic and training loop args (#13)

* [chat] refactor experience send logic

* [chat] refactor trainer

* [chat] refactor trainer

* [chat] refactor experience maker

* [chat] refactor pbar

* [chat] refactor example folder (#14)

* [chat] support quant (#15)

* [chat] add quant

* [chat] add quant example

* prompt example (#16)

* prompt example

* prompt load csv data

* remove legacy try

---------
Co-authored-by: csric <richcsr256@gmail.com>

* [chat] add mmmt dummy example and refactor experience sending (#17)

* [chat] add mmmt dummy example

* [chat] refactor naive strategy

* [chat] fix struck problem

* [chat] fix naive strategy

* [chat] optimize experience maker sending logic

* [chat] refactor sending assignment

* [chat] refactor performance evaluator (#18)

* Prompt Example & requires_grad state_dict & sharding state_dict (#19)

* prompt example

* prompt load csv data

* remove legacy try

* maker models require_grad set to False

* working on zero redundancy update

* mmmt_prompt example; naive strategy requires_grad state_dict & sharding; maker model requires_no_grad.

* remove legacy examples

* remove legacy examples

* remove replay buffer tp state. bad design

---------
Co-authored-by: csric <richcsr256@gmail.com>

* state_dict sending adapts to new unwrap function (#20)

* prompt example

* prompt load csv data

* remove legacy try

* maker models require_grad set to False

* working on zero redundancy update

* mmmt_prompt example; naive strategy requires_grad state_dict & sharding; maker model requires_no_grad.

* remove legacy examples

* remove legacy examples

* remove replay buffer tp state. bad design

* opt benchmark

* better script

* nothing

* [chat] strategy refactor unwrap model

* [chat] strategy refactor save model

* [chat] add docstr

* [chat] refactor trainer save model

* [chat] fix strategy typing

* [chat] refactor trainer save model

* [chat] update readme

* [chat] fix unit test

* working on lora reconstruction

* state_dict sending adapts to new unwrap function

* remove comments

---------
Co-authored-by: csric <richcsr256@gmail.com>
Co-authored-by: ver217 <lhx0217@gmail.com>

* [chat-ray] add readme (#21)

* add readme

* transparent graph

* add note background

---------
Co-authored-by: csric <richcsr256@gmail.com>

* [chat] get images from url (#22)

* Refactor/chat ray (#23)

* [chat] lora add todo

* [chat] remove unused pipeline strategy

* [chat] refactor example structure

* [chat] setup ci for ray

* [chat-ray] Support LoRA trainer. LoRA weights reconstruction. (#24)

* lora support prototype

* lora support

* 1mmt lora & remove useless code

---------
Co-authored-by: csric <richcsr256@gmail.com>

* [chat] fix test ci for ray

* [chat] fix test ci requirements for ray

* [chat] fix ray runtime env

* [chat] fix ray runtime env

* [chat] fix example ci docker args

* [chat] add debug info in trainer

* [chat] add nccl debug info

* [chat] skip ray test

* [doc] fix typo

---------
Co-authored-by: csric <59389055+CsRic@users.noreply.github.com>
Co-authored-by: csric <richcsr256@gmail.com>

b5f05663

05 Jun, 2023 1 commit
- support evaluation for english (#3880) · 57a6d768
  Yuanchen authored Jun 05, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  57a6d768
30 May, 2023 1 commit

[evaluation] improvement on evaluation (#3862) · 2506e275

Yuanchen authored May 30, 2023



* fix a bug when the config file contains one category but the answer file doesn't contains that category

* fix Chinese prompt file

* support gpt-3.5-turbo and gpt-4 evaluation

* polish and update README

* resolve pr comments

---------
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>

2506e275

25 May, 2023 1 commit

[nfc] fix typo colossalai/ applications/ (#3831) · e2d81eba

digger yu authored May 25, 2023

* fix typo colossalai/autochunk auto_parallel amp

* fix typo colossalai/auto_parallel nn utils etc.

* fix typo colossalai/auto_parallel autochunk fx/passes  etc.

* fix typo docs/

* change placememt_policy to placement_policy in docs/ and examples/

* fix typo colossalai/ applications/

e2d81eba

24 May, 2023 1 commit

[evaluation] add automatic evaluation pipeline (#3821) · 34966378

Yuanchen authored May 24, 2023



* add functions for gpt evaluation

* add automatic eval

Update eval.py

* using jload and modify the type of answers1 and answers2

* Update eval.py

Update eval.py

* Update evaluator.py

* support gpt evaluation

* update readme.md

update README.md

update READNE.md

modify readme.md

* add Chinese example for config, battle prompt and evaluation prompt file

* remove GPT-4 config

* remove sample folder

---------
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
Co-authored-by: Camille Zhong <44392324+Camille7777@users.noreply.github.com>

34966378

23 May, 2023 1 commit
- [NFC]fix typo colossalai/auto_parallel nn utils etc. (#3779) · 9265f2d4
  digger yu authored May 23, 2023
```
* fix typo colossalai/autochunk auto_parallel amp

* fix typo colossalai/auto_parallel nn utils etc.
```
  9265f2d4
22 May, 2023 1 commit
- [format] applied code formatting on changed files in pull request 3786 (#3787) · 62c7e67f
  github-actions[bot] authored May 22, 2023
```
Co-authored-by: github-actions <github-actions@github.com>
```
  62c7e67f
19 May, 2023 1 commit
- [chat] add performance and tutorial (#3786) · ad2cf58f
  binmakeswell authored May 19, 2023
  
  ad2cf58f
17 May, 2023 1 commit
- [chat] fix bugs in stage 3 training (#3759) · 05759839
  Yuanchen authored May 17, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  05759839
15 May, 2023 1 commit
- [NFC] fix typo applications/ and colossalai/ (#3735) · ad6460cf
  digger-yu authored May 15, 2023
  
  ad6460cf
10 May, 2023 2 commits

[CI] fix some spelling errors (#3707) · b7141c36

digger-yu authored May 10, 2023

* fix spelling error with examples/comminity/

* fix spelling error with tests/

* fix some spelling error with tests/ colossalai/ etc.

b7141c36

[chat] fix community example ray (#3719) · f7361ee1
MisterLin1995 authored May 10, 2023
```
Co-authored-by: jiangwen <zxl265370@antgroup.com>
```
f7361ee1

06 May, 2023 2 commits
- [chat] fix train_prompts.py gemini strategy bug (#3666) · 2da5d81d
  zhang-yi-chi authored May 06, 2023
```
* fix gemini strategy bug

* add comment

* add comment

* better solution
```
  2da5d81d
- fix some spelling error with applications/Chat/examples/ (#3692) · 65bdc315
  digger-yu authored May 06, 2023
```
* fix spelling error with examples/comminity/

* fix spelling error with example/
```
  65bdc315
05 May, 2023 2 commits

[chat] PPO stage3 doc enhancement (#3679) · 0f785cb1

Camille Zhong authored May 05, 2023

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

update roberta with coati

chat ci update

Revert "chat ci update"

This reverts commit 17ae7ae01fa752bd3289fc39069868fde99cf846.

* Update README.md

Update README.md

* update readme

* Update test_ci.sh

* update readme and add a script

update readme and add a script

modify readme

Update README.md

0f785cb1

[doc] fix chat spelling error (#3671) · 6650daeb

digger-yu authored May 05, 2023

* Update README.md

change "huggingaface" to "huggingface"

* Update README.md

change "Colossa-AI" to "Colossal-AI"

6650daeb

04 May, 2023 3 commits
- [chat] add opt attn kernel (#3655) · 7bd0bee8
  Hongxin Liu authored May 04, 2023
```
* [chat] add opt attn kernel

* [chat] disable xformer during fwd
```
  7bd0bee8
- Update generate_gpt35_answers.py · 8ba78587
  digger-yu authored May 04, 2023
```
fix spelling error with generate_gpt35_answers.py
```
  8ba78587
- fix spelling error · bfbf6505
  digger-yu authored May 04, 2023
```
fix spelling error with evaluate.py
```
  bfbf6505
28 Apr, 2023 4 commits
- [chat] typo accimulation_steps -> accumulation_steps (#3662) · 1a60dc07
  tanitna authored Apr 28, 2023
  
  1a60dc07
- [chat] set default zero2 strategy (#3667) · 268b3cd8
  binmakeswell authored Apr 28, 2023
```
* [chat] set default gemini strategy

* [chat] set default zero2 strategy

* [chat] set default zero2 strategy
```
  268b3cd8
- update readme · c1a35594
  Tong Li authored Apr 28, 2023
  
  c1a35594
- update documentation · ed3eaa69
  Tong Li authored Apr 28, 2023
  
  ed3eaa69
27 Apr, 2023 5 commits

update questions and readme · c4191173
Tong Li authored Apr 27, 2023

c4191173
remove unnecessary step and update readme · aa77ddae
Tong Li authored Apr 27, 2023

aa77ddae

[chat] refactor model save/load logic (#3654) · 842768a1

Hongxin Liu authored Apr 27, 2023

* [chat] strategy refactor unwrap model

* [chat] strategy refactor save model

* [chat] add docstr

* [chat] refactor trainer save model

* [chat] fix strategy typing

* [chat] refactor trainer save model

* [chat] update readme

* [chat] fix unit test

842768a1

[chat] remove lm model class (#3653) · 6ef70114

Hongxin Liu authored Apr 27, 2023

* [chat] refactor lora

* [chat] remove lm class

* [chat] refactor save model

* [chat] refactor train sft

* [chat] fix ci

* [chat] fix ci

6ef70114

[Doc] enhancement on README.md for chat examples (#3646) · 8bccb72c

Camille Zhong authored Apr 27, 2023

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

update roberta with coati

chat ci update

Revert "chat ci update"

This reverts commit 17ae7ae01fa752bd3289fc39069868fde99cf846.

* Update README.md

Update README.md

* update readme

* Update test_ci.sh

8bccb72c

26 Apr, 2023 3 commits

[chat] refactor trainer (#3648) · 2a951955

Hongxin Liu authored Apr 26, 2023

* [chat] ppo trainer remove useless args

* [chat] update examples

* [chat] update benchmark

* [chat] update examples

* [chat] fix sft training with wandb

* [chat] polish docstr

2a951955

[chat] polish performance evaluator (#3647) · f8288315
Hongxin Liu authored Apr 26, 2023

f8288315

[gemini] accelerate inference (#3641) · 50793b35

Hongxin Liu authored Apr 26, 2023

* [gemini] support don't scatter after inference

* [chat] update colossalai strategy

* [chat] fix opt benchmark

* [chat] update opt benchmark

* [gemini] optimize inference

* [test] add gemini inference test

* [chat] fix unit test ci

* [chat] fix ci

* [chat] fix ci

* [chat] skip checkpoint test

50793b35

24 Apr, 2023 1 commit
- [Chat] Remove duplicate functions (#3625) · df309fc6
  ddobokki authored Apr 24, 2023
  
  df309fc6
22 Apr, 2023 1 commit
- [chat] fix enable single gpu training bug · 739cfe33
  zhang-yi-chi authored Apr 22, 2023
  
  739cfe33
20 Apr, 2023 2 commits
- [chat] polish code note typo (#3612) · d7bf2847
  digger-yu authored Apr 20, 2023
  
  d7bf2847
- Chat evaluate (#3608) · c4709d34
  Yuanchen authored Apr 20, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  c4709d34