Commits · 8bccb72c8d6b4ff21d3d596f0188c6280d8b29f6 · OpenDAS / ColossalAI

27 Apr, 2023 1 commit

[Doc] enhancement on README.md for chat examples (#3646) · 8bccb72c

Camille Zhong authored Apr 27, 2023

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

update roberta with coati

chat ci update

Revert "chat ci update"

This reverts commit 17ae7ae01fa752bd3289fc39069868fde99cf846.

* Update README.md

Update README.md

* update readme

* Update test_ci.sh

8bccb72c

26 Apr, 2023 3 commits

[chat] refactor trainer (#3648) · 2a951955

Hongxin Liu authored Apr 26, 2023

* [chat] ppo trainer remove useless args

* [chat] update examples

* [chat] update benchmark

* [chat] update examples

* [chat] fix sft training with wandb

* [chat] polish docstr

2a951955

[chat] polish performance evaluator (#3647) · f8288315
Hongxin Liu authored Apr 26, 2023

f8288315

[gemini] accelerate inference (#3641) · 50793b35

Hongxin Liu authored Apr 26, 2023

* [gemini] support don't scatter after inference

* [chat] update colossalai strategy

* [chat] fix opt benchmark

* [chat] update opt benchmark

* [gemini] optimize inference

* [test] add gemini inference test

* [chat] fix unit test ci

* [chat] fix ci

* [chat] fix ci

* [chat] skip checkpoint test

50793b35

24 Apr, 2023 1 commit
- [Chat] Remove duplicate functions (#3625) · df309fc6
  ddobokki authored Apr 24, 2023
  
  df309fc6
22 Apr, 2023 1 commit
- [chat] fix enable single gpu training bug · 739cfe33
  zhang-yi-chi authored Apr 22, 2023
  
  739cfe33
20 Apr, 2023 2 commits
- [chat] polish code note typo (#3612) · d7bf2847
  digger-yu authored Apr 20, 2023
  
  d7bf2847
- Chat evaluate (#3608) · c4709d34
  Yuanchen authored Apr 20, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  c4709d34
18 Apr, 2023 3 commits

[coati] fix install cmd (#3592) · 5a79cffd
binmakeswell authored Apr 18, 2023

5a79cffd
reconstruct chat trainer and fix training script (#3588) · 1ec0d386
Yuanchen authored Apr 18, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
1ec0d386

Update test_ci.sh · 36a519b4

Camille Zhong authored Mar 22, 2023

update

Update test_ci.sh

Update test_ci.sh

Update test_ci.sh

Update test_ci.sh

Update test_ci.sh

Update test_ci.sh

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update test_ci.sh

Update test_ci.sh

update

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

update ci

Update test_ci.sh

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

Update test_ci.sh

Update test_ci.sh

Update run_chatgpt_examples.yml

Update test_ci.sh

Update test_ci.sh

Update test_ci.sh

update test ci

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

Update test_ci.sh

Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

update roberta with coati

chat ci update

Revert "chat ci update"

This reverts commit 17ae7ae01fa752bd3289fc39069868fde99cf846.

[test]chat_update_ci

Update test_ci.sh

Update test_ci.sh

test

Update gpt_critic.py

Update gpt_critic.py

Update run_chatgpt_unit_tests.yml

update test ci

update

update

update

update

Update test_ci.sh

update

Update test_ci.sh

Update test_ci.sh

Update run_chatgpt_examples.yml

Update run_chatgpt_examples.yml

36a519b4

17 Apr, 2023 4 commits

fix: fix sft (#3568) · 7788e0b0
tingfeng cao authored Apr 17, 2023

7788e0b0
[coati] add costom model suppor tguide (#3579) · 6b1a39b1
Fazzie-Maqianli authored Apr 17, 2023

6b1a39b1
[chat] update reward model sh (#3578) · cc1eec2f
binmakeswell authored Apr 17, 2023

cc1eec2f

[chatgpt] Detached PPO Training (#3195) · e3551443

csric authored Apr 17, 2023



* run the base

* working on dist ppo

* sync

* detached trainer

* update detached trainer. no maker update function

* facing init problem

* 1 maker 1 trainer detached run. but no model update

* facing cuda problem

* fix save functions

* verified maker update

* nothing

* add ignore

* analyize loss issue

* remove some debug codes

* facing 2m1t stuck issue

* 2m1t verified

* do not use torchrun

* working on 2m2t

* working on 2m2t

* initialize strategy in ray actor env

* facing actor's init order issue

* facing ddp model update issue (need unwarp ddp)

* unwrap ddp actor

* checking 1m2t stuck problem

* nothing

* set timeout for trainer choosing. It solves the stuck problem!

* delete some debug output

* rename to sync with upstream

* rename to sync with upstream

* coati rename

* nothing

* I am going to detach the replaybuffer from trainer and make it a Ray Actor. Two benefits: 1. support TP trainer. 2. asynchronized buffer operations

* experience_maker_holder performs target-revolving _send_experience() instead of length comparison.

* move code to ray subfolder

* working on pipeline inference

* apply comments

---------
Co-authored-by: csric <richcsr256@gmail.com>

e3551443

13 Apr, 2023 2 commits

[chat] ChatGPT train prompts on ray example (#3309) · 1a809edd

MisterLin1995 authored Apr 13, 2023



* [feat][chatgpt]train prompts on ray example

* [fix]simplify code

* [fix]remove depreciated parameter

* [fix]add dependencies

* [fix]method calling

* [fix]experience maker

* [fix]missing loss function

* [fix]init optimizer

* [feat]add usage comment

* [fix]rename files

* [fix]add readme

* [fix]file path

* [fix]move directory

---------
Co-authored-by: jiangwen <zxl265370@antgroup.com>

1a809edd

[chat] polish tutorial doc (#3551) · 535b8964

binmakeswell authored Apr 13, 2023

* [chat] clean up duplicate tutorial

* [chat] clean up duplicate tutorial

* [chat] clean up duplicate tutorial

* [chat] clean up duplicate tutorial

535b8964

12 Apr, 2023 1 commit
- [chat]add examples of training with limited resources in chat readme (#3536) · 7182ac2a
  Yuanchen authored Apr 12, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  7182ac2a
11 Apr, 2023 1 commit
- [chat]: add vf_coef argument for PPOTrainer (#3318) · e6a132a4
  zhang-yi-chi authored Apr 11, 2023
  
  e6a132a4
10 Apr, 2023 2 commits
- [chat] add zero2 cpu strategy for sft training (#3520) · 89fd10a1
  ver217 authored Apr 10, 2023
  
  89fd10a1
- [Chat Community] Update README.md (fixed#3487) (#3506) · 635d0a1b
  NatalieC323 authored Apr 10, 2023
```
* Update README.md

* Update README.md

* Update README.md

* Update README.md

---------
Co-authored-by: Fazzie-Maqianli <55798671+Fazziekey@users.noreply.github.com>
```
  635d0a1b
07 Apr, 2023 1 commit

[coati] Fix LlamaCritic (#3475) · a7ca2972

gongenlei authored Apr 07, 2023



* mv LlamaForCausalLM to LlamaModel

* rm unused imports

---------
Co-authored-by: gongenlei <gongenlei@baidu.com>

a7ca2972

06 Apr, 2023 7 commits

[chat] fix stage3 PPO sample sh command (#3477) · 891b8e7f
binmakeswell authored Apr 06, 2023

891b8e7f
add community example dictionary (#3465) · 6afeb120
Fazzie-Maqianli authored Apr 06, 2023

6afeb120

[test] refactor tests with spawn (#3452) · 80eba05b

Frank Lee authored Apr 06, 2023

* [test] added spawn decorator

* polish code

* polish code

* polish code

* polish code

* polish code

* polish code

80eba05b

[Chat]Add Peft support & fix the ptx bug (#3433) · 62f4e2eb

YY Lin authored Apr 06, 2023

* Update ppo.py

Fix the bug of fetching wrong batch data

* Add peft model support in SFT and Prompts training

In stage-1 and stage-3, the peft model supports are added. So the trained artifacts will be only a small lora additions instead of the whole bunch of files.

* Delete test_prompts.txt

* Delete test_pretrained.txt

* Move the peft stuffs to a community folder.

* Move the demo sft to community

* delete dirty files

* Add instructions to install peft using source

* Remove Chinese comments

* remove the Chinese comments

62f4e2eb

[chat]fix save_model(#3377) · 73afb635
Dr-Corgi authored Apr 06, 2023
```
The function save_model should be a part of PPOTrainer.
```
73afb635
[chat]fix readme (#3429) · 57a3c4db
kingkingofall authored Apr 06, 2023
```
* fix stage 2

fix stage 2

* add torch
```
57a3c4db

[Chat] fix the tokenizer "int too big to convert" error in SFT training (#3453) · 72cb4dd4

Camille Zhong authored Apr 06, 2023

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

* Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

* Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

* Update test_ci.sh

* Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

* Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

* Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

* Update test_ci.sh

* Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

* update roberta with coati

* chat ci update

* Revert "chat ci update"

This reverts commit 17ae7ae01fa752bd3289fc39069868fde99cf846.

* [Chat] fix the tokenizer "int too big to convert" error in SFT training

fix the tokenizer error during SFT training using Bloom and OPT

72cb4dd4

05 Apr, 2023 1 commit
- fix save_model indent error in ppo trainer (#3450) · b9231390
  Yuanchen authored Apr 05, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
  b9231390
04 Apr, 2023 3 commits

fix save_model inin naive and ddp strategy (#3436) · 773955ab
Yuanchen authored Apr 04, 2023
```
Co-authored-by: Yuanchen Xu <yuanchen.xu00@gmail.com>
```
773955ab

[zero] reorganize zero/gemini folder structure (#3424) · 26b7aac0

ver217 authored Apr 04, 2023

* [zero] refactor low-level zero folder structure

* [zero] fix legacy zero import path

* [zero] fix legacy zero import path

* [zero] remove useless import

* [zero] refactor gemini folder structure

* [zero] refactor gemini folder structure

* [zero] refactor legacy zero import path

* [zero] refactor gemini folder structure

* [zero] refactor gemini folder structure

* [zero] refactor gemini folder structure

* [zero] refactor legacy zero import path

* [zero] fix test import path

* [zero] fix test

* [zero] fix circular import

* [zero] update import

26b7aac0

[chat]fix sft training for bloom, gpt and opt (#3418) · b09adff7
Yuanchen authored Apr 04, 2023
```
fix sft training for bloom, gpt and opt 
```
b09adff7

03 Apr, 2023 1 commit

[chatgpt] add pre-trained model RoBERTa for RLHF stage 2 & 3 (#3223) · 30412866

Camille Zhong authored Apr 03, 2023

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

* Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

* Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

* add test for reward model training

* Update test_ci.sh

* Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

* Add RoBERTa for RLHF Stage 2 & 3 (test)

RoBERTa for RLHF Stage 2 & 3 (still in testing)

* Revert "Add RoBERTa for RLHF Stage 2 & 3 (test)"

This reverts commit 06741d894dcbe958acd4e10d771f22275e20e368.

* Add RoBERTa for RLHF stage 2 & 3

1. add roberta folder under model folder
2. add  roberta option in train_reward_model.py
3. add some test in testci

* Update test_ci.sh

* Revert "Update test_ci.sh"

This reverts commit 9c7352b81766f3177d31eeec0ec178a301df966a.

* update roberta with coati

30412866

30 Mar, 2023 1 commit
- [chat] correcting a few obvious typos and grammars errors (#3338) · 82132f4e
  Andrew authored Mar 29, 2023
  
  82132f4e
29 Mar, 2023 5 commits
- [doc] added authors to the chat application (#3307) · 0fbadce7
  Fazzie-Maqianli authored Mar 29, 2023
  
  0fbadce7
- Polish readme link (#3306) · b5128936
  BlueRum authored Mar 29, 2023
  
  b5128936
- [format] applied code formatting on changed files in pull request 3300 (#3302) · cb413ccf
  github-actions[bot] authored Mar 29, 2023
```
Co-authored-by: github-actions <github-actions@github.com>
```
  cb413ccf
- [doc] add ColossalChat news (#3304) · 31c78f2b
  binmakeswell authored Mar 29, 2023
```
* [doc] add ColossalChat news

* [doc] add ColossalChat news
```
  31c78f2b
- [application] updated the README (#3301) · e235a246
  Frank Lee authored Mar 29, 2023
```
* [application] updated the README

* polish code
```
  e235a246