ColossalAI/applications/Chat/tests
Wenhao Chen 7b9b86441f
[chat]: update rm, add wandb and fix bugs (#4471)
* feat: modify forward fn of critic and reward model

* feat: modify calc_action_log_probs

* to: add wandb in sft and rm trainer

* feat: update train_sft

* feat: update train_rm

* style: modify type annotation and add warning

* feat: pass tokenizer to ppo trainer

* to: modify trainer base and maker base

* feat: add wandb in ppo trainer

* feat: pass tokenizer to generate

* test: update generate fn tests

* test: update train tests

* fix: remove action_mask

* feat: remove unused code

* fix: fix wrong ignore_index

* fix: fix mock tokenizer

* chore: update requirements

* revert: modify make_experience

* fix: fix inference

* fix: add padding side

* style: modify _on_learn_batch_end

* test: use mock tokenizer

* fix: use bf16 to avoid overflow

* fix: fix workflow

* [chat] fix gemini strategy

* [chat] fix

* sync: update colossalai strategy

* fix: fix args and model dtype

* fix: fix checkpoint test

* fix: fix requirements

* fix: fix missing import and wrong arg

* fix: temporarily skip gemini test in stage 3

* style: apply pre-commit

* fix: temporarily skip gemini test in stage 1&2

---------

Co-authored-by: Mingyan Jiang <1829166702@qq.com>
2023-09-20 15:53:58 +08:00
..
__init__.py [Coati] first commit (#3283) 2023-03-28 20:25:36 +08:00
test_benchmarks.sh [chat] fix bugs and add unit tests (#4213) 2023-08-02 10:17:36 +08:00
test_checkpoint.py [chat]: update rm, add wandb and fix bugs (#4471) 2023-09-20 15:53:58 +08:00
test_dataset.py [chat]: update rm, add wandb and fix bugs (#4471) 2023-09-20 15:53:58 +08:00
test_experience.py [chat]: update rm, add wandb and fix bugs (#4471) 2023-09-20 15:53:58 +08:00
test_inference.sh [chat] fix bugs and add unit tests (#4213) 2023-08-02 10:17:36 +08:00
test_models.py [chat]: update rm, add wandb and fix bugs (#4471) 2023-09-20 15:53:58 +08:00
test_train.sh [chat]: update rm, add wandb and fix bugs (#4471) 2023-09-20 15:53:58 +08:00