Commit Graph

13 Commits (845ea7214ebac42f71a34ef1c36d4716f2d4668c)

Author SHA1 Message Date
YeAnbang b3594d4d68 fix orpo cross entropy loss
5 months ago
YeAnbang e7a8634636 fix eval
5 months ago
pre-commit-ci[bot] 8a9721bafe [pre-commit.ci] auto fixes from pre-commit.com hooks
5 months ago
YeAnbang d888c3787c add benchmark for sft, dpo, simpo, orpo. Add benchmarking result. Support lora with gradient checkpoint
5 months ago
YeAnbang 82aecd6374 add SimPO
5 months ago
YeAnbang 84eab13078 update sft trainning script
6 months ago
YeAnbang 0d7ff10ea5 replace the customized dataloader setup with the build-in one
6 months ago
YeAnbang 0b4a33548c moupdate ci tests, st ci test cases passed, tp failed in generation for ppo, sp is buggy
6 months ago
YeAnbang 7e65b71815 run pre-commit
6 months ago
YeAnbang 929e1e3da4 upgrade ppo dpo rm script
6 months ago
YeAnbang 7a7e86987d upgrade colossal-chat support tp_group>1, add sp for sft
6 months ago
Hongxin Liu 7f8b16635b
[misc] refactor launch API and tensor constructor (#5666)
7 months ago
YeAnbang df5e9c53cf
[ColossalChat] Update RLHF V2 (#5286)
8 months ago