Commit Graph

13 Commits (ff14144d9cebc2d334de714f52d1b599ece0a116)

Author SHA1 Message Date
Tong Li 39e2597426
[ColossalChat] Add PP support (#6001)
3 months ago
YeAnbang ed97d3a5d3
[Chat] fix readme (#5989)
4 months ago
YeAnbang 30f4e31a33
[Chat] Fix lora (#5946)
4 months ago
YeAnbang 12fe8b5858 refactor evaluation
4 months ago
YeAnbang 09d5ffca1a add kto
4 months ago
YeAnbang e7a8634636 fix eval
5 months ago
YeAnbang 8aad064fe7 fix style
5 months ago
YeAnbang 82aecd6374 add SimPO
5 months ago
YeAnbang 0d7ff10ea5 replace the customized dataloader setup with the build-in one
6 months ago
YeAnbang 790e1362a6 merge
6 months ago
YeAnbang 0b4a33548c moupdate ci tests, st ci test cases passed, tp failed in generation for ppo, sp is buggy
6 months ago
YeAnbang 929e1e3da4 upgrade ppo dpo rm script
6 months ago
YeAnbang df5e9c53cf
[ColossalChat] Update RLHF V2 (#5286)
8 months ago