320 Commits (37e35230ff4666231dd65435b5f7b2a2fcfaf9e6)

Author SHA1 Message Date
Tong Li c650a906db
[Hotfix] Remove deprecated install (#6042) 3 months ago
Tong Li 0d3a85d04f
add fused norm (#6038) 3 months ago
Tong Li 4a68efb7da
[Colossal-LLaMA] Refactor latest APIs (#6030) 3 months ago
wangbluo dae39999d7 fix 3 months ago
flybird11111 0bc9a870c0
Update train_dpo.py 3 months ago
Wang Binluo eea37da6fa
[fp8] Merge feature/fp8_comm to main branch of Colossalai (#6016) 3 months ago
Tong Li 39e2597426
[ColossalChat] Add PP support (#6001) 3 months ago
pre-commit-ci[bot] 81272e9d00 [pre-commit.ci] auto fixes from pre-commit.com hooks 3 months ago
YeAnbang ed97d3a5d3
[Chat] fix readme (#5989) 3 months ago
Tong Li ad3fa4f49c
[Hotfix] README link (#5966) 4 months ago
flybird11111 0c10afd372
[FP8] rebase main (#5963) 4 months ago
YeAnbang 0b2d55c4ab Support overall loss, update KTO logging 4 months ago
Tong Li 19d1510ea2
[feat] Dist Loader for Eval (#5950) 4 months ago
Tong Li 1aeb5e8847
[hotfix] Remove unused plan section (#5957) 4 months ago
YeAnbang 66fbf2ecb7
Update README.md (#5958) 4 months ago
YeAnbang 30f4e31a33
[Chat] Fix lora (#5946) 4 months ago
YeAnbang 6fd9e86864 fix style 4 months ago
YeAnbang de1bf08ed0 fix style 4 months ago
YeAnbang 8a3ff4f315 fix style 4 months ago
zhurunhua ad35a987d3
[Feature] Add a switch to control whether the model checkpoint needs to be saved after each epoch ends (#5941) 4 months ago
YeAnbang 9688e19b32 remove real data path 4 months ago
YeAnbang b0e15d563e remove real data path 4 months ago
YeAnbang 12fe8b5858 refactor evaluation 4 months ago
YeAnbang c5f582f666 fix test data 4 months ago
zhurunhua 4ec17a7cdf
[FIX BUG] UnboundLocalError: cannot access local variable 'default_conversation' where it is not associated with a value (#5931) 4 months ago
YeAnbang d49550fb49 refactor tokenization 4 months ago
Tong Li f585d4e38e
[ColossalChat] Hotfix for ColossalChat (#5910) 4 months ago
YeAnbang 544b7a38a1 fix style, add kto data sample 4 months ago
YeAnbang 09d5ffca1a add kto 4 months ago
YeAnbang b3594d4d68 fix orpo cross entropy loss 4 months ago
YeAnbang 115c4cc5a4 hotfix citation 4 months ago
YeAnbang e7a8634636 fix eval 4 months ago
pre-commit-ci[bot] 8a9721bafe [pre-commit.ci] auto fixes from pre-commit.com hooks 5 months ago
YeAnbang f6ef5c3609 fix style 5 months ago
YeAnbang d888c3787c add benchmark for sft, dpo, simpo, orpo. Add benchmarking result. Support lora with gradient checkpoint 5 months ago
pre-commit-ci[bot] 7c2f79fa98
[pre-commit.ci] pre-commit autoupdate (#5572) 5 months ago
YeAnbang ff535204fe update transformers version 5 months ago
Haze188 416580b314
[MoE/ZeRO] Moe refactor with zero refactor (#5821) 5 months ago
YeAnbang a8af6ccb73 fix torch colossalai version 5 months ago
YeAnbang b117274074 fix colossalai, transformers version 5 months ago
YeAnbang afa53066ca fix colossalai, transformers version 5 months ago
YeAnbang 384c64057d fix colossalai, transformers version 5 months ago
YeAnbang 8aad064fe7 fix style 5 months ago
YeAnbang c8d1b4a968 add orpo 5 months ago
binmakeswell 4ccaaaab63
[doc] add GPU cloud playground (#5851) 5 months ago
YeAnbang f3de5a025c remove debug code 5 months ago
YeAnbang 0b2d6275c4 fix dataloader 5 months ago
YeAnbang 82aecd6374 add SimPO 5 months ago
YeAnbang 84eab13078 update sft trainning script 5 months ago
YeAnbang 2abdede1d7 fix readme 6 months ago