2370 Commits (34966378e81e1c885af9b50670552d3487a60b57)
 

Author SHA1 Message Date
digger-yu b7141c36dd
[CI] fix some spelling errors (#3707) 2 years ago
MisterLin1995 f7361ee1bd
[chat] fix community example ray (#3719) 2 years ago
jiangmingyan 20068ba188
[booster] add tests for ddp and low level zero's checkpointio (#3715) 2 years ago
Hongxin Liu 6552cbf8e1
[booster] fix no_sync method (#3709) 2 years ago
Hongxin Liu 3bf09efe74
[booster] update prepare dataloader method for plugin (#3706) 2 years ago
Hongxin Liu f83ea813f5
[example] add train resnet/vit with booster example (#3694) 2 years ago
YH 2629f9717d
[tensor] Refactor handle_trans_spec in DistSpecManager 2 years ago
zhang-yi-chi 2da5d81dec
[chat] fix train_prompts.py gemini strategy bug (#3666) 2 years ago
Hongxin Liu d556648885
[example] add finetune bert with booster example (#3693) 2 years ago
digger-yu 65bdc3159f
fix some spelling error with applications/Chat/examples/ (#3692) 2 years ago
Hongxin Liu d0915f54f4
[booster] refactor all dp fashion plugins (#3684) 2 years ago
digger-yu b49020c1b1
[CI] Update test_sharded_optim_with_sync_bn.py (#3688) 2 years ago
Tong Li b36e67cb2b
Merge pull request #3680 from digger-yu/digger-yu-patch-2 2 years ago
jiangmingyan 307894f74d
[booster] gemini plugin support shard checkpoint (#3610) 2 years ago
Camille Zhong 0f785cb1f3
[chat] PPO stage3 doc enhancement (#3679) 2 years ago
digger-yu 6650daeb0a
[doc] fix chat spelling error (#3671) 2 years ago
Hongxin Liu 7bd0bee8ea
[chat] add opt attn kernel (#3655) 2 years ago
digger-yu 8ba7858753
Update generate_gpt35_answers.py 2 years ago
digger-yu bfbf650588
fix spelling error 2 years ago
tanitna 1a60dc07a8
[chat] typo accimulation_steps -> accumulation_steps (#3662) 2 years ago
Tong Li 816add7e7f
Merge pull request #3656 from TongLi3701/chat/update_eval 2 years ago
binmakeswell 268b3cd80d
[chat] set default zero2 strategy (#3667) 2 years ago
Tong Li c1a355940e update readme 2 years ago
Tong Li ed3eaa6922 update documentation 2 years ago
Tong Li c419117329 update questions and readme 2 years ago
Tong Li aa77ddae33 remove unnecessary step and update readme 2 years ago
YH a22407cc02
[zero] Suggests a minor change to confusing variable names in the ZeRO optimizer. (#3173) 2 years ago
Hongxin Liu 842768a174
[chat] refactor model save/load logic (#3654) 2 years ago
Hongxin Liu 6ef7011462
[chat] remove lm model class (#3653) 2 years ago
Camille Zhong 8bccb72c8d
[Doc] enhancement on README.md for chat examples (#3646) 2 years ago
Hongxin Liu 2a951955ad
[chat] refactor trainer (#3648) 2 years ago
Hongxin Liu f8288315d9
[chat] polish performance evaluator (#3647) 2 years ago
Hongxin Liu 50793b35f4
[gemini] accelerate inference (#3641) 2 years ago
Hongxin Liu 4b3240cb59
[booster] add low level zero plugin (#3594) 2 years ago
digger-yu b9a8dff7e5
[doc] Fix typo under colossalai and doc(#3618) 2 years ago
Tong Li e1b0a78afa
Merge pull request #3621 from zhang-yi-chi/fix/chat-train-prompts-single-gpu 2 years ago
ddobokki df309fc6ab
[Chat] Remove duplicate functions (#3625) 2 years ago
Hongxin Liu 179558a87a
[devops] fix chat ci (#3628) 2 years ago
zhang-yi-chi 739cfe3360 [chat] fix enable single gpu training bug 2 years ago
digger-yu d7bf284706
[chat] polish code note typo (#3612) 2 years ago
Yuanchen c4709d34cf
Chat evaluate (#3608) 2 years ago
digger-yu 633bac2f58
[doc] .github/workflows/README.md (#3605) 2 years ago
digger-yu becd3b0f54
[doc] fix setup.py typo (#3603) 2 years ago
digger-yu 7570d9ae3d
[doc] fix op_builder/README.md (#3597) 2 years ago
Hongxin Liu 12eff9eb4c
[gemini] state dict supports fp16 (#3590) 2 years ago
github-actions[bot] d544ed4345
[bot] Automated submodule synchronization (#3596) 2 years ago
digger-yu d96567bb5d
[misc] op_builder/builder.py (#3593) 2 years ago
binmakeswell 5a79cffdfd
[coati] fix install cmd (#3592) 2 years ago
Yuanchen 1ec0d386a9
reconstruct chat trainer and fix training script (#3588) 2 years ago
Hongxin Liu dac127d0ee
[fx] fix meta tensor registration (#3589) 2 years ago