Commit Graph

2324 Commits (b37797ed3d3d6af294a095397b4bc135264b8c6a)
 

Author SHA1 Message Date
BlueRum c8b723d6c2
[chat]Update Readme (#3296)
2 years ago
ver217 73b542a124
[coati] inference supports profanity check (#3295)
2 years ago
ver217 ce2cafae76
[coati] add repetition_penalty for inference (#3294)
2 years ago
Fazzie-Maqianli a88ed0f83a
add limit (#3293)
2 years ago
Fazzie-Maqianli c5484281aa
[ColossalChat]add cite for datasets (#3292)
2 years ago
Fazzie-Maqianli ec7af22a43
fix image (#3288)
2 years ago
Fazzie-Maqianli 1f7d9afbf8
add example (#3286)
2 years ago
ver217 4905b21b94
[coati] fix inference output (#3285)
2 years ago
Fazzie-Maqianli bb6196e71a
remove chatgpt (#3284)
2 years ago
Fazzie-Maqianli b0ce5a1032
[Coati] first commit (#3283)
2 years ago
YuliangLiu0306 fd6add575d
[examples] polish AutoParallel readme (#3270)
2 years ago
HELSON 02b058032d
[fx] meta registration compatibility (#3253)
2 years ago
Frank Lee 73d3e4d309
[booster] implemented the torch ddd + resnet example (#3232)
2 years ago
YH 1a229045af
Add interface for colo tesnor dp size (#3227)
2 years ago
Hakjin Lee 1653063fce
[CI] Fix pre-commit workflow (#3238)
2 years ago
NatalieC323 280fcdc485
polish code (#3194)
2 years ago
YuliangLiu0306 4d5d8f98a4
[API] implement device mesh manager (#3221)
2 years ago
CsRic 052b03e83f
limit torch version (#3213)
2 years ago
binmakeswell d32ef94ad9
[doc] fix typo (#3222)
2 years ago
YuliangLiu0306 045afa3ea2
[hotfix] skip torchaudio tracing test (#3211)
2 years ago
ver217 78fd31f9c1
[chatgpt] add precision option for colossalai (#3233)
2 years ago
Fazzie-Maqianli bd39877da4
support instrcut training (#3230)
2 years ago
Camille Zhong 9bc702ab48
[doc] update chatgpt doc paper link (#3229)
2 years ago
Fazzie-Maqianli bbac6760e5
fix torch version (#3225)
2 years ago
Fazzie-Maqianli fa97a9cab4
[chatgpt] unnify datasets (#3218)
2 years ago
Fazzie-Maqianli 4fd4bd9d9a
[chatgpt] support instuct training (#3216)
2 years ago
Frank Lee cd142fbefa
[api] implemented the checkpoint io module (#3205)
2 years ago
ver217 f8289d4221
[lazyinit] combine lazy tensor with dtensor (#3204)
2 years ago
Yan Fang 189347963a
[auto] fix requirements typo for issue #3125 (#3209)
2 years ago
Yuanchen 9998d5ef64
[chatgpt]add reward model code for deberta (#3199)
2 years ago
Fazzie-Maqianli 1e1b9d2fea
[chatgpt]support llama (#3070)
2 years ago
Frank Lee e3ad88fb48
[booster] implemented the cluster module (#3191)
2 years ago
YuliangLiu0306 019a847432
[Analyzer] fix analyzer tests (#3197)
2 years ago
YuliangLiu0306 f57d34958b
[FX] refactor experimental tracer and adapt it with hf models (#3157)
2 years ago
pgzhang b429529365
[chatgpt] add supervised learning fine-tune code (#3183)
2 years ago
Frank Lee e7f3bed2d3
[booster] added the plugin base and torch ddp plugin (#3180)
2 years ago
NatalieC323 e5f668f280
[dreambooth] fixing the incompatibity in requirements.txt (#3190)
2 years ago
Zihao 18dbe76cae
[auto-parallel] add auto-offload feature (#3154)
2 years ago
YuliangLiu0306 258b43317c
[hotfix] layout converting issue (#3188)
2 years ago
YH 80aed29cd3
[zero] Refactor ZeroContextConfig class using dataclass (#3186)
2 years ago
YH 9d644ff09f
Fix docstr for zero statedict (#3185)
2 years ago
zbian 7bc0afc901 updated flash attention usage
2 years ago
Frank Lee 085e7f4eff
[test] fixed torchrec registration in model zoo (#3177)
2 years ago
NatalieC323 4e921cfbd6
[examples] Solving the diffusion issue of incompatibility issue#3169 (#3170)
2 years ago
Frank Lee a9b8402d93
[booster] added the accelerator implementation (#3159)
2 years ago
Frank Lee 1ad3a636b1
[test] fixed torchrec model test (#3167)
2 years ago
Saurav Maheshkar 20d1c99444
[refactor] update docs (#3174)
2 years ago
BlueRum 7548ca5a54
[chatgpt]Reward Model Training Process update (#3133)
2 years ago
ver217 1e58d31bb7
[chatgpt] fix trainer generate kwargs (#3166)
2 years ago
ver217 c474fda282
[chatgpt] fix ppo training hanging problem with gemini (#3162)
2 years ago