Commit Graph

2398 Commits (12c90db3f30b6d9013a32eee27ea04ec4d631ddc)
 

Author SHA1 Message Date
Hongxin Liu 12c90db3f3
[doc] add lazy init tutorial (#3922)
2 years ago
Frank Lee c622bb3630
Merge pull request #3915 from FrankLeeeee/update/develop
2 years ago
Hongxin Liu 9c88b6cbd1
[lazy] fix compatibility problem on torch 1.13 (#3911)
2 years ago
Hongxin Liu b5f0566363
[chat] add distributed PPO trainer (#3740)
2 years ago
Hongxin Liu 41fb7236aa
[devops] hotfix CI about testmon cache (#3910)
2 years ago
digger yu 0e484e6201
[nfc]fix typo colossalai/pipeline tensor nn (#3899)
2 years ago
Baizhou Zhang c1535ccbba
[doc] fix docs about booster api usage (#3898)
2 years ago
Hongxin Liu ec9bbc0094
[devops] improving testmon cache (#3902)
2 years ago
Yuanchen 57a6d7685c
support evaluation for english (#3880)
2 years ago
digger yu 1878749753
[nfc] fix typo colossalai/nn (#3887)
2 years ago
Hongxin Liu ae02d4e4f7
[bf16] add bf16 support (#3882)
2 years ago
jiangmingyan 07cb21142f
[doc]update moe chinese document. (#3890)
2 years ago
Liu Ziming 8065cc5fba
Modify torch version requirement to adapt torch 2.0 (#3896)
2 years ago
Hongxin Liu dbb32692d2
[lazy] refactor lazy init (#3891)
2 years ago
digger yu 70c8cdecf4
[nfc] fix typo colossalai/cli fx kernel (#3847)
2 years ago
jiangmingyan 281b33f362
[doc] update document of zero with chunk. (#3855)
2 years ago
jiangmingyan 5f79008c4a
[example] update gemini examples (#3868)
2 years ago
Yuanchen 2506e275b8
[evaluation] improvement on evaluation (#3862)
2 years ago
jiangmingyan b0474878bf
[doc] update nvme offload documents. (#3850)
2 years ago
Frank Lee ae959a72a5
[workflow] fixed workflow check for docker build (#3849)
2 years ago
Frank Lee d42b1be09d
[release] bump to v0.3.0 (#3830)
2 years ago
digger yu e2d81eba0d
[nfc] fix typo colossalai/ applications/ (#3831)
2 years ago
jiangmingyan a64df3fa97
[doc] update document of gemini instruction. (#3842)
2 years ago
Frank Lee 54e97ed7ea
[workflow] supported test on CUDA 10.2 (#3841)
2 years ago
wukong1992 3229f93e30
[booster] add warning for torch fsdp plugin doc (#3833)
2 years ago
Frank Lee 84500b7799
[workflow] fixed testmon cache in build CI (#3806)
2 years ago
digger yu 518b31c059
[docs] change placememt_policy to placement_policy (#3829)
2 years ago
digger yu e90fdb1000 fix typo docs/
2 years ago
Yuanchen 34966378e8
[evaluation] add automatic evaluation pipeline (#3821)
2 years ago
Frank Lee 05b8a8de58
[workflow] changed to doc build to be on schedule and release (#3825)
2 years ago
Yanming W 269150b6f4
[Docker] Fix a couple of build issues (#3691)
2 years ago
digger yu 7f8203af69
fix typo colossalai/auto_parallel autochunk fx/passes etc. (#3808)
2 years ago
jiangmingyan 725365f297
Merge pull request #3810 from jiangmingyan/amp
2 years ago
jiangmingyan 278fcbc444 [doc]fix
2 years ago
jiangmingyan 8aa1fb2c7f [doc]fix
2 years ago
Frank Lee 1e3b64f26c
[workflow] enblaed doc build from a forked repo (#3815)
2 years ago
Hongxin Liu 19d153057e
[doc] add warning about fsdp plugin (#3813)
2 years ago
wukong1992 6b305a99d6
[booster] torch fsdp fix ckpt (#3788)
2 years ago
jiangmingyan c425a69d52 [doc] add removed change of config.py
2 years ago
jiangmingyan 75272ef37b [doc] add removed warning
2 years ago
Mingyan Jiang a520610bd9 [doc] update amp document
2 years ago
Mingyan Jiang 1167bf5b10 [doc] update amp document
2 years ago
Mingyan Jiang 8c62e50dbb [doc] update amp document
2 years ago
digger yu 9265f2d4d7
[NFC]fix typo colossalai/auto_parallel nn utils etc. (#3779)
2 years ago
jiangmingyan e871e342b3
[API] add docstrings and initialization to apex amp, naive amp (#3783)
2 years ago
Frank Lee 615e2e5fc1
[test] fixed lazy init test import error (#3799)
2 years ago
Frank Lee ad93c736ea
[workflow] enable testing for develop & feature branch (#3801)
2 years ago
jiangmingyan ef02d7ef6d
[doc] update gradient accumulation (#3771)
2 years ago
Frank Lee f5c425c898
fixed the example docstring for booster (#3795)
2 years ago
Frank Lee 788e07dbc5
[workflow] fixed the docker build workflow (#3794)
2 years ago