2849 Commits (cabc1286ca4a2defffe8e74aaca18023620099f6)
 

Author SHA1 Message Date
ver217 52d055119b increase the timeout limit in CI temporarily 3 years ago
ver217 253e54d98a fix grad shape 3 years ago
Jiarui Fang ea2872073f [zero] global model data memory tracer (#360) 3 years ago
Jiarui Fang cb34cd384d [test] polish zero related unitest (#351) 3 years ago
HELSON 534e0bb118 Fixed import bug for no-tensorboard environment (#354) 3 years ago
HELSON c57e089824 [profile] added example for ProfilerContext (#349) 3 years ago
ver217 532ae79cb0 add test sharded optim with cpu adam (#347) 3 years ago
Jiarui Fang 10e2826426 move async memory to an individual directory (#345) 3 years ago
HELSON 425bb0df3f Added Profiler Context to manage all profilers (#340) 3 years ago
ver217 d0ae0f2215 [zero] update sharded optim v2 (#334) 3 years ago
ver217 2b8cddd40e skip bert in test engine 3 years ago
ver217 d41a9f12c6 install transformers in CI 3 years ago
ver217 f5f0ad266e fix bert unit test 3 years ago
jiaruifang 5663616921 polish code 3 years ago
jiaruifang d271f2596b polish engine unitest 3 years ago
jiaruifang 354c0f9047 polish code 3 years ago
jiaruifang 4d94cd513e adapting bert unitest interface 3 years ago
jiaruifang 7977422aeb add bert for unitest and sharded model is not able to pass the bert case 3 years ago
Frank Lee 3d5d64bd10 refactored grad scaler (#338) 3 years ago
Frank Lee 6a3188167c set criterion as optional in colossalai initialize (#336) 3 years ago
Jie Zhu 3213554cc2 [profiler] add adaptive sampling to memory profiler (#330) 3 years ago
ver217 1388671699 [zero] Update sharded model v2 using sharded param v2 (#323) 3 years ago
jiaruifang 799d105bb4 using pytest parametrize 3 years ago
jiaruifang dec24561cf show pytest parameterize 3 years ago
Jiarui Fang 11bddb6e55 [zero] update zero context init with the updated test utils (#327) 3 years ago
Frank Lee 6268446b81 [test] refactored testing components (#324) 3 years ago
HELSON 4f26fabe4f fixed strings in profiler outputs (#325) 3 years ago
Jiarui Fang de0468c7a8 [zero] zero init context (#321) 3 years ago
1SAA 73bff11288 Added profiler communication operations 3 years ago
binmakeswell d275b98b7d add badge and contributor list 3 years ago
LuGY a3269de5c9 [zero] cpu adam kernel (#288) 3 years ago
Jiarui Fang 90d3aef62c [zero] yet an improved sharded param (#311) 3 years ago
Jiarui Fang c9e7d9582d [zero] polish shard strategy (#310) 3 years ago
ver217 3092317b80 polish code 3 years ago
ver217 36f9a74ab2 fix sharded param hook and unit test 3 years ago
ver217 001ca624dd impl shard optim v2 and add unit test 3 years ago
Jiarui Fang 74f77e314b [zero] a shard strategy in granularity of tensor (#307) 3 years ago
Jiarui Fang 80364c7686 [zero] sharded tensor (#305) 3 years ago
Jie Zhu d344689274 [profiler] primary memory tracer 3 years ago
FrankLeeeee dfc3fafe89 update unit testing CI rules 3 years ago
FrankLeeeee bbbfe9b2c9 added compatibility CI and options for release ci 3 years ago
FrankLeeeee 115bcc0b41 added pypi publication CI and remove formatting CI 3 years ago
ver217 b105371ace rename shared adam to sharded optim v2 3 years ago
ver217 70814dc22f fix master params dtype 3 years ago
ver217 795210dd99 add fp32 master params in sharded adam 3 years ago
ver217 a109225bc2 add sharded adam 3 years ago
Jiarui Fang 8f74fbd9c9 polish license (#300) 3 years ago
Jiarui Fang e17e92c54d Polish sharded parameter (#297) 3 years ago
ver217 7aef75ca42 [zero] add sharded grad and refactor grad hooks for ShardedModel (#287) 3 years ago
Frank Lee 9afb5c8b2d fixed typo in ShardParam (#294) 3 years ago