3757 Commits (8e08c27e19d3f8dcfbae36dffcad0591c0cf9cfc)
 

Author SHA1 Message Date
ver217 b105371ace rename shared adam to sharded optim v2 3 years ago
ver217 70814dc22f fix master params dtype 3 years ago
ver217 795210dd99 add fp32 master params in sharded adam 3 years ago
ver217 a109225bc2 add sharded adam 3 years ago
Jiarui Fang 8f74fbd9c9 polish license (#300) 3 years ago
Jiarui Fang e17e92c54d Polish sharded parameter (#297) 3 years ago
ver217 7aef75ca42 [zero] add sharded grad and refactor grad hooks for ShardedModel (#287) 3 years ago
Frank Lee 9afb5c8b2d fixed typo in ShardParam (#294) 3 years ago
Frank Lee 27155b8513 added unit test for sharded optimizer (#293) 3 years ago
Frank Lee e17e54e32a added buffer sync to naive amp model wrapper (#291) 3 years ago
Jiarui Fang 8d653af408 add a common util for hooks registered on parameter. (#292) 3 years ago
Jie Zhu f867365aba bug fix: pass hook_list to engine (#273) 3 years ago
Jiarui Fang 5a560a060a Feature/zero (#279) 3 years ago
binmakeswell 08eccfe681 add community group and update issue template(#271) 3 years ago
Sze-qq 3312d716a0 update experimental visualization (#253) 3 years ago
binmakeswell 753035edd3 add Chinese README 3 years ago
1SAA 82023779bb Added TPExpert for special situation 3 years ago
HELSON 36b8477228 Fixed parameter initialization in FFNExpert (#251) 3 years ago
アマデウス e13293bb4c fixed CI dataset directory; fixed import error of 2.5d accuracy (#255) 3 years ago
1SAA 219df6e685 Optimized MoE layer and fixed some bugs; 3 years ago
zbian 3dba070580 fixed padding index issue for vocab parallel embedding layers; updated 3D linear to be compatible with examples in the tutorial 3 years ago
ver217 24f8583cc4 update setup info (#233) 3 years ago
github-actions b9f8521f8c Automated submodule synchronization 3 years ago
Frank Lee f5ca88ec97 fixed apex import (#227) 3 years ago
Frank Lee eb3fda4c28 updated readme and change log (#224) 3 years ago
ver217 578ea0583b update setup and workflow (#222) 3 years ago
Frank Lee 3a1a9820b0 fixed mkdir conflict and align yapf config with flake (#220) 3 years ago
Frank Lee 65e72983dc added flake8 config (#219) 3 years ago
アマデウス 9ee197d0e9 moved env variables to global variables; (#215) 3 years ago
Frank Lee b82d60be02 updated github action for develop branch (#214) 3 years ago
BoxiangW 7d15ec7fe2
Update github actions (#205) 3 years ago
github-actions[bot] 5420809f43
Automated submodule synchronization (#203) 3 years ago
Frank Lee fd570ab285
add changelog and contributing doc (#202) 3 years ago
Frank Lee 02f13fa9d1
add code quality badge (#201) 3 years ago
Frank Lee 812357d63c
fixed utils docstring and add example to readme (#200) 3 years ago
Frank Lee b9a761b9b8
added github action to synchronize submodule commits automatically (#193) 3 years ago
BoxiangW a2f1565672
Update GitHub action and pre-commit settings (#196) 3 years ago
Frank Lee 765db512b5
fixed ddp bug on torch 1.8 (#194) 3 years ago
Jiarui Fang 569357fea0
add pytorch hooks (#179) 3 years ago
ver217 708404d5f8
fix pipeline forward return tensors (#176) 3 years ago
WANG-CR 6fb550acdb update logo 3 years ago
HELSON 0f8c7f9804
Fixed docstring in colossalai (#171) 3 years ago
Frank Lee e2089c5c15
adapted for sequence parallel (#163) 3 years ago
Frank Lee a2e649da39
update readme (#168) 3 years ago
Frank Lee 9684bdce5c
fixed submodule url (#167) 3 years ago
BoxiangW bd4840f1f1
Update workflow files and README.md (#166) 3 years ago
ver217 1949d3a889
update doc requirements and rtd conf (#165) 3 years ago
Frank Lee be85a0f366 removed tutorial markdown and refreshed rst files for consistency 3 years ago
Frank Lee ca4ae52d6b
Set examples as submodule (#162) 3 years ago
binmakeswell 17ce8569a8
add logo at homepage, add forum in issue template (#161) 3 years ago