Commit Graph

2376 Commits (feature/elixir)
 

Author SHA1 Message Date
Frank Lee c173a69b3e
[elixir] refactored the chunk module (#3956)
1 year ago
Frank Lee 86ff5c152b
[elixir] moved simulator build to op_builder (#3939)
1 year ago
Frank Lee 3b58ff5c73
[elixir] updated readme (#3944)
1 year ago
Haichen Huang 1ee247a51c
update buffer size calculation (#3871)
2 years ago
Haichen Huang dbb9659099
[elixir] add elixir plugin and its unit test (#3865)
2 years ago
Haichen Huang 206280408a
[elixir] add elixir and its unit tests (#3835)
2 years ago
Yuanchen 34966378e8
[evaluation] add automatic evaluation pipeline (#3821)
2 years ago
Frank Lee 05b8a8de58
[workflow] changed to doc build to be on schedule and release (#3825)
2 years ago
Yanming W 269150b6f4
[Docker] Fix a couple of build issues (#3691)
2 years ago
digger yu 7f8203af69
fix typo colossalai/auto_parallel autochunk fx/passes etc. (#3808)
2 years ago
jiangmingyan 725365f297
Merge pull request #3810 from jiangmingyan/amp
2 years ago
jiangmingyan 278fcbc444 [doc]fix
2 years ago
jiangmingyan 8aa1fb2c7f [doc]fix
2 years ago
Frank Lee 1e3b64f26c
[workflow] enblaed doc build from a forked repo (#3815)
2 years ago
Hongxin Liu 19d153057e
[doc] add warning about fsdp plugin (#3813)
2 years ago
wukong1992 6b305a99d6
[booster] torch fsdp fix ckpt (#3788)
2 years ago
jiangmingyan c425a69d52 [doc] add removed change of config.py
2 years ago
jiangmingyan 75272ef37b [doc] add removed warning
2 years ago
Mingyan Jiang a520610bd9 [doc] update amp document
2 years ago
Mingyan Jiang 1167bf5b10 [doc] update amp document
2 years ago
Mingyan Jiang 8c62e50dbb [doc] update amp document
2 years ago
digger yu 9265f2d4d7
[NFC]fix typo colossalai/auto_parallel nn utils etc. (#3779)
2 years ago
jiangmingyan e871e342b3
[API] add docstrings and initialization to apex amp, naive amp (#3783)
2 years ago
Frank Lee 615e2e5fc1
[test] fixed lazy init test import error (#3799)
2 years ago
Frank Lee ad93c736ea
[workflow] enable testing for develop & feature branch (#3801)
2 years ago
jiangmingyan ef02d7ef6d
[doc] update gradient accumulation (#3771)
2 years ago
Frank Lee f5c425c898
fixed the example docstring for booster (#3795)
2 years ago
Frank Lee 788e07dbc5
[workflow] fixed the docker build workflow (#3794)
2 years ago
liuzeming 4d29c0f8e0
Fix/docker action (#3266)
2 years ago
github-actions[bot] 62c7e67f9f
[format] applied code formatting on changed files in pull request 3786 (#3787)
2 years ago
jiangmingyan fe1561a884
[doc] update gradient cliping document (#3778)
2 years ago
Yanjia0 d9393b85f1
[doc] add deprecated warning on doc Basics section (#3754)
2 years ago
Hongxin Liu 72688adb2f
[doc] add booster docstring and fix autodoc (#3789)
2 years ago
Hongxin Liu 3c07a2846e
[plugin] a workaround for zero plugins' optimizer checkpoint (#3780)
2 years ago
Hongxin Liu 60e6a154bc
[doc] add tutorial for booster checkpoint (#3785)
2 years ago
binmakeswell ad2cf58f50
[chat] add performance and tutorial (#3786)
2 years ago
Hongxin Liu b4788d63ed
[devops] fix doc test on pr (#3782)
2 years ago
digger yu 32f81f14d4
[NFC] fix typo colossalai/amp auto_parallel autochunk (#3756)
2 years ago
Hongxin Liu 21e29e2212
[doc] add tutorial for booster plugins (#3758)
2 years ago
Hongxin Liu 5ce6c9d86f
[doc] add tutorial for cluster utils (#3763)
2 years ago
Hongxin Liu 5452df63c5
[plugin] torch ddp plugin supports sharded model checkpoint (#3775)
2 years ago
jiangmingyan 2703a37ac9
[amp] Add naive amp demo (#3774)
2 years ago
jiangmingyan 48bd056761
[doc] update hybrid parallelism doc (#3770)
2 years ago
binmakeswell 15024e40d9
[auto] fix install cmd (#3772)
2 years ago
jiangmingyan d449525acf
[doc] update booster tutorials (#3718)
2 years ago
Yuanchen 05759839bd
[chat] fix bugs in stage 3 training (#3759)
2 years ago
Hongxin Liu 5dd573c6b6
[devops] fix ci for document check (#3751)
2 years ago
Hongxin Liu c03bd7c6b2
[devops] make build on PR run automatically (#3748)
2 years ago
digger yu 1baeb39c72
[NFC] fix typo with colossalai/auto_parallel/tensor_shard (#3742)
2 years ago
Ziyue Jiang 7386c6669d
[fix] Add init to fix import error when importing _analyzer (#3668)
2 years ago