16 Commits (cloud/coati)

Author SHA1 Message Date
Hongxin Liu 8accecd55b [legacy] move engine to legacy (#4560) 1 year ago
Frank Lee 80eba05b0a
[test] refactor tests with spawn (#3452) 2 years ago
ver217 933048ad3e
[test] reorganize zero/gemini tests (#3445) 2 years ago
ver217 26b7aac0be
[zero] reorganize zero/gemini folder structure (#3424) 2 years ago
Jiarui Fang 1e885329f4
[test] align model name with the file name. (#2045) 2 years ago
HELSON a088022efc
[moe] fix moe bugs (#1633) 2 years ago
Frank Lee 5a1a095b92
[test] refactored with the new rerun decorator (#763) 3 years ago
ver217 e396bb71f2
[zero] add tensor placement policies (#743) 3 years ago
HELSON 22c4b88d56
[zero] refactor ShardedParamV2 for convenience (#742) 3 years ago
Jiarui Fang 53cb584808
[utils] correct cpu memory used and capacity in the context of multi-process (#726) 3 years ago
HELSON b9b469ea50
[moe] add checkpoint for moe zero test (#729) 3 years ago
HELSON a9b8300d54
[zero] improve adaptability for not-shard parameters (#708) 3 years ago
HELSON ee112fe1da
[zero] adapt zero hooks for unsharded module (#699) 3 years ago
HELSON d7ecaf362b
[zero] fix init bugs in zero context (#686) 3 years ago
HELSON 055fbf5be6
[zero] adapt zero for unsharded paramters (Optimizer part) (#601) 3 years ago
HELSON e6d50ec107
[zero] adapt zero for unsharded parameters (#561) 3 years ago