3132 Commits (f5b7de38a487ab9b879dfef3619b66f0aa624ceb)
 

Author SHA1 Message Date
binmakeswell 6df844b8c4
[release] grok-1 314b inference (#5490) 8 months ago
Hongxin Liu 848a574c26
[example] add grok-1 inference (#5485) 8 months ago
binmakeswell d158fc0e64
[doc] update open-sora demo (#5479) 8 months ago
binmakeswell bd998ced03
[doc] release Open-Sora 1.0 with model weights (#5468) 8 months ago
flybird11111 5e16bf7980
[shardformer] fix gathering output when using tensor parallelism (#5431) 8 months ago
Hongxin Liu f2e8b9ef9f
[devops] fix compatibility (#5444) 8 months ago
digger yu 385e85afd4
[hotfix] fix typo s/keywrods/keywords etc. (#5429) 9 months ago
Camille Zhong da885ed540
fix tensor data update for gemini loss caluculation (#5442) 9 months ago
Hongxin Liu 8020f42630
[release] update version (#5411) 9 months ago
Camille Zhong 743e7fad2f
[colossal-llama2] add stream chat examlple for chat version model (#5428) 9 months ago
Youngon 68f55a709c
[hotfix] fix stable diffusion inference bug. (#5289) 9 months ago
hugo-syn c8003d463b
[doc] Fix typo s/infered/inferred/ (#5288) 9 months ago
digger yu 5e1c93d732
[hotfix] fix typo change MoECheckpintIO to MoECheckpointIO (#5335) 9 months ago
Dongruixuan Li a7ae2b5b4c
[eval-hotfix] set few_shot_data to None when few shot is disabled (#5422) 9 months ago
digger yu 049121d19d
[hotfix] fix typo change enabel to enable under colossalai/shardformer/ (#5317) 9 months ago
digger yu 16c96d4d8c
[hotfix] fix typo change _descrption to _description (#5331) 9 months ago
digger yu 70cce5cbed
[doc] update some translations with README-zh-Hans.md (#5382) 9 months ago
Luo Yihang e239cf9060
[hotfix] fix typo of openmoe model source (#5403) 9 months ago
MickeyCHAN e304e4db35
[hotfix] fix sd vit import error (#5420) 9 months ago
Hongxin Liu 070df689e6
[devops] fix extention building (#5427) 9 months ago
binmakeswell 822241a99c
[doc] sora release (#5425) 9 months ago
flybird11111 29695cf70c
[example]add gpt2 benchmark example script. (#5295) 9 months ago
Camille Zhong 4b8312c08e
fix sft single turn inference example (#5416) 9 months ago
binmakeswell a1c6cdb189 [doc] fix blog link 9 months ago
binmakeswell 5de940de32 [doc] fix blog link 9 months ago
Frank Lee 2461f37886
[workflow] added pypi channel (#5412) 9 months ago
Tong Li a28c971516
update requirements (#5407) 9 months ago
flybird11111 0a25e16e46
[shardformer]gather llama logits (#5398) 9 months ago
Frank Lee dcdd8a5ef7
[setup] fixed nightly release (#5388) 9 months ago
QinLuo bf34c6fef6
[fsdp] impl save/load shard model/optimizer (#5357) 9 months ago
Hongxin Liu d882d18c65
[example] reuse flash attn patch (#5400) 9 months ago
Hongxin Liu 95c21e3950
[extension] hotfix jit extension setup (#5402) 9 months ago
Stephan Kölker 5d380a1a21
[hotfix] Fix wrong import in meta_registry (#5392) 9 months ago
CZYCW b833153fd5
[hotfix] fix variable type for top_p (#5313) 9 months ago
Frank Lee 705a62a565
[doc] updated installation command (#5389) 9 months ago
yixiaoer 69e3ad01ed
[doc] Fix typo (#5361) 9 months ago
Hongxin Liu 7303801854
[llama] fix training and inference scripts (#5384) 9 months ago
Hongxin Liu adae123df3
[release] update version (#5380) 10 months ago
Frank Lee efef43b53c
Merge pull request #5372 from hpcaitech/exp/mixtral 10 months ago
Frank Lee 4c03347fc7
Merge pull request #5377 from hpcaitech/example/llama-npu 10 months ago
ver217 06db94fbc9 [moe] fix tests 10 months ago
Hongxin Liu 65e5d6baa5 [moe] fix mixtral optim checkpoint (#5344) 10 months ago
Hongxin Liu 956b561b54 [moe] fix mixtral forward default value (#5329) 10 months ago
Hongxin Liu b60be18dcc [moe] fix mixtral checkpoint io (#5314) 10 months ago
Hongxin Liu da39d21b71 [moe] support mixtral (#5309) 10 months ago
Hongxin Liu c904d2ae99 [moe] update capacity computing (#5253) 10 months ago
Xuanlei Zhao 7d8e0338a4 [moe] init mixtral impl 10 months ago
Hongxin Liu 084c91246c
[llama] fix memory issue (#5371) 10 months ago
Hongxin Liu c53ddda88f
[lr-scheduler] fix load state dict and add test (#5369) 10 months ago
Hongxin Liu eb4f2d90f9
[llama] polish training script and fix optim ckpt (#5368) 10 months ago