3054 Commits (641b1ee71a19e2337f3363620b228dd355835b04)
 

Author SHA1 Message Date
Hongxin Liu 641b1ee71a
[devops] remove post commit ci (#5566) 8 months ago
digger yu 341263df48
[hotfix] fix typo s/get_defualt_parser /get_default_parser (#5548) 8 months ago
digger yu a799ca343b
[fix] fix typo s/muiti-node /multi-node etc. (#5448) 8 months ago
Edenzzzz 15055f9a36
[hotfix] quick fixes to make legacy tutorials runnable (#5559) 8 months ago
Zhongkai Zhao 8e412a548e
[shardformer] Sequence Parallelism Optimization (#5533) 8 months ago
Edenzzzz 7e0ec5a85c
fix incorrect sharding without zero (#5545) 8 months ago
Wenhao Chen e614aa34f3
[shardformer, pipeline] add `gradient_checkpointing_ratio` and heterogenous shard policy for llama (#5508) 8 months ago
YeAnbang df5e9c53cf
[ColossalChat] Update RLHF V2 (#5286) 8 months ago
Yuanheng Zhao 36c4bb2893
[Fix] Grok-1 use tokenizer from the same pretrained path (#5532) 8 months ago
Insu Jang 00525f7772
[shardformer] fix pipeline forward error if custom layer distribution is used (#5189) 8 months ago
github-actions[bot] e6707a6e8d
[format] applied code formatting on changed files in pull request 5510 (#5517) 8 months ago
Hongxin Liu 19e1a5cf16
[shardformer] update colo attention to support custom mask (#5510) 8 months ago
Edenzzzz 9a3321e9f4
Merge pull request #5515 from Edenzzzz/fix_layout_convert 8 months ago
Edenzzzz 18edcd5368 Empty-Commit 8 months ago
Edenzzzz 61da3fbc52 fixed layout converter caching and updated tester 8 months ago
Rocky Duan cbe34c557c
Fix ColoTensorSpec for py11 (#5440) 8 months ago
Hongxin Liu a7790a92e8
[devops] fix example test ci (#5504) 8 months ago
Yuanheng Zhao 131f32a076
[fix] fix grok-1 example typo (#5506) 8 months ago
flybird11111 0688d92e2d
[shardformer]Fix lm parallel. (#5480) 8 months ago
binmakeswell 34e909256c
[release] grok-1 inference benchmark (#5500) 8 months ago
Wenhao Chen bb0a668fee
[hotfix] set return_outputs=False in examples and polish code (#5404) 8 months ago
Yuanheng Zhao 5fcd7795cd
[example] update Grok-1 inference (#5495) 8 months ago
binmakeswell 6df844b8c4
[release] grok-1 314b inference (#5490) 8 months ago
Hongxin Liu 848a574c26
[example] add grok-1 inference (#5485) 8 months ago
binmakeswell d158fc0e64
[doc] update open-sora demo (#5479) 8 months ago
binmakeswell bd998ced03
[doc] release Open-Sora 1.0 with model weights (#5468) 8 months ago
flybird11111 5e16bf7980
[shardformer] fix gathering output when using tensor parallelism (#5431) 8 months ago
Hongxin Liu f2e8b9ef9f
[devops] fix compatibility (#5444) 8 months ago
digger yu 385e85afd4
[hotfix] fix typo s/keywrods/keywords etc. (#5429) 8 months ago
Camille Zhong da885ed540
fix tensor data update for gemini loss caluculation (#5442) 9 months ago
Hongxin Liu 8020f42630
[release] update version (#5411) 9 months ago
Camille Zhong 743e7fad2f
[colossal-llama2] add stream chat examlple for chat version model (#5428) 9 months ago
Youngon 68f55a709c
[hotfix] fix stable diffusion inference bug. (#5289) 9 months ago
hugo-syn c8003d463b
[doc] Fix typo s/infered/inferred/ (#5288) 9 months ago
digger yu 5e1c93d732
[hotfix] fix typo change MoECheckpintIO to MoECheckpointIO (#5335) 9 months ago
Dongruixuan Li a7ae2b5b4c
[eval-hotfix] set few_shot_data to None when few shot is disabled (#5422) 9 months ago
digger yu 049121d19d
[hotfix] fix typo change enabel to enable under colossalai/shardformer/ (#5317) 9 months ago
digger yu 16c96d4d8c
[hotfix] fix typo change _descrption to _description (#5331) 9 months ago
digger yu 70cce5cbed
[doc] update some translations with README-zh-Hans.md (#5382) 9 months ago
Luo Yihang e239cf9060
[hotfix] fix typo of openmoe model source (#5403) 9 months ago
MickeyCHAN e304e4db35
[hotfix] fix sd vit import error (#5420) 9 months ago
Hongxin Liu 070df689e6
[devops] fix extention building (#5427) 9 months ago
binmakeswell 822241a99c
[doc] sora release (#5425) 9 months ago
flybird11111 29695cf70c
[example]add gpt2 benchmark example script. (#5295) 9 months ago
Camille Zhong 4b8312c08e
fix sft single turn inference example (#5416) 9 months ago
binmakeswell a1c6cdb189 [doc] fix blog link 9 months ago
binmakeswell 5de940de32 [doc] fix blog link 9 months ago
Frank Lee 2461f37886
[workflow] added pypi channel (#5412) 9 months ago
Tong Li a28c971516
update requirements (#5407) 9 months ago
flybird11111 0a25e16e46
[shardformer]gather llama logits (#5398) 9 months ago