ColossalAI

Commit Graph

Author	SHA1	Message	Date
Hongxin Liu	8accecd55b	[legacy] move engine to legacy (#4560 ) * [legacy] move engine to legacy * [example] fix seq parallel example * [example] fix seq parallel example * [test] test gemini pluging hang * [test] test gemini pluging hang * [test] test gemini pluging hang * [test] test gemini pluging hang * [test] test gemini pluging hang * [example] update seq parallel requirements	1 year ago
Hongxin Liu	89fe027787	[legacy] move trainer to legacy (#4545 ) * [legacy] move trainer to legacy * [doc] update docs related to trainer * [test] ignore legacy test	1 year ago
binmakeswell	8d7b02290f	[doc] add llama2 benchmark (#4604 ) * [doc] add llama2 benchmark * [doc] add llama2 benchmark	1 year ago
Tian Siyuan	f1ae8c9104	[example] change accelerate version (#4431 ) Co-authored-by: Siyuan Tian <siyuant@vmware.com> Co-authored-by: Hongxin Liu <lhx0217@gmail.com>	1 year ago
ChengDaqi2023	8e2e1992b8	[example] update streamlit 0.73.1 to 1.11.1 (#4386 )	1 year ago
Hongxin Liu	0b00def881	[example] add llama2 example (#4527 ) * [example] transfer llama-1 example * [example] fit llama-2 * [example] refactor scripts folder * [example] fit new gemini plugin * [cli] fix multinode runner * [example] fit gemini optim checkpoint * [example] refactor scripts * [example] update requirements * [example] update requirements * [example] rename llama to llama2 * [example] update readme and pretrain script * [example] refactor scripts	1 year ago
Hongxin Liu	27061426f7	[gemini] improve compatibility and add static placement policy (#4479 ) * [gemini] remove distributed-related part from colotensor (#4379) * [gemini] remove process group dependency * [gemini] remove tp part from colo tensor * [gemini] patch inplace op * [gemini] fix param op hook and update tests * [test] remove useless tests * [test] remove useless tests * [misc] fix requirements * [test] fix model zoo * [test] fix model zoo * [test] fix model zoo * [test] fix model zoo * [test] fix model zoo * [misc] update requirements * [gemini] refactor gemini optimizer and gemini ddp (#4398) * [gemini] update optimizer interface * [gemini] renaming gemini optimizer * [gemini] refactor gemini ddp class * [example] update gemini related example * [example] update gemini related example * [plugin] fix gemini plugin args * [test] update gemini ckpt tests * [gemini] fix checkpoint io * [example] fix opt example requirements * [example] fix opt example * [example] fix opt example * [example] fix opt example * [gemini] add static placement policy (#4443) * [gemini] add static placement policy * [gemini] fix param offload * [test] update gemini tests * [plugin] update gemini plugin * [plugin] update gemini plugin docstr * [misc] fix flash attn requirement * [test] fix gemini checkpoint io test * [example] update resnet example result (#4457) * [example] update bert example result (#4458) * [doc] update gemini doc (#4468) * [example] update gemini related examples (#4473) * [example] update gpt example * [example] update dreambooth example * [example] update vit * [example] update opt * [example] update palm * [example] update vit and opt benchmark * [hotfix] fix bert in model zoo (#4480) * [hotfix] fix bert in model zoo * [test] remove chatglm gemini test * [test] remove sam gemini test * [test] remove vit gemini test * [hotfix] fix opt tutorial example (#4497) * [hotfix] fix opt tutorial example * [hotfix] fix opt tutorial example	1 year ago
Tian Siyuan	ff836790ae	[doc] fix a typo in examples/tutorial/auto_parallel/README.md (#4430 ) Co-authored-by: Siyuan Tian <siyuant@vmware.com>	1 year ago
binmakeswell	089c365fa0	[doc] add Series A Funding and NeurIPS news (#4377 ) * [doc] add Series A Funding and NeurIPS news * [kernal] fix mha kernal * [CI] skip moe * [CI] fix requirements	1 year ago
caption	16c0acc01b	[hotfix] update gradio 3.11 to 3.34.0 (#4329 )	1 year ago
binmakeswell	ef4b99ebcd	add llama example CI	1 year ago
binmakeswell	7ff11b5537	[example] add llama pretraining (#4257 )	1 year ago
github-actions[bot]	4e9b09c222	Automated submodule synchronization (#4217 ) Co-authored-by: github-actions <github-actions@github.com>	1 year ago
digger yu	2d40759a53	fix #3852 path error (#4058 )	1 year ago
Jianghai	31dc302017	[examples] copy resnet example to image (#4090 ) * copy resnet example * add pytest package * skip test_ci * skip test_ci * skip test_ci	1 year ago
Baizhou Zhang	4da324cd60	[hotfix]fix argument naming in docs and examples (#4083 )	1 year ago
LuGY	160c64c645	[example] fix bucket size in example of gpt gemini (#4028 )	1 year ago
Baizhou Zhang	b3ab7fbabf	[example] update ViT example using booster api (#3940 )	1 year ago
Liu Ziming	e277534a18	Merge pull request #3905 from MaruyamaAya/dreambooth [example] Adding an example of training dreambooth with the new booster API	1 year ago
digger yu	33eef714db	fix typo examples and docs (#3932 )	1 year ago
Maruyama_Aya	9b5e7ce21f	modify shell for check	1 year ago
digger yu	407aa48461	fix typo examples/community/roberta (#3925 )	1 year ago
Maruyama_Aya	730a092ba2	modify shell for check	1 year ago
Maruyama_Aya	49567d56d1	modify shell for check	1 year ago
Maruyama_Aya	039854b391	modify shell for check	1 year ago
Baizhou Zhang	e417dd004e	[example] update opt example using booster api (#3918 )	1 year ago
Maruyama_Aya	cf4792c975	modify shell for check	1 year ago
Maruyama_Aya	c94a33579b	modify shell for check	2 years ago
Liu Ziming	b306cecf28	[example] Modify palm example with the new booster API (#3913 ) * Modify torch version requirement to adapt torch 2.0 * modify palm example using new booster API * roll back * fix port * polish * polish	2 years ago
wukong1992	a55fb00c18	[booster] update bert example, using booster api (#3885 )	2 years ago
Maruyama_Aya	4fc8bc68ac	modify file path	2 years ago
Maruyama_Aya	b4437e88c3	fixed port	2 years ago
Maruyama_Aya	79c9f776a9	fixed port	2 years ago
Maruyama_Aya	d3379f0be7	fixed model saving bugs	2 years ago
Maruyama_Aya	b29e1f0722	change directory	2 years ago
Maruyama_Aya	1c1f71cbd2	fixing insecure hash function	2 years ago
Maruyama_Aya	b56c7f4283	update shell file	2 years ago
Maruyama_Aya	176010f289	update performance evaluation	2 years ago
Maruyama_Aya	25447d4407	modify path	2 years ago
Maruyama_Aya	60ec33bb18	Add a new example of Dreambooth training using the booster API	2 years ago
jiangmingyan	5f79008c4a	[example] update gemini examples (#3868 ) * [example]update gemini examples * [example]update gemini examples	2 years ago
digger yu	518b31c059	[docs] change placememt_policy to placement_policy (#3829 ) * fix typo colossalai/autochunk auto_parallel amp * fix typo colossalai/auto_parallel nn utils etc. * fix typo colossalai/auto_parallel autochunk fx/passes etc. * fix typo docs/ * change placememt_policy to placement_policy in docs/ and examples/	2 years ago
github-actions[bot]	62c7e67f9f	[format] applied code formatting on changed files in pull request 3786 (#3787 ) Co-authored-by: github-actions <github-actions@github.com>	2 years ago
binmakeswell	ad2cf58f50	[chat] add performance and tutorial (#3786 )	2 years ago
binmakeswell	15024e40d9	[auto] fix install cmd (#3772 )	2 years ago
digger-yu	b7141c36dd	[CI] fix some spelling errors (#3707 ) * fix spelling error with examples/comminity/ * fix spelling error with tests/ * fix some spelling error with tests/ colossalai/ etc.	2 years ago
Hongxin Liu	3bf09efe74	[booster] update prepare dataloader method for plugin (#3706 ) * [booster] add prepare dataloader method for plug * [booster] update examples and docstr	2 years ago
Hongxin Liu	f83ea813f5	[example] add train resnet/vit with booster example (#3694 ) * [example] add train vit with booster example * [example] update readme * [example] add train resnet with booster example * [example] enable ci * [example] enable ci * [example] add requirements * [hotfix] fix analyzer init * [example] update requirements	2 years ago
Hongxin Liu	d556648885	[example] add finetune bert with booster example (#3693 )	2 years ago
digger-yu	b9a8dff7e5	[doc] Fix typo under colossalai and doc(#3618 ) * Fixed several spelling errors under colossalai * Fix the spelling error in colossalai and docs directory * Cautious Changed the spelling error under the example folder * Update runtime_preparation_pass.py revert autograft to autograd * Update search_chunk.py utile to until * Update check_installation.py change misteach to mismatch in line 91 * Update 1D_tensor_parallel.md revert to perceptron * Update 2D_tensor_parallel.md revert to perceptron in line 73 * Update 2p5D_tensor_parallel.md revert to perceptron in line 71 * Update 3D_tensor_parallel.md revert to perceptron in line 80 * Update README.md revert to resnet in line 42 * Update reorder_graph.py revert to indice in line 7 * Update p2p.py revert to megatron in line 94 * Update initialize.py revert to torchrun in line 198 * Update routers.py change to detailed in line 63 * Update routers.py change to detailed in line 146 * Update README.md revert random number in line 402	2 years ago

1 2 3 4 5 ...

301 Commits (8accecd55bf1a5aaaeb4b84c06fac0d63850fd5e)