ColossalAI

Commit Graph

Author	SHA1	Message	Date
zbian	653b0a620e	added skip_bias_add for non-tp linear	2022-11-09 15:41:08 +08:00
DouJS	f586887a90	[NFC] polish colossalai/nn/layer/colossalai_layer/dropout.py code style (#1568 )	2022-09-08 22:11:04 +08:00
Ofey Chan	7cc052f6c0	[NFC] polish colossalai/nn/layer/colossalai_layer/linear.py (#1556 )	2022-09-08 22:11:04 +08:00
アマデウス	b8899e0905	[TP] allow layernorm without bias (#750 )	2022-04-14 11:43:56 +08:00
Liang Bowen	828e465622	[hotfix] Raise messages for indivisible batch sizes with tensor parallelism (#622 )	2022-04-02 16:12:04 +08:00
アマデウス	cd13b63832	[model checkpoint] reworked unified layers for ease of save/load states (#593 )	2022-04-01 16:49:56 +08:00
Ziyue Jiang	763dc325f1	[TP] Add gather_out arg to Linear (#541 )	2022-03-30 09:35:46 +08:00
Liang Bowen	ec5086c49c	Refactored docstring to google style	2022-03-29 17:17:47 +08:00
アマデウス	9ee197d0e9	moved env variables to global variables; (#215 ) added branch context; added vocab parallel layers; moved split_batch from load_batch to tensor parallel embedding layers; updated gpt model; updated unit test cases; fixed few collective communicator bugs	2022-02-15 11:31:13 +08:00
HELSON	0f8c7f9804	Fixed docstring in colossalai (#171 )	2022-01-21 10:44:30 +08:00
BoxiangW	4a3d3446b0	Update layer integration documentations (#108 ) Update the documentations of layer integration Update _log_hook.py Update _operation.py	2022-01-10 18:05:58 +08:00
ver217	96780e6ee4	Optimize pipeline schedule (#94 ) * add pipeline shared module wrapper and update load batch * added model parallel process group for amp and clip grad (#86) * added model parallel process group for amp and clip grad * update amp and clip with model parallel process group * remove pipeline_prev/next group (#88) * micro batch offload * optimize pipeline gpu memory usage * pipeline can receive tensor shape (#93) * optimize pipeline gpu memory usage * fix grad accumulation step counter * rename classes and functions Co-authored-by: Frank Lee <somerlee.9@gmail.com>	2021-12-30 15:56:46 +08:00
アマデウス	01a80cd86d	Hotfix/Colossalai layers (#92 ) * optimized 1d layer apis; reorganized nn.layer modules; fixed tests * fixed 2.5d runtime issue * reworked split batch, now called in trainer.schedule.load_batch Co-authored-by: BoxiangW <45734921+BoxiangW@users.noreply.github.com>	2021-12-29 23:32:10 +08:00

13 Commits (4ee311c0262dfbca9b5da7e18f04dd8f1f23fe4c)