Frank Lee
|
250be4d31e
|
[utils] integrated colotensor with lazy init context (#1324)
* [utils] integrated colotensor with lazy init context
* polish code
* polish code
* polish code
|
2 years ago |
Frank Lee
|
659a740738
|
[workflow] roll back to use torch 1.11 for unit testing (#1325)
|
2 years ago |
Frank Lee
|
4d5dbf48a6
|
[workflow] fixed trigger condition for 8-gpu unit test (#1323)
|
2 years ago |
YuliangLiu0306
|
e8acf55e8b
|
[fx] add balanced policy v2 (#1251)
* [CLI] add CLI launcher
* Revert "[CLI] add CLI launcher"
This reverts commit df7e6506d4 .
* [fx] add balanced policy v2
* add unittest
|
2 years ago |
XYE
|
ca2d3f284f
|
[fx] Add unit test and fix bugs for transform_mlp_pass (#1299)
* add test and fix bugs
* add functions back
* add comments
|
2 years ago |
HELSON
|
1b41686461
|
[hotfix] fix unit test test_module_spec (#1321)
|
2 years ago |
Jiarui Fang
|
9e4c6449b0
|
[checkpoint] add ColoOptimizer checkpointing (#1316)
|
2 years ago |
Frank Lee
|
7c2634f4b3
|
[workflow] updated release bdist workflow (#1318)
* [workflow] updated release bdist workflow
* polish workflow
* polish workflow
|
2 years ago |
github-actions[bot]
|
869cf3d3b8
|
Automated submodule synchronization (#1319)
Co-authored-by: github-actions <github-actions@github.com>
|
2 years ago |
Frank Lee
|
efdc240f1f
|
[workflow] disable SHM for compatibility CI on rtx3080 (#1315)
|
2 years ago |
ver217
|
7c70bfbefa
|
[hotfix] fix PipelineSharedModuleGradientHandler (#1314)
|
2 years ago |
Jiarui Fang
|
85f933b58b
|
[Optimizer] Remove useless ColoOptimizer (#1312)
|
2 years ago |
Frank Lee
|
c9c37dcc4d
|
[workflow] updated pytorch compatibility test (#1311)
|
2 years ago |
Jiarui Fang
|
9f10524313
|
[Optimizer] polish the init method of ColoOptimizer (#1310)
|
2 years ago |
HELSON
|
36086927e1
|
[hotfix] fix ColoTensor GPT2 unitest (#1309)
|
2 years ago |
Jiarui Fang
|
3ef3791a3b
|
[checkpoint] add test for bert and hotfix save bugs (#1297)
|
2 years ago |
Jiarui Fang
|
bd71e2a88b
|
[hotfix] add missing file (#1308)
|
2 years ago |
Frank Lee
|
4f4d8c3656
|
[fx] added apex normalization to patched modules (#1300)
* [fx] added apex normalization to patched modules
* remove unused imports
|
2 years ago |
Jiarui Fang
|
4165eabb1e
|
[hotfix] remove potiential circle import (#1307)
* make it faster
* [hotfix] remove circle import
|
2 years ago |
github-actions[bot]
|
6f2f9eb214
|
Automated submodule synchronization (#1305)
Co-authored-by: github-actions <github-actions@github.com>
|
2 years ago |
YuliangLiu0306
|
93a75433df
|
[hotfix] skip some unittest due to CI environment. (#1301)
|
2 years ago |
lucasliunju
|
339520c6e0
|
[NFC] polish build_colossalai_wheel.py code style (#1306)
|
2 years ago |
HELSON
|
260a55804a
|
[hotfix] fix shape error in backward when using ColoTensor (#1298)
|
2 years ago |
runluo
|
f83c4d6597
|
[NFC] polish colossalai/nn/layer/wrapper/pipeline_wrapper.py code style (#1303)
|
2 years ago |
binmakeswell
|
7696cead8d
|
Recover kernal files
|
2 years ago |
XYE
|
e83b2ce853
|
[NFC] polish colossalai/nn/layer/vanilla/layers.py code style (#1295)
|
2 years ago |
Liping233
|
1000a41fd5
|
[NFC] polish colossalai/nn/layer/vanilla/__init__.py code style (#1293)
|
2 years ago |
Maruyama_Aya
|
87f679aeae
|
[NFC] polish colossalai/kernel/cuda_native/csrc/kernels/include/kernels.h code style (#1291)
|
2 years ago |
Wangbo Zhao(黑色枷锁)
|
552667825b
|
[NFC] polish colossalai/nn/layer/parallel_1d/layers.py code style (#1290)
|
2 years ago |
doubleHU
|
d6f5ef8860
|
[NFC] polish colossalai/kernel/cuda_native/csrc/kernels/transform_kernels.cu code style (#1286)
|
2 years ago |
Ziheng Qin
|
6d6c01e94d
|
[NFC] polish colossalai/__init__.py code style (#1285)
|
2 years ago |
Jiatong Han
|
38e3ccd1e9
|
[NFC] polish colossalai/nn/layer/parallel_sequence/layers.py code style (#1280)
Co-authored-by: JThh <jiatong.han@u.nus.edu>
|
2 years ago |
Boyuan Yao
|
b414eaa5db
|
[NFC] polish colossalai/nn/optimizer/lamb.py code style (#1275)
|
2 years ago |
yuxuan-lou
|
5f6ab35d25
|
Hotfix/format (#1274)
* [NFC] Polish colossalai/kernel/cuda_native/csrc/multi_tensor_lamb.cu code style. (#937)
* [NFC] polish colossalai/kernel/cuda_native/csrc/kernels/include/cuda_util.h code style
* [NFC] polish colossalai/kernel/cuda_native/csrc/scaled_masked_softmax.cpp code style
Co-authored-by: BoxiangW <45734921+BoxiangW@users.noreply.github.com>
|
2 years ago |
Super Daniel
|
52d145a342
|
[NFC] polish colossalai/nn/lr_scheduler/onecycle.py code style (#1269)
|
2 years ago |
Geng Zhang
|
0e06f62160
|
[NFC] polish colossalai/nn/layer/parallel_sequence/_operation.py code style (#1266)
|
2 years ago |
binmakeswell
|
c95e18cdb9
|
[NFC] polish colossalai/kernel/cuda_native/csrc/scaled_upper_triang_masked_softmax.h code style (#1270)
|
2 years ago |
xyupeng
|
94bfd35184
|
[NFC] polish colossalai/builder/builder.py code style (#1265)
|
2 years ago |
DouJS
|
db13f96333
|
[NFC] polish colossalai/kernel/cuda_native/csrc/multi_tensor_apply.cuh code style (#1264)
|
2 years ago |
shenggan
|
5d7366b144
|
[NFC] polish colossalai/kernel/cuda_native/csrc/scaled_masked_softmax.h code style (#1263)
|
2 years ago |
Zangwei Zheng
|
197a2c89e2
|
[NFC] polish colossalai/communication/collective.py (#1262)
|
2 years ago |
ziyu huang
|
f1cafcc73a
|
[NFC] polish colossalai/kernel/cuda_native/csrc/kernels/dropout_kernels.cu code style (#1261)
Co-authored-by: “Arsmart123 <202476410arsmart@gmail.com>
|
2 years ago |
Sze-qq
|
f8b9aaef47
|
[NFC] polish colossalai/kernel/cuda_native/csrc/type_shim.h code style (#1260)
|
2 years ago |
superhao1995
|
f660152c73
|
[NFC] polish colossalai/nn/layer/parallel_3d/_operation.py code style (#1258)
Co-authored-by: Research <research@soccf-snr3-017.comp.nus.edu.sg>
|
2 years ago |
Thunderbeee
|
9738fb0f78
|
[NFC] polish colossalai/nn/lr_scheduler/__init__.py (#1255)
code style
|
2 years ago |
Kai Wang (Victor Kai)
|
50f2ad213f
|
[NFC] polish colossalai/engine/ophooks/utils.py code style (#1256)
|
2 years ago |
Ofey Chan
|
2dd4d556fb
|
[NFC] polish colossalai/nn/init.py code style (#1292)
|
2 years ago |
Jiarui Fang
|
556b9b7e1a
|
[hotfix] Dist Mgr gather torch version (#1284)
* make it faster
* [hotfix] torchvison fx tests
* [hotfix] rename duplicated named test_gpt.py
* [hotfix] dist mgr torch version
|
2 years ago |
Frank Lee
|
7e8114a8dd
|
[hotfix] skipped unsafe test cases (#1282)
|
2 years ago |
Jiarui Fang
|
79fe7b027a
|
[hotfix] test model unittest hotfix (#1281)
|
2 years ago |