1029 Commits (868afdb31191ef7b3fa48d6fa71e7758c8707786)

Author SHA1 Message Date
YuliangLiu0306 ecd643f1e4
[test] add torchrec models to test model zoo (#3139) 2 years ago
ver217 14a115000b
[tests] model zoo add torchaudio models (#3138) 2 years ago
Frank Lee 6d48eb0560
[test] added transformers models to test model zoo (#3135) 2 years ago
Frank Lee a674c63348
[test] added torchvision models to test model zoo (#3132) 2 years ago
HELSON 1216d1e7bd
[tests] diffuser models in model zoo (#3136) 2 years ago
YuliangLiu0306 2eca4cd376
[DTensor] refactor dtensor with new components (#3089) 2 years ago
Frank Lee 86ac782d7c
[test] added timm models to test model zoo (#3129) 2 years ago
Xuanlei Zhao 30dd13c450
[autochunk] support complete benchmark (#3121) 2 years ago
Super Daniel fff98f06ed
[analyzer] a minimal implementation of static graph analyzer (#2852) 2 years ago
Xuanlei Zhao 10c61de2f7
[autochunk] support vit (#3084) 2 years ago
YuliangLiu0306 8e4e8601b7
[DTensor] implement layout converter (#3055) 2 years ago
Xuanlei Zhao 2ca9728cbb
[autochunk] refactor chunk memory estimation (#2762) 2 years ago
YuliangLiu0306 29386a54e6
[DTensor] refactor CommSpec (#3034) 2 years ago
YuliangLiu0306 4269196c79
[hotfix] skip auto checkpointing tests (#3029) 2 years ago
YuliangLiu0306 cd2b0eaa8d
[DTensor] refactor sharding spec (#2987) 2 years ago
YuliangLiu0306 e414e4092b
[DTensor] implementation of dtensor (#2946) 2 years ago
YuliangLiu0306 197d0bf4ed
[autoparallel] apply repeat block to reduce solving time (#2912) 2 years ago
YuliangLiu0306 819e25d8b1
[hotfix] fix autoparallel compatibility test issues (#2754) 2 years ago
YuliangLiu0306 0f392d7403
[autoparallel] find repeat blocks (#2854) 2 years ago
Boyuan Yao c7764d3f22
[autoparallel] Patch meta information of `torch.where` (#2822) 2 years ago
Boyuan Yao fcc4097efa
[autoparallel] Patch meta information of `torch.tanh()` and `torch.nn.Dropout` (#2773) 2 years ago
Boyuan Yao 7ea6bc7f69
[autoparallel] Patch tensor related operations meta information (#2789) 2 years ago
HELSON 56ddc9ca7a
[hotfix] add correct device for fake_param (#2796) 2 years ago
Boyuan Yao a2b43e393d
[autoparallel] Patch meta information of `torch.nn.Embedding` (#2760) 2 years ago
YuliangLiu0306 1dc003c169
[autoparallel] distinguish different parallel strategies (#2699) 2 years ago
YuliangLiu0306 21d6a48f4d
[autoparallel] add shard option (#2696) 2 years ago
YuliangLiu0306 cb2c6a2415
[autoparallel] refactor runtime pass (#2644) 2 years ago
YuliangLiu0306 0b2a738393
[autoparallel] remove deprecated codes (#2664) 2 years ago
YuliangLiu0306 7fa6be49d2
[autoparallel] test compatibility for gemini and auto parallel (#2700) 2 years ago
Boyuan Yao 40c916b192
[autoparallel] Patch meta information of `torch.nn.functional.softmax` and `torch.nn.Softmax` (#2674) 2 years ago
HELSON 8213f89fd2
[gemini] add fake_release_chunk for keep-gathered chunk in the inference mode (#2671) 2 years ago
Boyuan Yao 0385b26ebf
[autoparallel] Patch meta information of `torch.nn.LayerNorm` (#2647) 2 years ago
YuliangLiu0306 37df666f38
[autoparallel] refactor handlers which reshape input tensors (#2615) 2 years ago
YuliangLiu0306 cb3d1bef62
[autoparallel] adapt autoparallel tests with latest api (#2626) 2 years ago
Boyuan Yao 90a9fdd91d
[autoparallel] Patch meta information of `torch.matmul` (#2584) 2 years ago
oahzxl 6ba8364881
[autochunk] support diffusion for autochunk (#2621) 2 years ago
oahzxl c4b15661d7
[autochunk] add benchmark for transformer and alphafold (#2543) 2 years ago
oahzxl 05671fcb42
[autochunk] support multi outputs chunk search (#2538) 2 years ago
oahzxl 63199c6687
[autochunk] support transformer (#2526) 2 years ago
Frank Lee b55deb0662
[workflow] only report coverage for changed files (#2524) 2 years ago
HELSON b528eea0f0
[zero] add zero wrappers (#2523) 2 years ago
HELSON 077a5cdde4
[zero] fix gradient clipping in hybrid parallelism (#2521) 2 years ago
HELSON 707b11d4a0
[gemini] update ddp strict mode (#2518) 2 years ago
HELSON 2d1a7dfe5f
[zero] add strict ddp mode (#2508) 2 years ago
oahzxl c04f183237
[autochunk] support parsing blocks (#2506) 2 years ago
oahzxl 72341e65f4
[auto-chunk] support extramsa (#3) (#2504) 2 years ago
oahzxl ecccc91f21
[autochunk] support autochunk on evoformer (#2497) 2 years ago
HELSON d565a24849
[zero] add unit testings for hybrid parallelism (#2486) 2 years ago
oahzxl 4953b4ace1
[autochunk] support evoformer tracer (#2485) 2 years ago
YuliangLiu0306 67e1912b59
[autoparallel] support origin activation ckpt on autoprallel system (#2468) 2 years ago