2350 Commits (f5c425c89874f2500600be71b3c9aadad2da822f)
 

Author SHA1 Message Date
Fazzie-Maqianli fa97a9cab4
[chatgpt] unnify datasets (#3218) 2 years ago
Fazzie-Maqianli 4fd4bd9d9a
[chatgpt] support instuct training (#3216) 2 years ago
Frank Lee cd142fbefa
[api] implemented the checkpoint io module (#3205) 2 years ago
ver217 f8289d4221
[lazyinit] combine lazy tensor with dtensor (#3204) 2 years ago
Yan Fang 189347963a
[auto] fix requirements typo for issue #3125 (#3209) 2 years ago
Yuanchen 9998d5ef64
[chatgpt]add reward model code for deberta (#3199) 2 years ago
Fazzie-Maqianli 1e1b9d2fea
[chatgpt]support llama (#3070) 2 years ago
Frank Lee e3ad88fb48
[booster] implemented the cluster module (#3191) 2 years ago
YuliangLiu0306 019a847432
[Analyzer] fix analyzer tests (#3197) 2 years ago
YuliangLiu0306 f57d34958b
[FX] refactor experimental tracer and adapt it with hf models (#3157) 2 years ago
pgzhang b429529365
[chatgpt] add supervised learning fine-tune code (#3183) 2 years ago
Frank Lee e7f3bed2d3
[booster] added the plugin base and torch ddp plugin (#3180) 2 years ago
NatalieC323 e5f668f280
[dreambooth] fixing the incompatibity in requirements.txt (#3190) 2 years ago
Zihao 18dbe76cae
[auto-parallel] add auto-offload feature (#3154) 2 years ago
YuliangLiu0306 258b43317c
[hotfix] layout converting issue (#3188) 2 years ago
YH 80aed29cd3
[zero] Refactor ZeroContextConfig class using dataclass (#3186) 2 years ago
YH 9d644ff09f
Fix docstr for zero statedict (#3185) 2 years ago
zbian 7bc0afc901 updated flash attention usage 2 years ago
Frank Lee 085e7f4eff
[test] fixed torchrec registration in model zoo (#3177) 2 years ago
NatalieC323 4e921cfbd6
[examples] Solving the diffusion issue of incompatibility issue#3169 (#3170) 2 years ago
Frank Lee a9b8402d93
[booster] added the accelerator implementation (#3159) 2 years ago
Frank Lee 1ad3a636b1
[test] fixed torchrec model test (#3167) 2 years ago
Saurav Maheshkar 20d1c99444
[refactor] update docs (#3174) 2 years ago
BlueRum 7548ca5a54
[chatgpt]Reward Model Training Process update (#3133) 2 years ago
ver217 1e58d31bb7
[chatgpt] fix trainer generate kwargs (#3166) 2 years ago
ver217 c474fda282
[chatgpt] fix ppo training hanging problem with gemini (#3162) 2 years ago
ver217 6ae8ed0407
[lazyinit] add correctness verification (#3147) 2 years ago
binmakeswell 3c01280a56
[doc] add community contribution guide (#3153) 2 years ago
Frank Lee ed19290560
[booster] implemented mixed precision class (#3151) 2 years ago
YuliangLiu0306 ecd643f1e4
[test] add torchrec models to test model zoo (#3139) 2 years ago
ver217 14a115000b
[tests] model zoo add torchaudio models (#3138) 2 years ago
Frank Lee 6d48eb0560
[test] added transformers models to test model zoo (#3135) 2 years ago
Frank Lee a674c63348
[test] added torchvision models to test model zoo (#3132) 2 years ago
HELSON 1216d1e7bd
[tests] diffuser models in model zoo (#3136) 2 years ago
Saurav Maheshkar 1a46e71e07
[docker] Add opencontainers image-spec to `Dockerfile` (#3006) 2 years ago
YuliangLiu0306 2eca4cd376
[DTensor] refactor dtensor with new components (#3089) 2 years ago
ver217 ed8f60b93b
[lazyinit] refactor lazy tensor and lazy init ctx (#3131) 2 years ago
Frank Lee 86ac782d7c
[test] added timm models to test model zoo (#3129) 2 years ago
BlueRum 23cd5e2ccf
[chatgpt]update ci (#3087) 2 years ago
Frank Lee 169ed4d24e
[workflow] purged extension cache before GPT test (#3128) 2 years ago
Xuanlei Zhao 30dd13c450
[autochunk] support complete benchmark (#3121) 2 years ago
BlueRum 68577fbc43
[chatgpt]Fix examples (#3116) 2 years ago
BlueRum 0672b5afac
[chatgpt] fix lora support for gpt (#3113) 2 years ago
github-actions[bot] 0aa92c0409
Automated submodule synchronization (#3105) 2 years ago
Jeff Rasley 453f7ae5a0
prevent op_builder being installed in site-pkgs (#3104) 2 years ago
hiko2MSP 191daf7411
[chatgpt] type miss of kwargs (#3107) 2 years ago
binmakeswell 145ccfd7d1
[doc] add Intel cooperation for biomedicine (#3108) 2 years ago
BlueRum c9dd036592
[chatgpt] fix lora save bug (#3099) 2 years ago
binmakeswell 018936a3f3
[tutorial] update notes for TransformerEngine (#3098) 2 years ago
Kirthi Shankar Sivamani 65a4dbda6c
[NVIDIA] Add FP8 example using TE (#3080) 2 years ago