Commit Graph

2455 Commits (822c3d4d66d2d74cb7c7080abed6a207602dddfd)
 

Author SHA1 Message Date
hiko2MSP 191daf7411
[chatgpt] type miss of kwargs (#3107)
2 years ago
binmakeswell 145ccfd7d1
[doc] add Intel cooperation for biomedicine (#3108)
2 years ago
BlueRum c9dd036592
[chatgpt] fix lora save bug (#3099)
2 years ago
binmakeswell 018936a3f3
[tutorial] update notes for TransformerEngine (#3098)
2 years ago
Kirthi Shankar Sivamani 65a4dbda6c
[NVIDIA] Add FP8 example using TE (#3080)
2 years ago
Frank Lee 26db1cb57b
[release] v0.2.7 (#3094)
2 years ago
Fazzie-Maqianli 02ae80bf9c
[chatgpt]add flag of action mask in critic(#3086)
2 years ago
Frank Lee 95a36eae63
[kernel] added kernel loader to softmax autograd function (#3093)
2 years ago
Super Daniel fff98f06ed
[analyzer] a minimal implementation of static graph analyzer (#2852)
2 years ago
Fazzie-Maqianli 5d5f475d75
[diffusers] fix ci and docker (#3085)
2 years ago
Frank Lee 3213347b49
[doc] fixed typos in docs/README.md (#3082)
2 years ago
Xuanlei Zhao 10c61de2f7
[autochunk] support vit (#3084)
2 years ago
Camille Zhong e58a3c804c
Fix the version of lightning and colossalai in Stable Diffusion environment requirement (#3073)
2 years ago
YuliangLiu0306 8e4e8601b7
[DTensor] implement layout converter (#3055)
2 years ago
Frank Lee 89aa7926ac
[release] v0.2.6 (#3057)
2 years ago
Frank Lee 416a50dbd7
[doc] moved doc test command to bottom (#3075)
2 years ago
Frank Lee 91ccf97514
[workflow] fixed doc build trigger condition (#3072)
2 years ago
Frank Lee f19b49e164
[booster] init module structure and definition (#3056)
2 years ago
github-actions[bot] faa8526b85
Automated submodule synchronization (#3062)
2 years ago
binmakeswell 360674283d
[example] fix redundant note (#3065)
2 years ago
Tomek af3888481d
[example] fixed opt model downloading from huggingface
2 years ago
Xuanlei Zhao 2ca9728cbb
[autochunk] refactor chunk memory estimation (#2762)
2 years ago
wenjunyang b51bfec357
[chatgpt] change critic input as state (#3042)
2 years ago
ramos 2ef855c798
support shardinit option to avoid OPT OOM initializing problem (#3037)
2 years ago
YuliangLiu0306 29386a54e6
[DTensor] refactor CommSpec (#3034)
2 years ago
Frank Lee ea0b52c12e
[doc] specified operating system requirement (#3019)
2 years ago
ver217 378d827c6b
[doc] update nvme offload doc (#3014)
2 years ago
Fazzie-Maqianli c21b11edce
change nn to models (#3032)
2 years ago
YuliangLiu0306 4269196c79
[hotfix] skip auto checkpointing tests (#3029)
2 years ago
Frank Lee 8fedc8766a
[workflow] supported conda package installation in doc test (#3028)
2 years ago
Frank Lee 2cd6ba3098
[workflow] fixed the post-commit failure when no formatting needed (#3020)
2 years ago
Frank Lee 2e427ddf42
[revert] recover "[refactor] restructure configuration files (#2977)" (#3022)
2 years ago
github-actions[bot] e86d9bb2e1
[format] applied code formatting on changed files in pull request 3025 (#3026)
2 years ago
YuliangLiu0306 cd2b0eaa8d
[DTensor] refactor sharding spec (#2987)
2 years ago
Ziyue Jiang 400f63012e
[pipeline] Add Simplified Alpa DP Partition (#2507)
2 years ago
Super Daniel b42d3d28ed
[fx] remove depreciated algorithms. (#2312) (#2313)
2 years ago
BlueRum 55dcd3051a
[chatgpt] fix readme (#3025)
2 years ago
LuGY 287d60499e
[chatgpt] Add saving ckpt callback for PPO (#2880)
2 years ago
BlueRum e588703454
[chatgpt]fix inference model load (#2988)
2 years ago
github-actions[bot] 82503a96f2
[format] applied code formatting on changed files in pull request 2997 (#3008)
2 years ago
binmakeswell 52a5078988
[doc] add ISC tutorial (#2997)
2 years ago
Saurav Maheshkar 35c8f4ce47
[refactor] restructure configuration files (#2977)
2 years ago
ver217 823f3b9cf4
[doc] add deepspeed citation and copyright (#2996)
2 years ago
Frank Lee e0a1c1321c
[doc] added reference to related works (#2994)
2 years ago
Yasyf Mohamedali 19fa0e57f6
Remove extraneous comma (#2993)
2 years ago
Frank Lee 3a5d93bc2c
[kernel] cached the op kernel and fixed version check (#2886)
2 years ago
ver217 0ff8406b00
[chatgpt] allow shard init and display warning (#2986)
2 years ago
BlueRum f5ca0397dd
[chatgpt] fix lora gemini conflict in RM training (#2984)
2 years ago
ver217 19ad49fb3b
[chatgpt] making experience support dp (#2971)
2 years ago
github-actions[bot] 827a0af8cc
Automated submodule synchronization (#2982)
2 years ago