InternLM

Commit Graph

Author	SHA1	Message	Date
ytxiong	853becfb6e	feat(): support fp32 training (#155 ) support float32 training * fix lint * add adaptation in model/utils.py * remove some unnecessary code * fix lint * feat(optim): add support for fp32 zero * Revert "Merge pull request #2 from SolenoidWGT/fp32_zero" This reverts commit `53fc50b0e5`, reversing changes made to `40f24d0a73`. revert commit * merge develop * Update utils.py * support fp32 in zero optimizer * modify the dtype --------- Co-authored-by: wangguoteng.p <wangguoteng925@qq.com>	2023-08-04 16:05:30 +08:00
huangting4201	66a23e326a	feat(utils/evaluation.py): support evaluate (#154 ) * style(internlm): fix lint error * feat(utils/logger.py): support uniscale logger * fix(utils/logger.py): fix import circular error * feat(train.py): support dashboard metric panel and fix ci train config * fix(ci_scripts/train/slurm_train.sh): fix ci train error * fix(ci_scripts/train/torchrun.sh): fix ci train error * feat(utils/evaluation.py): support evaluate on validation dataset * fix(utils/evaluation.py): fix demo error * fix(ci_scripts/train/ci_7B_sft.py): fix ci train error * feat(initialize/launch.py): set default value for valid_bsz and valid_every * fix(ci_scripts/train): restore ci update * docs(configs/7B_sft.py): update comment for config * fix(config.json): delete config.json * fix evaluation bug in scheduler when use_flash_attn=False * feat(scheduler/no_pipeline_scheduler.py): support micro_bsz>1 in no pp * modify the jugement in pp and no-pp scheduler * modify the data_process_func in evaluation * fix bugs when use_flash_attn=False * rename symbol * feat(configs/7B_sft.py): change para valid_bsz to valid_micro_num * feat(scheduler/no_pipeline_scheduler.py): update para set _grad_accum_batch_size --------- Co-authored-by: 黄婷 <huangting3@CN0014010744M.local> Co-authored-by: huangting.p <huangting@sensetime.com> Co-authored-by: yingtongxiong <974106207@qq.com>	2023-08-02 19:03:59 +08:00
ytxiong	307c4741d1	fix(initialize/launch.py): set default value for use_flash_attn (#158 ) * add default for use_flash_attn * fix lint	2023-08-01 16:03:06 +08:00
huangting4201	0d3d27cdf4	feat(utils/writer.py): support tensorboard writer (#63 ) * feat(utils/writer.py): support tensorboard writer * feat(utils/writer.py): add class comment --------- Co-authored-by: 黄婷 <huangting3@CN0014010744M.local>	2023-07-21 15:53:24 +08:00
Sun Peng	912fc8f8aa	doc: update the training examples (#27 ) * doc: update the training examples * update README * change all "++++" log * Update pylint * solve lint err	2023-07-07 15:54:09 +08:00
Sun Peng	fa7337b37b	initial commit	2023-07-06 12:55:23 +08:00

6 Commits (853becfb6ef4ad7c7b27b335943c6f06f5b6b51c)