github-actions[bot]
|
4fb4a22a72
|
[format] applied code formatting on changed files in pull request 5234 (#5235)
Co-authored-by: github-actions <github-actions@github.com>
|
11 months ago |
binmakeswell
|
b9b32b15e6
|
[doc] add Colossal-LLaMA-2-13B (#5234)
* [doc] add Colossal-LLaMA-2-13B
* [doc] add Colossal-LLaMA-2-13B
* [doc] add Colossal-LLaMA-2-13B
|
11 months ago |
Camille Zhong
|
915b4652f3
|
[doc] Update README.md of Colossal-LLAMA2 (#5233)
* Update README.md
* Update README.md
|
11 months ago |
Tong Li
|
d992b55968
|
[Colossal-LLaMA-2] Release Colossal-LLaMA-2-13b-base model (#5224)
* update readme
* update readme
* update link
* update
* update readme
* update
* update
* update
* update title
* update example
* update example
* fix content
* add conclusion
* add license
* update
* update
* update version
* fix minor
|
11 months ago |
Yuanchen
|
b397104438
|
[Colossal-Llama-2] Add finetuning Colossal-Llama-2 example (#4878)
* Add finetuning Colossal-Llama-2 example
* Add finetuning Colossal-Llama-2 example 2
* Add finetuning Colossal-Llama-2 example and support NEFTuning
* Add inference example and refine neftune
* Modify readme file
* update the imports
---------
Co-authored-by: Xu Yuanchen <yuanchen.xu00@gmail.com>
Co-authored-by: Camille Zhong <44392324+Camille7777@users.noreply.github.com>
|
12 months ago |
digger yu
|
9110406a47
|
fix typo change JOSNL TO JSONL etc. (#5116)
|
1 year ago |
digger yu
|
d5661f0f25
|
[nfc] fix typo change directoty to directory (#5111)
|
1 year ago |
github-actions[bot]
|
a41cf88e9b
|
[format] applied code formatting on changed files in pull request 4908 (#4918)
Co-authored-by: github-actions <github-actions@github.com>
|
1 year ago |
Zian(Andy) Zheng
|
7768afbad0
|
Update flash_attention_patch.py
To be compatible with the new change in the Transformers library, where a new argument 'padding_mask' was added to forward function of attention layer.
https://github.com/huggingface/transformers/pull/25598
|
1 year ago |
Camille Zhong
|
652adc2215
|
Update README.md
|
1 year ago |
Camille Zhong
|
afe10a85fd
|
Update README.md
|
1 year ago |
Camille Zhong
|
3043d5d676
|
Update modelscope link in README.md
add modelscope link
|
1 year ago |
Yuanchen
|
1fa8c5e09f
|
Update Qwen-7B results (#4821)
Co-authored-by: Xu Yuanchen <yuanchen.xu00@gmail.com>
|
1 year ago |
Chandler-Bing
|
b6cf0aca55
|
[hotfix] change llama2 Colossal-LLaMA-2 script filename (#4800)
change filename:
pretraining.py -> trainin.py
there is no file named pretraing.py. wrong writing
|
1 year ago |
Tong Li
|
8cbce6184d
|
update
|
1 year ago |
Tong Li
|
bd014673b0
|
update readme
|
1 year ago |
binmakeswell
|
d512a4d38d
|
[doc] add llama2 domain-specific solution news (#4789)
* [doc] add llama2 domain-specific solution news
|
1 year ago |
Tong Li
|
74aa7d964a
|
initial commit: add colossal llama 2 (#4784)
|
1 year ago |