Commit Graph

1 Commits (eea37da6fa62c35b8ec35607fe9fdf1e14287df3)

Author SHA1 Message Date
Runyu Lu e37ee2fb65
[Feat]Tensor Model Parallel Support For Inference (#5563)
* tensor parallel support naive source

* [fix]precision, model load and refactor the framework

* add tp unit test

* docstring

* fix do_sample
2024-04-18 16:56:46 +08:00