PaddlePaddle / PaddlePaddle/FastDeploy
some bugs when enable LLama2-7b with PaddleNLP
Open
@Jiang-Jia-Jun is already working on this.
Since Jun 18, 2025.
- Dominant language
- Python
- Stars
- 3.7k
- Forks
- 756
- Avg merge
- 19h 28m
- Merged PRs (30d)
- 4
Description
- 分布式环境的初始化
- AutoTokenizer的选择和get_pad_id的返回
- worker最好添加一个参数传给PredictorArgument,同时load完模型后初始化kv cache
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.