PaddlePaddle / PaddlePaddle/FastDeploy

some bugs when enable LLama2-7b with PaddleNLP

Open
#2,627 0 comments 0 reactions 1 assignee View on GitHub

@Jiang-Jia-Jun is already working on this.

Since Jun 18, 2025.

Dominant language
Python
Stars
3.7k
Forks
756
Avg merge
19h 28m
Merged PRs (30d)
4

Description

  1. 分布式环境的初始化
    Image
  2. AutoTokenizer的选择和get_pad_id的返回
    Image
  3. worker最好添加一个参数传给PredictorArgument,同时load完模型后初始化kv cache
    Image

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.