vllm-project / vllm-project/production-stack

Question-Is there any solution to load a local model?

Open
#304 12 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

question
Dominant language
Python
Stars
2.6k
Forks
503
Avg merge
4d 17h
Merged PRs (30d)
8

Description

Is any reference or method to load a local model with vLLM Production Stack?
I have tried some methods, but it seems that it cannot access my local model path.
Can anyone provide some suggestions on which configurations I should modify? I would really appreciate it!

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

The report names no configuration file, manifest, entry point, or test; start by locating the Production Stack configuration that defines the model path and checking how that path is exposed to the deployed service. Done means a documented configuration can load the local model with reproducible validation, but the attempted configuration and deployment details are needed before a newcomer can begin.

Written by the indexing model from the issue text.

Assessment

Tech stack
kubernetes, python
Domain
infrastructure, machine-learning
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.