InternLM/lmdeploy
View on GitHubLMDeploy is a toolkit for compressing, deploying, and serving LLMs.
- Stars
- 8.1k
- Forks
- 748
- Open beginner issues
- 0
- Indexed issues
- 546
- Avg merge
- 6d 2h
- Merged PRs (30d)
- 54
- Dominant language
- Python
- License
- Apache-2.0
- Last GitHub push
- Sep 16, 2026
- Latest indexed
- Sep 17, 2026
- Contributing guide
- Contributing guide
- Code of conduct
- No code of conduct
- Beginner labels
- No beginner labels indexed
-
[Bug] Does lmdeploy support deploying glm-5 on A800 or A100, or are there any plans to support it? Open
InternLM/lmdeploy#4407 · 10 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4428 · 0 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4436 · 5 comments · 0 reactions · 1 assignee ·
-
[Bug] MXFP4 Bug Open
InternLM/lmdeploy#4440 · 1 comment · 0 reactions · 1 assignee ·
-
[Feature] Integrate Mooncake Transfer Engine with TurboMind for PD disaggregation and remote KV Open
InternLM/lmdeploy#4443 · 5 comments · 2 reactions · 0 assignees ·
-
InternLM/lmdeploy#4464 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4475 · 0 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4479 · 3 comments · 1 reaction · 2 assignees ·
-
InternLM/lmdeploy#4484 · 3 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4489 · 3 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4491 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4492 · 2 comments · 0 reactions · 1 assignee ·
-
backlog
InternLM/lmdeploy#4527 · 0 comments · 0 reactions · 0 assignees ·
-
planned feature
InternLM/lmdeploy#4530 · 3 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4553 · 2 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4556 · 12 comments · 1 reaction · 1 assignee ·
-
InternLM/lmdeploy#4558 · 3 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4562 · 3 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4612 · 1 comment · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4658 · 0 comments · 1 reaction · 0 assignees ·
-
InternLM/lmdeploy#4673 · 0 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4686 · 0 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4698 · 2 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4720 · 0 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4750 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4761 · 1 comment · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4775 · 2 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4824 · 9 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4832 · 2 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4865 · 0 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4879 · 0 comments · 0 reactions · 1 assignee ·
-
[Bug] pytorch backend silently hangs under concurrent decoding with qwen3_5_mtp speculative decoding Open
InternLM/lmdeploy#4883 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4889 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4899 · 0 comments · 4 reactions · 0 assignees ·
-
InternLM/lmdeploy#4905 · 2 comments · 0 reactions · 0 assignees ·
-
[Bug] Mooncake external KV keys omit KV format and weights lineage, with no default tenant isolation Open
InternLM/lmdeploy#4930 · 2 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4933 · 2 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4944 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4953 · 1 comment · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4958 · 3 comments · 0 reactions · 0 assignees ·
-
InternLM/lmdeploy#4960 · 0 comments · 0 reactions · 1 assignee ·
-
[Bug] Openawaiting response
InternLM/lmdeploy#4962 · 2 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4965 · 0 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4967 · 0 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4971 · 2 comments · 0 reactions · 1 assignee ·
-
InternLM/lmdeploy#4972 · 0 comments · 0 reactions · 1 assignee ·