vllm-project / vllm-project/aibrix

How to run AIBrix kvcache offload on npu,like 910B

Open
#1,476 3 comments 0 reactions 1 assignee Claimed by @DwyaneShi View on GitHub
area/kv-cache kind/documentation
Dominant language
Go
Stars
5.1k
Forks
694
Avg merge
1d 19h
Merged PRs (30d)
98

Description

### 🚀 Feature Description and Motivation

image: aibrix-container-registry-cn-beijing.cr.volces.com/aibrix/vllm-openai-aibrix-kvcache:v0.9.1-20250724 can only run on nvidia GPU,now,i want to run on npu,like 910B. how can i do that? which need vllm-ascend.

### Use Case

dist kvcache is very valuable in large model infer,but i don't know how to run it on npu.

### Proposed Solution

_No response_

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.