vllm-project / vllm-project/aibrix

Do LLM Cache Support V100 hardware?

Open
#791 4 comments 0 reactions 0 assignees View on GitHub
area/kv-cache kind/support
Dominant language
Go
Stars
5.1k
Forks
694
Avg merge
1d 19h
Merged PRs (30d)
98

Description

I using V100 gpu to testing deploy Distributed KV Cache exmaple, unfortunately it's failed, because requires flash attention backend.
![Image](https://github.com/user-attachments/assets/997a8957-dd17-46fc-95b8-f4bc5e32356f)

Contributor guide

Open the contributing guide

Research direction

Start by reproducing the Distributed KV Cache example on V100 hardware and inspect the flash attention backend requirement shown in the report. Determine whether V100 is supported; done means either the example works on V100 or the compatibility limitation and required hardware are clearly documented.

Written by the indexing model from the issue text.

Assessment

Domain
ai, distributed-systems
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.