alipay / alipay/PainlessInferenceAcceleration
How the performance VS vLLM inference(vLLM vs Lookahead)
Open
- Dominant language
- Python
- Stars
- 370
- Forks
- 23
- PR merge metrics
- No merged PRs in 30d
Description
In the benchmark comparison results, could we add a comparison with VLLM to see the acceleration effects?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.