deepseek-ai / deepseek-ai/DeepSeek-VL2

the inference is very slow

Open
#21 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
5.4k
Forks
1.8k
PR merge metrics
No merged PRs in 30d

Description

Hi all, when I try deepseek vl2 model at one A800 card, the inference time is about 3-4 minutes, is it correct?

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.