deepseek-ai / deepseek-ai/FlashMLA

Could the model trained in A100 use FlashMLA in H100 for inference?

Open
#40 3 comments 0 reactions 0 assignees View on GitHub
Dominant language
C++
Stars
12.9k
Forks
1.2k
Avg merge
4h 20m
Merged PRs (30d)
2

Description

Hi Author,
Thanks for the release. I would like a question as title.
I have trained the model in A100, now I hope to use FlashMLA to speed up inference in H1200. Is it possible? If yes, do I need to do some changes? Any suggestion would be appreciated, thanks

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.