deepseek-ai / deepseek-ai/FlashMLA
Could the model trained in A100 use FlashMLA in H100 for inference?
Open
- Dominant language
- C++
- Stars
- 12.9k
- Forks
- 1.2k
- Avg merge
- 4h 20m
- Merged PRs (30d)
- 2
Description
Hi Author,
Thanks for the release. I would like a question as title.
I have trained the model in A100, now I hope to use FlashMLA to speed up inference in H1200. Is it possible? If yes, do I need to do some changes? Any suggestion would be appreciated, thanks
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.