deepseek-ai / deepseek-ai/FlashMLA
how to test sparse attention in test_flash_mla_decoding
Open
- Dominant language
- C++
- Stars
- 12.9k
- Forks
- 1.2k
- Avg merge
- 4h 20m
- Merged PRs (30d)
- 2
Description
It says "the kernel does not require the block_table parameter." why in test_flash_mla_decoding.py, it still needs block_table.
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.