deepseek-ai / deepseek-ai/FlashMLA

how to test sparse attention in test_flash_mla_decoding

Open
#116 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
C++
Stars
12.9k
Forks
1.2k
Avg merge
4h 20m
Merged PRs (30d)
2

Description

It says "the kernel does not require the block_table parameter." why in test_flash_mla_decoding.py, it still needs block_table.

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.