Support 1 * 128 and 128 * 128 block-wise quant?
オープン
@szkarpinski がすでに取り組んでいます。
2025年8月6日 から。
enhancement
- 主要言語
- Cython
- スター
- 601
- フォーク
- 46
- PR マージ指標
- 30日以内にマージされた PR はありません
説明
In the CUDA 12.9 cuBLASLt documentation, I noticed support for 1×128 and 128×128 block-wise quantization methods. However, I found that nvmath-python currently lacks bindings for this type of quantize approach. I wonder do we have any plan for support this approach?
https://docs.nvidia.com/cuda/cublas/index.html#cublasltmatmulmatrixscale-t
コントリビューションガイド
はじめの一歩
- issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
- 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
- リポジトリをフォークし、ブランチを切って変更します。
- issue 番号を参照したプルリクエストを送ります。
評価
この issue はまだ評価されていません。