NVIDIA / NVIDIA/CUDALibrarySamples
LtNvfp4Matmul‘s scale is NULL!
Open
@rsdubtso is already working on this.
Since Jul 27, 2025.
cuBLASLt
- Dominant language
- Cuda
- Stars
- 2.5k
- Forks
- 478
- PR merge metrics
- No merged PRs in 30d
Description
When testing the LtNvfp4Matmul example on an SM_101 architecture GPU, I encountered the following runtime error:
"[cublasLt][702814][Error][cublasLtMatmulAlgoGetHeuristic] The scaling mode for A scale (VEC16_UE4M3) requires a custom pointer for A scale
cuBLAS API failed with status 7"
I noticed that the a_scale parameter being passed is NULL. Could you please advise how to modify this?
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.