NVIDIA / NVIDIA/CUDALibrarySamples

LtNvfp4Matmul‘s scale is NULL!

Open
#274 1 comment 0 reactions 1 assignee View on GitHub

@rsdubtso is already working on this.

Since Jul 27, 2025.

cuBLASLt
Dominant language
Cuda
Stars
2.5k
Forks
478
PR merge metrics
No merged PRs in 30d

Description

When testing the LtNvfp4Matmul example on an SM_101 architecture GPU, I encountered the following runtime error:
"[cublasLt][702814][Error][cublasLtMatmulAlgoGetHeuristic] The scaling mode for A scale (VEC16_UE4M3) requires a custom pointer for A scale
cuBLAS API failed with status 7"

I noticed that the a_scale parameter being passed is NULL. Could you please advise how to modify this?

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.