NVIDIA / NVIDIA/Model-Optimizer
Pre-Quantized Checkpoints: Gemma 4 models
Open
@yueshen2016 is already working on this.
Since Apr 16, 2026.
feature request
- Dominant language
- Python
- Stars
- 3.8k
- Forks
- 604
- Avg merge
- 2d 8h
- Merged PRs (30d)
- 142
Description
Detailed description of the requested feature
It would be great to have gemma 4 models in your published optimized models (i.e. https://huggingface.co/collections/nvidia/inference-optimized-checkpoints-with-model-optimizer).
Timeline
Soon, but not blocking.
Describe alternatives you've considered
Target hardware/use case
For me, a 5080 for local inference. But it would be broadly useful for anyone running gemma 4 models.
Contributor guide
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.