intel / intel/auto-round

[Feature]: Support Nemotron 3 Super

Open
#1,546 4 comments 0 reactions 1 assignee Claimed by @yiliu30 View on GitHub
enhancement
Dominant language
Python
Stars
1.6k
Forks
175
Avg merge
1d 18h
Merged PRs (30d)
99

Description

### Feature Description

Submitting request, per:

```
2026-03-14 19:21:04 INFO __main__.py L586: start to quantize ./NVIDIA-Nemotron-3-Super-120B-A12B-BF16
2026-03-14 19:21:04 WARNING base.py L304: This MoE model has not been optimized by AutoRound yet, which may result in high RAM usage, Please consider submitting an issue to https://github.com/intel/auto-round/issues
```

### Motivation and Use Case

High RAM usage is bad, when the whole point is to quantize a model so that it fits smaller systems.

### Alternatives Considered

_No response_

### Definition of Done

_No response_

### Additional Context

_No response_

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.