intel / intel/auto-round

[Feature]: the ram usage of quantizing gemma-4-26B-A4B-it is too high

Open
#2,210 0 comments 0 reactions 1 assignee Claimed by @n1ck-guo View on GitHub
enhancement
Dominant language
Python
Stars
1.6k
Forks
175
Avg merge
1d 18h
Merged PRs (30d)
99

Description

### Feature Description

https://github.com/intel/auto-round/issues/2203#issuecomment-5367654320

### Motivation and Use Case

~

### Alternatives Considered

_No response_

### Definition of Done

_No response_

### Additional Context

_No response_

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.