intel / intel/neural-compressor

[Feature] support MXFP W4A8 evaluation on benchmark/vllm-qdq-plugin for accuracy simulation purpose.

Open
#2,557 0 comments 0 reactions 1 assignee Claimed by @yiliu30 View on GitHub
Dominant language
Python
Stars
2.7k
Forks
322
Avg merge
4d 9h
Merged PRs (30d)
25

Description

This issue has no description.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.