OpenMOSS / OpenMOSS/MOSS-Audio
Hi, does this series of model have a quantized version like NVFP4 or FP8?
Open
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 665
- Forks
- 45
- PR merge metrics
- No merged PRs in 30d
Description
Hi, while quantizing the model is simple enough, I'm wondering if the MOSS's fork of SGLang supports quantization like NVFP4 & FP8 for efficent deployment of the model?
Thanks for the info in advanced
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
No files, tests, or entry points are named. Start by checking the MOSS SGLang fork's quantization documentation and deployment entry points for existing NVFP4 or FP8 support. Done means documenting whether either format is supported for this model, or defining the implementation scope if neither is available.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Quiet
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100