microsoft / microsoft/onnxruntime
[Documentation] Running genai-directml-quantized models with metacommands disabled
Open
documentation
ep:DML
quantization
- Dominant language
- C++
- Stars
- 21.9k
- Forks
- 4.2k
- Avg merge
- 4d 11h
- Merged PRs (30d)
- 184
Description
### Describe the documentation issue
Is there currently any way to run models quantized by GenAI on DirectML with metacommands turned off?
### Page / URL
_No response_
Contributor guide
Research direction
No documentation page, file, test, or entry point is specified. Start by locating the documentation for genai-directml-quantized models and the metacommands setting, then determine whether a supported procedure exists; the documentation is done when it clearly answers how to run these models with metacommands disabled.
Written by the indexing model from the issue text.
Assessment
- Domain
- documentation, machine-learning
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100