Have you conducted experimental comparisons on DeepSeek-R1-Distill-Qwen-32B?
Open
- Dominant language
- Python
- Stars
- 190
- Forks
- 6
- PR merge metrics
- No merged PRs in 30d
Description
In Table 1 of the paper, the results for **OREAL-7B**, **OREAL-DSR1-Distill-Qwen-7B**, and **OREAL-32B** are provided, **but** there are no results for **OREAL-DSR1-Distill-Qwen-32B**. Is the RL performance on this model not good?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.