InternLM / InternLM/InternLM-Math
Model evaluation on minif2f fails?
未关闭
- 主要语言
- Python
- 星标
- 550
- 派生
- 39
- PR 合并指标
- 30 天内没有已合并 PR
描述
When I evaluate InternLM2-Math-Plus-7b in minif2f through this code, it fails. The model only generates one line "Here is the predicted next tactic:" without any tactics. If I let the model continue generating until I get a tactic each time, I only get a pass rate 38.1% instead of 43.4%.
贡献指南
这个仓库没有索引到贡献指南
评估
这个 Issue 还没有评估数据。