open-compass / open-compass/VLMEvalKit
两次相同的测试结果,同一个评估模型,结果差异很大
Open
@PhoenixZ810 is already working on this.
Since Feb 25, 2025.
- Dominant language
- Python
- Stars
- 4.4k
- Forks
- 768
- Avg merge
- 1d 10h
- Merged PRs (30d)
- 17
Description
上面两个结果是一致的,对过了,但是评估模型给的分数差异比较大,用的是MiniCPM-V-2_6,是评估模型哦能力不行吗?因为是内网用不了openai,推荐用哪个本地LLM
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.