open-compass / open-compass/opencompass

[Feature] 请问主观评测脚本支持用本地模型作为judge模型吗?

Open
#2,159 1 comment 0 reactions 1 assignee View on GitHub

@MaiziXiao is already working on this.

Since Jun 18, 2025.

Dominant language
Python
Stars
7.5k
Forks
869
Avg merge
17h 52m
Merged PRs (30d)
13

Description

Describe the feature

examples/eval_subjective.py
在这个文件中,我把judge_models改为了vllmwithchattemplate的形式,似乎并不能正常评测,alpaca eval的最终输出结果为空。
请问主观评测脚本支持用本地模型作为judge模型吗?

Will you implement it?
  • I would like to implement this feature and create a PR!

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.