DAMO-NLP-SG / DAMO-NLP-SG/Auto-Arena-LLMs

when all winners in every battle is model A (or model B or tie), the Evaluating will get error at the function of "compute_mle_elo"

Open
#4 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
45
Forks
1
PR merge metrics
No merged PRs in 30d

Description

when all winners in every battle is model A (or model B or tie), the Evaluating will get error at the line 180 of the function "compute_mle_elo" in utils/score_utils.py. It is possible for the situation that all winners are same type, really hope a way to solve the problem.

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.