Which version of gpt-4 is used to generate the mt-bench scores on lmsys leaderboard?
Open
Nobody has claimed this yet.
- Dominant language
- Python
- Stars
- 39.5k
- Forks
- 4.8k
- PR merge metrics
- No merged PRs in 30d
Description
Which version of gpt-4 is used to generate the mt-bench scores on lmsys leaderboard https://chat.lmsys.org/?leaderboard ? Is it gpt-4-0613 or gpt-4-0314?
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Research direction
Open the linked LMSYS leaderboard and inspect the mt-bench evaluation configuration or documentation. Confirm whether the judge model is gpt-4-0613 or gpt-4-0314, then document the identified version in the issue or relevant project documentation.
Written by the indexing model from the issue text.
Assessment
- Domain
- analytics, machine-learning
- Issue type
- Documentation
- Difficulty
- 1/5
- Estimated time
- Under an hour
- Activity status
- Stale
- Clarity
- Clearly specified
- Newbie friendliness
- 25/100