huggingface / huggingface/lighteval

[FT] Add tool usage benchmarks

Open
#256 1 comment 0 reactions 0 assignees View on GitHub
feature
Dominant language
Python
Stars
2.5k
Forks
555
Avg merge
1d 6h
Merged PRs (30d)
1

Description

## Issue encountered
Lighteval does not allow evaluating models on tool usage.

## Solution/Feature
Add benchmarks for tool usage

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.