huggingface / huggingface/lighteval

[EVAL] Add RULER for evaluating long context

Open
#726 0 comments 0 reactions 1 assignee Claimed by @NathanHB View on GitHub
new-task science-team
Dominant language
Python
Stars
2.5k
Forks
555
Avg merge
1d 6h
Merged PRs (30d)
1

Description

## Evaluation short description
- Evaluate long context for lLM and figure out the real context size

## Evaluation metadata
Provide all available
- Paper url:
- Github URL: https://github.com/NVIDIA/RULER
- Dataset url:

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.