OpenEuroLLM / OpenEuroLLM/Taskboard
Warmup evals
Open
@timurcarstensen is already working on this.
Since Sep 23, 2025.
- Dominant language
- No language data
- Stars
- 3
- Forks
- 0
- PR merge metrics
- No merged PRs in 30d
Description
Some of the open-sci experiments with variable warmup period are here
https://huggingface.co/collections/open-sci/open-sci-ref-001-c4-6880bb9ec71ae334ab2d4bf2
Executed on 1.3B scale, for 50B and 300B token budget.
The task is to evaluate those checkpoints with our eval collection and figure out whether
- long warmup 25000 for 300B is indeed better performing than 20000, 10000
- short warmup of 1000 is critical for 50B (longer warmup worse)
Contributor guide
No contributing guide indexed for this repository
First steps
- Read the whole issue, then the project's contributing guide.
- Comment on the issue to say you are picking it up — it saves two people doing the same work.
- Fork the repository and make your change on a branch.
- Open a pull request that references the issue number.
Assessment
This issue has not been assessed yet.