alibaba / alibaba/EasyNLP

choose best checkpoint according to metrics in train_and_evaluate

Open
#91 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
2.2k
Forks
257
PR merge metrics
No merged PRs in 30d

Description

This issue has no description.

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by locating the train_and_evaluate entry point and determine which metrics it currently produces and how checkpoints are saved or selected. Clarify which metric should decide the best checkpoint, including direction and comparison behavior, then define done as selecting the expected checkpoint consistently and covering the behavior with the relevant existing tests.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.