ChenRocks / ChenRocks/UNITER

How to judge the convergence of the pre-training model?

Open
#89 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
799
Forks
111
PR merge metrics
No merged PRs in 30d

Description

How to measure the loss weight of different pre-training tasks? Which task's loss determines the model training convergence?

Contributor guide

No contributing guide indexed for this repository

Research direction

The issue names no files, tests, or entry points to inspect. Start by reviewing the repository's pre-training task and loss definitions, then establish which convergence criterion and loss weighting guidance would answer the question; done would be a documented, reproducible explanation.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
ai, machine-learning
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.