bytedance / bytedance/1d-tokenizer

reconstruction loss for the first stage

Open
#43 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
1.2k
Forks
70
PR merge metrics
No merged PRs in 30d

Description

Thank you very much for your work, I switched to a new dataset for retraining and I'm wondering what is probably a reasonable amount of reconstruction loss for the first stage. What was it roughly when you trained?

Contributor guide

No contributing guide indexed for this repository

Research direction

No file, notebook, test, or implementation entry point is named. Start by locating the first-stage training and reconstruction-loss reporting in the repository, then determine whether the project records a reference value for the original dataset. Done would require documenting a supported comparison or explaining that no fixed target is available.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook
Domain
machine-learning
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.