babysor / babysor/MockingBird

请问作者大大,encoder如何训练?

Open
#996 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
36.9k
Forks
5.2k
PR merge metrics
No merged PRs in 30d

Description

看了知乎链接的教程,尝试训练encoder
![Screenshot 2024-04-29 184437](https://github.com/babysor/MockingBird/assets/51695272/901047ec-0d6a-4dd8-842a-6d813eae230c)
练了半天,这结果似乎没什么变化
![wei_umap_038600](https://github.com/babysor/MockingBird/assets/51695272/b924f344-9ab0-429e-bfaa-ea6dde7c791f)
![wei_umap_038700](https://github.com/babysor/MockingBird/assets/51695272/7d0b5ba1-9105-487b-8be5-d50ae50bd88d)
数据自建的,有2个多G
问题1:这种情况是正常的吗?如果不正常是什么原因造成的?
问题2:根据知乎上的说法“**实测了一次 训练synthesizer时,4000左右step就能attention收敛,22k step的时候loss就到0.35了,可以很快进行finetune,算是超越预期。**”,训练synthesizer时,如何把encoder加入?

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing the referenced Zhihu tutorial and reproducing encoder training with the reported self-built dataset, comparing the two UMAP screenshots and training behavior. Then trace the repository's encoder and synthesizer training workflows to determine whether encoder outputs are already included; done means documenting the expected result, likely causes of no change, and the supported integration steps.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
machine-learning
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.