bytedance / bytedance/MegaTTS3

子模块中音频和文本实现音素对齐模型怎么用?

Open
#72 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
6.1k
Forks
474
PR merge metrics
No merged PRs in 30d

Description

这个部分在ReadME中阐述了,但不知道怎么使用!如何输入文字和音频实现对齐!麻烦作者给个使用方式感谢!!!

Contributor guide

No contributing guide indexed for this repository

Research direction

Start with the README section describing the audio-and-text phoneme alignment model and trace how the relevant submodule is invoked. Document the expected text and audio inputs, the usage steps, and what output demonstrates successful alignment.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
documentation
Issue type
Documentation
Difficulty
2/5
Estimated time
1-3 hours
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
30/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.