bytedance / bytedance/MegaTTS3
Feature request: Allow generating subtitles/word-time-stamps along
Open
- Dominant language
- Python
- Stars
- 6.1k
- Forks
- 474
- PR merge metrics
- No merged PRs in 30d
Description
Not only generate speech audio, but the subtitles/word-time-stamp along.
Contributor guide
No contributing guide indexed for this repository
Research direction
The issue names no files, tests, or entry points. Start by reviewing how the existing speech audio generation works and how its output is represented. Done means the project generates subtitles or word timestamps alongside the speech audio, with an agreed output format and tests covering the result.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- audio-video-rtc
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100