huggingface / huggingface/transformers
vits support?
Open
New model
- Dominant language
- Python
- Stars
- 166k
- Forks
- 34.6k
- Avg merge
- 3d 9h
- Merged PRs (30d)
- 281
Description
add support for vits, an tts transformer based e2e model.
https://github.com/jaywalnut310/vits
Contributor guide
Research direction
The issue names no Transformers files, tests, or entry points. Start by reviewing the linked VITS project and existing audio or speech model implementations in Transformers to determine the integration scope. Done would mean VITS support is implemented and covered by the project's relevant model and inference tests.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python, pytorch
- Domain
- audio-video-rtc, machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 20/100