huggingface / huggingface/transformers

vits support?

Open
#15,535 1 comment 1 reaction 0 assignees View on GitHub
New model
Dominant language
Python
Stars
166k
Forks
34.6k
Avg merge
3d 9h
Merged PRs (30d)
281

Description

add support for vits, an tts transformer based e2e model.

https://github.com/jaywalnut310/vits

Contributor guide

Open the contributing guide

Research direction

The issue names no Transformers files, tests, or entry points. Start by reviewing the linked VITS project and existing audio or speech model implementations in Transformers to determine the integration scope. Done would mean VITS support is implemented and covered by the project's relevant model and inference tests.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
audio-video-rtc, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.