huggingface / huggingface/diffusers
[Looking for community contribution] support Wan 2.2 S2V: an audio-driven cinematic video generation model
- Lingua principale
- Python
- Stelle
- 34.5k
- Fork
- 7.3k
- Merge medio
- 3g 3h
- PR unite (30g)
- 91
Descrizione
We're super excited about the Wan 2.2 S2V (Speech-to-Video) model and want to get it integrated into Diffusers! This would be an amazing addition, and we're looking for experienced community contributors to help make this happen.
- **Project Page**: https://humanaigc.github.io/wan-s2v-webpage/
- **Source Code**: https://github.com/Wan-Video/Wan2.2#run-speech-to-video-generation
- **Model Weights**: https://huggingface.co/Wan-AI/Wan2.2-S2V-14B
This is a priority for us, so we will try review fast and actively collabrate with you throughout the process :)
Guida per i contributori
Apri la guida per i contributori
Direzione di ricerca
Inizia con le istruzioni per la generazione speech-to-video del codice sorgente di Wan2.2 e con i pesi del modello Wan-AI/Wan2.2-S2V-14B. Determina i requisiti di integrazione per Diffusers. Il completamento è definito dal supporto per Wan 2.2 S2V in Diffusers e dalla revisione da parte dei maintainer.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Valutazione
- Stack tecnologico
- python, pytorch
- Ambito
- machine-learning
- Tipo di issue
- Funzionalità
- Difficoltà
- 5/5
- Tempo stimato
- Più di una settimana
- Stato di attività
- Attiva
- Chiarezza
- Abbastanza chiara
- Idoneità per principianti
- 35/100