huggingface / huggingface/diffusers

[Looking for community contribution] support Wan 2.2 S2V: an audio-driven cinematic video generation model

Open
#12,257 7 comments 1 reaction 0 assignees View on GitHub
contributions-welcome Good second issue help wanted
Dominant language
Python
Stars
34.5k
Forks
7.3k
Avg merge
3d 3h
Merged PRs (30d)
91

Description

We're super excited about the Wan 2.2 S2V (Speech-to-Video) model and want to get it integrated into Diffusers! This would be an amazing addition, and we're looking for experienced community contributors to help make this happen.

- **Project Page**: https://humanaigc.github.io/wan-s2v-webpage/
- **Source Code**: https://github.com/Wan-Video/Wan2.2#run-speech-to-video-generation
- **Model Weights**: https://huggingface.co/Wan-AI/Wan2.2-S2V-14B

This is a priority for us, so we will try review fast and actively collabrate with you throughout the process :)

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.