huggingface / huggingface/diffusers
Support for InfiniteTalk
- 主要言語
- Python
- スター
- 34.5k
- フォーク
- 7.3k
- 平均マージ
- 3日 3時間
- マージ済み PR(30日)
- 91
説明
### Model/Pipeline/Scheduler description
https://huggingface.co/MeiGen-AI/InfiniteTalk is a wonderful audio driven video generation model and can also support infinite frame , which is based on wan2.1. The demo and user's workflow is also awesome. some examples: https://www.runninghub.cn/ai-detail/1958438624956203010
### Open source status
- [x] The model implementation is available.
- [x] The model weights are available (Only relevant if addition is not a scheduler).
### Provide useful links for the implementation
https://huggingface.co/MeiGen-AI/InfiniteTalk
https://github.com/MeiGen-AI/InfiniteTalk
コントリビューションガイド
調査の方向性
まず InfiniteTalk のモデルカードとリンク先の実装を読み、その後、それらのアーキテクチャと要件を diffusers に既存する Wan 2.1 のサポートと比較します。モデルが統合され、文書化された音声駆動の動画生成ワークフローを実証する検証が用意されていれば完了です。
索引モデルが issue の本文から書いたものです。
評価
- 技術スタック
- python, pytorch
- 領域
- machine-learning
- issue の種類
- 機能追加
- 難易度
- 5/5
- 見積もり時間
- 1週間以上
- 活発さ
- 活発
- 明瞭さ
- 説明が足りない
- 初心者へのやさしさ
- 35/100