huggingface / huggingface/diffusers
[Add] VEnhancer - the interpolation and upscaler for CogVideoX-5b
- Lingua principale
- Python
- Stelle
- 34.5k
- Fork
- 7.3k
- Merge medio
- 3g 3h
- PR unite (30g)
- 91
Descrizione
### Model/Pipeline/Scheduler description
VEnhancer, a generative space-time enhancement framework that can improve the existing T2V results.
https://github.com/Vchitect/VEnhancer
### Open source status
- [X] The model implementation is available.
- [X] The model weights are available (Only relevant if addition is not a scheduler).
### Provide useful links for the implementation
_No response_
Guida per i contributori
Apri la guida per i contributori
Direzione di ricerca
Start by reading the VEnhancer implementation and model-weight links in the issue, then compare them with the repository's existing CogVideoX-5b pipeline integrations. The work is complete when VEnhancer is available as an interpolation and upscaling addition for CogVideoX-5b, with behavior matching the upstream implementation.
Scritto dal modello di indicizzazione a partire dal testo della issue.
Valutazione
- Stack tecnologico
- python, pytorch
- Ambito
- machine-learning
- Tipo di issue
- Funzionalità
- Difficoltà
- 5/5
- Tempo stimato
- Più di una settimana
- Stato di attività
- Ferma
- Chiarezza
- Da chiarire
- Idoneità per principianti
- 20/100