huggingface / huggingface/diffusers

[Add] VEnhancer - the interpolation and upscaler for CogVideoX-5b

Aperta
#9,303 3 commenti 1 reazione 0 assegnatari Vedi su GitHub
stale
Lingua principale
Python
Stelle
34.5k
Fork
7.3k
Merge medio
3g 3h
PR unite (30g)
91

Descrizione

### Model/Pipeline/Scheduler description

VEnhancer, a generative space-time enhancement framework that can improve the existing T2V results.

https://github.com/Vchitect/VEnhancer

### Open source status

- [X] The model implementation is available.
- [X] The model weights are available (Only relevant if addition is not a scheduler).

### Provide useful links for the implementation

_No response_

Guida per i contributori

Apri la guida per i contributori

Direzione di ricerca

Start by reading the VEnhancer implementation and model-weight links in the issue, then compare them with the repository's existing CogVideoX-5b pipeline integrations. The work is complete when VEnhancer is available as an interpolation and upscaling addition for CogVideoX-5b, with behavior matching the upstream implementation.

Scritto dal modello di indicizzazione a partire dal testo della issue.

Valutazione

Stack tecnologico
python, pytorch
Ambito
machine-learning
Tipo di issue
Funzionalità
Difficoltà
5/5
Tempo stimato
Più di una settimana
Stato di attività
Ferma
Chiarezza
Da chiarire
Idoneità per principianti
20/100

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.