work on llava-next LLaVA-Video-7B-Qwen2 ?
Open
- Dominant language
- Python
- Stars
- 702
- Forks
- 34
- PR merge metrics
- No merged PRs in 30d
Description
I'm currently studying about LLaVA-Video-7B-Qwen2, it uses vision model : siglip-so400m-patch14-384, can you share how to switch vision model to use MLCD-ViT-B-32-224px ?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.