deepglint / deepglint/unicom

work on llava-next LLaVA-Video-7B-Qwen2 ?

Open
#117 0 comments 0 reactions 1 assignee Claimed by @anxiangsir View on GitHub
Dominant language
Python
Stars
702
Forks
34
PR merge metrics
No merged PRs in 30d

Description

I'm currently studying about LLaVA-Video-7B-Qwen2, it uses vision model : siglip-so400m-patch14-384, can you share how to switch vision model to use MLCD-ViT-B-32-224px ?

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.