DAMO-NLP-SG / DAMO-NLP-SG/VideoLLaMA3
How to Perform Inference with Fine-Tuned VideoLLaMA3?
Open
- Dominant language
- Jupyter Notebook
- Stars
- 1.2k
- Forks
- 89
- PR merge metrics
- No merged PRs in 30d
Description
Thanks for your great work! I have a question: After fine-tuning VideoLLaMA3-2B, how can I use the model for inference? The fine-tuned model has different weights compared to the original VideoLLaMA3-2B, as it includes a vision encoder. Should I use the same configuration file as the original VideoLLaMA3-2B?
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.