DAMO-NLP-SG / DAMO-NLP-SG/VideoLLaMA2
Webvid-10M (40% sampling)
Open
- Dominant language
- Python
- Stars
- 1.3k
- Forks
- 90
- PR merge metrics
- No merged PRs in 30d
Description
Hello,
After going through the paper, I understood that 40% of video-text pairs are used from webvid-10M dataset. Can you please provide me the rationale, or, point me in the direction which helps me understand how these 40% of video are picked.
Contributor guide
No contributing guide indexed for this repository
Assessment
This issue has not been assessed yet.