facebookresearch / facebookresearch/ImageBind

Generating a video and text as output given a input voice

Open
#23 0 comments 6 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
9.1k
Forks
842
PR merge metrics
No merged PRs in 30d

Description

Hi @likethesky @Celebio @neuhaus @colesbury ,
Thanks for the great work and paving way for the multimodal AI research. I am new to multimodal AI.I only worked on computer vision before. I have a small query. How we can make use of Imagebind to create a video and Video Captions(subtitles) as outputs given an input audio in another language ? Just curious to apply Imagebind in different applications .

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.