ByteDance-Seed / ByteDance-Seed/Depth-Anything-3

How to Input Multi-View Videos Simultaneously?

Open
#150 3 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
6.3k
Forks
702
PR merge metrics
No merged PRs in 30d

Description

I noticed that the video demonstrates the capability to synchronize and process multi-view autonomous driving footage with multi-view fusion. However, I haven’t found a way to input multiple video perspectives at once in the CLI. Could you please guide me on how to ensure multi-view videos can be processed simultaneously?

As shown in the demo in the image below:

Image

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by inspecting the repository's CLI entry point and the existing video-input handling, then compare those paths with the multi-view fusion demo shown in the issue. Determine whether simultaneous multi-view input is already supported or requires a defined interface; done should mean the supported input procedure is clear and verified, or the requested capability is implemented and tested.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
cli, computer-vision
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
30/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.