ByteDance-Seed / ByteDance-Seed/Depth-Anything-3

About DA3-Long

Open
#132 8 comments 2 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
6.3k
Forks
702
PR merge metrics
No merged PRs in 30d

Description

Hi, thanks for your great work on Depth Anything 3!

I have two questions:

1. When will the code for DA3-Long be released?

I saw the mention of the “long” version in the paper and demos, and I’m wondering if there is an expected timeline for releasing the code or pretrained models for DA3-Long.

2. How to obtain metric-scaled point clouds when using DA3-Long without the pose-conditioned mode?

If I don’t use the pose-conditioned mode, but I do have prior poses or trajectories that already contain metric scale (e.g., from GPS/INS/visual odometry), is there a way to enforce or recover metric scale in the depth predictions produced by DA3-Long?

In other words, is there a recommended way to combine DA3-Long with externally scaled poses so that the resulting point clouds also maintain real-world scale?

Thanks!

Contributor guide

No contributing guide indexed for this repository

Research direction

Review the paper and demos referenced in the issue alongside the current project documentation and release information. A useful outcome would be a maintainer-confirmed DA3-Long release status and clear guidance on whether externally scaled poses can produce metric-scaled point clouds without pose-conditioned mode.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
computer-vision, machine-learning
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.