ByteDance-Seed / ByteDance-Seed/Depth-Anything-3

Finetune on ImageNet3D dataset

Open
#275 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
6.3k
Forks
702
PR merge metrics
No merged PRs in 30d

Description

Hi,

Thanks for the great work! Recently I noticed there is a ImageNet3D dataset (https://github.com/Rawmantic/Dataset3D), which annoted with pose, normal and depth. I'm able to improve foundationstereo greatly with such dataset, but for depthanythingv3, I'm wondering how to finetune it with improved foundationstereo label, or use that dataset directly. Is there any training script and data preparation I can refer to?

Thanks!

Contributor guide

No contributing guide indexed for this repository

Research direction

No repository files, tests, or entry points are named. Start by locating the existing training and data-preparation workflows, then determine how ImageNet3D data or improved FoundationStereo labels would be represented and consumed. Done would require a working finetuning path or documented scripts for preparing and training on the requested data.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
computer-vision, machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Needs clarification
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.