ByteDance-Seed / ByteDance-Seed/Depth-Anything-3
Finetune on ImageNet3D dataset
- Dominant language
- Python
- Stars
- 6.3k
- Forks
- 702
- PR merge metrics
- No merged PRs in 30d
Description
Hi,
Thanks for the great work! Recently I noticed there is a ImageNet3D dataset (https://github.com/Rawmantic/Dataset3D), which annoted with pose, normal and depth. I'm able to improve foundationstereo greatly with such dataset, but for depthanythingv3, I'm wondering how to finetune it with improved foundationstereo label, or use that dataset directly. Is there any training script and data preparation I can refer to?
Thanks!
Contributor guide
No contributing guide indexed for this repository
Research direction
No repository files, tests, or entry points are named. Start by locating the existing training and data-preparation workflows, then determine how ImageNet3D data or improved FoundationStereo labels would be represented and consumed. Done would require a working finetuning path or documented scripts for preparing and training on the requested data.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python
- Domain
- computer-vision, machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Active
- Clarity
- Needs clarification
- Newbie friendliness
- 35/100