facebookresearch / facebookresearch/dinov2

multi-object retrieval

Open
#175 1 comment 2 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
13.3k
Forks
1.3k
PR merge metrics
No merged PRs in 30d

Description

As far as I know, most of the current image retrieval is for a single object or several objects. But for the field of autonomous driving, a picture taken is multi-object, for example, a picture contains people, cars, fences, buildings, trees, etc. Only using [cls_token] will lose a lot of object information. Are there any suggestions for multi-object retrieval?

Contributor guide

Open the contributing guide

Research direction

The issue names no files, tests, or entry points. Start by locating the existing image-retrieval path and how [cls_token] is used, then determine whether the proposed multi-object case has an agreed implementation direction; done would require a concrete, validated approach rather than general suggestions.

Written by the indexing model from the issue text.

Assessment

Tech stack
pytorch
Domain
autonomous-driving, computer-vision, search
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.