facebookresearch / facebookresearch/sam2

Finetune video predictor on custom data

Open
#305 13 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
19.9k
Forks
2.5k
PR merge metrics
No merged PRs in 30d

Description

I was able to finetune SAM on custom image segmentation dataset which trained the mask decoder and got very high IoU. Then I used the same model weights to test video predictor but didn't see improvement. I assume it has to do with memory encoder. How can I finetune the whole model? Didn't want to reinvent the wheel if someone has done it.

Contributor guide

Open the contributing guide

Research direction

The issue names no file, test, or entry point. Start with the repository's example notebooks and available checkpoint or training guidance, then determine whether whole-model fine-tuning for the video predictor is supported; done means a documented or implemented path that addresses custom-data video prediction.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook
Domain
machine-learning
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.