facebookresearch / facebookresearch/sam2

Why not training on dynamic input sizes?

Open
#309 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
19.9k
Forks
2.5k
PR merge metrics
No merged PRs in 30d

Description

Since many datasets, real world videos, might not have too large resolution. Why not consider traing the model with dynamic sizes? Like find the closest shape mutiple to 32? Since the Hiera can support rectangle and dynamic input.

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.