facebookresearch / facebookresearch/SlowFast
MaskFeat training support ?
- Dominant language
- Python
- Stars
- 7.4k
- Forks
- 1.3k
- PR merge metrics
- No merged PRs in 30d
Description
thanks for the great project, I just want to know is there any plan to release MaskFeat pre-trained model or MViT training code. :smile_cat:
> Our results on standard video benchmarks are groundbreaking: MaskFeat pre-trained MViT-L [56] gets 86.7%
top-1 accuracy on Kinetics-400 [51] without using any external data, greatly surpassing the best prior number of this
kind by +5.2%, and also methods using large-scale image
datasets, e.g., IN-21K and JFT-300M [80]. When transferring to downstream tasks, MaskFeat gets unprecedented
results of 38.8 mAP on action detection (AVA [42]) and
75.0% top-1 accuracy on human-object interaction classification (SSv2 [40]). When generalized to the image domain,
MaskFeat also obtains competitive 84.0% top-1 with ViT-B
and 85.7% with ViT-L using only ImageNet-1K [24].
Our code will be available in PyTorchVideo1,2[29, 30].
Contributor guide
Research direction
The issue names no files, tests, or entry points. First check the current SlowFast and PyTorchVideo code and release status for MaskFeat pre-trained models and MViT training support. Done would need a defined implementation or release scope, plus documented verification of the requested support.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python, pytorch
- Domain
- machine-learning
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 15/100