google-deepmind / google-deepmind/kinetics-i3d

Inflating pre-trained models

Open
#109 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
1.8k
Forks
467
PR merge metrics
No merged PRs in 30d

Description

Hello All,

Upon reading the paper, I see that the authors have repeated the same image numerous time to create a "boring" video. What I fail to understand is the following, upon creating the "boring" video is the I3D model trained from scratch to create pre-trained weights? Or does there exists a schema wherein the 2-D pre-trained weights are converted to 3D in nature to initialize the network?

Contributor guide

Open the contributing guide

Research direction

Start by reviewing the paper's description of the repeated-image video and the repository's I3D model context. Determine whether the pre-trained weights are learned from scratch or initialized by converting 2-D weights to 3-D, then document the evidence and clarify the training path for readers.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
computer-vision, machine-learning
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.