ChenRocks / ChenRocks/UNITER

Vocabulary and single image-question pair prediction

Open
#46 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
799
Forks
111
PR merge metrics
No merged PRs in 30d

Description

1. Is the vocabulary available that takes the words of the questions and converts them to 'input_ids'?
2. Is there a function that does this for an input question?
3. Is there a code that take a single image-question pair and predicts the answer?

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by locating the repository's Python entry points for vocabulary handling and image-question inference, then check whether they expose the requested single-pair workflow. Done means identifying or adding a clear path for question-to-input_ids conversion and single image-question answer prediction, with usage documented for verification.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
computer-vision, machine-learning
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.