pytorch / pytorch/audio

Add high-level IO function for image and video

Open
#3,327 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Python
Stars
2.9k
Forks
799
Avg merge
58m
Merged PRs (30d)
3

Description

Leveraging StreamReader we can load video and images. We should add functions

torchaudio.io.load_image
torchaudio.io.load_video
torchaudio.io.load_audio

which are thin wrapper around StreamReader.

(and perhaps save versions)

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by locating StreamReader and the existing torchaudio.io entry points, then trace how audio, image, and video data are currently loaded. The issue is complete when load_image, load_video, and load_audio provide thin, consistent wrappers around StreamReader; the optional save functions need scope clarification before implementation.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
audio-video-rtc
Issue type
Feature
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Mostly clear
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.