faroit / faroit/python_audio_loading_benchmark

Benchmark DALI decoding for pytorch and tensorflow

Open
#11 1 comment 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
152
Forks
11
PR merge metrics
No merged PRs in 30d

Description

I totally missed that [DALI added audio decoding ops](https://docs.nvidia.com/deeplearning/dali/user-guide/docs/supported_ops.html#nvidia.dali.ops.AudioDecoder). As far as I can see, the ops seems to be based on sndfile and do not support GPU decoding. I wonder if this would still improve single threaded ETL speed on pytorch or tensorflow.

pinging DALI devs @mzient @szalpal as they might have benchmarked the op against [tfio](https://www.tensorflow.org/io/api_docs/python/tfio/audio/decode_wav) or [torchaudio](https://pytorch.org/audio/) which are also sndfile based.

Contributor guide

No contributing guide indexed for this repository

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.