facebookresearch / facebookresearch/perception_models

Reproduce the audio zero-shot classification results

Open
#121 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
Jupyter Notebook
Stars
2.4k
Forks
162
PR merge metrics
No merged PRs in 30d

Description

Thank you for your excellent work and for open-sourcing this project! 🙌

I'm trying to reproduce the audio zero-shot classification results and was wondering if you could share a few quick details:
- The prompt template used (e.g., `"This is a sound of {}"`)
- The exact class label list (and how they were formatted)
- Any audio/text preprocessing steps

If there's a config snippet or eval script lying around, that'd be awesome too! 😄
No worries if it's not handy—just thought I'd ask. Really appreciate your help!

Contributor guide

Open the contributing guide

Assessment

This issue has not been assessed yet.

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.