asteroid-team / asteroid-team/asteroid

Inference script to hear samples from AVSpeech Model ?

Open
#689 0 comments 0 reactions 0 assignees View on GitHub
question
Dominant language
Python
Stars
2.6k
Forks
450
PR merge metrics
No merged PRs in 30d

Description

I have implemented the [AVSpeech](https://github.com/asteroid-team/asteroid/tree/master/egs/avspeech/looking-to-listen
) training and eval. How can hear out the sound and check how the model performs? Also, what is the ideal SNR for the model?

How can I find the code to infer to hear the outputs? The checkpoint is `best.pth`?

Contributor guide

Open the contributing guide

Research direction

Start with egs/avspeech/looking-to-listen and the existing AVSpeech training and evaluation setup. Trace how the best.pth checkpoint is used, then identify the documented path for producing audio outputs and checking SNR; done means a newcomer can run inference and listen to the samples.

Written by the indexing model from the issue text.

Assessment

Tech stack
python, pytorch
Domain
machine-learning
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.