asteroid-team / asteroid-team/asteroid
Inference script to hear samples from AVSpeech Model ?
- Dominant language
- Python
- Stars
- 2.6k
- Forks
- 450
- PR merge metrics
- No merged PRs in 30d
Description
I have implemented the [AVSpeech](https://github.com/asteroid-team/asteroid/tree/master/egs/avspeech/looking-to-listen
) training and eval. How can hear out the sound and check how the model performs? Also, what is the ideal SNR for the model?
How can I find the code to infer to hear the outputs? The checkpoint is `best.pth`?
Contributor guide
Research direction
Start with egs/avspeech/looking-to-listen and the existing AVSpeech training and evaluation setup. Trace how the best.pth checkpoint is used, then identify the documented path for producing audio outputs and checking SNR; done means a newcomer can run inference and listen to the samples.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- python, pytorch
- Domain
- machine-learning
- Issue type
- Documentation
- Difficulty
- 4/5
- Estimated time
- 3-5 days
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100