What is the procedure of training one far-distance ASR model for noise environment
- Dominant language
- Shell
- Stars
- 15.5k
- Forks
- 5.4k
- PR merge metrics
- No merged PRs in 30d
Description
Hi,
As we all know, recipes in kaldi is mostly for quiet environment. Now I want to train a ASR model with **chain model** for noise environment. my question is as follows:
1) Does i need to add noise segments to clean wav in GMM training stages? Namely, only clean wav can be used in GMM training stages, is right ?
2) the procedure of adding noise segments appeared in after GMM training and before chain model training,is right?. And, we need to use 'steps/align_fmmr_lats.sh' for realigning. is right ?
Please tell me the whole procedure of training one far-distance ASR model with chain model for noise environment. Thank you
Contributor guide
No contributing guide indexed for this repository
Research direction
Start by reviewing the Kaldi recipes and the steps/align_fmmr_lats.sh script named in the issue. The requested outcome is a complete, documented chain-model training procedure for far-distance speech in noise, including answers about clean or noisy GMM stages and when realignment occurs.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- shell
- Domain
- machine-learning
- Issue type
- Documentation
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 18/100