kaldi-asr / kaldi-asr/kaldi

What is the procedure of training one far-distance ASR model for noise environment

Open
#4,702 3 comments 0 reactions 0 assignees View on GitHub
kaldi10-TODO stale
Dominant language
Shell
Stars
15.5k
Forks
5.4k
PR merge metrics
No merged PRs in 30d

Description

Hi,
As we all know, recipes in kaldi is mostly for quiet environment. Now I want to train a ASR model with **chain model** for noise environment. my question is as follows:
1) Does i need to add noise segments to clean wav in GMM training stages? Namely, only clean wav can be used in GMM training stages, is right ?
2) the procedure of adding noise segments appeared in after GMM training and before chain model training,is right?. And, we need to use 'steps/align_fmmr_lats.sh' for realigning. is right ?
Please tell me the whole procedure of training one far-distance ASR model with chain model for noise environment. Thank you

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reviewing the Kaldi recipes and the steps/align_fmmr_lats.sh script named in the issue. The requested outcome is a complete, documented chain-model training procedure for far-distance speech in noise, including answers about clean or noisy GMM stages and when realignment occurs.

Written by the indexing model from the issue text.

Assessment

Tech stack
shell
Domain
machine-learning
Issue type
Documentation
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
18/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.