[Discussion] Why is it useful to train just masks of stft?
Open
question
- Dominant language
- Python
- Stars
- 28.4k
- Forks
- 3.1k
- PR merge metrics
- No merged PRs in 30d
Description
Spleeter trains many masks of stft to split songs. But why does it work? Is is possible to get a better model if I just input stft feature and output stft of each instrument?
Contributor guide
Assessment
This issue has not been assessed yet.