deezer / deezer/spleeter

why 2stems training still need drums,bass and other but accompaniment.wav?

Open
#671 1 comment 0 reactions 0 assignees View on GitHub
question
Dominant language
Python
Stars
28.4k
Forks
3.1k
PR merge metrics
No merged PRs in 30d

Description

the training seems endless eventhough i set only 1 song.
i use 2stems config and i tried to remove 'drums_path','bass_path','other_path' from the .csv file then the trainging won't start. so i added 'accompaniment_path' to the .csv file the training can get to start, but there's no record about loading accompaniment.wav. **Although the instrument list only contains vocals and accompaniment, the training still loaded drums.wav bass.wav other.wav.** It just ignore my accompaniment.wav file anyway.

my command: **nvidia-docker run -v /media/danshafaker/软件专用/SpleeterGUI:/base researchdeezer/spleeter:3.7-gpu-2stems train -p /base/configs/2stems/base_config.json -d /base/xunlian**

here're my configs

![2021-10-20 01-10-28 的屏幕截图](https://user-images.githubusercontent.com/43251345/137960545-2344f483-1bf5-4b7f-9389-f86caac4ece3.png)

{
"train_csv":
"/base/configs/musdb_train.csv",
"validation_csv": "/base/configs/musdb_validation.csv",
"model_dir": "/base/2stems",
"mix_name": "mix",
"instrument_list": ["vocals", "accompaniment"],
"sample_rate":44100,
"frame_length":4096,
"frame_step":1024,
"T":512,
"F":1024,
"n_channels":2,
"separation_exponent":2,
"mask_extension":"zeros",
"learning_rate": 1e-4,
"batch_size":4,
"training_cache":"/base/cache/training_cache",
"validation_cache":"/base/cache/validation_cache",
"train_max_steps": 1000000,
"throttle_secs":300,
"random_seed":0,
"save_checkpoints_steps":150,
"save_summary_steps":5,
"model":{
"type":"unet.unet",
"params":{}
}
}

**traincsv**
mix_path,vocals_path,drums_path,bass_path,other_path,accompaniment_path,duration
train/A Classic Education - NightOwl/mixture.wav,train/A Classic Education - NightOwl/vocals.wav,train/A Classic Education - NightOwl/drums.wav,train/A Classic Education - NightOwl/bass.wav,train/A Classic Education - NightOwl/other.wav,train/A Classic Education - NightOwl/accompaniment.wav,171.247166

**validatecsv**
mix_path,vocals_path,drums_path,bass_path,other_path,accompaniment_path,duration
train/ANiMAL - Rockshow/mixture.wav,train/ANiMAL - Rockshow/vocals.wav,train/ANiMAL - Rockshow/drums.wav,train/ANiMAL - Rockshow/bass.wav,train/ANiMAL - Rockshow/other.wav,train/ANiMAL - Rockshow/accompaniment.wav,165.511837

is it possible training the 2stems model only with the original song and the instrumental ?

Contributor guide

Open the contributing guide

Research direction

Start with the train command and compare configs/2stems/base_config.json against configs/musdb_train.csv and the validation CSV shown in the report. Trace how instrument_list and CSV column names are used during training, then verify that a two-stem run loads vocals.wav and accompaniment.wav rather than the drums, bass, and other paths. Done means training starts with the supplied two-stem data and logs the expected files.

Written by the indexing model from the issue text.

Assessment

Tech stack
docker, python, tensorflow
Domain
audio-video-rtc, machine-learning
Issue type
Bug
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Mostly clear
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.