EveryVoiceTTS / EveryVoiceTTS/EveryVoice

For improving EveryVoice models, we could look at reporting phoneme distribution in dataset

Open
#420 0 comments 0 reactions 0 assignees View on GitHub
enhancement
Dominant language
Python
Stars
45
Forks
4
Avg merge
1d 8h
Merged PRs (30d)
14

Description

We could have tools that report the distribution of phonemes in the dataset

Contributor guide

Open the contributing guide

Research direction

The issue names no files, tests, or entry points. Start by locating the dataset-related tooling and how phonemes are extracted, then determine the expected report format. Done means a tool can report phoneme distribution for a dataset, with tests covering the resulting output.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
data, machine-learning
Issue type
Feature
Difficulty
3/5
Estimated time
1-2 days
Activity status
Stale
Clarity
Mostly clear
Newbie friendliness
35/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.