nvidia-riva / nvidia-riva/tutorials

How to use a model that I've downloaded?

Open
#109 0 comments 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
Jupyter Notebook
Stars
175
Forks
56
PR merge metrics
No merged PRs in 30d

Description

I've downloaded the models from https://catalog.ngc.nvidia.com/orgs/nvidia/teams/tao/models/speechsynthesis_en_us_fastpitch_ipa and https://catalog.ngc.nvidia.com/orgs/nvidia/teams/tao/models/speechsynthesis_en_us_hifigan_ipa. But I've no idea to use it. The tutorials just give how to generate speech with Riva TTS APIs. Can you give me a tutorial on how to generate speech using the downloaded models?

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Review the existing tutorials that generate speech with the Riva TTS APIs, then compare them with the downloaded FastPitch IPA and HiFi-GAN IPA models from the linked NGC pages. Determine the documented workflow for loading those models and generating speech, and add a tutorial that demonstrates the complete process from downloaded models to generated audio.

Written by the indexing model from the issue text.

Assessment

Tech stack
jupyter-notebook
Domain
documentation, machine-learning
Issue type
Documentation
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.