espnet / espnet/espnet_onnx

tts onnx gpu inference time problems

Open
#70 2 comments 0 reactions 0 assignees View on GitHub
Dominant language
Python
Stars
169
Forks
25
PR merge metrics
No merged PRs in 30d

Description

hi, i met a problem when onnx inference on gpu
1. onnx inference on gpu slower than onnx cpu inference much time and sometimes faster than gpu pt inference(2 times acceleration)
2. when i inference same text twice or more, inference achieves 2 time acceleration compare to gpu pt inference
any advicec?
thanks

Contributor guide

No contributing guide indexed for this repository

Research direction

Start by reproducing the reported GPU and CPU inference timings for the same text, including repeated inferences, and compare them with GPU PyTorch timings. The issue does not name a source file or test; done means identifying the cause of the inconsistent performance and documenting or validating a measurable improvement.

Written by the indexing model from the issue text.

Assessment

Tech stack
python
Domain
machine-learning, performance
Issue type
Bug
Difficulty
4/5
Estimated time
3-5 days
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.