OpenNMT / OpenNMT/CTranslate2

Performance comparison with TensorRT-LLM

Open
#1,896 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

Dominant language
C++
Stars
4.7k
Forks
536
Avg merge
12h 12m
Merged PRs (30d)
4

Description

Hi, I would like to know if anyone has compared the performance data of some models, such as Whisper, with TensorRT-LLM on NVIDIA machines.

Contributor guide

Open the contributing guide

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

No file, test, or entry point is mentioned. First clarify the models, NVIDIA hardware, TensorRT-LLM version, metrics, and expected comparison scope, then identify the repository's existing benchmarking entry point; done means producing reproducible performance data against the requested TensorRT-LLM baseline.

Written by the indexing model from the issue text.

Assessment

Tech stack
cpp
Domain
machine-learning, performance
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.