NVIDIA

NVIDIA/TensorRT-LLM

View on GitHub

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

Stars
14.7k
Forks
2.8k
Open beginner issues
21
Indexed issues
596
Avg merge
2d 23h
Merged PRs (30d)
489
Dominant language
Python
License
No license data
Last GitHub push
Sep 19, 2026
Latest indexed
Sep 19, 2026
Contributing guide
Contributing guide
Code of conduct
Code of conduct
Beginner labels
No beginner labels indexed
596 issues indexed so far Loading issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.