Repository Issues

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

View on GitHub
Stars
 (14,296 stars)
Forks
 (2,637 forks)
Indexed issues
 (0 indexed issues)
open beginner issues
 (0 open beginner issues)
Latest indexed
Aug 4, 2026
Last GitHub push
Aug 4, 2026
License
Other
Contributing guide
Contributing guide
Code of conduct
Code of conduct
Dominant language
Python
PR merge metrics
 (Avg merge 7d 17h) (705 merged PRs in 30d)
Beginner labels
No beginner labels indexed

Issues

0 open indexed issues

No open indexed issues found for this repository.