仓库议题

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

在 GitHub 查看
星标
 (14,296 个星标)
派生
 (2,637 个派生)
已索引议题
 (0 个已索引议题)
个开放新手议题
 (0 个开放的新手议题)
最近索引
2026年8月4日
最近 GitHub push
2026年8月4日
许可证
Other
贡献指南
贡献指南
行为准则
行为准则
主要语言
Python
PR 合并指标
 (平均合并 7天 17小时) (30 天内合并 705 个 PR)
新手标签
没有已索引的新手标签

议题

0 个已索引议题

此仓库没有已索引议题。