倉庫議題

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.

在 GitHub 查看
星標
 (14,393 顆星)
分叉
 (2,664 個分叉)
已索引議題
 (0 個已索引議題)
個開放新手議題
 (0 個開放的新手議題)
最近索引
2026年8月11日
最近 GitHub push
2026年8月16日
授權條款
Other
貢獻指南
貢獻指南
行為準則
行為準則
主要語言
Python
PR 合併指標
 (平均合併 7天 17小時) (30 天內合併 705 個 PR)
新手標籤
沒有已索引的新手標籤

議題

0 個開放索引議題

此倉庫沒有開放的已索引議題。