TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
active 2024-12-24 → 2025-02-08 (UTC)
Activity over time
Daily event counts in the loaded window
Line chart, 47 days from 2024-12-24 to 2025-02-08. Pushes: 23 total, peak 8 in a day. Pull requests: 7 total, peak 2 in a day. Issues: 0 total, peak 0 in a day. Comments: 0 total, peak 0 in a day. Stars: 0 total, peak 0 in a day.
- Pushes
- Pull requests
- Issues
- Comments
- Stars
Top contributors
Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity
Recent activity
Latest issues, pull requests and releases
- Pull request#1yingcanw2025-02-08 12:06
- Pull request#1yingcanw2025-02-08 12:05
- Pull request#3yingcanw2025-02-07 06:04
- Pull request#2yingcanw2025-01-02 07:44
- Pull request#2jershi4252025-01-02 06:52
- Pull request#1yingcanw2024-12-24 14:20
- Pull request#1jershi4252024-12-24 13:46
Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 0 stars here means stars gained during the window, not the repo's star count.