TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
active 2024-05-15 → 2024-06-12 (UTC)
Activity over time
Daily event counts in the loaded window
Line chart, 29 days from 2024-05-15 to 2024-06-12. Pushes: 3 total, peak 1 in a day. Pull requests: 2 total, peak 2 in a day. Issues: 0 total, peak 0 in a day. Comments: 0 total, peak 0 in a day. Stars: 0 total, peak 0 in a day.
- Pushes
- Pull requests
- Issues
- Comments
- Stars
Top contributors
Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity
| Contributor | Contributions | Pushes | PRs | Comments |
|---|---|---|---|---|
| WilliamEricCheung | 5 | 3 | 2 | 0 |
Recent activity
Latest issues, pull requests and releases
- Pull request#1WilliamEricCheung2024-06-12 02:51
- Pull request#1WilliamEricCheung2024-06-12 02:50
Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 0 stars here means stars gained during the window, not the repo's star count.