Automated vLLM server parameter tuning tool. Finds optimal max-num-seqs and max-num-batched-tokens to maximize throughput. Includes presets for Llama/Qwen/Mixtral, batch processing, result analysis with visualizations, and environment checks. Supports TPU/GPU with latency constraints.
active 2025-10-30 → 2025-11-23 (UTC)
Activity over time
Daily event counts in the loaded window
Line chart, 25 days from 2025-10-30 to 2025-11-23. Pushes: 0 total, peak 0 in a day. Pull requests: 1 total, peak 1 in a day. Issues: 0 total, peak 0 in a day. Comments: 0 total, peak 0 in a day. Stars: 1 total, peak 1 in a day.
- Pushes
- Pull requests
- Issues
- Comments
- Stars
Top contributors
Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity
| Contributor | Contributions | Pushes | PRs | Comments |
|---|---|---|---|---|
| massif-01 | 1 | 0 | 1 | 0 |
Recent activity
Latest issues, pull requests and releases
- Pull request#1massif-012025-10-30 10:25
Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 1 stars here means stars gained during the window, not the repo's star count.