Skip to content

Dockerized LLM inference server with constrained output (JSON mode), built on top of vLLM. Faster, cheaper and without rate limits. Compare the quality and latency to your current LLM API provider.

active 2024-02-132024-07-26 (UTC)

Partial coverage20,526 / 26,282 hourly files (78%) · 2 absent upstream · 5,753 failed, retryable2023-08-152026-08-14 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
47
Pushes
18
Pull requests
6
Issues
0
Stars
20
Forks
0

Activity over time

Daily event counts in the loaded window

Line chart, 165 days from 2024-02-13 to 2024-07-26. Pushes: 18 total, peak 9 in a day. Pull requests: 6 total, peak 6 in a day. Issues: 0 total, peak 0 in a day. Comments: 0 total, peak 0 in a day. Stars: 20 total, peak 7 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
Pierre-LouisBJT181440
oulianov6420

Recent activity

Latest issues, pull requests and releases

  • Pull request#3oulianov2024-02-17 02:49
  • Pull request#3oulianov2024-02-17 02:49
  • Pull request#2Pierre-LouisBJT2024-02-17 02:38
  • Pull request#2Pierre-LouisBJT2024-02-17 02:38
  • Pull request#1Pierre-LouisBJT2024-02-17 02:19
  • Pull request#1Pierre-LouisBJT2024-02-17 02:19

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 20 stars here means stars gained during the window, not the repo's star count.