Skip to content

Disaggregated serving system for Large Language Models (LLMs).

active 2024-04-262026-04-27 (UTC)

Partial coverage22,895 / 26,312 hourly files (87%) · 2 absent upstream · 3,416 failed, retryable2023-08-152026-08-15 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
811
Pushes
21
Pull requests
9
Issues
51
Stars
536
Forks
67

Activity over time

Daily event counts in the loaded window

Line chart, 732 days from 2024-04-26 to 2026-04-27. Pushes: 21 total, peak 4 in a day. Pull requests: 9 total, peak 2 in a day. Issues: 51 total, peak 3 in a day. Comments: 118 total, peak 19 in a day. Stars: 536 total, peak 6 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
interestingLSY316118
PKUFlyingPig287313
GindaChen12803
YLSnowy11006
KylinC9015
William12github7006
RobertLou6005
LordEdison6006
FuHaoTHU6003
TZHelloWorld5003
Dreamer-HIT4003
Liaukx4012
67lc4003
vhch3011
wangguanggg3001
ddqspace-xyz3003
gursimar3001
irasin3002
Toseic3011
hyuenmin-choi3002

Recent activity

Latest issues, pull requests and releases

  • Issue#66Sabiha12252025-06-27 04:31
    'cudaMemcpy(ith_context_req_token_index.ptr, ith_context_req_token_index_cpu, sizeof(int32_t) * (batch_size+1), cudaMemcpyHostToDevice)'
  • Issue comment#21YitaoYuan2025-05-21 06:14
    SwitfTransformer compilation fails with ambiguous conversion error at PyTorch 24.05 container.
  • Issue#65Qiu-Jianrong2025-04-28 07:28
    关于论文引用的小小问题
  • Issue comment#58Liaukx2025-04-24 06:16
    How to use DistServe with ray?
  • Issue comment#58Haoyanlong2025-04-24 06:06
    How to use DistServe with ray?
  • Issue#64Haoyanlong2025-04-23 09:25
    初始化的时候, 报错
  • Issue comment#63LordEdison2025-04-17 10:44
    the inference result sentence is null
  • Issue comment#59Mrxiangli2025-04-15 13:48
    How to handle preemption cases
  • Issue comment#50Dreamer-HIT2025-04-08 14:35
    How to independently measure the performance of the Prefill phase and the Decode phase?
  • Issue comment#63Dreamer-HIT2025-04-08 14:30
    the inference result sentence is null
  • Issue comment#63LordEdison2025-04-08 14:20
    the inference result sentence is null
  • Issue comment#63LordEdison2025-04-08 14:18
    the inference result sentence is null
  • Issue comment#59Dreamer-HIT2025-04-08 13:17
    How to handle preemption cases
  • Issue#63Dreamer-HIT2025-04-08 13:15
    the inference result sentence is null
  • Pull request#62interestingLSY2025-04-06 15:54
  • Issue comment#62interestingLSY2025-04-06 15:53
    Fix problem: use offline.py and Llama-2-7b-hf in a local directory
  • Issue comment#10sjlgaga2025-04-06 10:15
    decoder.embed_tokens.weight.pt not found
  • Pull request#62sjlgaga2025-04-06 10:13
  • Issue comment#40LordEdison2025-03-26 07:47
    模型推理结果混乱,怎么解决。
  • Issue comment#40LordEdison2025-03-17 12:49
    模型推理结果混乱,怎么解决。
  • Issue#61HarryWu992025-03-17 09:41
    How to guarantee that kv cache transmission finished
  • Issue comment#25LordEdison2025-03-09 17:52
    Swift transformers cmak build 一直循序
  • Pull request#60Liaukx2025-02-28 03:46
  • Issue comment#10Liaukx2025-02-28 02:45
    decoder.embed_tokens.weight.pt not found
  • Issue#59Mrxiangli2025-02-26 22:48
    How to handle preemption cases

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 536 stars here means stars gained during the window, not the repo's star count.