Skip to content

alipay/PainlessInferenceAcceleration

View on GitHub ↗Related repositories →

Retrieval LLM

active 2023-12-192026-02-26 (UTC)

Complete coverage26,602 / 26,602 hourly files (100%) · 2 absent upstream2023-08-152026-08-27 (UTC)
Events
575
Pushes
39
Pull requests
6
Issues
61
Stars
355
Forks
21

Activity over time

Daily event counts in the loaded window

Line chart, 801 days from 2023-12-19 to 2026-02-26. Pushes: 39 total, peak 6 in a day. Pull requests: 6 total, peak 2 in a day. Issues: 61 total, peak 10 in a day. Comments: 80 total, peak 11 in a day. Stars: 355 total, peak 39 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
zheyishine6829226
chenliangjyj3110015
jivanph9006
nrmer7000
May-Yaha6004
AGI-Jarvis6004
shinerdeng5003
ZipECHO5005
20650374214000
Vithulep4001
learning-chip4003
janelu94002
xiningnlp4001
snippetzero2001
dafen122001
yuenyu12000
gigigigig12001
MeJerry2152001
Mewo5182020
bhpugongying1000

Recent activity

Latest issues, pull requests and releases

  • Issue comment#45zlH5182025-11-08 17:13
    fix: add cleanup method for clean the processes
  • Pull request#40hjyai942025-09-28 07:30
  • Issue comment#39zheyishine2025-04-23 06:17
    Looking forward to the flood paper since offline is sort of being ignored currently, when will u post it?
  • Issue#39cvtower2025-04-21 03:08
    Looking forward to the flood paper since offline is sort of being ignored currently, when will u post it?
  • Pull request#38Mewo5182025-04-17 05:14
  • Pull request#38Mewo5182025-04-17 05:08
  • Issue#37bhpugongying2025-04-16 03:39
    您好,请问lookahead现在支持Qwen2.5吗?
  • Issue comment#36zheyishine2025-04-01 11:31
    I'm unable to replicate the experimental results presented in the paper.
  • Issue comment#36gigigigig12025-03-28 06:42
    I'm unable to replicate the experimental results presented in the paper.
  • Issue#36gigigigig12025-03-28 02:55
    I'm unable to replicate the experimental results presented in the paper.
  • Issue#3520650374212025-03-15 10:47
    triton.runtime.errors.OutOfResources
  • Issue#3420650374212025-03-15 10:47
    'LLM' object has no attribute 'stream_generate'
  • Issue comment#35zheyishine2025-03-14 07:22
    triton.runtime.errors.OutOfResources
  • Issue#3520650374212025-03-12 11:53
    triton.runtime.errors.OutOfResources
  • Issue#3420650374212025-03-12 07:59
    'LLM' object has no attribute 'stream_generate'
  • Pull request#33zheyishine2025-03-08 09:48
  • Issue comment#32Vithulep2025-03-07 09:02
    Output of llma-2-7b model is repetative.
  • Issue#9zheyishine2025-03-06 13:10
    In the benchmark studies, how are the draft tokens generated?
  • Issue comment#9zheyishine2025-03-06 13:10
    In the benchmark studies, how are the draft tokens generated?
  • Issue#22zheyishine2025-03-06 13:09
    Changing naive attention to SDPA gives wrong result for batched llama example
  • Issue#13zheyishine2025-03-06 13:09
    Consider Support CodeLlama?
  • Issue#23zheyishine2025-03-06 13:08
    AntRAG
  • Issue comment#20zheyishine2025-03-06 13:08
    是否支持Qwen 1.5?
  • Issue#20zheyishine2025-03-06 13:08
    是否支持Qwen 1.5?
  • Issue#24zheyishine2025-03-06 13:07
    Do lookahead and repetition_penalty conflict?

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 355 stars here means stars gained during the window, not the repo's star count.