Skip to content

Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"

Python · active 2025-05-282026-05-30 (UTC)

Complete coverage26,631 / 26,631 hourly files (100%) · 2 absent upstream2023-08-152026-08-28 (UTC)
Events
605
Pushes
45
Pull requests
10
Issues
34
Stars
417
Forks
53

Activity over time

Daily event counts in the loaded window

Line chart, 368 days from 2025-05-28 to 2026-05-30. Pushes: 45 total, peak 4 in a day. Pull requests: 10 total, peak 1 in a day. Issues: 34 total, peak 3 in a day. Comments: 40 total, peak 3 in a day. Stars: 417 total, peak 25 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

Recent activity

Latest issues, pull requests and releases

  • Pull request#77hills-code2026-05-27 03:27
  • Issue comment#62Tomingz2026-04-24 09:31
    [v2] Reproducibility: release paper training script, preprocessing, and exact hyperparameters
  • Pull request#74SensenGao2026-04-17 01:38
  • Issue#64Laurence-Wu2026-01-16 00:49
    batch size > 1 and temperature != 0.0 not supported by the generate function file
  • Issue#55zhengli972026-01-05 07:19
    Error when changing the block size to 32
  • Issue#62yeonjoon-jung012026-01-05 06:37
    [v2] Reproducibility: release paper training script, preprocessing, and exact hyperparameters
  • Issue#61lizhuo0082025-12-30 15:56
    Why does Dream’s cache “global update” always update from current_block_start instead of using a sampling strategy?
  • Issue comment#59yongtang20252025-12-30 08:49
    请问为什么在eval_gsm8k.sh中,dual cache+paralle的steps参数是传入length,而不是steps
  • Issue comment#54littlestone1112025-12-29 16:22
    Question Regarding the Llada Parallel decoding in GSM8K
  • Issue#59zqc2142025-12-29 05:02
    请问为什么在eval_gsm8k.sh中,dual cache+paralle的steps参数是传入length,而不是steps
  • Issue comment#55zhengli972025-12-26 07:08
    Error when changing the block size to 32
  • Issue comment#55hills-code2025-12-25 13:59
    Error when changing the block size to 32
  • Issue comment#54hills-code2025-12-25 12:51
    Question Regarding the Llada Parallel decoding in GSM8K
  • Issue comment#23AndyJi12025-12-25 09:07
    Does the evaluation of MBPP require postprocessing, as in HumanEval?
  • Issue comment#53Agrim-Jain2025-12-19 02:47
    Unable to reproduce results for HumanEval dataset
  • Issue#53Agrim-Jain2025-12-18 14:55
    Unable to reproduce results for HumanEval dataset
  • Issue#52quaternior2025-12-11 05:24
    About warm-up stage and cache reuse stage in dual cache+parallel
  • Pull request#51NickCheng09212025-12-10 06:51
  • Issue#37hills-code2025-11-28 09:08
    Error when using the README code
  • Issue#42hills-code2025-11-28 09:07
    parralel_question
  • Issue comment#34hills-code2025-11-28 09:05
    fast-dllm-v2 mbpp and humaneval evaluation
  • Issue#38hills-code2025-11-28 09:04
    Bug: Changing threshold in Fast-dLLM v2 MMLU Inference Has No Effect
  • Issue comment#38hills-code2025-11-28 09:04
    Bug: Changing threshold in Fast-dLLM v2 MMLU Inference Has No Effect
  • Issue comment#36lihe072025-11-27 09:14
    fast-dllm-v2中,dual cache的实际作用?
  • Issue comment#20hills-code2025-11-25 13:11
    Dream parallel decoding implementation

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 417 stars here means stars gained during the window, not the repo's star count.