Skip to content

Step-Audio 2 is an end-to-end multi-modal large language model designed for industry-strength audio understanding and speech conversation.

active 2025-07-252026-05-25 (UTC)

Partial coverage11,258 / 11,987 hourly files (94%) · 2 absent upstream · 726 failed, retryable2025-03-292026-08-10 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
987
Pushes
19
Pull requests
8
Issues
62
Stars
754
Forks
45

Activity over time

Daily event counts in the loaded window

Line chart, 307 days from 2025-07-23 to 2026-05-25. Pushes: 19 total, peak 7 in a day. Pull requests: 8 total, peak 2 in a day. Issues: 62 total, peak 7 in a day. Comments: 98 total, peak 11 in a day. Stars: 754 total, peak 115 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
petronny5713134
lilyzlt7001
yangxueruivs5104
SimZhou5005
weedge4021
runninging4004
dadafsfsdf4001
wulaoshi3001
JinmingChe3001
heheda1663002
e1ijah13003
FlyTia3003
vlordier3021
Qoboty3002
ranck6263002
Sidak082002
yuekaizhang2011
PressEtoRace2001
GallonDeng2001
levi-li-PG2002

Recent activity

Latest issues, pull requests and releases

  • Issue#86npathak132026-03-01 07:07
    Step-Audio-2-mini running on AMD RDNA 4 (gfx1151 / Strix Halo) via ROCm — build fixes and Dockerfile
  • Issue comment#84zh199909062026-02-22 02:22
    Add asynchronous support for audio processing
  • Issue#83nano-micro2026-01-22 03:07
    推理输出的时候只有文本没有语音
  • Issue#82LiXinYuECNU2026-01-02 06:03
    关于sft step-audio2系列模型的训练数据格式
  • Issue#81Pangkaiyuyu2025-12-24 10:24
    ASR LORA
  • Issue#80shanhaidexiamo2025-12-16 09:46
    Floating point exception (core dumped)
  • Issue#79jujunchen2025-12-05 07:05
    gradio 的版本是多少?
  • Issue#55ZhikangNiu2025-12-02 02:36
    Is the think process of the think model limited to Chinese only?
  • Issue comment#70yuekaizhang2025-11-21 06:57
    Add TensorRT Token2wav
  • Issue#76mingyi4562025-11-16 12:44
    Flash attention not supported?
  • Issue#75chenggangqcg2025-11-07 08:44
    模型输出:无限个嗯,嗯,嗯,嗯,嗯,嗯,嗯
  • Issue#71hashiting2025-11-03 08:56
    流式语音输出
  • Issue comment#71petronny2025-10-31 11:56
    流式语音输出
  • Issue comment#58SimZhou2025-10-22 10:44
    Performance issues
  • Issue comment#58SimZhou2025-10-22 10:32
    Performance issues
  • Issue comment#58SimZhou2025-10-22 10:04
    Performance issues
  • Issue comment#58petronny2025-10-22 09:58
    今天一天测试(对接了FreeSwitch上),试了 三种模型 step-1o-audio ,step-audio-2,step-audio-2-mini 都不行。
  • Issue comment#58SimZhou2025-10-22 09:51
    今天一天测试(对接了FreeSwitch上),试了 三种模型 step-1o-audio ,step-audio-2,step-audio-2-mini 都不行。
  • Issue comment#58petronny2025-10-22 09:42
    今天一天测试(对接了FreeSwitch上),试了 三种模型 step-1o-audio ,step-audio-2,step-audio-2-mini 都不行。
  • Issue comment#58SimZhou2025-10-22 09:06
    今天一天测试(对接了FreeSwitch上),试了 三种模型 step-1o-audio ,step-audio-2,step-audio-2-mini 都不行。
  • Pull request#70yuekaizhang2025-10-22 03:02
  • Issue#69lilyzlt2025-10-15 09:06
    How to Support Longer Audio in the vLLM Implementation?
  • Issue#67Hooyoung-for-AI2025-10-08 14:59
    微调代码请求 finetune code request
  • Issue#66wwfcnu2025-09-30 03:22
    vllm环境
  • Issue comment#63petronny2025-09-29 03:24
    vllm版本只支持API推理吗?

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 754 stars here means stars gained during the window, not the repo's star count.