Skip to content

EMNLP MAIN 2025 StepSearch: Igniting LLMs Search Ability via Step-Wise Proximal Policy Optimization

active 2025-05-232026-04-02 (UTC)

Complete coverage26,466 / 26,466 hourly files (100%) · 2 absent upstream2023-08-152026-08-21 (UTC)
Events
71
Pushes
11
Pull requests
0
Issues
9
Stars
30
Forks
4

Activity over time

Daily event counts in the loaded window

Line chart, 315 days from 2025-05-23 to 2026-04-02. Pushes: 11 total, peak 6 in a day. Pull requests: 0 total, peak 0 in a day. Issues: 9 total, peak 3 in a day. Comments: 15 total, peak 7 in a day. Stars: 30 total, peak 4 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
zxh20001117207010
Zillwang4400
CarnegieBin4004
wxt6251000
sunhaonlp1000
boardman01000
ZijunSong1000
dawson-chen1000
luxuriance191001

Recent activity

Latest issues, pull requests and releases

  • Issue comment#7zxh200011172025-12-05 09:30
    Question about the “support_docs” field in data processing for step reward
  • Issue#9wxt6252025-11-06 16:33
    Question about retrive
  • Issue#7ZijunSong2025-10-30 03:29
    Question about the “support_docs” field in data processing for step reward
  • Issue comment#6luxuriance192025-09-28 08:12
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6CarnegieBin2025-09-27 13:44
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6CarnegieBin2025-09-27 07:21
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6zxh200011172025-09-26 09:52
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6zxh200011172025-09-26 09:13
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6CarnegieBin2025-09-26 08:45
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6zxh200011172025-09-26 08:45
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6CarnegieBin2025-09-26 08:21
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6zxh200011172025-09-26 04:50
    Why are your model’s test results even weaker than Search-R1?
  • Issue comment#6zxh200011172025-09-26 04:42
    Why are your model’s test results even weaker than Search-R1?
  • Issue#5zxh200011172025-09-22 17:48
    qusetion of train.sh
  • Issue#4zxh200011172025-09-22 17:48
    论文里的r_overall是不是少写了一个部分?
  • Issue#3zxh200011172025-09-22 17:48
    Is the dataset prepared for open-sourcing?
  • Issue comment#5zxh200011172025-09-22 17:47
    qusetion of train.sh
  • Issue#5boardman02025-09-20 03:24
    qusetion of train.sh
  • Issue comment#4zxh200011172025-09-11 13:38
    论文里的r_overall是不是少写了一个部分?
  • Issue#4Victoriaheiheihei2025-07-25 03:22
    论文里的r_overall是不是少写了一个部分?
  • Issue#2sunhaonlp2025-06-19 08:41
    Concerns regarding ZeroSearch Reproduction Setup
  • Issue comment#2zxh200011172025-06-18 08:59
    Concerns regarding ZeroSearch Reproduction Setup
  • Issue comment#1zxh200011172025-05-29 09:34
    关于图5a中的一点疑问
  • Issue#1dawson-chen2025-05-27 06:12
    关于图5a中的一点疑问

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 30 stars here means stars gained during the window, not the repo's star count.