Skip to content

bigcode-project/bigcode-evaluation-harness

View on GitHub ↗Related repositories →

Adjusting the big code harness to serve the specialized need of properly training and evaluating quantized PEFT methods

active 2024-11-242026-05-15 (UTC)

Partial coverage12,915 / 15,023 hourly files (86%) · 2 absent upstream · 2,102 failed, retryable2024-11-232026-08-11 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
144
Pushes
1
Pull requests
11
Issues
5
Stars
95
Forks
27

Activity over time

Daily event counts in the loaded window

Line chart, 538 days from 2024-11-24 to 2026-05-15. Pushes: 1 total, peak 1 in a day. Pull requests: 11 total, peak 2 in a day. Issues: 5 total, peak 1 in a day. Comments: 3 total, peak 1 in a day. Stars: 95 total, peak 4 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
arjunguha4120
gameofby2020
zawedcvg2020
dsocek1001
yxchng1000
ggcr1010
neehar181010
TK-21st1010
pengzhangzhi1000
showlibia1000
3542466951000
akashgokul1001
ahmedashrafy1010
seldereyy1001
sad-mathematician1010

Recent activity

Latest issues, pull requests and releases

  • Issue#3153542466952025-09-09 09:10
    Wrong Hugging Face link for Spider dataset in bigcode-evaluation-harness/docs/README.md line 434
  • Issue#311showlibia2025-08-17 12:12
    Improve pass@1 Score on Humaneval
  • Pull request#314arjunguha2025-07-15 10:45
  • Issue#224arjunguha2025-07-01 09:16
    Multiple-E Go test file name suffix does not contain _test.go
  • Pull request#225arjunguha2025-07-01 09:16
  • Issue comment#224seldereyy2025-06-30 21:22
    Multiple-E Go test file name suffix does not contain _test.go
  • Pull request#310sad-mathematician2025-04-20 19:56
  • Pull request#309zawedcvg2025-04-02 03:38
  • Pull request#309zawedcvg2025-04-02 03:38
  • Issue#307pengzhangzhi2025-03-26 01:43
    how to add new model?
  • Issue comment#131akashgokul2025-02-13 06:09
    'HumanEval' object has no attribute 'dataset'
  • Pull request#301TK-21st2025-01-25 15:11
  • Issue#300yxchng2025-01-21 07:57
    is there a benchmark page on the benchmark results evaluated using bigcode-evaluation-harness
  • Pull request#298ggcr2025-01-11 23:28
  • Pull request#296neehar182024-12-17 22:35
  • Pull request#292ahmedashrafy2024-12-08 23:25
  • Issue comment#281dsocek2024-12-04 23:57
    add support for hpu devices
  • Pull request#291gameofby2024-12-04 02:32
  • Pull request#286gameofby2024-12-04 01:42

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 95 stars here means stars gained during the window, not the repo's star count.