AutoEvals is a tool for quickly and easily evaluating AI model outputs using best practices.
active 2025-03-17 → 2026-07-17 (UTC)
Activity over time
Daily event counts in the loaded window
Line chart, 488 days from 2025-03-17 to 2026-07-17. Pushes: 99 total, peak 15 in a day. Pull requests: 37 total, peak 9 in a day. Issues: 7 total, peak 1 in a day. Comments: 58 total, peak 9 in a day. Stars: 182 total, peak 4 in a day.
- Pushes
- Pull requests
- Issues
- Comments
- Stars
Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.
Top contributors
Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity
| Contributor | Contributions | Pushes | PRs | Comments |
|---|---|---|---|---|
| Qard | 58 | 34 | 12 | 7 |
| ibolmo | 43 | 20 | 7 | 6 |
| github-actions[bot] | 33 | 0 | 0 | 33 |
| cpinn | 22 | 15 | 2 | 3 |
| manugoyal | 11 | 4 | 7 | 0 |
| ankrgyl | 11 | 4 | 2 | 3 |
| delner | 6 | 6 | 0 | 0 |
| clutchski | 5 | 1 | 1 | 1 |
| AbhiPrasad | 4 | 3 | 0 | 0 |
| CLowbrow | 4 | 1 | 2 | 0 |
| graphite-app[bot] | 3 | 0 | 0 | 1 |
| choochootrain | 3 | 1 | 1 | 0 |
| erin2722 | 3 | 3 | 0 | 0 |
| CodingCanuck | 2 | 0 | 0 | 1 |
| alexr17 | 2 | 2 | 0 | 0 |
| justcodebruh | 2 | 1 | 1 | 0 |
| holdenmatt | 1 | 0 | 0 | 1 |
| zzzev | 1 | 0 | 0 | 0 |
| wong-codaio | 1 | 0 | 0 | 1 |
| davhao | 1 | 0 | 0 | 0 |
Recent activity
Latest issues, pull requests and releases
- Issue comment#191github-actions[bot]2026-06-08 21:41Move from legacy proxy to gateway
- Issue comment#192wong-codaio2026-06-02 13:00[Deps] Pin braintrust <0.13 to unbreak Python CI
- Issue comment#186RheagalFire2026-04-21 16:39feat: add [ LiteLLM AI Gateway ] for provider independence
- Issue comment#184github-actions[bot]2026-04-03 21:54ci(publish-js): align pnpm version with mise
- Issue comment#183github-actions[bot]2026-04-02 15:57chore: Publish python via trusted publishing and unify release process
- Issue comment#182github-actions[bot]2026-04-02 01:10Add pnpm enforcement and config
- Issue comment#182github-actions[bot]2026-04-01 01:06Add pnpm enforcement and config
- Pull request#181ibolmo2026-04-01 00:16
- Pull request#177ankrgyl2026-02-28 23:19
- Issue#176fizcogar2026-02-25 12:35ContextPrecision does not meet the specification
- Pull request#175CLowbrow2026-02-20 21:26
- Issue comment#173github-actions[bot]2026-02-18 23:28[WIP] Get autoevals to work with trace scoring
- Pull request#173CLowbrow2026-02-18 23:27
- Pull request#171cpinn2026-02-12 18:14
- Issue comment#168github-actions[bot]2026-01-29 22:53Thread injection in trace scorers (js)
- Pull request#168ankrgyl2026-01-29 22:53
- Issue comment#165github-actions[bot]2026-01-19 21:12Add reasoningEffort/reasoning_effort parameter support
- Issue#132Qard2026-01-19 21:11Add reasoning effort support
- Issue comment#168github-actions[bot]2026-01-18 00:50Thread injection in trace scorers (js)
- Pull request#165Qard2026-01-14 20:22
- Issue comment#162github-actions[bot]2026-01-14 03:37Allow setting embedding model in AnswerCorrectness
- Pull request#162Qard2026-01-14 03:37
- Issue comment#164github-actions[bot]2026-01-14 01:48Add models configuration object to init()
- Issue#101Qard2026-01-13 18:20Supported scores
- Issue comment#158github-actions[bot]2026-01-13 18:16Fix CJS dependency issue
Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 182 stars here means stars gained during the window, not the repo's star count.