Skip to content

This repo contains fork designed for Personal.ai Agent Training. Its the evaluation code for the paper "MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI"

active 2023-11-282026-06-01 (UTC)

Complete coverage26,350 / 26,350 hourly files (100%) · 2 absent upstream2023-08-152026-08-16 (UTC)
Events
884
Pushes
81
Pull requests
37
Issues
109
Stars
450
Forks
47

Activity over time

Daily event counts in the loaded window

Line chart, 917 days from 2023-11-28 to 2026-06-01. Pushes: 81 total, peak 4 in a day. Pull requests: 37 total, peak 8 in a day. Issues: 109 total, peak 9 in a day. Comments: 152 total, peak 6 in a day. Stars: 450 total, peak 34 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
NipElement9457622
xiangyue96077211923
drogozhang6312036
Zheng04283211714
insafim20009
Qzy5686003
teasgen5003
Rubics-Xuan5005
Makefunof5003
wvangansbeke5004
XiongweiWu5001
mckinziebrandon5001
SweetGUOguo3001
Xiaolong-RRL3001
fxmeng3001
lucasmgomez3001
shannany06063001
ycszen3012
dchichkov2001
beichenzbc2000

Recent activity

Latest issues, pull requests and releases

  • Issue comment#82NipElement2026-05-10 19:52
    Question about leaderboard settings
  • Issue#82llsj142026-05-10 11:00
    Question about leaderboard settings
  • Issue#80NipElement2026-04-07 19:53
    dataset overlap
  • Issue comment#80manglu0972026-04-07 14:10
    dataset overlap
  • Issue comment#80NipElement2026-04-06 22:02
    dataset overlap
  • Issue#79NipElement2026-02-12 10:00
    EvalAI提交评测
  • Issue comment#79NipElement2026-02-11 10:25
    EvalAI提交评测
  • Issue comment#78NipElement2026-01-23 06:34
    EvalAI leaderboard提交后一直在submited
  • Issue#78drogozhang2026-01-21 03:53
    EvalAI leaderboard提交后一直在submited
  • Issue#77drogozhang2025-09-26 17:42
    Need help evaluating...
  • Issue#77blazgocompany2025-07-25 13:55
    Need help evaluating...
  • Issue comment#76drogozhang2025-05-19 16:12
    Batch 支持
  • Issue comment#76remember000002025-05-19 15:37
    Batch 支持
  • Issue comment#74drogozhang2025-05-19 14:56
    Key error
  • Issue#74drogozhang2025-05-19 14:56
    Key error
  • Issue comment#76drogozhang2025-05-19 14:55
    Batch 支持
  • Issue#76drogozhang2025-05-19 14:55
    Batch 支持
  • Issue#76remember000002025-05-19 08:25
    Batch 支持
  • Issue comment#75drogozhang2025-05-12 00:37
    Does MMU support scoring based on image categories?
  • Issue#75drogozhang2025-05-12 00:37
    Does MMU support scoring based on image categories?
  • Issue comment#75drogozhang2025-05-06 20:28
    Does MMU support scoring based on image categories?
  • Issue comment#74drogozhang2025-05-06 20:26
    Key error
  • Issue#75Blissy-322025-05-06 07:18
    Does MMU support scoring based on image categories?
  • Issue#74manglu0972025-05-02 09:36
    Key error
  • Issue comment#73Zheng04282025-04-15 03:12
    🤗 Transformers multimodal models eval on MMMU_Pro

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 450 stars here means stars gained during the window, not the repo's star count.