Skip to content

多模态推理模型及训练方法

active 2025-05-212026-06-14 (UTC)

Partial coverage11,815 / 13,165 hourly files (90%) · 2 absent upstream · 1,346 failed, retryable2025-02-082026-08-10 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
954
Pushes
13
Pull requests
25
Issues
62
Stars
607
Forks
31

Activity over time

Daily event counts in the loaded window

Line chart, 390 days from 2025-05-21 to 2026-06-14. Pushes: 13 total, peak 3 in a day. Pull requests: 25 total, peak 5 in a day. Issues: 62 total, peak 3 in a day. Comments: 193 total, peak 9 in a day. Stars: 607 total, peak 36 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
JaaackHongggg433139
dependabot[bot]2801711
ChenShawn25968
lky-violet11009
w-zhih10005
mengzchen8005
Alisaleelee6003
dddraxxx6005
RenlyH5005
wcpcp5005
FolSpark4004
zapqqqwe4002
EugeneLiu014004
lbj00jay4003
Daisy-Zhang3002
Copilot3003
qinguangming19993003
FloSophorae3002
sfc-gh-bzhai3001
chunyu-li3003

Recent activity

Latest issues, pull requests and releases

  • Issue#137liuchaohu2026-02-09 11:16
    The evaluation differences of the Vstar dataset
  • Issue comment#31Michael49332026-01-27 12:11
    我得到了一个很奇怪的答案,这个答案合理么?
  • Issue#136Tristan-Liu-132026-01-21 13:12
    请问下huggingface的模型怎么输出思维链呢?
  • Issue comment#122Liac-li2025-12-04 01:27
    Qwen2.5VL 3B training with lora OOM on H100
  • Issue comment#117JaaackHongggg2025-11-25 14:00
    模型推理慢
  • Issue comment#103JaaackHongggg2025-11-25 13:48
    Model Response Score
  • Issue comment#99JaaackHongggg2025-11-25 13:44
    About the training data
  • Issue comment#80JaaackHongggg2025-11-25 13:32
    Does the Model Memorize Answers on the V* Benchmark?
  • Issue#133snow-like-kk2025-11-13 13:15
    How can I remove flash-attn?
  • Issue comment#131ckx5067720992025-11-03 08:25
    Using DeepEyes checkpoint fine-tuning, many <im_start> are repeatedly output when rollout.
  • Issue#130Cyyyyyyyyyyyyyyyyyyyyyyyyyyyyyyyy2025-10-31 07:50
    how to deploy to vllm and the deepspeed version unknown
  • Issue#125pspdada2025-10-07 10:17
    Cannot find VISUAL_DATASET_TRAIN_0_8 file
  • Issue comment#16pspdada2025-09-30 12:11
    request for evalution/inference code
  • Issue comment#103Jam1ezhang2025-09-27 08:52
    Model Response Score
  • Issue#123yuyi03122025-09-19 07:23
    How to get the tool_call curve if wandb is offline?
  • Issue comment#113chunyu-li2025-09-12 10:17
    ray cpu memory OOM/cpu memory leak
  • Issue comment#113chunyu-li2025-09-11 12:45
    ray cpu memory OOM/cpu memory leak
  • Issue comment#113woshitff2025-09-11 12:23
    ray cpu memory OOM/cpu memory leak
  • Issue comment#113chunyu-li2025-09-11 12:14
    ray cpu memory OOM/cpu memory leak
  • Issue#122Liac-li2025-09-11 05:25
    Qwen2.5VL 3B training with lora OOM on H100
  • Issue comment#99dearjohn122025-09-09 06:50
    About the training data
  • Issue comment#102lbj00jay2025-09-09 06:34
    cannot import name 'process_image' from 'verl.utils.dataset.rl_dataset'
  • Issue comment#78wcpcp2025-09-08 07:51
    感觉论文偏假,根本复现不出来,工具调用是0
  • Issue comment#70lbj00jay2025-09-08 02:06
    How do I use these saved models?
  • Issue comment#113woshitff2025-09-05 13:11
    ray cpu memory OOM/cpu memory leak

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 607 stars here means stars gained during the window, not the repo's star count.