Skip to content

Training Medusa with QLoRA + FSDP

active 2024-01-142026-06-25 (UTC)

Complete coverage26,668 / 26,668 hourly files (100%) · 2 absent upstream2023-08-152026-08-30 (UTC)
Events
2.2K
Pushes
260
Pull requests
64
Issues
46
Stars
1.5K
Forks
194

Activity over time

Daily event counts in the loaded window

Line chart, 894 days from 2024-01-14 to 2026-06-25. Pushes: 260 total, peak 12 in a day. Pull requests: 64 total, peak 9 in a day. Issues: 46 total, peak 3 in a day. Comments: 133 total, peak 16 in a day. Stars: 1,454 total, peak 282 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Stars, PRs, issues and forks are under-captured in the later part of this window. GH Archive progressively stopped capturing non-push events during 2026 — −95% or worse by the end of the window. Every series here except Pushes fades for that reason, so a decline above reflects the archive, not this repository. Pushes stay reliable throughout, so read them, and the contributor counts derived from them, as the real signal. Data health has the measurements.

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
KeremTurgutlu20218486
warner-benjamin63331710
johnowhitaker62292010
hsb1995220020
austinvhuang16435
geronimi73150311
sanipanwala130013
iseesaw10009
griff46929900
jph007104
rationalism6003
bilalghanem6004
jeromeku6014
Xynonners5004
ehartford5003
chwenjun2255021
Pugio4003
catid4002
mrgohlke3001
deepankarsharma3011

Recent activity

Latest issues, pull requests and releases

  • Issue comment#76Ruiwen5052025-05-19 14:18
    Bitsandbytes error while run train.py
  • Issue comment#76matthewdouglas2025-05-13 19:57
    Bitsandbytes error while run train.py
  • Issue comment#76Ruiwen5052025-05-12 14:31
    Bitsandbytes error while run train.py
  • Issue comment#76matthewdouglas2025-05-10 00:42
    Bitsandbytes error while run train.py
  • Issue comment#76geronimi732025-05-08 20:05
    Bitsandbytes error while run train.py
  • Issue#76Ruiwen5052025-05-08 14:52
    Bitsandbytes error while run train.py
  • Issue#75yzhang1232025-03-25 15:32
    NotImplementedError: c10d::broadcast_: at
  • Issue#74alon-123452025-01-29 16:11
    Training using FSDP, qLoRa on multinode
  • Pull request#62geronimi732025-01-09 10:22
  • Pull request#71chwenjun2252024-12-27 06:15
  • Issue comment#73ghsama2024-12-13 11:06
    Out of memory while ConvertingTheStateDict - How to split across GPUs?
  • Issue comment#45mrgohlke2024-11-06 19:14
    nan when the input length is large
  • Issue#73mrgohlke2024-10-19 17:40
    Out of memory while ConvertingTheStateDict - How to split across GPUs?
  • Issue#73mrgohlke2024-10-19 17:20
    Out of memory while ConvertingTheStateDict - How to split across GPUs?
  • Pull request#72chrismrutherford2024-09-13 09:57
  • Pull request#71chwenjun2252024-08-31 16:37
  • Issue#70chwenjun2252024-08-31 16:14
    `Converting the State Dict.ipynb` - Runtine error because Unexpected keys
  • Issue comment#70chwenjun2252024-08-31 16:13
    `Converting the State Dict.ipynb` - Runtine error because Unexpected keys
  • Issue#70chwenjun2252024-08-31 15:23
    `Converting the State Dict.ipynb` - Runtine error because Unexpected keys
  • Issue#69asmith262024-08-31 15:05
    How to fine-tune a Vision Language Model (VLM)?
  • Issue#68BenjaminBossan2024-08-15 10:00
    DoRA training not taking dropout or alpha into account
  • Issue comment#60williambarberjr2024-07-13 22:59
    Request for Scripts to Merge QDoRA Adapters with Base Model for vLLM Inference
  • Issue comment#60lochuynh14122024-06-18 18:41
    Request for Scripts to Merge QDoRA Adapters with Base Model for vLLM Inference
  • Pull request#67austinvhuang2024-06-11 03:37
  • Issue comment#66jph002024-06-09 08:20
    Add profiling to train.py

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 1,454 stars here means stars gained during the window, not the repo's star count.