Skip to content

An Easy-to-use, Scalable and High-performance RLHF Framework (Support 70B+ full tuning & LoRA & Mixtral & KTO)

Python · active 2023-11-062024-07-17 (UTC)

Partial coverage18,857 / 24,253 hourly files (78%) · 2 absent upstream · 5,393 failed, retryable2023-11-062026-08-12 (UTC)— sampled evenly across the window, so rankings and trends hold; absolute counts scale up.
Events
2.4K
Pushes
483
Pull requests
103
Issues
167
Stars
953
Forks
83

Activity over time

Daily event counts in the loaded window

Line chart, 255 days from 2023-11-06 to 2024-07-17. Pushes: 483 total, peak 33 in a day. Pull requests: 103 total, peak 4 in a day. Issues: 167 total, peak 6 in a day. Comments: 440 total, peak 36 in a day. Stars: 953 total, peak 40 in a day.

  • Pushes
  • Pull requests
  • Issues
  • Comments
  • Stars

Top contributors

Pushes, PRs, issues, reviews and comments — stars and forks excluded, so this is contribution rather than popularity

ContributorContributionsPushesPRsComments
hijkzzz78046861191
wuxibin898110950
catqaq345513
karthik-nexusflow220019
paulcx200015
karthik19967829170014
mickelliu160410
ZiyiLiubird12005
mgerstgrasser12017
victorShawFan11006
louieworth9007
hehebamei9006
zhanghaoie8006
eyuansu628005
LSC5278004
NZ997007
mickel-liu7004
yangzhipeng11087002
tianhao-nexusflow7006
ZhaofengWu7004

Recent activity

Latest issues, pull requests and releases

  • Issue comment#361hijkzzz2024-07-16 02:47
    support remote rm api for ppo and ppo ray
  • Issue comment#361catqaq2024-07-16 02:36
    support remote rm api for ppo and ppo ray
  • Releasehijkzzz2024-07-16 01:46
    Release v0.3.6
  • Issue comment#361hijkzzz2024-07-15 23:47
    support remote rm api for ppo and ppo ray
  • Issue#359hijkzzz2024-07-15 00:10
    Possible minor bug
  • Issue#359ZhaofengWu2024-07-14 22:34
    Possible minor bug
  • Issue#263hijkzzz2024-07-14 09:14
    [Baseline] LLaMA2-7B RLHF training curves
  • Releasehijkzzz2024-07-13 23:26
    Release v0.3.5
  • Issue comment#358hijkzzz2024-07-13 23:15
    vllm engine not working
  • Issue comment#358babu1112024-07-13 13:50
    vllm engine not working
  • Issue#358babu1112024-07-13 13:49
    vllm engine not working
  • Issue#357babu1112024-07-13 13:47
    process groups for actor and vllm engine
  • Issue comment#357hijkzzz2024-07-12 11:58
    process groups for actor and vllm engine
  • Issue#357babu1112024-07-12 10:07
    process groups for actor and vllm engine
  • Issue comment#354hijkzzz2024-07-12 07:49
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#354hehebamei2024-07-12 06:41
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#354hehebamei2024-07-12 06:41
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#354hijkzzz2024-07-12 05:32
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#354hehebamei2024-07-12 04:24
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#354hijkzzz2024-07-12 03:38
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#355hijkzzz2024-07-12 03:37
    load transformers' issue
  • Issue#355chauncygu2024-07-12 03:36
    load transformers' issue
  • Issue#354hehebamei2024-07-12 02:52
    请问下rm 模型训练大概需要什么级别的显卡,需要几张?
  • Issue comment#353hijkzzz2024-07-11 23:13
    会不会支持异步生成训练
  • Issue#351hijkzzz2024-07-10 23:59
    Qwen2-1.5b模型做sft微调报错padding_side='right'

Totals cover only the window loaded into ClickHouse and count events, not GitHub's lifetime totals — 953 stars here means stars gained during the window, not the repo's star count.