← Developers
Open on GitHub 


#2442 8
xiaoxigua
@xiaoxigua999
9.95k
Weighted score
1.30k
Contributions
3
Repos
Top repos
Ranked AI repos3
3691810
OpenRLHF/OpenRLHF
+1
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
10.1k· Python· Model Development
12743482
TideDra/lmm-r1
+0
Extend OpenRLHF to support LMM RL training for reproduction of DeepSeek-R1 on multimodal tasks.
847· Python· Model Development
58822117
TsinghuaC3I/MARTI
+1
[ICLR 2026] A Framework for LLM-based Multi-Agent Reinforced Training and Inference
567· Python· Model Development