OpenRLHF

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

9.9k
Stars
+1.6k
Gained
19.1%
Growth
Python
Language

🎭 Best For

🏷️ Topics & Ecosystem

large-language-models proximal-policy-optimization raylib reinforcement-learning reinforcement-learning-from-human-feedback transformers visual-language-models vllm

📊 Activity

Latest commit: 2026-08-13. Over the past 285 days, this repository gained 1.6k stars (+19.1% growth). Activity data is based on daily RepoPi snapshots of the GitHub repository.