OpenRLHF

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

10.1k
Stars
+1.7k
Gained
20.6%
Growth
Python
Language

🎭 Best For

🏷️ Topics & Ecosystem

large-language-models proximal-policy-optimization raylib reinforcement-learning reinforcement-learning-from-human-feedback transformers visual-language-models vllm

📊 Activity

Latest commit: 2026-10-05. Over the past 330 days, this repository gained 1.7k stars (+20.6% growth). Activity data is based on daily RepoPi snapshots of the GitHub repository.