OpenRLHF
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
🎭 Best For
🏷️ Topics & Ecosystem
large-language-models
proximal-policy-optimization
raylib
reinforcement-learning
reinforcement-learning-from-human-feedback
transformers
visual-language-models
vllm
📊 Activity
Latest commit: 2026-08-13. Over the past 285 days, this repository gained 1.6k stars (+19.1% growth). Activity data is based on daily RepoPi snapshots of the GitHub repository.