O
OpenRLHF
OpenRLHF/OpenRLHF
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
★9.9kstars
Python
Apache-2.0
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A , built with Python open-source project in the Data & Infrastructure category, core strengths: large-language-models/proximal-policy-optimization
Who made it?Maintained by OpenRLHF team, 9.9K⭐ on GitHub, #202 out of 3133 in Data & Infrastructure
Why does it exist?The OpenRLHF team recognized that existing large-language-models tools in Data & Infrastructure were hard to use. OpenRLHF was designed to make proximal-policy-optimization more accessible.
What can it do?Key use cases: raylib, reinforcement-learning, reinforcement-learning-from-human-feedback
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/OpenRLHF/OpenRLHF | 官网 https://openrlhf.readthedocs.io/
🔗 github.com/OpenRLHF/OpenRLHF | 官网 https://openrlhf.readthedocs.io/
Topics
large-language-modelsproximal-policy-optimizationraylibreinforcement-learningreinforcement-learning-from-human-feedbacktransformersvisual-language-modelsvllm