Skip to content
@OpenRLHF

OpenRLHF

Open-sourced Reinforcment Learning from Human Feedback

Pinned Loading

  1. OpenRLHF OpenRLHF Public

    An Easy-to-use, Scalable and High-performance RLHF Framework based on Ray (PPO & GRPO & REINFORCE++ & vLLM & Ray & Dynamic Sampling & Async Agentic RL)

    Python 7.5k 734

  2. OpenRLHF-M OpenRLHF-M Public

    An Easy-to-use, Scalable and High-performance RLHF Framework designed for Multimodal Models.

    Python 138 7

  3. OpenRLHF-Docs OpenRLHF-Docs Public

    3 4

Repositories

Showing 3 of 3 repositories
  • OpenRLHF Public

    An Easy-to-use, Scalable and High-performance RLHF Framework based on Ray (PPO & GRPO & REINFORCE++ & vLLM & Ray & Dynamic Sampling & Async Agentic RL)

    OpenRLHF/OpenRLHF’s past year of commit activity
    Python 7,511 Apache-2.0 734 260 17 Updated Jul 28, 2025
  • OpenRLHF-Docs Public
    OpenRLHF/OpenRLHF-Docs’s past year of commit activity
    3 4 0 0 Updated Jul 27, 2025
  • OpenRLHF-M Public

    An Easy-to-use, Scalable and High-performance RLHF Framework designed for Multimodal Models.

    OpenRLHF/OpenRLHF-M’s past year of commit activity
    Python 138 Apache-2.0 7 7 1 Updated Apr 7, 2025

Top languages

Python

Most used topics

Loading…