← Explore
TOPIC

#reinforcement-learning-from-human-feedback

Open source repositories tagged with #reinforcement-learning-from-human-feedback, ranked by health score.

OpenRLHF
OpenRLHF/OpenRLHF
Python
88
health

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

10.0k