Back to List
Notice:This resource is provided by a third-party author. Please review the code with AI tools or manually before use to ensure security and compatibility.
PythonOpenRLHF/OpenRLHF

OpenRLHF

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & TIS & vLLM & Ray & Async RL)

89.2/100
9.1KForks: 888
View on GitHubHomepage →
Loading report...

Similar Projects

safe-rlhf

53

Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback

Python1.6K

ml-engineering

74

Machine Learning Engineering Open Book

Python17.3K

LLM-RL-Visualized

70

🌟100+ 原创 LLM / RL 原理图📚,《大模型算法》作者巨献!💥(100+ LLM/RL Algorithm Maps )

Python3.7K

llm-guard

60

The Security Toolkit for LLM Interactions

Python2.6K
Back to List