# OpenRLHF

> Scalable RLHF framework supporting 70B+ PPO full tuning, iterative DPO, and LoRA

- Category: [MLOps & Training](https://ailandscape.org/category/mlops-training) › Fine-tuning & RLHF
- Homepage: https://openrlhf.readthedocs.io/
- Repository: https://github.com/OpenRLHF/OpenRLHF
- Added to the landscape: 2026-03-18

## Similar tools in Fine-tuning & RLHF

- [Axolotl](https://ailandscape.org/tool/axolotl): Fine-tuning framework for LLMs supporting LoRA, QLoRA, DeepSpeed, and multi-GPU
- [TRL](https://ailandscape.org/tool/trl): Library for training transformer LMs with RLHF, SFT, reward modeling, and PPO
- [unsloth](https://ailandscape.org/tool/unsloth): 2-5x faster LLM fine-tuning with 70% less memory

---

Part of [AI Landscape](https://ailandscape.org), an open map of the AI ecosystem. Web page: https://ailandscape.org/tool/openrlhf · Index for AI assistants: https://ailandscape.org/llms.txt
