Skip to content
AI Landscape

About

Library for training transformer LMs with RLHF, SFT, reward modeling, and PPO

See something off? Suggest an edit

Related in Fine-tuning & RLHF