34 repos
Reinforcement learning from human feedback (RLHF) systems and reward model development for large language models. The cluster centers on training and evaluating reward models that score LLM outputs, a critical component of aligning language models with human preferences. While the most connected repos focus on Beaver reward model variants across different versions and cost optimization, the broader cluster encompasses transformer-based models, reinforcement learning infrastructure, and vLLM serving tooling needed to build end-to-end RLHF pipelines.