19 repos
Libraries, frameworks, and research implementations for training language models and vision-language agents using reinforcement learning techniques, with emphasis on reasoning tasks and R1-style approaches. This cluster covers RL training pipelines, reasoning-focused model development, and vision-language action systems that leverage policy optimization and reward modeling. Repositories include foundational RL training frameworks, reasoning-specific implementations, and practical applications of RL to multimodal agent tasks.