4 repos
GAD-cell/vlm-grpo
An implementation of GRPO for Unsloth's VLMs training
84
98 commits
cudnah124/ChartQA
Vision-language AI for chart question answering using Qwen3-VL with SFT and GRPO training
6
16 commits
EvolvingLMMs-Lab/ParaVT
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning
58
5 commits
sanchit97/chartrl
ChartRL - Improving Chart understanding through GRPO and RLVR
7
53 commits