* 🐙 GitHub Repo: waltonfuture/MM-UPT
0
5 commits
1 linked in READMEs
updated Jun 4, 2025
4 commits
1 commits
WaltonFuture/GeoQA-8K-direct-synthesizing
This dataset supports the unsupervised post-training of multi-modal large language models (MLLMs)…
1
WaltonFuture/geometry3k-in-context-synthesizing
This dataset is used for unsupervised post-training of multi-modal large language models (MLLMs).…
2
WaltonFuture/geometry3k-direct-synthesizing
This dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO…
WaltonFuture/MMR1-direct-synthesizing
WaltonFuture/MMR1-in-context-synthesizing
This dataset is designed for unsupervised post-training of Multi-Modal Large Language Models…
WaltonFuture/GEOQA_R1V_Train_8K
leonardPKU/GEOQA_R1V_Train_8K
Processed from [Geo170K](https://huggingface.co/datasets/Luckyjhg/Geo170K), we filtered out the…
14
Ricky06662/VisionReasoner_multi_object_1k_840
This dataset is used in [VisionReasoner: Unified Visual Perception and Reasoning via Reinforcement…
3
waltonfuture/MM-UPT
[NeurIPS 2025] First SFT, Second RL, Third UPT: Continual Improving Multi-Modal LLM Reasoning via…
89