Here we provide a pure-text version of GameQA, encompassing some appropriate games. (See https://github.com/tongjingqi/Code2Logic/issues/2)
This is the first work, to the best of our knowledge, that leverages game code to synthesize multimodal reasoning data for training VLMs. Furthermore, when trained with a GRPO strategy solely on GameQA (synthesized via our proposed Code2Logic approach), multiple cutting-edge open-source models exhibit significantly enhanced out-of-domain generalization.
[π Paper] [π» Code] [π€ GameQA-140K Dataset] [π€ GameQA-InternVL3-8B ] [π€ GameQA-Qwen2.5-VL-7B] [π€ GameQA-LLaVA-OV-7B ]

5 commits
Here we provide a pure-text version of GameQA, encompassing some appropriate games. (See https://github.com/tongjingqi/Code2Logic/issues/2)
This is the first work, to the best of our knowledge, that leverages game code to synthesize multimodal reasoning data for training VLMs. Furthermore, when trained with a GRPO strategy solely on GameQA (synthesized via our proposed Code2Logic approach), multiple cutting-edge open-source models exhibit significantly enhanced out-of-domain generalization.
[π Paper] [π» Code] [π€ GameQA-140K Dataset] [π€ GameQA-InternVL3-8B ] [π€ GameQA-Qwen2.5-VL-7B] [π€ GameQA-LLaVA-OV-7B ]

5 commits