Zhiqiang007/MathV360K

Dataset

Overview

35

6 commits

3 linked in READMEs

updated Jun 27, 2024

See the code

README

Overview

MathV360K is proposed by Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models, which consists 40K images from 24 datasets and 360K question-answer pairs. MathV360K is used to enhance the multimodal mathematical reasoning capabilities of MLLMs, achieving 46.6% accuracy on MathVista benchmark and 15.69% accuracy on MathVision dataset.

Paper or resources for more information: [Paper] [Code] [Model]

Source Data

source_data.jpg

Contributors

Zhiqiang007

6 commits

Zhiqiang007/MathV360K

Dataset

Overview

35

6 commits

3 linked in READMEs

updated Jun 27, 2024

See the code

README

Overview

MathV360K is proposed by Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models, which consists 40K images from 24 datasets and 360K question-answer pairs. MathV360K is used to enhance the multimodal mathematical reasoning capabilities of MLLMs, achieving 46.6% accuracy on MathVista benchmark and 15.69% accuracy on MathVision dataset.

Paper or resources for more information: [Paper] [Code] [Model]

Source Data

source_data.jpg

Contributors

Zhiqiang007

6 commits