shiym2000/Med-2E3-M3D

Model

1

stars

24

commits

1

linked in READMEs

Feb 16, 2026

updated

conversational
image-text-to-text

README

Med-2E3-M3D

Introduction

A 3D medical LVLM, Med-2E3, trained on 3D CT volumes and English medical texts (M3D-Cap & M3D-VQA), enabling tasks such as report generation and medical VQA.

Config
3D Image encoderGoodBaiBai88/M3D-CLIP
2D Image encodergoogle/siglip-large-patch16-256
ConnectorTG-IS scoring module
LLMQwen/Qwen2.5-3B-Instruct
Image resolution32*256*256
Sequence length768

Quickstart

Please refer to Med-2E3.

Citation

@inproceedings{shi2025med,
  title={Med-2e3: A 2d-enhanced 3d medical multimodal large language model},
  author={Shi, Yiming and Zhu, Xun and Wang, Kaiwen and Hu, Ying and Guo, Chenyi and Li, Miao and Wu, Ji},
  booktitle={2025 IEEE International Conference on Bioinformatics and Biomedicine (BIBM)},
  pages={2754--2759},
  year={2025},
  organization={IEEE}
}

Contributors

shiym2000

24 commits

shiym2000/Med-2E3-M3D

Model

1

stars

24

commits

1

linked in READMEs

Feb 16, 2026

updated

conversational
image-text-to-text

README

Med-2E3-M3D

Introduction

A 3D medical LVLM, Med-2E3, trained on 3D CT volumes and English medical texts (M3D-Cap & M3D-VQA), enabling tasks such as report generation and medical VQA.

Config
3D Image encoderGoodBaiBai88/M3D-CLIP
2D Image encodergoogle/siglip-large-patch16-256
ConnectorTG-IS scoring module
LLMQwen/Qwen2.5-3B-Instruct
Image resolution32*256*256
Sequence length768

Quickstart

Please refer to Med-2E3.

Citation

@inproceedings{shi2025med,
  title={Med-2e3: A 2d-enhanced 3d medical multimodal large language model},
  author={Shi, Yiming and Zhu, Xun and Wang, Kaiwen and Hu, Ying and Guo, Chenyi and Li, Miao and Wu, Ji},
  booktitle={2025 IEEE International Conference on Bioinformatics and Biomedicine (BIBM)},
  pages={2754--2759},
  year={2025},
  organization={IEEE}
}

Contributors

shiym2000

24 commits