7
stars
5
commits
4
linked in READMEs
Feb 27, 2026
updated
| Model | HuggingFace |
|---|---|
| MediX-R1-2B | MBZUAI/MediX-R1-2B |
| MediX-R1-8B | MBZUAI/MediX-R1-8B |
| MediX-R1-30B | MBZUAI/MediX-R1-30B |
If you use MediX-R1 in your research, please cite our work as follows:
@misc{mullappilly2026medixr1openendedmedical,
title={MediX-R1: Open Ended Medical Reinforcement Learning},
author={Sahal Shaji Mullappilly and Mohammed Irfan Kurpath and Omair Mohamed and Mohamed Zidan and Fahad Khan and Salman Khan and Rao Anwer and Hisham Cholakkal},
year={2026},
eprint={2602.23363},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2602.23363},
}
This project is released for research purposes only under CC-BY-NC-SA 4.0 License. It is not intended for clinical or commercial use.
Users are urged to employ MediX-R1 responsibly, especially when applying its outputs in real-world medical scenarios. It is imperative to verify the model's advice with qualified healthcare professionals and not rely on it for medical diagnoses or treatment decisions.
We are thankful to EasyR1 (a fork of veRL) for their open-source RL training framework.
This work was partially supported with NVIDIA Academic Grant 2025 and MBZUAI-IITD Research Collaboration Seed Grant.
We are grateful to MBZUAI for compute and support.
3 commits
2 commits
7
stars
5
commits
4
linked in READMEs
Feb 27, 2026
updated
| Model | HuggingFace |
|---|---|
| MediX-R1-2B | MBZUAI/MediX-R1-2B |
| MediX-R1-8B | MBZUAI/MediX-R1-8B |
| MediX-R1-30B | MBZUAI/MediX-R1-30B |
If you use MediX-R1 in your research, please cite our work as follows:
@misc{mullappilly2026medixr1openendedmedical,
title={MediX-R1: Open Ended Medical Reinforcement Learning},
author={Sahal Shaji Mullappilly and Mohammed Irfan Kurpath and Omair Mohamed and Mohamed Zidan and Fahad Khan and Salman Khan and Rao Anwer and Hisham Cholakkal},
year={2026},
eprint={2602.23363},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2602.23363},
}
This project is released for research purposes only under CC-BY-NC-SA 4.0 License. It is not intended for clinical or commercial use.
Users are urged to employ MediX-R1 responsibly, especially when applying its outputs in real-world medical scenarios. It is imperative to verify the model's advice with qualified healthcare professionals and not rely on it for medical diagnoses or treatment decisions.
We are thankful to EasyR1 (a fork of veRL) for their open-source RL training framework.
This work was partially supported with NVIDIA Academic Grant 2025 and MBZUAI-IITD Research Collaboration Seed Grant.
We are grateful to MBZUAI for compute and support.
3 commits
2 commits