MBZUAI/medix-rl-data

Dataset

7

stars

5

commits

4

linked in READMEs

Feb 27, 2026

updated

biology
medical
Browse cluster: Biomedical Data and Multimodal Learning

README

MediX-R1: Open-Ended Medical Reinforcement Learning

MediX-R1

MediX-R1

Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI), UAE

Website Paper HuggingFace Leaderboard


Model Zoo

ModelHuggingFace
MediX-R1-2BMBZUAI/MediX-R1-2B
MediX-R1-8BMBZUAI/MediX-R1-8B
MediX-R1-30BMBZUAI/MediX-R1-30B

Citation

If you use MediX-R1 in your research, please cite our work as follows:

@misc{mullappilly2026medixr1openendedmedical,
      title={MediX-R1: Open Ended Medical Reinforcement Learning}, 
      author={Sahal Shaji Mullappilly and Mohammed Irfan Kurpath and Omair Mohamed and Mohamed Zidan and Fahad Khan and Salman Khan and Rao Anwer and Hisham Cholakkal},
      year={2026},
      eprint={2602.23363},
      archivePrefix={arXiv},
      primaryClass={cs.CV},
      url={https://arxiv.org/abs/2602.23363}, 
}

License

This project is released for research purposes only under CC-BY-NC-SA 4.0 License. It is not intended for clinical or commercial use.

Users are urged to employ MediX-R1 responsibly, especially when applying its outputs in real-world medical scenarios. It is imperative to verify the model's advice with qualified healthcare professionals and not rely on it for medical diagnoses or treatment decisions.


Acknowledgements

We are thankful to EasyR1 (a fork of veRL) for their open-source RL training framework.

This work was partially supported with NVIDIA Academic Grant 2025 and MBZUAI-IITD Research Collaboration Seed Grant.

We are grateful to MBZUAI for compute and support.

Contributors

sahalshajim

3 commits

k-m-irfan

2 commits

MBZUAI/medix-rl-data

Dataset

7

stars

5

commits

4

linked in READMEs

Feb 27, 2026

updated

biology
medical
Browse cluster: Biomedical Data and Multimodal Learning

README

MediX-R1: Open-Ended Medical Reinforcement Learning

MediX-R1

MediX-R1

Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI), UAE

Website Paper HuggingFace Leaderboard


Model Zoo

ModelHuggingFace
MediX-R1-2BMBZUAI/MediX-R1-2B
MediX-R1-8BMBZUAI/MediX-R1-8B
MediX-R1-30BMBZUAI/MediX-R1-30B

Citation

If you use MediX-R1 in your research, please cite our work as follows:

@misc{mullappilly2026medixr1openendedmedical,
      title={MediX-R1: Open Ended Medical Reinforcement Learning}, 
      author={Sahal Shaji Mullappilly and Mohammed Irfan Kurpath and Omair Mohamed and Mohamed Zidan and Fahad Khan and Salman Khan and Rao Anwer and Hisham Cholakkal},
      year={2026},
      eprint={2602.23363},
      archivePrefix={arXiv},
      primaryClass={cs.CV},
      url={https://arxiv.org/abs/2602.23363}, 
}

License

This project is released for research purposes only under CC-BY-NC-SA 4.0 License. It is not intended for clinical or commercial use.

Users are urged to employ MediX-R1 responsibly, especially when applying its outputs in real-world medical scenarios. It is imperative to verify the model's advice with qualified healthcare professionals and not rely on it for medical diagnoses or treatment decisions.


Acknowledgements

We are thankful to EasyR1 (a fork of veRL) for their open-source RL training framework.

This work was partially supported with NVIDIA Academic Grant 2025 and MBZUAI-IITD Research Collaboration Seed Grant.

We are grateful to MBZUAI for compute and support.

Contributors

sahalshajim

3 commits

k-m-irfan

2 commits