The codebase and some introductions of FineMed.
Jupyter Notebook
31
81 commits
updated Sep 11, 2025
This repo includes the codebase and some introductions of FineMed, as described in the paper FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training.
We have uploaded the datasets to huggingface:
SFT Datasets: https://huggingface.co/datasets/hongzhouyu/FineMed-SFT
DPO Dataset: https://huggingface.co/datasets/hongzhouyu/FineMed-DPO
We have uploaded the models to huggingface:
FineMedLM: https://huggingface.co/hongzhouyu/FineMedLM
FineMedLM-o1: https://huggingface.co/hongzhouyu/FineMedLM-o1
If you want to reproduce our research, please run the following code in sequence:
Synthetic Data
Qwen_med_cls
Training
TTT
If you are interested in FineMed, feel free to email me at hzyu24@m.fudan.edu.cn!
If FineMed or this repository is useful in your own research, you can use the following BibTeX entry:
@misc{yu2025finemedlmo1enhancingmedicalknowledge,
title={FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training},
author={Hongzhou Yu and Tianhao Cheng and Yingwen Wang and Wen He and Qing Wang and Ying Cheng and Yuejie Zhang and Rui Feng and Xiaobo Zhang},
year={2025},
eprint={2501.09213},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2501.09213},
}
Jupyter Notebook
93.0%
Python
7.0%
The codebase and some introductions of FineMed.
Jupyter Notebook
31
81 commits
updated Sep 11, 2025
This repo includes the codebase and some introductions of FineMed, as described in the paper FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training.
We have uploaded the datasets to huggingface:
SFT Datasets: https://huggingface.co/datasets/hongzhouyu/FineMed-SFT
DPO Dataset: https://huggingface.co/datasets/hongzhouyu/FineMed-DPO
We have uploaded the models to huggingface:
FineMedLM: https://huggingface.co/hongzhouyu/FineMedLM
FineMedLM-o1: https://huggingface.co/hongzhouyu/FineMedLM-o1
If you want to reproduce our research, please run the following code in sequence:
Synthetic Data
Qwen_med_cls
Training
TTT
If you are interested in FineMed, feel free to email me at hzyu24@m.fudan.edu.cn!
If FineMed or this repository is useful in your own research, you can use the following BibTeX entry:
@misc{yu2025finemedlmo1enhancingmedicalknowledge,
title={FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training},
author={Hongzhou Yu and Tianhao Cheng and Yingwen Wang and Wen He and Qing Wang and Ying Cheng and Yuejie Zhang and Rui Feng and Xiaobo Zhang},
year={2025},
eprint={2501.09213},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2501.09213},
}
Jupyter Notebook
93.0%
Python
7.0%