SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation.
The repository's model_checkpoints directory contains checkpoints for both student and teacher models. Each model is available in three variants:
The utils_checkpoint directory includes additional checkpoints for auxiliary components, such as the audio compression model.
@inproceedings{saito2025soundctm,
title={Sound{CTM}: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation},
author={Koichi Saito and Dongjun Kim and Takashi Shibuya and Chieh-Hsin Lai and Zhi Zhong and Yuhta Takida and Yuki Mitsufuji},
booktitle={The Thirteenth International Conference on Learning Representations},
year={2025},
url={https://openreview.net/forum?id=KrK6zXbjfO}
}
12 commits
SoundCTM: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation.
The repository's model_checkpoints directory contains checkpoints for both student and teacher models. Each model is available in three variants:
The utils_checkpoint directory includes additional checkpoints for auxiliary components, such as the audio compression model.
@inproceedings{saito2025soundctm,
title={Sound{CTM}: Unifying Score-based and Consistency Models for Full-band Text-to-Sound Generation},
author={Koichi Saito and Dongjun Kim and Takashi Shibuya and Chieh-Hsin Lai and Zhi Zhong and Yuhta Takida and Yuki Mitsufuji},
booktitle={The Thirteenth International Conference on Learning Representations},
year={2025},
url={https://openreview.net/forum?id=KrK6zXbjfO}
}
12 commits