zhoujiaming777/DIFFA

Model

0

stars

8

commits

2

linked in READMEs

Aug 19, 2025

updated

pytorch

README

DIFFA: Large Language Diffusion Models Can Listen and Understand

arXiv deploy Github

DIFFA is the first diffusion-based large audio-language model for spoken language understanding.
It combines a frozen diffusion LLM with dual adapters (semantic + acoustic) to enhance audio perception and reasoning.

Contributors

zhoujiaming777/DIFFA

Model

0

stars

8

commits

2

linked in READMEs

Aug 19, 2025

updated

pytorch

README

DIFFA: Large Language Diffusion Models Can Listen and Understand

arXiv deploy Github

DIFFA is the first diffusion-based large audio-language model for spoken language understanding.
It combines a frozen diffusion LLM with dual adapters (semantic + acoustic) to enhance audio perception and reasoning.

Contributors