Fetching the paper…

MiDashengLM: Efficient Audio Understanding with General Audio Captions · Around