Fetching the paper…

Parameter Efficient Audio Captioning With Faithful Guidance Using Audio-text Shared Latent Representation · Around