Fetching the paper…

JTAV: Jointly Learning Social Media Content Representation by Fusing Textual, Acoustic, and Visual Features · Around