Fetching the paper…

Gaining Extra Supervision via Multi-task learning for Multi-Modal Video Question Answering · Around