2020

QuerYD: A video dataset with high-quality text and audio narrations

Oncescu, Andreea-Maria, Henriques, João F., Liu, Yang et al.

Understand

We introduce QuerYD, a new large-scale dataset for retrieval and event localisation in video.

  • A unique feature of our dataset is the availability of two audio tracks for each video: the original audio, and a high-quality spoken description of the visual content.
  • The dataset is based on YouDescribe, a volunteer project that assists visually-impaired people by attaching voiced narrations to existing YouTube videos.
  • This ever-growing collection of videos contains highly detailed, temporally aligned audio and text annotations.

Reading the bibliography…