Fetching the paper…
Reading the bibliography…
The YLI Multimedia Event Detection corpus is a public-domain index of videos with annotations and computed features, specialized for research in multimedia event detection (MED), i.e., automatically identifying what's happening in a video by analyzing the audio and visual content.
Human categorization
Eleanor Rosch · 1977
Earlier work this paper cites.
Principles of categorization
Eleanor Rosch · 1977
Earlier work this paper cites.
Categorization of natural objects
Carolyn Mervis and Eleanor Rosch · 1981
Earlier work this paper cites.
Frames and the semantics of understanding
Charles J. Fillmore · 1985
Earlier work this paper cites.
Women, Fire, and Dangerous Things: What Categories Reveal about the Mind
George Lakoff · 1987
Earlier work this paper cites.
Folksonomies: Tidying up tags
Marieke Guy and Emma Tonkin · 2006
Earlier work this paper cites.
Probabilistic linear discriminant analysis
Sergey Ioffe · 2006
Earlier work this paper cites.
Evaluation campaigns and TRECVid
Alan F. Smeaton, Paul Over, and Wessel Kraaij · 2006
Earlier work this paper cites.
Support vector machines versus fast scoring in the low-dimensional total variability space for speaker verification
Najim Dehak, Réda Dehak, Patrick Kenny, Niko Brümmer, Pierre Ouellet, and Pierre Dumouchel · 2009
Earlier work this paper cites.
Visual Concept Learning from User-Tagged Web Video
Adrian Ulges · 2009
Earlier work this paper cites.
A comparative study of Flickr tags and index terms in a general image collection
Abebe Rorissa · 2010
Cited alongside, same era.
Visual concept learning from weakly labeled web videos
Adrian Ulges, Damian Borth, and Thomas M Breuel · 2010
Cited alongside, same era.
Discriminatively trained probabilistic Linear Discriminant Analysis for speaker verification
Lukáš Burget, Oldřich Plchot, Sandro Cumani, Ondřej Glembek, Pavel Matějka, and Niko Brümmer · 2011
Cited alongside, same era.
2011 TRECVID multimedia event detection track
National Institute of Standards and Technology · 2011
Cited alongside, same era.
SRI-Sarnoff AURORA system at TRECVID 2012: Multimedia event detection and recounting
Hui Cheng, Jingen Liu, Saad Ali, Omar Javed, Qian Yu, Amir Tamrakar, Ajay Divakaran, Harpreet S. Sawhney, R. Manmatha, James Allan, Alex Hauptmann, Mubarak Shah, Subhabrata Bhattacharya, Afshin Dehghan, Gerald Friedland, Benjamín Martinez Elizalde, Trevor Darrell, Michael Witbrock, and Jon Curtis · 2012
Cited alongside, same era.
An i-vector representation of acoustic environments for audio-based video event detection on user generated content
Benjamin Elizalde, Howard Lei, and Gerald Friedland · 2013
Later among the works it cites.
SRI-Sarnoff AURORA system at TRECVID 2013: Multimedia event detection and recounting
Jingen Liu, Hui Cheng, Omar Javed, Qian Yu, Ishani Chakraborty, Weiyu Zhang, Ajay Divakaran, Harpreet S. Sawhney, James Allan, R. Manmatha, John Foley, Mubarak Shah, Afshin Dehghan, Michael Witbrock, Jon Curtis, and Gerald Friedland · 2013
Later among the works it cites.
The Placing Task: A large-scale geo-estimation challenge for social-media videos and images
Jaeyoung Choi, Bart Thomee, Gerald Friedland, Liangliang Cao, Karl Ni, Damian Borth, Benjamin Elizalde, Luke Gottlieb, Carmen Carrano, Roger Pearce, and Doug Poland · 2014
Later among the works it cites.
Caffe: Convolutional architecture for fast feature embedding
Yangqing Jia, Evan Shelhamer, Jeff Donahue, Sergey Karayev, Jonathan Long, Ross B. Girshick, Sergio Guadarrama, and Trevor Darrell · 2014
Later among the works it cites.
The Yahoo-Livermore-ICSI (YLI) multimedia feature set
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
TRECVID 2011 - an overview of the goals, tasks, data, evaluation mechanisms, and metrics
Paul Over, George Awad, Jonathan Fiscus, Brian Antonishek, Martial Michel, Alan Smeaton, Wessel Kraaij, and Georges Quénot · 2012
Cited alongside, same era.
Creating HAVIC: Heterogeneous audio visual Internet collection
Stephanie Strassel, Amanda Morris, Jonathan Fiscus, Christopher Caruso, Haejoong Lee, Paul Over, James Fiumara, Barbara Shaw, Brian Antonishek, and Martial Michel · 2012
Cited alongside, same era.
Linking visual concept detection with viewer demographics
Adrian Ulges, Markus Koch, and Damian Borth · 2012
Cited alongside, same era.
Large-scale visual sentiment ontology and detectors using adjective noun pairs
Damian Borth, Rongrong Ji, Tao Chen, Thomas Breuel, and Shih-Fu Chang · 2013
Cited alongside, same era.
Karl S. Ni, Carmen C. Carrano, Doug N. Poland, Benjamin M. Elizalde, Gerald Friedland, Luke R. Gottlieb, and Damian S. Borth · 2014
Later among the works it cites.
Social Event Detection at MediaEval: a three-year retrospect of tasks and results
Georgios Petkos, Symeon Papadopoulos, Vasileios Mezaris, Raphael Troncy, Philipp Cimiano, Timo Reuter, and Yiannis Kompatsiaris · 2014
Later among the works it cites.
Yahoo! Webscope dataset YFCC-100M
Yahoo Labs · 2014
Later among the works it cites.
Audio-based multimedia event detection with DNNs and sparse sampling
Khalid Ashraf, Benjamin Elizalde, Forrest Iandola, Matthew Moskewicz, Julia Bernd, Gerald Friedland, and Kurt Keutzer · 2015
Closest in time.
The new data and new challenges in multimedia research
Bart Thomee, David A. Shamma, Benjamin Elizalde, Gerald Friedland, Karl Ni, Douglas Poland, Damian Borth, and Li-Jia Li · 2015
Closest in time.