Fetching the paper…
Reading the bibliography…
Episodic memory retrieval enables wearable cameras to recall objects or events previously observed in video.
Episodic memory: From mind to brain
Endel Tulving · 2002
Earlier work this paper cites.
Max-margin early event detectors
Minh Hoai and Fernando De la Torre · 2014
Earlier work this paper cites.
Online action detection
Roeland De Geest, Efstratios Gavves, Amir Ghodrati, Zhenyang Li, Cees Snoek, and Tinne Tuytelaars · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2016
Earlier work this paper cites.
Quo vadis, action recognition? a new model and the kinetics dataset
João Carreira and Andrew Zisserman · 2017
Earlier work this paper cites.
RED: Reinforced encoder-decoder networks for action anticipation
Jiyang Gao, Zhenheng Yang, and Ram Nevatia · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie · 2017
Earlier work this paper cites.
Encouraging lstms to anticipate actions very early
Mohammad Sadegh Aliakbarian, Fatemeh Sadat Saleh, Mathieu Salzmann, Basura Fernando, Lars Petersson, and Lars Andersson · 2017
Earlier work this paper cites.
Slowfast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Earlier work this paper cites.
Know your surroundings: Exploiting scene information for object tracking
Goutam Bhat, Martin Danelljan, Luc Van Gool, and Radu Timofte · 2020
Earlier work this paper cites.
Understanding human hands in contact at internet scale
Dandan Shan, Jiaqi Geng, Michelle Shu, and David F Fouhey · 2020
Earlier work this paper cites.
Siam r-cnn: Visual tracking by re-detection
Paul Voigtlaender, Jonathon Luiten, Philip HS Torr, and Bastian Leibe · 2020
Earlier work this paper cites.
Is space-time attention all you need for video understanding?
Gedas Bertasius, Heng Wang, and Lorenzo Torresani · 2021
Earlier work this paper cites.
Is first person vision challenging for object tracking?
Matteo Dunnhofer, Antonino Furnari, Giovanni Maria Farinella, and Christian Micheloni · 2021
Earlier work this paper cites.
Rolling-unrolling lstms for action anticipation from first-person video
Antonino Furnari and Giovanni Farinella · 2021
Earlier work this paper cites.
Anticipative video transformer
Rohit Girdhar and Kristen Grauman · 2021
Earlier work this paper cites.
Hota: A higher order metric for evaluating multi-object tracking
Jonathon Luiten, Aljosa Osep, Patrick Dendorfer, Philip Torr, Andreas Geiger, Laura Leal-Taixé, and Bastian Leibe · 2021
Earlier work this paper cites.
Predicting the future from first person (egocentric) vision: A survey
Ivan Rodin, Antonino Furnari, Dimitrios Mavroedis, and Giovanni Maria Farinella · 2021
Cited alongside, same era.
Learning spatio-temporal transformer for visual tracking
Bin Yan, Houwen Peng, Jianlong Fu, Dong Wang, and Huchuan Lu · 2021
Cited alongside, same era.
Bot-sort: Robust associations multi-pedestrian tracking
Nir Aharon, Roy Orfaig, and Ben-Zion Bobrovsky · 2022
Cited alongside, same era.
Where did i leave my keys?-episodic-memory-based question answering on egocentric videos
Leonard Bärmann and Alex Waibel · 2022
Cited alongside, same era.
Internvideo-ego4d: A pack of champion solutions to ego4d challenges
Guo Chen, Sen Xing, Zhe Chen, Yi Wang, Kunchang Li, Yizhuo Li, Yi Liu, Jiahao Wang, Yin-Dong Zheng, Bingkun Huang, et al · 2022
Cited alongside, same era.
The wisdom of crowds: Temporal progressive attention for early action prediction
Alexandros Stergiou and Dima Damen · 2023
Later among the works it cites.
Egotracks: a long-term egocentric visual object tracking dataset
Hao Tang, Kevin J Liang, Kristen Grauman, Matt Feiszli, and Weiyao Wang · 2023
Later among the works it cites.
Where is my wallet? modeling object proposal sets for egocentric visual query localization
Mengmeng Xu, Yanghao Li, Cheng-Yang Fu, Bernard Ghanem, Tao Xiang, and Juan-Manuel Pérez-Rúa · 2023
Later among the works it cites.
Egoobjects: A large-scale egocentric dataset for fine-grained object understanding
Chenchen Zhu, Fanyi Xiao, Andrés Alvarado, Yasmine Babaei, Jiabo Hu, Hichem El-Mohri, Sean Culatana, Roshan Sumbaly, and Zhicheng Yan · 2023
Later among the works it cites.
Yolo-world: Real-time open-vocabulary object detection
Tianheng Cheng, Lin Song, Yixiao Ge, Wenyu Liu, Xinggang Wang, and Ying Shan · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100
Dima Damen, Hazel Doughty, Giovanni Maria Farinella, Antonino Furnari, Evangelos Kazakos, Jian Ma, Davide Moltisanti, Jonathan Munro, Toby Perrett, Will Price, and Michael Wray · 2022
Cited alongside, same era.
Ego4d: Around the World in 3,000 Hours of Egocentric Video
Kristen Grauman et al · 2022
Cited alongside, same era.
Reler@ zju-alibaba submission to the ego4d natural language queries challenge 2022
Naiyuan Liu, Xiaohan Wang, Xiaobo Li, Yi Yang, and Yueting Zhuang · 2022
Cited alongside, same era.
Negative frames matter in egocentric visual query 2d localization
Mengmeng Xu, Cheng-Yang Fu, Yanghao Li, Bernard Ghanem, Juan-Manuel Perez-Rua, and Tao Xiang · 2022
Cited alongside, same era.
Real-time online video detection with temporal smoothing transformers
Yue Zhao and Philipp Krähenbühl · 2022
Cited alongside, same era.
Miniroad: Minimal rnn framework for online action detection
Joungbin An, Hyolim Kang, Su Ho Han, Ming-Hsuan Yang, and Seon Joo Kim · 2023
Cited alongside, same era.
Visual object tracking in first person vision
Matteo Dunnhofer, Antonino Furnari, Giovanni Maria Farinella, and Christian Micheloni · 2023
Cited alongside, same era.
Videoagent: A memory-augmented multimodal agent for video understanding
Yue Fan, Xiaojian Ma, Rujie Wu, Yuntao Du, Jiaqi Li, Zhi Gao, and Qing Li · 2024
Closest in time.
Objectnlq@ ego4d episodic memory challenge 2024
Yisen Feng, Haoyu Zhang, Yuquan Xie, Zaijing Li, Meng Liu, and Liqiang Nie · 2024
Closest in time.
Amego: Active memory from long egocentric videos
Gabriele Goletto, Tushar Nagarajan, Giuseppe Averta, and Dima Damen · 2024
Closest in time.
Matching anything by segmenting anything
Siyuan Li, Lei Ke, Martin Danelljan, Luigi Piccinelli, Mattia Segu, Luc Van Gool, and Fisher Yu · 2024
Closest in time.
End-to-end temporal action detection with 1b parameters across 1000 frames
Shuming Liu, Chen-Lin Zhang, Chen Zhao, and Bernard Ghanem · 2024
Closest in time.
Egovideo: Exploring egocentric foundation model and downstream adaptation
Baoqi Pei, Guo Chen, Jilan Xu, Yuping He, Yicheng Liu, Kanghua Pan, Yifei Huang, Yali Wang, Tong Lu, Limin Wang, et al · 2024
Closest in time.
An outlook into the future of egocentric vision
Chiara Plizzari, Gabriele Goletto, Antonino Furnari, Siddhant Bansal, Francesco Ragusa, Giovanni Maria Farinella, Dima Damen, and Tatiana Tommasi · 2024
Closest in time.
Yolov10: Real-time end-to-end object detection. arxiv 2024
A Wang, H Chen, L Liu, K Chen, Z Lin, J Han, and G Ding · 2024
Closest in time.
Detrs beat yolos on real-time object detection
Yian Zhao, Wenyu Lv, Shangliang Xu, Jinman Wei, Guanzhong Wang, Qingqing Dang, Yi Liu, and Jie Chen · 2024
Closest in time.
Is tracking really more challenging in egocentric first person vision?
Matteo Dunnhofer, Zaira Manigrasso, and Christian Micheloni · 2025
Closest in time.
Prvql: Progressive knowledge-guided refinement for robust egocentric visual query localization
Bing Fan, Yunhe Feng, Yapeng Tian, Yuewei Lin, Yan Huang, and Heng Fan · 2025
Closest in time.
Spatial cognition from egocentric video: Out of sight, not out of mind
Chiara Plizzari, Shubham Goel, Toby Perrett, Jacob Chalk, Angjoo Kanazawa, and Dima Damen · 2025
Closest in time.