Fetching the paper…
Reading the bibliography…
While Large Language Models (LLMs) have recently shown impressive results in reasoning tasks, their application to pedestrian trajectory prediction remains challenging due to two key limitations: insufficient use of visual information and the difficulty of predicting entire trajectories.
Social force model for pedestrian dynamics
Dirk Helbing and Peter Molnar. 1995 · 1995
Earlier work this paper cites.
Crowds by example. In Computer graphics forum , Vol. 26. Wiley Online Library, 655–664
Alon Lerner, Yiorgos Chrysanthou, and Dani Lischinski. 2007 · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition . Ieee, 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
You’ll never walk alone: Modeling social behavior for multi-target tracking. In 2009 IEEE 12th international conference on computer vision . IEEE, 261–268
Stefano Pellegrini, Andreas Ess, Konrad Schindler, and Luc Van Gool. 2009 · 2009
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation. In Medical image computing and computer-assisted intervention–MICCAI 2015: 18th international conference, Munich, Germany, October 5-9, 2015, proceedings, part III 18 . Springer, 234–241
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015 · 2015
Earlier work this paper cites.
Social lstm: Human trajectory prediction in crowded spaces. In Proceedings of the IEEE conference on computer vision and pattern recognition . 961–971
Alexandre Alahi, Kratarth Goel, Vignesh Ramanathan, Alexandre Robicquet, Li Fei-Fei, and Silvio Savarese. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin. 2018 · 2018
Earlier work this paper cites.
Social gan: Socially acceptable trajectories with generative adversarial networks. In Proceedings of the IEEE conference on computer vision and pattern recognition . 2255–2264
Agrim Gupta, Justin Johnson, Li Fei-Fei, Silvio Savarese, and Alexandre Alahi. 2018 · 2018
Earlier work this paper cites.
Social ways: Learning multi-modal distributions of pedestrian trajectories with gans. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops . 0–0
Javad Amirian, Jean-Bernard Hayet, and Julien Pettré. 2019 · 2019
Earlier work this paper cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Earlier work this paper cites.
Sophie: An attentive gan for predicting paths compliant to social and physical constraints. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 1349–1358
Amir Sadeghian, Vineet Kosaraju, Ali Sadeghian, Noriaki Hirose, Hamid Rezatofighi, and Silvio Savarese. 2019 · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom B Brown. 2020 · 2020
Earlier work this paper cites.
Goal-gan: Multimodal trajectory prediction based on goal position estimation. In Proceedings of the Asian Conference on Computer Vision
Patrick Dendorfer, Aljosa Osep, and Laura Leal-Taixé. 2020 · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy. 2020 · 2020
Earlier work this paper cites.
It is not the journey but the destination: Endpoint conditioned trajectory prediction. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part II 16 . Springer, 759–776
Karttikeya Mangalam, Harshayu Girase, Shreyas Agarwal, Kuan-Hui Lee, Ehsan Adeli, Jitendra Malik, and Adrien Gaidon. 2020 · 2020
Earlier work this paper cites.
Social-stgcnn: A social spatio-temporal graph convolutional neural network for human trajectory prediction. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 14424–14432
Abduallah Mohamed, Kun Qian, Mohamed Elhoseiny, and Christian Claudel. 2020 · 2020
Earlier work this paper cites.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Earlier work this paper cites.
Trajectron++: Dynamically-feasible trajectory forecasting with heterogeneous data. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XVIII 16 . Springer, 683–700
Tim Salzmann, Boris Ivanovic, Punarjay Chakravarty, and Marco Pavone. 2020 · 2020
Cited alongside, same era.
Spatio-temporal graph transformer networks for pedestrian trajectory prediction. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XII 16 . Springer, 507–523
Cunjun Yu, Xiao Ma, Jiawei Ren, Haiyu Zhao, and Shuai Yi. 2020 · 2020
Cited alongside, same era.
Mg-gan: A multi-generator model preventing out-of-distribution samples in pedestrian trajectory prediction. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 13158–13167
Patrick Dendorfer, Sven Elflein, and Laura Leal-Taixé. 2021 · 2021
Cited alongside, same era.
Street life and pedestrian activities in smart cities: opportunities and challenges for computational urban science
Zhuangyuan Fan and Becky PY Loo. 2021 · 2021
Cited alongside, same era.
Eigentrajectory: Low-rank descriptors for multi-modal trajectory forecasting. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 10017–10029
Inhwan Bae, Jean Oh, and Hae-Gon Jeon. 2023 · 2023
Later among the works it cites.
Evolutionary-scale prediction of atomic-level protein structure with a language model
Zeming Lin, Halil Akin, Roshan Rao, Brian Hie, Zhongkai Zhu, Wenting Lu, Nikita Smetanin, Robert Verkuil, Ori Kabeli, Yaniv Shmueli, et al · 2023
Later among the works it cites.
Leapfrog diffusion model for stochastic trajectory prediction. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 5517–5526
Weibo Mao, Chenxin Xu, Qi Zhu, Siheng Chen, and Yanfeng Wang. 2023 · 2023
Later among the works it cites.
Interaction-aware trajectory prediction for autonomous vehicle based on LSTM-MLP model. In Proceedings of KES-STS International Symposium . Springer, 91–99
Zhiwei Meng, Jiaming Wu, Sumin Zhang, Rui He, and Bing Ge. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
From goals, waypoints & paths to long term human trajectory forecasting. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 15233–15242
Karttikeya Mangalam, Yang An, Harshayu Girase, and Jitendra Malik. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision. In International conference on machine learning . PMLR, 8748–8763
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Semantics-STGCNN: A semantics-guided spatial-temporal graph convolutional network for multi-class trajectory prediction. In 2021 IEEE International Conference on Systems, Man, and Cybernetics (SMC) . IEEE, 2959–2966
Ben A Rainbow, Qianhui Men, and Hubert PH Shum. 2021 · 2021
Cited alongside, same era.
A review of deep learning-based methods for pedestrian trajectory prediction
Bogdan Ilie Sighencea, Rare · 2021
Cited alongside, same era.
Agentformer: Agent-aware transformers for socio-temporal multi-agent forecasting. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 9813–9823
Ye Yuan, Xinshuo Weng, Yanglan Ou, and Kris M Kitani. 2021 · 2021
Cited alongside, same era.
Non-probability sampling network for stochastic human trajectory prediction. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6477–6487
Inhwan Bae, Jin-Hwi Park, and Hae-Gon Jeon. 2022 · 2022
Cited alongside, same era.
Exploring visual prompts for adapting large-scale models
Hyojin Bahng, Ali Jahanian, Swami Sankaranarayanan, and Phillip Isola. 2022 · 2022
Cited alongside, same era.
Goal-driven self-attentive recurrent networks for trajectory prediction. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 2518–2527
Luigi Filippo Chiara, Pasquale Coscia, Sourav Das, Simone Calderara, Rita Cucchiara, and Lamberto Ballan. 2022 · 2022
Cited alongside, same era.
Trace and pace: Controllable pedestrian animation via guided trajectory diffusion. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 13756–13766
Davis Rempe, Zhengyi Luo, Xue Bin Peng, Ye Yuan, Kris Kitani, Karsten Kreis, Sanja Fidler, and Or Litany. 2023 · 2023
Later among the works it cites.
What does clip know about a red circle? visual prompt engineering for vlms. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 11987–11997
Aleksandar Shtedritski, Christian Rupprecht, and Andrea Vedaldi. 2023 · 2023
Later among the works it cites.
D-STGCN: Dynamic Pedestrian Trajectory Prediction Using Spatio-Temporal Graph Convolutional Networks
Bogdan Ilie Sighencea, Ion Rare · 2023
Later among the works it cites.
A survey of reasoning with foundation models
Jiankai Sun, Chuanyang Zheng, Enze Xie, Zhengying Liu, Ruihang Chu, Jianing Qiu, Jiaqi Xu, Mingyu Ding, Hongyang Li, Mengzhe Geng, et al · 2023
Later among the works it cites.
Eqmotion: Equivariant multi-agent motion prediction with invariant interaction reasoning. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 1410–1420
Chenxin Xu, Robby T Tan, Yuhong Tan, Siheng Chen, Yu Guang Wang, Xinchao Wang, and Yanfeng Wang. 2023 · 2023
Later among the works it cites.
DNAGPT: a generalized pretrained tool for multiple DNA sequence analysis tasks
Daoan Zhang, Weitong Zhang, Bing He, Jianguo Zhang, Chenchen Qin, and Jianhua Yao. 2023 · 2023
Later among the works it cites.
Can Language Beat Numerical Regression? Language-Based Multimodal Trajectory Prediction. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
Inhwan Bae, Junoh Lee, and Hae-Gon Jeon. 2024 · 2024
Later among the works it cites.
ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 12914–12923
Mu Cai, Haotian Liu, Siva Karthik Mustikovela, Gregory P Meyer, Yuning Chai, Dennis Park, and Yong Jae Lee. 2024 · 2024
Later among the works it cites.
Dice: Diverse diffusion model with scoring for trajectory prediction. In 2024 IEEE Intelligent Vehicles Symposium (IV) . IEEE, 3023–3029
Younwoo Choi, Ray Coden Mercurius, Soheil Mohamad Alizadeh Shabestary, and Amir Rasouli. 2024 · 2024
Later among the works it cites.
Time-LLM: Time series forecasting by reprogramming large language models. In International Conference on Learning Representations (ICLR)
Ming Jin, Shiyu Wang, Lintao Ma, Zhixuan Chu, James Y Zhang, Xiaoming Shi, Pin-Yu Chen, Yuxuan Liang, Yuan-Fang Li, Shirui Pan, and Qingsong Wen. 2024 · 2024
Later among the works it cites.
Human activity recognition via temporal fusion contrastive learning
Inkyung Kim, Juwan Lim, and Jaekoo Lee. 2024 · 2024
Later among the works it cites.
Can large language models reason about medical questions?
Valentin Liévin, Christoffer Egeberg Hother, Andreas Geert Motzfeldt, and Ole Winther. 2024 · 2024
Later among the works it cites.
Remoteclip: A vision language foundation model for remote sensing
Fan Liu, Delong Chen, Zhangqingyun Guan, Xiaocong Zhou, Jiale Zhu, Qiaolin Ye, Liyong Fu, and Jun Zhou. 2024 · 2024
Later among the works it cites.
Cpt: Colorful prompt tuning for pre-trained vision-language models
Yuan Yao, Ao Zhang, Zhengyan Zhang, Zhiyuan Liu, Tat-Seng Chua, and Maosong Sun. 2024 · 2024
Later among the works it cites.