Fetching the paper…
Reading the bibliography…
In this paper, we present Change3D, a framework that reconceptualizes the change detection and captioning tasks through video modeling.
Image change detection algorithms: a systematic survey
Richard J Radke, Srinivas Andra, Omar Al-Kofahi, and Badrinath Roysam · 2005
Earlier work this paper cites.
Remote sensing change detection tools for natural resource managers: Understanding concepts and tradeoffs in the design of landscape monitoring projects
Robert E Kennedy, Philip A Townsend, John E Gross, Warren B Cohen, Paul Bolstad, YQ Wang, and Phyllis Adams · 2009
Earlier work this paper cites.
3d convolutional neural networks for human action recognition
Shuiwang Ji, Wei Xu, Ming Yang, and Kai Yu · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
Learning spatiotemporal features with 3d convolutional networks
Du Tran, Lubomir Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
V-net: Fully convolutional neural networks for volumetric medical image segmentation
Fausto Milletari, Nassir Navab, and Seyed-Ahmad Ahmadi · 2016
Earlier work this paper cites.
Hollywood in homes: Crowdsourcing data collection for activity understanding
Gunnar A Sigurdsson, Gül Varol, Xiaolong Wang, Ali Farhadi, Ivan Laptev, and Abhinav Gupta · 2016
Earlier work this paper cites.
Temporal segment networks: Towards good practices for deep action recognition
Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao, Dahua Lin, Xiaoou Tang, and Luc Van Gool · 2016
Earlier work this paper cites.
Quo vadis, action recognition? a new model and the kinetics dataset
Joao Carreira and Andrew Zisserman · 2017
Earlier work this paper cites.
The” something something” video database for learning and evaluating visual common sense
Raghav Goyal, Samira Ebrahimi Kahou, Vincent Michalski, Joanna Materzynska, Susanne Westphal, Heuna Kim, Valentin Haenel, Ingo Fruend, Peter Yianilos, Moritz Mueller-Freitag, et al · 2017
Earlier work this paper cites.
The kinetics human action video dataset
Will Kay, Joao Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, et al · 2017
Earlier work this paper cites.
Learning spatio-temporal representation with pseudo-3d residual networks
Zhaofan Qiu, Ting Yao, and Tao Mei · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Fully convolutional siamese networks for change detection
Rodrigo Caye Daudt, Bertr Le Saux, and Alexandre Boulch · 2018
Earlier work this paper cites.
Urban change detection for multispectral earth observation using convolutional neural networks
Rodrigo Caye Daudt, Bertr Le Saux, Alexandre Boulch, and Yann Gousseau · 2018
Earlier work this paper cites.
Ava: A video dataset of spatio-temporally localized atomic visual actions
Chunhui Gu, Chen Sun, David A Ross, Carl Vondrick, Caroline Pantofaru, Yeqing Li, Sudheendra Vijayanarasimhan, George Toderici, Susanna Ricco, Rahul Sukthankar, et al · 2018
Earlier work this paper cites.
Squeeze-and-excitation networks
Jie Hu, Li Shen, and Gang Sun · 2018
Earlier work this paper cites.
Fully convolutional networks for multisource building extraction from an open aerial and satellite imagery data set
Shunping Ji, Shiqing Wei, and Meng Lu · 2018
Earlier work this paper cites.
A closer look at spatiotemporal convolutions for action recognition
Du Tran, Heng Wang, Lorenzo Torresani, Jamie Ray, Yann LeCun, and Manohar Paluri · 2018
Earlier work this paper cites.
Non-local neural networks
Xiaolong Wang, Ross Girshick, Abhinav Gupta, and Kaiming He · 2018
Earlier work this paper cites.
Cbam: Convolutional block attention module
Sanghyun Woo, Jongchan Park, Joon-Young Lee, and In So Kweon · 2018
Earlier work this paper cites.
Multitask learning for large-scale semantic change detection
Rodrigo Caye Daudt, Bertrand Le Saux, Alexandre Boulch, and Yann Gousseau · 2019
Earlier work this paper cites.
Slowfast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Cited alongside, same era.
Creating xbd: A dataset for assessing building damage from satellite imagery
Ritwik Gupta, Bryce Goodman, Nirav Patel, Ricky Hosfelt, Sandra Sajeev, Eric Heim, Jigar Doshi, Keane Lucas, Howie Choset, and Matthew Gaston · 2019
Cited alongside, same era.
Stm: Spatiotemporal and motion encoding for action recognition
Boyuan Jiang, MengMeng Wang, Weihao Gan, Wei Wu, and Junjie Yan · 2019
Cited alongside, same era.
Tsm: Temporal shift module for efficient video understanding
Ji Lin, Chuang Gan, and Song Han · 2019
Cited alongside, same era.
Robust change captioning
Dong Huk Park, Trevor Darrell, and Anna Rohrbach · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al · 2019
Icif-net: Intra-scale cross-interaction and inter-scale feature fusion network for bitemporal remote sensing images change detection
Yuchao Feng, Honghui Xu, Jiawei Jiang, Hao Liu, and Jianwei Zheng · 2022
Later among the works it cites.
Change captioning: A new paradigm for multitemporal remote sensing image analysis
Genc Hoxha, Seloua Chouaf, Farid Melgani, and Youcef Smara · 2022
Later among the works it cites.
Transition is a process: Pair-to-video change detection networks for very high resolution remote sensing images
Manhui Lin, Guangyi Yang, and Hongyan Zhang · 2022
Later among the works it cites.
Land-cover change detection using multi-temporal modis ndvi data
Ross S Lunetta, Joseph F Knight, Jayantha Ediriwickrema, John G Lyon, and L Dorsey Worthy · 2022
Later among the works it cites.
Vitpose: Simple vision transformer baselines for human pose estimation
Yufei Xu, Jing Zhang, Qiming Zhang, and Dacheng Tao · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Video classification with channel-separated convolutional networks
Du Tran, Heng Wang, Lorenzo Torresani, and Matt Feiszli · 2019
Cited alongside, same era.
A spatial-temporal attention-based method and a new dataset for remote sensing image change detection
Hao Chen and Zhenwei Shi · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
X3d: Expanding architectures for efficient video recognition
Christoph Feichtenhofer · 2020
Cited alongside, same era.
Tea: Temporal excitation and aggregation for action recognition
Yan Li, Bin Ji, Xintian Shi, Jianguo Zhang, Bin Kang, and Limin Wang · 2020
Cited alongside, same era.
Building change detection for remote sensing images using a dual-task constrained deep siamese convolutional network model
Yi Liu, Chao Pang, Zongqian Zhan, Xiaomeng Zhang, and Xue Yang · 2020
Cited alongside, same era.
Swinsunet: Pure transformer network for remote sensing image change detection
Cui Zhang, Liejun Wang, Shuli Cheng, and Yongming Li · 2022
Later among the works it cites.
Spatially and semantically enhanced siamese network for semantic change detection in high-resolution remote sensing images
Manqi Zhao, Zifei Zhao, Shuai Gong, Yunfei Liu, Jian Yang, Xiong Xiong, and Shengyang Li · 2022
Later among the works it cites.
Changemask: Deep multi-task encoder-transformer-decoder architecture for semantic change detection
Zhuo Zheng, Yanfei Zhong, Shiqi Tian, Ailong Ma, and Liangpei Zhang · 2022
Later among the works it cites.
Vision transformer adapter for dense predictions
Zhe Chen, Yuchen Duan, Wenhai Wang, Junjun He, Tong Lu, Jifeng Dai, and Yu Qiao · 2023
Later among the works it cites.
Mtscd-net: A network based on multi-task learning for semantic change detection of bitemporal remote sensing images
Fengzhi Cui and Jie Jiang · 2023
Later among the works it cites.
Changer: Feature interaction is what you need for change detection
Sheng Fang, Kaiyu Li, and Zhe Li · 2023
Later among the works it cites.
Vct: Visual change transformer for remote sensing image change detection
Bo Jiang, Zitian Wang, Xixi Wang, Ziyan Zhang, Lan Chen, Xiao Wang, and Bin Luo · 2023
Later among the works it cites.
Temporal-agnostic change region proposal for semantic change detection
Shiqi Tian, Xicheng Tan, Ailong Ma, Zhuo Zheng, Liangpei Zhang, and Yanfei Zhong · 2023
Later among the works it cites.
Fully convolutional change detection framework with generative adversarial network for unsupervised, weakly supervised and regional supervised change detection
Chen Wu, Bo Du, and Liangpei Zhang · 2023
Later among the works it cites.
A triple-branch hybrid attention network with bitemporal feature joint refinement for remote sensing image semantic change detection
Hao Chang, Peijin Wang, Wenhui Diao, Guangluan Xu, and Xian Sun · 2024
Later among the works it cites.
Changemamba: Remote sensing change detection with spatio-temporal state space model
Hongruixuan Chen, Jian Song, Chengxi Han, Junshi Xia, and Naoto Yokoya · 2024
Later among the works it cites.
Joint spatio-temporal modeling for semantic change detection in remote sensing images
Lei Ding, Jing Zhang, Haitao Guo, Kai Zhang, Bing Liu, and Lorenzo Bruzzone · 2024
Later among the works it cites.
Froster: Frozen clip is a strong teacher for open-vocabulary action recognition
Xiaohu Huang, Hao Zhou, Kun Yao, and Kai Han · 2024
Later among the works it cites.
A decoder-focused multi-task network for semantic change detection
Zhe Li, Xiaoxin Wang, Sheng Fang, Jianli Zhao, Shuqi Yang, and Wen Li · 2024
Later among the works it cites.
Eatder: Edge-assisted adaptive transformer detector for remote sensing change detection
Jingjing Ma, Junyi Duan, Xu Tang, Xiangrong Zhang, and Licheng Jiao · 2024
Later among the works it cites.
Pcdasnet: Position-constrained differential attention siamese network for building damage assessment
Jiaqi Wang, Haonan Guo, Xin Su, Li Zheng, and Qiangqiang Yuan · 2024
Later among the works it cites.
Vitmatte: Boosting image matting with pre-trained plain vision transformers
Jingfeng Yao, Xinggang Wang, Shusheng Yang, and Baoyuan Wang · 2024
Later among the works it cites.
Single-stream extractor network with contrastive pre-training for remote sensing change captioning
Qing Zhou, Junyu Gao, Yuan Yuan, and Qi Wang · 2024
Later among the works it cites.
Changevit: Unleashing plain vision transformers for change detection
Duowang Zhu, Xiaohu Huang, Haiyan Huang, Zhenfeng Shao, and Qimin Cheng · 2024
Later among the works it cites.