Fetching the paper…
Reading the bibliography…
Matching cross-modality features between images and point clouds is a fundamental problem for image-to-point cloud registration.
A solution for the best rotation to relate two sets of vectors
Wolfgang Kabsch · 1976
Earlier work this paper cites.
The opencv library
Gary Bradski · 2000
Earlier work this paper cites.
Computer vision: a modern approach
David A Forsyth and Jean Ponce · 2002
Earlier work this paper cites.
Softposit: Simultaneous pose and correspondence determination
Philip David, Daniel Dementhon, Ramani Duraiswami, and Hanan Samet · 2004
Earlier work this paper cites.
Epnp: An accurate o (n) solution to the p n p problem
Vincent Lepetit, Francesc Moreno-Noguer, and Pascal Fua · 2009
Earlier work this paper cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2010
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2011
Earlier work this paper cites.
Unsupervised feature learning for 3d scene labeling
Kevin Lai, Liefeng Bo, and Dieter Fox · 2014
Earlier work this paper cites.
Joint embeddings of shapes and images via cnn image purification
Yangyan Li, Hao Su, Charles Ruizhongtai Qi, Noa Fish, Daniel Cohen-Or, and Leonidas J Guibas · 2015
Earlier work this paper cites.
Single-image depth perception in the wild
Weifeng Chen, Zhao Fu, Dawei Yang, and Jia Deng · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Scannet: Richly-annotated 3d reconstructions of indoor scenes
Angela Dai, Angel X Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner · 2017
Earlier work this paper cites.
Feature pyramid networks for object detection
Tsung-Yi Lin, Piotr Dollár, Ross Girshick, Kaiming He, Bharath Hariharan, and Serge Belongie · 2017
Earlier work this paper cites.
1 year, 1000 km: The oxford robotcar dataset
Will Maddern, Geoffrey Pascoe, Chris Linegar, and Paul Newman · 2017
Earlier work this paper cites.
Sparsity invariant cnns
Jonas Uhrig, Nick Schneider, Lukas Schneider, Uwe Franke, Thomas Brox, and Andreas Geiger · 2017
Earlier work this paper cites.
3dmatch: Learning local geometric descriptors from rgb-d reconstructions
Andy Zeng, Shuran Song, Matthias Nießner, Matthew Fisher, Jianxiong Xiao, and Thomas Funkhouser · 2017
Earlier work this paper cites.
Superpoint: Self-supervised interest point detection and description
Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2018
Earlier work this paper cites.
Difnet: Semantic segmentation by diffusion networks
Peng Jiang, Fanglin Gu, Yunhai Wang, Changhe Tu, and Baoquan Chen · 2018
Earlier work this paper cites.
In defense of classical image processing: Fast depth completion on the cpu
Jason Ku, Ali Harakeh, and Steven L Waslander · 2018
Earlier work this paper cites.
Monocular relative depth perception with web stereo data supervision
Ke Xian, Chunhua Shen, Zhiguo Cao, Hao Lu, Yang Xiao, Ruibo Li, and Zhenbo Luo · 2018
Earlier work this paper cites.
3dtnet: Learning local features using 2d and 3d cues
Xiaoxia Xing, Yinghao Cai, Tao Lu, Shaojun Cai, Yiping Yang, and Dayong Wen · 2018
Earlier work this paper cites.
Open3d: A modern library for 3d data processing
Qian-Yi Zhou, Jaesik Park, and Vladlen Koltun · 2018
Earlier work this paper cites.
The alignment of the spheres: Globally-optimal spherical mixture alignment for camera pose estimation
Dylan Campbell, Lars Petersson, Laurent Kneip, Hongdong Li, and Stephen Gould · 2019
Earlier work this paper cites.
Fully convolutional geometric features
Christopher Choy, Jaesik Park, and Vladlen Koltun · 2019
Earlier work this paper cites.
2d3d-matchnet: Learning to match keypoints across 2d image and 3d point cloud
Mengdan Feng, Sixing Hu, Marcelo H Ang, and Gim Hee Lee · 2019
Earlier work this paper cites.
Generative modeling by estimating gradients of the data distribution
Yang Song and Stefano Ermon · 2019
Earlier work this paper cites.
Enforcing geometric constraints of virtual normal for depth prediction
Wei Yin, Yifan Liu, Chunhua Shen, and Youliang Yan · 2019
Earlier work this paper cites.
Mapillary planet-scale depth dataset
Manuel López Antequera, Pau Gargallo, Markus Hofinger, Samuel Rota Bulò, Yubin Kuang, and Peter Kontschieder · 2020
Earlier work this paper cites.
Unsupervised multi-modal image registration via geometry preserving image-to-image translation
Moab Arar, Yiftach Ginger, Dov Danon, Amit H Bermano, and Daniel Cohen-Or · 2020
Cited alongside, same era.
Oasis: A large-scale dataset for single image 3d in the wild
Weifeng Chen, Shengyi Qian, David Fan, Noriyuki Kojima, Max Hamilton, and Jia Deng · 2020
Cited alongside, same era.
Registration of large-scale terrestrial laser scanner point clouds: A review and benchmark
Zhen Dong, Fuxun Liang, Bisheng Yang, Yusheng Xu, Yufu Zang, Jianping Li, Yuan Wang, Wenxia Dai, Hongchao Fan, Juha Hyyppä, et al · 2020
Cited alongside, same era.
Deep learning for 3d point clouds: A survey
Yulan Guo, Hanyun Wang, Qingyong Hu, Hao Liu, Li Liu, and Mohammed Bennamoun · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Learning 2d-3d correspondences to solve the blind perspective-n-point problem
Geometric transformer for fast and robust point cloud registration
Zheng Qin, Hao Yu, Changjian Wang, Yulan Guo, Yuxing Peng, and Kai Xu · 2022
Later among the works it cites.
Corri2p: Deep image-to-point cloud registration via dense correspondence
Siyu Ren, Yiming Zeng, Junhui Hou, and Xiaodong Chen · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, et al · 2022
Later among the works it cites.
Semantic diffusion network for semantic segmentation
Haoru Tan, Sitong Wu, and Jimin Pi · 2022
Later among the works it cites.
Diffusion models for implicit image segmentation ensembles
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Liu Liu, Dylan Campbell, Hongdong Li, Dingfu Zhou, Xibin Song, and Ruigang Yang · 2020
Cited alongside, same era.
Lcd: Learned cross-domain descriptors for 2d-3d matching
Quang-Hieu Pham, Mikaela Angelina Uy, Binh-Son Hua, Duc Thanh Nguyen, Gemma Roig, and Sai-Kit Yeung · 2020
Cited alongside, same era.
Superglue: Learning feature matching with graph neural networks
Paul-Edouard Sarlin, Daniel DeTone, Tomasz Malisiewicz, and Andrew Rabinovich · 2020
Cited alongside, same era.
Circle loss: A unified perspective of pair similarity optimization
Yifan Sun, Changmao Cheng, Yuhan Zhang, Chi Zhang, Liang Zheng, Zhongdao Wang, and Yichen Wei · 2020
Cited alongside, same era.
Structure-guided ranking loss for single image depth prediction
Ke Xian, Jianming Zhang, Oliver Wang, Long Mai, Zhe Lin, and Zhiguo Cao · 2020
Cited alongside, same era.
Segdiff: Image segmentation with diffusion probabilistic models
Tomer Amit, Tal Shaharbany, Eliya Nachmani, and Lior Wolf · 2021
Cited alongside, same era.
Label-efficient semantic segmentation with diffusion models
Dmitry Baranchuk, Ivan Rubachev, Andrey Voynov, Valentin Khrulkov, and Artem Babenko · 2021
Cited alongside, same era.
Julia Wolleb, Robin Sandkühler, Florentin Bieder, Philippe Valmaggia, and Philippe C Cattin · 2022
Later among the works it cites.
New crfs: Neural window fully-connected crfs for monocular depth estimation
Weihao Yuan, Xiaodong Gu, Zuozhuo Dai, Siyu Zhu, and Ping Tan · 2022
Later among the works it cites.
Nice-slam: Neural implicit scalable encoding for slam
Zihan Zhu, Songyou Peng, Viktor Larsson, Weiwei Xu, Hujun Bao, Zhaopeng Cui, Martin R Oswald, and Marc Pollefeys · 2022
Later among the works it cites.
Zoedepth: Zero-shot transfer by combining relative and metric depth
Shariq Farooq Bhat, Reiner Birkl, Diana Wofk, Peter Wonka, and Matthias Müller · 2023
Closest in time.
Diffusiondepth: Diffusion denoising approach for monocular depth estimation
Yiqun Duan, Xianda Guo, and Zheng Zhu · 2023
Closest in time.
Towards zero-shot scale-aware monocular depth estimation
Vitor Guizilini, Igor Vasiljevic, Dian Chen, Rares Ambrus, and Adrien Gaidon · 2023
Closest in time.
Unsupervised semantic correspondence using stable diffusion
Eric Hedlin, Gopal Sharma, Shweta Mahajan, Hossam Isack, Abhishek Kar, Andrea Tagliasacchi, and Kwang Moo Yi · 2023
Closest in time.
Ep2p-loc: End-to-end 3d point to 2d pixel localization for large-scale visual localization
Minjung Kim, Junseo Koo, and Gunhee Kim · 2023
Closest in time.
2d3d-matr: 2d-3d matching transformer for detection-free registration between images and point clouds
Minhao Li, Zheng Qin, Zhirui Guo, Renjiao Yi, Chengyang Zhu, and Kai Xu · 2023
Closest in time.
Se-calib: Semantic edges based lidar-camera boresight online calibration in urban scenes
Youqi Liao, Jianping Li, Shuhao Kang, Qiang Li, Guifang Zhu, Shenghai Yuan, Zhen Dong, and Bisheng Yang · 2023
Closest in time.
Syncdreamer: Generating multiview-consistent images from a single-view image
Yuan Liu, Cheng Lin, Zijiao Zeng, Xiaoxiao Long, Lingjie Liu, Taku Komura, and Wenping Wang · 2023
Closest in time.
Chong Mou, Xintao Wang, Liangbin Xie, Jian Zhang, Zhongang Qi, Ying Shan, and Xiaohu Qie · 2023
Closest in time.
Dinov2: Learning robust visual features without supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, et al · 2023
Closest in time.
Orienternet: Visual localization in 2d public maps with neural matching
Paul-Edouard Sarlin, Daniel DeTone, Tsun-Yi Yang, Armen Avetisyan, Julian Straub, Tomasz Malisiewicz, Samuel Rota Bulò, Richard Newcombe, Peter Kontschieder, and Vasileios Balntas · 2023
Closest in time.
Emergent correspondence from image diffusion
Luming Tang, Menglin Jia, Qianqian Wang, Cheng Perng Phoo, and Bharath Hariharan · 2023
Closest in time.
Plug-and-play diffusion features for text-driven image-to-image translation
Narek Tumanyan, Michal Geyer, Shai Bagon, and Tali Dekel · 2023
Closest in time.
Argoverse 2: Next generation datasets for self-driving perception and forecasting
Benjamin Wilson, William Qi, Tanmay Agarwal, John Lambert, Jagjeet Singh, Siddhesh Khandelwal, Bowen Pan, Ratnesh Kumar, Andrew Hartnett, Jhony Kaesemodel Pontes, et al · 2023
Closest in time.
Cfi2p: Coarse-to-fine cross-modal correspondence learning for image-to-point cloud registration
Gongxin Yao, Yixin Xuan, Yiwei Chen, and Yu Pan · 2023
Closest in time.
Metric3d: Towards zero-shot metric 3d prediction from a single image
Wei Yin, Chi Zhang, Hao Chen, Zhipeng Cai, Gang Yu, Kaixuan Wang, Xiaozhi Chen, and Chunhua Shen · 2023
Closest in time.
A tale of two features: Stable diffusion complements dino for zero-shot semantic correspondence
Junyi Zhang, Charles Herrmann, Junhwa Hur, Luisa Polania Cabrera, Varun Jampani, Deqing Sun, and Ming-Hsuan Yang · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang and Maneesh Agrawala · 2023
Closest in time.
Differentiable registration of images and lidar point clouds with voxelpoint-to-pixel matching
Junsheng Zhou, Baorui Ma, Wenyuan Zhang, Yi Fang, Yu-Shen Liu, and Zhizhong Han · 2023
Closest in time.