Fetching the paper…
Reading the bibliography…
Tactility provides crucial support and enhancement for the perception and interaction capabilities of both humans and robots.
Tactile sensing—from humans to humanoids
Ravinder S Dahiya, Giorgio Metta, Maurizio Valle, and Giulio Sandini. 2009 · 2009
Earlier work this paper cites.
Coding and use of tactile signals from the fingertips in object manipulation tasks
Roland S Johansson and J Randall Flanagan. 2009 · 2009
Earlier work this paper cites.
Microgeometry capture using an elastomeric sensor
Micah K Johnson, Forrester Cole, Alvin Raj, and Edward H Adelson. 2011 · 2011
Earlier work this paper cites.
Moving object detection using keypoints reference model
Wan Mimi Diyana Wan Zaki, Aini Hussain, and Mohamed Hedayati. 2011 · 2011
Earlier work this paper cites.
The feeling of success: Does touch sensing help predict grasp outcomes?
Roberto Calandra, Andrew Owens, Manu Upadhyaya, Wenzhen Yuan, Justin Lin, Edward H Adelson, and Sergey Levine. 2017 · 2017
Earlier work this paper cites.
Improved gelsight tactile sensor for measuring geometry and slip
Siyuan Dong, Wenzhen Yuan, and Edward H Adelson. 2017 · 2017
Earlier work this paper cites.
Shape-independent hardness estimation using deep learning and a gelsight tactile sensor
Wenzhen Yuan, Chenzhuo Zhu, Andrew Owens, Mandayam A Srinivasan, and Edward H Adelson. 2017b · 2017
Earlier work this paper cites.
See, feel, act: Hierarchical learning for complex manipulation skills with multisensory fusion
Nima Fazeli, Miquel Oller, Jiajun Wu, Zheng Wu, Joshua B Tenenbaum, and Alberto Rodriguez. 2019 · 2019
Earlier work this paper cites.
Connecting touch and vision via cross-modal prediction
Yunzhu Li, Jun-Yan Zhu, Russ Tedrake, and Antonio Torralba. 2019 · 2019
Earlier work this paper cites.
Learning to identify object instances by touch: Tactile recognition via multimodal matching
Justin Lin, Roberto Calandra, and Sergey Levine. 2019 · 2019
Earlier work this paper cites.
Self-attention based visual-tactile fusion learning for predicting grasp outcomes
Shaowei Cui, Rui Wang, Junhang Wei, Jingyi Hu, and Shuo Wang. 2020 · 2020
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al. 2020 · 2020
Earlier work this paper cites.
Supervised autoencoder joint learning on heterogeneous tactile sensory data: Improving material classification performance
Ruihan Gao, Tasbolat Taunyazov, Zhiping Lin, and Yan Wu. 2020 · 2020
Earlier work this paper cites.
Digit: A novel design for a low-cost compact high-resolution tactile sensor with application to in-hand manipulation
Mike Lambeta, Po-Wei Chou, Stephen Tian, Brian Yang, Benjamin Maloon, Victoria Rose Most, Dave Stroud, Raymond Santos, Ahmad Byagowi, Gregg Kammerer, et al. 2020 · 2020
Earlier work this paper cites.
Objectfolder: A dataset of objects with implicit visual, auditory, and tactile representations
Ruohan Gao, Yen-Yu Chang, Shivani Mall, Li Fei-Fei, and Jiajun Wu. 2021 · 2021
Earlier work this paper cites.
Generation of gelsight tactile images for sim2real learning
Daniel Fernandes Gomes, Paolo Paoletti, and Shan Luo. 2021 · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J Hu, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, Weizhu Chen, et al. 2021 · 2021
Cited alongside, same era.
Openclip
Gabriel Ilharco, Mitchell Wortsman, Ross Wightman, Cade Gordon, Nicholas Carlini, Rohan Taori, Achal Dave, Vaishaal Shankar, Hongseok Namkoong, John Miller, Hannaneh Hajishirzi, Ali Farhadi, and Ludwig Schmidt. 2021 · 2021
Cited alongside, same era.
Scaling up visual and vision-language representation learning with noisy text supervision
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc Le, Yun-Hsuan Sung, Zhen Li, and Tom Duerig. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021 · 2021
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Later among the works it cites.
iquery: Instruments as queries for audio-visual sound separation
Jiaben Chen, Renrui Zhang, Dongze Lian, Jiaqi Yang, Ziyao Zeng, and Jianbo Shi. 2023 · 2023
Later among the works it cites.
The objectfolder benchmark: Multisensory learning with neural and real objects
Ruohan Gao, Yiming Dou, Hao Li, Tanmay Agarwal, Jeannette Bohg, Yunzhu Li, Li Fei-Fei, and Jiajun Wu. 2023 · 2023
Later among the works it cites.
Imagebind: One embedding space to bind them all
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, and Ishan Misra. 2023 · 2023
Later among the works it cites.
Ziyu Guo, Renrui Zhang, Xiangyang Zhu, Yiwen Tang, Xianzheng Ma, Jiaming Han, Kexin Chen, Peng Gao, Xianzhi Li, Hongsheng Li, et al. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Flamingo: a visual language model for few-shot learning
Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, Arthur Mensch, Katherine Millican, Malcolm Reynolds, et al. 2022 · 2022
Cited alongside, same era.
Objectfolder 2.0: A multisensory object dataset for sim2real transfer
Ruohan Gao, Zilin Si, Yen-Yu Chang, Samuel Clarke, Jeannette Bohg, Li Fei-Fei, Wenzhen Yuan, and Jiajun Wu. 2022 · 2022
Cited alongside, same era.
Audioclip: Extending clip to image, text and audio
Andrey Guzhov, Federico Raue, Jörn Hees, and Andreas Dengel. 2022 · 2022
Cited alongside, same era.
Visuotactile-rl: Learning multimodal manipulation policies with deep reinforcement learning
Johanna Hansen, Francois Hogan, Dmitriy Rivkin, David Meger, Michael Jenkin, and Gregory Dudek. 2022 · 2022
Cited alongside, same era.
Self-supervised visuo-tactile pretraining to locate and follow garment features
Justin Kerr, Huang Huang, Albert Wilcox, Ryan Hoque, Jeffrey Ichnowski, Roberto Calandra, and Ken Goldberg. 2022 · 2022
Cited alongside, same era.
LAION-5b: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade W Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, Patrick Schramowski, Srivatsa R Kundurthy, Katherine Crowson, Ludwig Schmidt, Robert Kaczmarczyk, and Jenia Jitsev. 2022 · 2022
Cited alongside, same era.
Taxim: An example-based simulation model for gelsight tactile sensors
Zilin Si and Wenzhen Yuan. 2022 · 2022
Cited alongside, same era.
Later among the works it cites.
Vit-lens-2: Gateway to omni-modal intelligence
Weixian Lei, Yixiao Ge, Kun Yi, Jianfeng Zhang, Difei Gao, Dylan Sun, Yuying Ge, Ying Shan, and Mike Zheng Shou. 2023 · 2023
Later among the works it cites.
General in-hand object rotation with vision and touch
Haozhi Qi, Brent Yi, Sudharshan Suresh, Mike Lambeta, Yi Ma, Roberto Calandra, and Jitendra Malik. 2023 · 2023
Later among the works it cites.
Midastouch: Monte-carlo inference over distributions across sliding touch
Sudharshan Suresh, Zilin Si, Stuart Anderson, Michael Kaess, and Mustafa Mukadam. 2023 · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al. 2023 · 2023
Later among the works it cites.
Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding
Le Xue, Mingfei Gao, Chen Xing, Roberto Martín-Martín, Jiajun Wu, Caiming Xiong, Ran Xu, Juan Carlos Niebles, and Silvio Savarese. 2023 · 2023
Later among the works it cites.
Generating visual scenes from touch
Fengyu Yang, Jiacheng Zhang, and Andrew Owens. 2023 · 2023
Later among the works it cites.
Bin Zhu, Bin Lin, Munan Ning, Yang Yan, Jiaxi Cui, WANG HongFa, Yatian Pang, Wenhao Jiang, Junwu Zhang, Zongwei Li, et al. 2023 · 2023
Later among the works it cites.
Multimodal visual-tactile representation learning through self-supervised contrastive pre-training
Vedant Dave, Fotios Lygerakis, and Elmar Rückert. 2024 · 2024
Closest in time.
Openshape: Scaling up 3d shape representation towards open-world understanding
Minghua Liu, Ruoxi Shi, Kaiming Kuang, Yinhao Zhu, Xuanlin Li, Shizhong Han, Hong Cai, Fatih Porikli, and Hao Su. 2024 · 2024
Closest in time.