Fetching the paper…
Reading the bibliography…
The ability to associate touch with other modalities has huge implications for humans and computational systems.
Relations between two sets of variates
Harold Hotelling · 1936
Earlier work this paper cites.
Hand movements: A window into haptic object recognition
Susan J. Lederman and Roberta L. Klatzky · 1987
Earlier work this paper cites.
The sense of touch
Paul R Manske · 1999
Earlier work this paper cites.
Canonical partial least squares and continuum power regression
Sijmen de Jong, Barry M. Wise, and N. L. Ricker · 2001
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
Linda Smith and Michael Gasser · 2005
Earlier work this paper cites.
Force and tactile sensors
Mark R. Cutkosky, Robert D. Howe, and William R. Provancher · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Retrographic sensing for the measurement of surface texture and shape
Micah K Johnson and Edward H Adelson · 2009
Earlier work this paper cites.
Tutorial review haptic perception: A tutorial
Susan J. Lederman and R. L. Klatzky · 2009
Earlier work this paper cites.
Microgeometry capture using an elastomeric sensor
Micah K. Johnson, Forrester Cole, Alvin Raj, and Edward H. Adelson · 2011
Earlier work this paper cites.
Tactile sensing in dexterous robot hands - review
Zhanat Kappasov, Juan Antonio Corrales, and Véronique Perdereau · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Touch: The science of the hand, heart, and mind
David J Linden · 2016
Earlier work this paper cites.
See, hear, and read: Deep aligned representations
Yusuf Aytar, Carl Vondrick, and Antonio Torralba · 2017
Earlier work this paper cites.
The feeling of success: Does touch sensing help predict grasp outcomes?
Roberto Calandra, Andrew Owens, Manu Upadhyaya, Wenzhen Yuan, Justin Lin, Edward H Adelson, and Sergey Levine · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros · 2017
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Connecting look and feel: Associating the visual and tactile properties of physical materials
Wenzhen Yuan, Shaoxiong Wang, Siyuan Dong, and Edward H. Adelson · 2017
Earlier work this paper cites.
More than a feeling: Learning to grasp and regrasp using vision and touch
Roberto Calandra, Andrew Owens, Dinesh Jayaraman, Justin Lin, Wenzhen Yuan, Jitendra Malik, Edward H. Adelson, and Sergey Levine · 2018
Earlier work this paper cites.
Vitac: Feature sharing between vision and tactile sensing for cloth texture recognition
Shan Luo, Wenzhen Yuan, Edward H. Adelson, Anthony G. Cohn, and Raul Fuentes · 2018
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals · 2018
Earlier work this paper cites.
Learning sight from sound: Ambient sound provides supervision for visual learning
Andrew Owens, Jiajun Wu, Josh H McDermott, William T Freeman, and Antonio Torralba · 2018
Earlier work this paper cites.
Unsupervised feature learning via non-parametric instance discrimination
Zhirong Wu, Yuanjun Xiong, Stella X Yu, and Dahua Lin · 2018
Earlier work this paper cites.
Learning an action-conditional model for haptic texture generation
Negin Heravi, Wenzhen Yuan, Allison M. Okamura, and Jeannette Bohg · 2019
Earlier work this paper cites.
Why is there so much more research on vision than on any other sensory modality?
Fabian Hutmacher · 2019
Earlier work this paper cites.
Making sense of vision and touch: Learning multimodal representations for contact-rich tasks
Michelle A. Lee, Yuke Zhu, Peter Zachares, Matthew Tan, Krishna Parasuram Srinivasan, Silvio Savarese, Fei-Fei Li, Animesh Garg, and Jeannette Bohg · 2019
Earlier work this paper cites.
Connecting touch and vision via cross-modal prediction
Yunzhu Li, Jun-Yan Zhu, Russ Tedrake, and Antonio Torralba · 2019
Earlier work this paper cites.
Learning to identify object instances by touch: Tactile recognition via multimodal matching
Justin Lin, Roberto Calandra, and Sergey Levine · 2019
Earlier work this paper cites.
Real-time soft body 3d proprioception via deep vision-based sensing
Ruoyu Wang, Shiheng Wang, Songyu Du, Erdong Xiao, Wenzhen Yuan, and Chen Feng · 2019
Earlier work this paper cites.
Deep supervised cross-modal retrieval
Liangli Zhen, Peng Hu, Xu Wang, and Dezhong Peng · 2019
Earlier work this paper cites.
Simulation of vision-based tactile sensors using physics based rendering
Arpit Agarwal, Tim Man, and Wenzhen Yuan · 2020
Earlier work this paper cites.
Labelling unlabelled videos from scratch with multi-modal self-supervision
Yuki Asano, Mandela Patrick, Christian Rupprecht, and Andrea Vedaldi · 2020
Earlier work this paper cites.
Spatio-temporal attention model for tactile texture recognition
Guanqun Cao, Yi Zhou, Danushka Bollegala, and Shan Luo · 2020
Earlier work this paper cites.
Supervised autoencoder joint learning on heterogeneous tactile sensory data: Improving material classification performance
Ruihan Gao, Tasbolat Taunyazov, Zhiping Lin, and Y. Wu · 2020
Earlier work this paper cites.
Digit: A novel design for a low-cost compact high-resolution tactile sensor with application to in-hand manipulation
Mike Lambeta, Po wei Chou, Stephen Tian, Brian Yang, Benjamin Maloon, Victoria Rose Most, Dave Stroud, Raymond Santos, Ahmad Byagowi, Gregg Kammerer, Dinesh Jayaraman, and Roberto Calandra · 2020
Earlier work this paper cites.
Goal-driven robotic pushing using tactile and proprioceptive feedback
John Lloyd and Nathan F. Lepora · 2020
Earlier work this paper cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2020
Cited alongside, same era.
Fast texture classification using tactile neural coding and spiking neural network
Tasbolat Taunyazov, Yansong Chua, Ruihan Gao, Harold Soh, and Y. Wu · 2020
Cited alongside, same era.
Contrastive multiview coding
Yonglong Tian, Dilip Krishnan, and Phillip Isola · 2020
Cited alongside, same era.
Tacto: A fast, flexible, and open-source simulator for high-resolution vision-based tactile sensors
Shaoxiong Wang, Mike Lambeta, Po wei Chou, and Roberto Calandra · 2020
Cited alongside, same era.
Multimodal perception for dexterous manipulation
Guanqun Cao and Shan Luo · 2021
Cited alongside, same era.
Towards learning to play piano with dexterous hands and touch
Huazhe Xu, Yuping Luo, Shaoxiong Wang, Trevor Darrell, and Roberto Calandra · 2022
Later among the works it cites.
Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding
Le Xue, Mingfei Gao, Chen Xing, Roberto Mart’in-Mart’in, Jiajun Wu, Caiming Xiong, Ran Xu, Juan Carlos Niebles, and Silvio Savarese · 2022
Later among the works it cites.
Sparse and complete latent organization for geospatial semantic segmentation
Fengyu Yang and Chenyang Ma · 2022
Later among the works it cites.
Unified contrastive learning in image-text-label space
Jianwei Yang, Chunyuan Li, Pengchuan Zhang, Bin Xiao, Ce Liu, Lu Yuan, and Jianfeng Gao · 2022
Later among the works it cites.
Xihang Yu, Sangli Teng, Theodor Chakhachiro, Wenzhe Tong, Tingjun Li, Tzu-Yuan Lin, Sarah Koehler, Manuel Ahumada, Jeffrey M Walls, and Maani Ghaffari · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Guanqun Cao, Jiaqi Jiang, Chen Lu, Daniel Fernandes Gomes, and Shan Luo · 2021
Cited alongside, same era.
Tactile sim-to-real policy transfer via real-to-sim image translation
Alex Church, John Lloyd, Raia Hadsell, and Nathan F. Lepora · 2021
Cited alongside, same era.
Virtex: Learning visual representations from textual annotations
Karan Desai and Justin Johnson · 2021
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Cited alongside, same era.
On explainability and sensor-adaptability of a robot tactile texture representation using a two-stage recurrent networks
Ruihan Gao, Tian Tian, Zhiping Lin, and Y. Wu · 2021
Cited alongside, same era.
Tactile image-to-image disentanglement of contact geometry from motion-induced shear
Anupam K. Gupta, Laurence Aitchison, and Nathan F. Lepora · 2021
Cited alongside, same era.
Robotic perception of object properties using tactile sensing
Jiaqi Jiang and Shan Luo · 2021
Cited alongside, same era.
Later among the works it cites.
Can language understand depth?
Renrui Zhang, Ziyao Zeng, Ziyu Guo, and Yafeng Li · 2022
Later among the works it cites.
Pointclip v2: Prompting clip and gpt for powerful 3d open-world learning
Xiangyang Zhu, Renrui Zhang, Bowei He, Ziyu Guo, Ziyao Zeng, Zipeng Qin, Shanghang Zhang, and Peng Gao · 2022
Later among the works it cites.
Learning to taste: A multimodal wine dataset
Thoranna Bender, Simon Møe Sørensen, Alireza Kashani, K Eldjarn Hjorleifsson, Grethe Hyldig, Søren Hauberg, Serge Belongie, and Frederik Warburg · 2023
Later among the works it cites.
Visit-bench: A benchmark for vision-language instruction following inspired by real-world use
Yonatan Bitton, Hritik Bansal, Jack Hessel, Rulin Shao, Wanrong Zhu, Anas Awadalla, Josh Gardner, Rohan Taori, and Ludwig Schimdt · 2023
Later among the works it cites.
Vis2hap: Vision-based haptic rendering by cross-modal generation
Guanqun Cao, Jiaqi Jiang, Ningtao Mao, Danushka Bollegala, Min Li, and Shan Luo · 2023
Later among the works it cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, 2023
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E. Gonzalez, Ion Stoica, and Eric P. Xing · 2023
Later among the works it cites.
Instructblip: Towards general-purpose vision-language models with instruction tuning
Wenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong, Junqi Zhao, Weisheng Wang, Boyang Albert Li, Pascale Fung, and Steven C. H. Hoi · 2023
Later among the works it cites.
Conditional generation of audio from video via foley analogies
Yuexi Du, Ziyang Chen, Justin Salamon, Bryan Russell, and Andrew Owens · 2023
Later among the works it cites.
Clap learning audio concepts from natural language supervision
Benjamin Elizalde, Soham Deshmukh, Mahmoud Al Ismail, and Huaming Wang · 2023
Later among the works it cites.
Self-supervised video forensics by audio-visual anomaly detection
Chao Feng, Ziyang Chen, and Andrew Owens · 2023
Later among the works it cites.
Imagebind: One embedding space to bind them all
Rohit Girdhar, Alaaeldin El-Nouby, Zhuang Liu, Mannat Singh, Kalyan Vasudev Alwala, Armand Joulin, and Ishan Misra · 2023
Later among the works it cites.
Daniel Fernandes Gomes, Paolo Paoletti, and Shan Luo · 2023
Later among the works it cites.
Ziyu Guo, Renrui Zhang, Xiangyang Zhu, Yiwen Tang, Xianzheng Ma, Jiaming Han, Kexin Chen, Peng Gao, Xianzhi Li, Hongsheng Li, et al · 2023
Later among the works it cites.
Dexterity from touch: Self-supervised pre-training of tactile representations with robotic play, 2023
Irmak Guzey, Ben Evans, Soumith Chintala, and Lerrel Pinto · 2023
Later among the works it cites.
Learning to read braille: Bridging the tactile reality gap with diffusion models
Carolina Higuera, Byron Boots, and Mustafa Mukadam · 2023
Later among the works it cites.
Learn from incomplete tactile data: Tactile representation learning with masked autoencoders
Jiaqi Jiang, Danushka Bollegala, Shan Luo, et al · 2023
Later among the works it cites.
Self-supervised visuo-tactile pretraining to locate and follow garment features
Justin Kerr, Huang Huang, Albert Wilcox, Ryan Hoque, Jeffrey Ichnowski, Roberto Calandra, and Ken Goldberg · 2023
Later among the works it cites.
In-hand manipulation of unknown objects with tactile sensing for insertion
Marion Lepert, Chaoyi Pan, Shenli Yuan, Rika Antonova, and Jeannette Bohg · 2023
Later among the works it cites.
Improved baselines with visual instruction tuning
Haotian Liu, Chunyuan Li, Yuheng Li, and Yong Jae Lee · 2023
Later among the works it cites.
General in-hand object rotation with vision and touch
Haozhi Qi, Brent Yi, Sudharshan Suresh, Mike Lambeta, Y. Ma, Roberto Calandra, and Jitendra Malik · 2023
Later among the works it cites.
Sound to visual scene generation by audio-to-visual latent alignment
Kim Sung-Bin, Arda Senocak, Hyunwoo Ha, Andrew Owens, and Tae-Hyun Oh · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al · 2023
Later among the works it cites.
Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation
Yusong Wu, Ke Chen, Tianyu Zhang, Yuchen Hui, Taylor Berg-Kirkpatrick, and Shlomo Dubnov · 2023
Later among the works it cites.
Visual-tactile sensing for in-hand object reconstruction
Wenqiang Xu, Zhenjun Yu, Han Xue, Ruolin Ye, Siqiong Yao, and Cewu Lu · 2023
Later among the works it cites.
Generating visual scenes from touch
Fengyu Yang, Jiacheng Zhang, and Andrew Owens · 2023
Later among the works it cites.
Rotating without seeing: Towards in-hand dexterity through touch
Zhao-Heng Yin, Binghao Huang, Yuzhe Qin, Qifeng Chen, and Xiaolong Wang · 2023
Later among the works it cites.
Mimictouch: Learning human’s control strategy with multi-modal tactile feedback
Kelin Yu, Yunhai Han, Matthew Zhu, and Ye Zhao · 2023
Later among the works it cites.
Investigating vision foundational models for tactile representation learning
Ben Zandonati, Ruohan Wang, Ruihan Gao, and Y. Wu · 2023
Later among the works it cites.
Llama-adapter: Efficient fine-tuning of language models with zero-init attention
Renrui Zhang, Jiaming Han, Aojun Zhou, Xiangfei Hu, Shilin Yan, Pan Lu, Hongsheng Li, Peng Gao, and Yu Qiao · 2023
Later among the works it cites.
Exif as language: Learning cross-modal associations between images and camera metadata
Chenhao Zheng, Ayush Shrivastava, and Andrew Owens · 2023
Later among the works it cites.
Touching a nerf: Leveraging neural radiance fields for tactile sensory data generation
Shaohong Zhong, Alessandro Albini, Oiwi Parker Jones, Perla Maiolino, and Ingmar Posner · 2023
Later among the works it cites.