Fetching the paper…
Reading the bibliography…
The ability to associate touch with sight is essential for tasks that require physically interacting with objects in the world.
Hand movements: A window into haptic object recognition
Susan J Lederman and Roberta L Klatzky · 1987
Earlier work this paper cites.
Learning classification with unlabeled data
Virginia R de Sa · 1994
Earlier work this paper cites.
On seeing stuff: the perception of materials by humans and machines
Edward H Adelson · 2001
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, Alan C Bovik, Hamid R Sheikh, and Eero P Simoncelli · 2004
Earlier work this paper cites.
Realistic materials in computer graphics
Hendrik PA Lensch, Michael Goesele, Yung-Yu Chuang, Tim Hawkins, Steve Marschner, Wojciech Matusik, and Gero Mueller · 2005
Earlier work this paper cites.
Force and Tactile Sensors
Mark R. Cutkosky, Robert D. Howe, and William R. Provancher · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Retrographic sensing for the measurement of surface texture and shape
Micah K Johnson and Edward H Adelson · 2009
Earlier work this paper cites.
Microgeometry capture using an elastomeric sensor
Micah Johnson, Forrester Cole, Alvin Raj, and Edward Adelson · 2011
Earlier work this paper cites.
Multimodal deep learning
Jiquan Ngiam, Aditya Khosla, Mingyu Kim, Juhan Nam, Honglak Lee, and Andrew Y Ng · 2011
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Robotic learning of haptic adjectives through physical interaction
Vivian Chu, Ian McMahon, Lorenzo Riano, Craig G McDonald, Qin He, Jorge Martinez Perez-Tejada, Michael Arrigo, Trevor Darrell, and Katherine J Kuchenbecker · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Earlier work this paper cites.
Tactile sensing in dexterous robot hands — review
Zhanat Kappassov, Juan-Antonio Corrales, and Véronique Perdereau · 2015
Earlier work this paper cites.
Spatio-temporal video autoencoder with differentiable memory
Viorica Patraucean, Ankur Handa, and Roberto Cipolla · 2015
Earlier work this paper cites.
Proton: A visuo-haptic data acquisition system for robotic learning of surface properties
Alex Burka, Siyao Hu, Stuart Helgeson, Shweta Krishnan, Yang Gao, Lisa Anne Hendricks, Trevor Darrell, and Katherine J Kuchenbecker · 2016
Earlier work this paper cites.
Unsupervised learning for physical interaction through video prediction
Chelsea Finn, Ian Goodfellow, and Sergey Levine · 2016
Earlier work this paper cites.
Perceptual losses for real-time style transfer and super-resolution, 2016
Justin Johnson, Alexandre Alahi, and Li Fei-Fei · 2016
Earlier work this paper cites.
Learning visual features from large weakly supervised data
Armand Joulin, Laurens van der Maaten, Allan Jabri, and Nicolas Vasilache · 2016
Earlier work this paper cites.
Visually indicated sounds
Andrew Owens, Phillip Isola, Josh McDermott, Antonio Torralba, Edward H. Adelson, and William T. Freeman · 2016
Earlier work this paper cites.
Ambient sound provides supervision for visual learning
Andrew Owens, Jiajun Wu, Josh H McDermott, William T Freeman, and Antonio Torralba · 2016
Earlier work this paper cites.
Generative adversarial text to image synthesis
Scott Reed, Zeynep Akata, Xinchen Yan, Lajanugen Logeswaran, Bernt Schiele, and Honglak Lee · 2016
Earlier work this paper cites.
Look, listen and learn
Relja Arandjelovic and Andrew Zisserman · 2017
Earlier work this paper cites.
Stochastic variational video prediction
Mohammad Babaeizadeh, Chelsea Finn, Dumitru Erhan, Roy H Campbell, and Sergey Levine · 2017
Earlier work this paper cites.
Proton 2: Increasing the sensitivity and portability of a visuo-haptic surface interaction recorder
Alex Burka, Abhinav Rajvanshi, Sarah Allen, and Katherine J Kuchenbecker · 2017
Earlier work this paper cites.
The feeling of success: Does touch sensing help predict grasp outcomes?
Roberto Calandra, Andrew Owens, Manu Upadhyaya, Wenzhen Yuan, Justin Lin, Edward H Adelson, and Sergey Levine · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros · 2017
Cited alongside, same era.
Deep predictive coding networks for video prediction and unsupervised learning
William Lotter, G. Kreiman, and David D. Cox · 2017
Cited alongside, same era.
High-resolution image synthesis and semantic manipulation with conditional gans, 2017
Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu, Andrew Tao, Jan Kautz, and Bryan Catanzaro · 2017
Cited alongside, same era.
Gelsight: High-resolution robot tactile sensors for estimating geometry and force
Wenzhen Yuan, Siyuan Dong, and Edward Adelson · 2017
Cited alongside, same era.
Gelsight: High-resolution robot tactile sensors for estimating geometry and force
Wenzhen Yuan, Siyuan Dong, and Edward H Adelson · 2017
Cited alongside, same era.
Manipulation by feel: Touch-based control with deep predictive models
Stephen Tian, Frederik Ebert, Dinesh Jayaraman, Mayur Mudigonda, Chelsea Finn, Roberto Calandra, and Sergey Levine · 2019
Later among the works it cites.
High fidelity video prediction with large stochastic recurrent neural networks, 2019
Ruben Villegas, Arkanath Pathak, Harini Kannan, Dumitru Erhan, Quoc V. Le, and Honglak Lee · 2019
Later among the works it cites.
Few-shot video-to-video synthesis
Ting-Chun Wang, Ming-Yu Liu, Andrew Tao, Guilin Liu, Jan Kautz, and Bryan Catanzaro · 2019
Later among the works it cites.
Sound2sight: Generating visual dynamics from sound and context
Moitreya Chatterjee and Anoop Cherian · 2020
Later among the works it cites.
Self-attention based visual-tactile fusion learning for predicting grasp outcomes
Shaowei Cui, Rui Wang, Junhang Wei, Jingyi Hu, and Shuo Wang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Shape-independent hardness estimation using deep learning and a gelsight tactile sensor
Wenzhen Yuan, Chenzhuo Zhu, Andrew Owens, Mandayam A Srinivasan, and Edward H Adelson · 2017
Cited alongside, same era.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros · 2017
Cited alongside, same era.
Instrumentation, data, and algorithms for visually understanding haptic surface properties
Alexander L Burka · 2018
Cited alongside, same era.
More than a feeling: Learning to grasp and regrasp using vision and touch
Roberto Calandra, Andrew Owens, Dinesh Jayaraman, Justin Lin, Wenzhen Yuan, Jitendra Malik, Edward H. Adelson, and Sergey Levine · 2018
Cited alongside, same era.
Stochastic video generation with a learned prior
Emily L. Denton and Rob Fergus · 2018
Cited alongside, same era.
Image generation from scene graphs
Justin Johnson, Agrim Gupta, and Li Fei-Fei · 2018
Cited alongside, same era.
Co-training of audio and video representations from self-supervised temporal synchronization
Bruno Korbar, Du Tran, and Lorenzo Torresani · 2018
Cited alongside, same era.
Materialgan: Reflectance capture using a generative svbrdf model
Yu Guo, Cameron Smith, Miloš Hašan, Kalyan Sunkavalli, and Shuang Zhao · 2020
Later among the works it cites.
Digit: A novel design for a low-cost compact high-resolution tactile sensor with application to in-hand manipulation
Mike Lambeta, Po wei Chou, Stephen Tian, Brian Yang, Benjamin Maloon, Victoria Rose Most, Dave Stroud, Raymond Santos, Ahmad Byagowi, Gregg Kammerer, Dinesh Jayaraman, and Roberto Calandra · 2020
Later among the works it cites.
Contrastive learning for conditional image synthesis
Taesung Park, Alexei A. Efros, Richard Zhang, and Jun-Yan Zhu · 2020
Later among the works it cites.
Grasping in the wild: Learning 6dof closed-loop grasping from low-cost demonstrations
Shuran Song, Andy Zeng, Johnny Lee, and Thomas Funkhouser · 2020
Later among the works it cites.
Contrastive multiview coding
Yonglong Tian, Dilip Krishnan, and Phillip Isola · 2020
Later among the works it cites.
Future video synthesis with object motion prediction
Yue Wu, Rongrong Gao, Jaesik Park, and Qifeng Chen · 2020
Later among the works it cites.
David Bau, Alex Andonian, Audrey Cui, YeonHwan Park, Ali Jahanian, Aude Oliva, and Antonio Torralba · 2021
Later among the works it cites.
Virtex: Learning visual representations from textual annotations
Karan Desai and Justin Johnson · 2021
Later among the works it cites.
Objectfolder: A dataset of objects with implicit visual, auditory, and tactile representations
Ruohan Gao, Yen-Yu Chang, Shivani Mall, Li Fei-Fei, and Jiajun Wu · 2021
Later among the works it cites.
Sound-guided semantic image manipulation
Seung Hyun Lee, Wonseok Roh, Wonmin Byeon, Sang Ho Yoon, Chan Young Kim, Jinkyu Kim, and Sangpil Kim · 2021
Later among the works it cites.
Audio-visual instance discrimination with cross-modal agreement
Pedro Morgado, Nuno Vasconcelos, and Ishan Misra · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Later among the works it cites.
Dynamic modeling of hand-object interactions via tactile sensing
Qiang Zhang, Yunzhu Li, Yiyue Luo, Wan Shou, Michael Foshey, Junchi Yan, Joshua B Tenenbaum, Wojciech Matusik, and Antonio Torralba · 2021
Later among the works it cites.
Adversarial single-image svbrdf estimation with hybrid training
Xilong Zhou and Nima Khademi Kalantari · 2021
Later among the works it cites.
Objectfolder 2.0: A multisensory object dataset for sim2real transfer
Ruohan Gao, Zilin Si, Yen-Yu Chang, Samuel Clarke, Jeannette Bohg, Li Fei-Fei, Wenzhen Yuan, and Jiajun Wu · 2022
Closest in time.
Comparing correspondences: Video prediction with correspondence-wise losses
Daniel Geng, Max Hamilton, and Andrew Owens · 2022
Closest in time.
Matformer: A generative model for procedural materials
Paul Guerrero, Milos Hasan, Kalyan Sunkavalli, Radomir Mech, Tamy Boubekeur, and Niloy Mitra · 2022
Closest in time.
Learning visual styles from audio-visual associations
Tingle Li, Yichen Liu, Andrew Owens, and Hang Zhao · 2022
Closest in time.
Tacto: A fast, flexible, and open-source simulator for high-resolution vision-based tactile sensors
Shaoxiong Wang, Mike Lambeta, Po-Wei Chou, and Roberto Calandra · 2022
Closest in time.
Virdo: Visio-tactile implicit representations of deformable objects
Youngsun Wi, Pete Florence, Andy Zeng, and Nima Fazeli · 2022
Closest in time.