Fetching the paper…
Reading the bibliography…
Manipulation by feel: Touch-based control with deep predictive models
Stephen Tian, Frederik Ebert, Dinesh Jayaraman, Mayur Mudigonda, Chelsea Finn, Roberto Calandra, and Sergey Levine · 1903
Earlier work this paper cites.
PHYRE: A New Benchmark for Physical Reasoning
Anton Bakhtin, Laurens van der Maaten, Justin Johnson, Laura Gustafson, and Ross Girshick · 1908
Earlier work this paper cites.
Clevrer: Collision events for video representation and reasoning
Kexin Yi, Chuang Gan, Yunzhu Li, Pushmeet Kohli, Jiajun Wu, Antonio Torralba, and Joshua B Tenenbaum · 1910
Earlier work this paper cites.
Piqa: Reasoning about physical commonsense in natural language
Yonatan Bisk, Rowan Zellers, Jianfeng Gao, Yejin Choi, et al · 1911
Earlier work this paper cites.
Tactile sensing: New directions, new challenges
Mark H. Lee · 2000
Earlier work this paper cites.
Teaching cameras to feel: Estimating tactile physical properties of surfaces from images
Matthew Purri and Kristin Dana · 2004
Earlier work this paper cites.
Tactile sensing in intelligent robotic manipulation—a review
Johan Tegin and Jan Wikander · 2005
Earlier work this paper cites.
Exploring Relationships between Touch Perception and Surface Physical Properties
X. Chen, Fei Shao, Cathy Barnes, Tom Childs, and Brian Henson · 2009
Earlier work this paper cites.
Event-driven visual-tactile sensing and learning for robots
Tasbolat Taunyazov, Weicong Sng, Hian Hian See, Brian Lim, Jethro Kuan, Abdul Fatir Ansari, Benjamin CK Tee, and Harold Soh · 2009
Earlier work this paper cites.
Tactual perception of material properties
Wouter M. Bergmann Tiest · 2010
Earlier work this paper cites.
Sensing and recognizing surface textures using a gelsight sensor
Rui Li and Edward H. Adelson · 2013
Earlier work this paper cites.
Deep learning for tactile understanding from visual and haptic data
Yang Gao, Lisa Anne Hendricks, Katherine J Kuchenbecker, and Trevor Darrell · 2016
Earlier work this paper cites.
Gaussian error linear units (gelus)
Dan Hendrycks and Kevin Gimpel · 2016
Earlier work this paper cites.
Haptic exploration
Roberta Klatzky and Catherine L Reed · 2016
Earlier work this paper cites.
Physics 101: Learning Physical Object Properties from Unlabeled Videos
Jiajun Wu, Joseph J Lim, Hongyi Zhang, Joshua B Tenenbaum, and William T Freeman · 2016
Earlier work this paper cites.
Estimating object hardness with a gelsight touch sensor
Wenzhen Yuan, Mandayam A. Srinivasan, and Edward H. Adelson · 2016
Earlier work this paper cites.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Exploring Tactile Perceptual Dimensions Using Materials Associated with Sensory Vocabulary
Maki Sakamoto and Junji Watanabe · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Gelsight: High-resolution robot tactile sensors for estimating geometry and force
Wenzhen Yuan, Siyuan Dong, and Edward H. Adelson · 2017
Cited alongside, same era.
Active clothing material perception using tactile sensing and deep learning
Wenzhen Yuan, Yuchen Mo, Shaoxiong Wang, and Edward H Adelson · 2018
Cited alongside, same era.
Deep visuo-tactile learning: Estimation of tactile properties from images
Kuniyuki Takahashi and Jethro Tan · 2019
Cited alongside, same era.
Dark, beyond deep: A paradigm shift to cognitive ai with humanlike common sense
Yixin Zhu, Tao Gao, Lifeng Fan, Siyuan Huang, Mark Edmonds, Hangxin Liu, Feng Gao, Chi Zhang, Siyuan Qi, Ying Nian Wu, et al · 2020
Cited alongside, same era.
Learn from Incomplete Tactile Data: Tactile Representation Learning with Masked Autoencoders
Guanqun Cao, Jiaqi Jiang, Danushka Bollegala, and Shan Luo · 2023
Later among the works it cites.
Minigpt-v2: large language model as a unified interface for vision-language multi-task learning
Jun Chen, Deyao Zhu, Xiaoqian Shen, Xiang Li, Zechun Liu, Pengchuan Zhang, Raghuraman Krishnamoorthi, Vikas Chandra, Yunyang Xiong, and Mohamed Elhoseiny · 2023
Later among the works it cites.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality, March 2023
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E. Gonzalez, Ion Stoica, and Eric P. Xing · 2023
Later among the works it cites.
Ar2-d2: Training a robot without a robot
Jiafei Duan, Yi Ru Wang, Mohit Shridhar, Dieter Fox, and Ranjay Krishna · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Prost: Physical reasoning of objects through space and time
Stéphane Aroca-Ouellette, Cory Paik, Alessandro Roncone, and Katharina Kann · 2021
Cited alongside, same era.
Space: A simulator for physical interactions and causal learning in 3d environments
Jiafei Duan, Samson Yu, and Cheston Tan · 2021
Cited alongside, same era.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Extended tactile perception: Vibration sensing through tools and grasped objects
Tasbolat Taunyazov, Luar Shui Song, Eugene Lim, Hian Hian See, David Lee, Benjamin CK Tee, and Harold Soh · 2021
Cited alongside, same era.
Flamingo: a visual language model for few-shot learning
Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, Arthur Mensch, Katherine Millican, Malcolm Reynolds, et al · 2022
Cited alongside, same era.
Objectfolder 2.0: A multisensory object dataset for sim2real transfer
Ruohan Gao, Zilin Si, Yen-Yu Chang, Samuel Clarke, Jeannette Bohg, Li Fei-Fei, Wenzhen Yuan, and Jiajun Wu · 2022
Cited alongside, same era.
Lei Li, Jingjing Xu, Qingxiu Dong, Ce Zheng, Xu Sun, Lingpeng Kong, and Qi Liu · 2023
Later among the works it cites.
Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Muhammad Maaz, Hanoona Rasheed, Salman Khan, and Fahad Shahbaz Khan · 2023
Later among the works it cites.
Benchmarks for Physical Reasoning AI
Andrew Melnik, Robin Schiewer, Moritz Lange, Andrei Muresanu, Mozhgan Saeidi, Animesh Garg, and Helge Ritter · 2023
Later among the works it cites.
Fine-tuned clip models are efficient video learners
Hanoona Rasheed, Muhammad Uzair Khattak, Muhammad Maaz, Salman Khan, and Fahad Shahbaz Khan · 2023
Later among the works it cites.
Mehmet Saygin Seyfioglu, Wisdom O Ikezogwo, Fatemeh Ghezloo, Ranjay Krishna, and Linda Shapiro · 2023
Later among the works it cites.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al · 2023
Later among the works it cites.
Vary: Scaling up the Vision Vocabulary for Large Vision-Language Models
Haoran Wei, Lingyu Kong, Jinyue Chen, Liang Zhao, Zheng Ge, Jinrong Yang, Jianjian Sun, Chunrui Han, and Xiangyu Zhang · 2023
Later among the works it cites.
Next-gpt: Any-to-any multimodal llm
Shengqiong Wu, Hao Fei, Leigang Qu, Wei Ji, and Tat-Seng Chua · 2023
Later among the works it cites.
Investigating Vision Foundational Models for Tactile Representation Learning
Ben Zandonati, Ruohan Wang, Ruihan Gao, and Yan Wu · 2023
Later among the works it cites.
Egoobjects: A large-scale egocentric dataset for fine-grained object understanding
Chenchen Zhu, Fanyi Xiao, Andrés Alvarado, Yasmine Babaei, Jiabo Hu, Hichem El-Mohri, Sean Chang, Roshan Sumbaly, and Zhicheng Yan · 2023
Later among the works it cites.
A touch, vision, and language dataset for multimodal alignment
Letian Fu, Gaurav Datta, Huang Huang, William Chung-Ho Panitch, Jaimyn Drake, Joseph Ortiz, Mustafa Mukadam, Mike Lambeta, Roberto Calandra, and Ken Goldberg · 2024
Closest in time.
Multiply: A multisensory object-centric embodied large language model in 3d world
Yining Hong, Zishuo Zheng, Peihao Chen, Yian Wang, Junyan Li, and Chuang Gan · 2024
Closest in time.
Binding touch to everything: Learning unified multimodal tactile representations
Fengyu Yang, Chao Feng, Ziyang Chen, Hyoungseob Park, Daniel Wang, Yiming Dou, Ziyao Zeng, Xien Chen, Rit Gangopadhyay, Andrew Owens, et al · 2024
Closest in time.