Fetching the paper…
Reading the bibliography…
Multisensory object-centric perception, reasoning, and interaction have been a key research topic in recent years.
Ray tracing volume densities
J. T. Kajiya and B. P. Von Herzen · 1984
Earlier work this paper cites.
The reviewing of object files: Object-specific integration of information
D. Kahneman, A. Treisman, and B. J. Gibbs · 1992
Earlier work this paper cites.
The development of embodied cognition: Six lessons from babies
L. Smith and M. Gasser · 2005
Earlier work this paper cites.
Core knowledge
E. S. Spelke and K. D. Kinzler · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Interactive learning of the acoustic properties of household objects
J. Sinapov, M. Wiemer, and A. Stoytchev · 2009
Earlier work this paper cites.
A new approach to cross-modal multimedia retrieval
N. Rasiwasia, J. Costa Pereira, E. Coviello, G. Doyle, G. R. Lanckriet, R. Levy, and N. Vasconcelos · 2010
Earlier work this paper cites.
Haptic feature extraction from a biomimetic tactile sensor: force, contact location and curvature
N. Wettels and G. E. Loeb · 2011
Earlier work this paper cites.
Example-guided physically based modal sound synthesis
Z. Ren, H. Yeh, and M. C. Lin · 2013
Earlier work this paper cites.
Bigbird:(big) berkeley instance recognition dataset
A. Singh, J. Sha, K. S. Narayan, T. Achim, and P. Abbeel · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick · 2014
Earlier work this paper cites.
3d shapenets: A deep representation for volumetric shapes
Z. Wu, S. Song, A. Khosla, F. Yu, L. Zhang, X. Tang, and J. Xiao · 2015
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
A. X. Chang, T. Funkhouser, L. Guibas, P. Hanrahan, Q. Huang, Z. Li, S. Savarese, M. Savva, S. Song, H. Su, et al · 2015
Earlier work this paper cites.
The ycb object and model set: Towards common benchmarks for manipulation research
B. Calli, A. Singh, A. Walsman, S. Srinivasa, P. Abbeel, and A. M. Dollar · 2015
Earlier work this paper cites.
Trust region policy optimization
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz · 2015
Earlier work this paper cites.
Acoustics based terrain classification for legged robots
J. Christie and N. Kottege · 2016
Earlier work this paper cites.
The curious robot: Learning visual representations via physical interactions
L. Pinto, D. Gandhi, Y. Han, Y.-L. Park, and A. Gupta · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
The feeling of success: Does touch sensing help predict grasp outcomes?
R. Calandra, A. Owens, M. Upadhyaya, W. Yuan, J. Lin, E. H. Adelson, and S. Levine · 2017
Cited alongside, same era.
Generative modeling of audible shapes for object perception
Z. Zhang, J. Wu, Q. Li, Z. Huang, J. Traer, J. H. McDermott, J. B. Tenenbaum, and W. T. Freeman · 2017
Cited alongside, same era.
Gelsight: High-resolution robot tactile sensors for estimating geometry and force
W. Yuan, S. Dong, and E. H. Adelson · 2017
Cited alongside, same era.
Virtualhome: Simulating household activities via programs
X. Puig, K. Ra, M. Boben, J. Li, T. Wang, S. Fidler, and A. Torralba · 2018
Cited alongside, same era.
Learning audio feedback for estimating amount and flow of granular material
S. Clarke, T. Rhodes, C. G. Atkeson, and O. Kroemer · 2018
Cited alongside, same era.
More than a feeling: Learning to grasp and regrasp using vision and touch
igibson, a simulation environment for interactive tasks in large realisticscenes
B. Shen, F. Xia, C. Li, R. Martín-Martín, L. Fan, G. Wang, S. Buch, C. D’Arpino, S. Srivastava, L. P. Tchapmi, et al · 2020
Later among the works it cites.
RoboTHOR: An Open Simulation-to-Real Embodied AI Platform
M. Deitke, W. Han, A. Herrasti, A. Kembhavi, E. Kolve, R. Mottaghi, J. Salvador, D. Schwenk, E. VanderBilt, M. Wallingford, L. Weihs, M. Yatskar, and A. Farhadi · 2020
Later among the works it cites.
Threedworld: A platform for interactive multi-modal physical simulation
C. Gan, J. Schwartz, S. Alter, M. Schrimpf, J. Traer, J. De Freitas, J. Kubilius, A. Bhandwaldar, N. Haber, M. Sano, et al · 2020
Later among the works it cites.
robosuite: A modular simulation framework and benchmark for robot learning
Y. Zhu, J. Wong, A. Mandlekar, and R. Martín-Martín · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Calandra, A. Owens, D. Jayaraman, J. Lin, W. Yuan, J. Malik, E. H. Adelson, and S. Levine · 2018
Cited alongside, same era.
Stochastic prediction of multi-agent interactions from partial observations
C. Sun, P. Karlsson, J. Wu, J. B. Tenenbaum, and K. Murphy · 2019
Cited alongside, same era.
Objectnet: A large-scale bias-controlled dataset for pushing the limits of object recognition models
A. Barbu, D. Mayo, J. Alverio, W. Luo, C. Wang, D. Gutfreund, J. Tenenbaum, and B. Katz · 2019
Cited alongside, same era.
Habitat: A Platform for Embodied AI Research
M. Savva, A. Kadian, O. Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, D. Parikh, and D. Batra · 2019
Cited alongside, same era.
Occupancy networks: Learning 3d reconstruction in function space
L. Mescheder, M. Oechsle, M. Niemeyer, S. Nowozin, and A. Geiger · 2019
Cited alongside, same era.
Deepsdf: Learning continuous signed distance functions for shape representation
J. J. Park, P. Florence, J. Straub, R. Newcombe, and S. Lovegrove · 2019
Cited alongside, same era.
Scene representation networks: Continuous 3d-structure-aware neural scene representations
V. Sitzmann, M. Zollhöfer, and G. Wetzstein · 2019
Cited alongside, same era.
S. James, Z. Ma, D. R. Arrojo, and A. J. Davison · 2020
Later among the works it cites.
Soundspaces: Audio-visual navigaton in 3d environments
C. Chen, U. Jain, C. Schissler, S. V. A. Gari, Z. Al-Halah, V. K. Ithapu, P. Robinson, and K. Grauman · 2020
Later among the works it cites.
Digit: A novel design for a low-cost compact high-resolution tactile sensor with application to in-hand manipulation
M. Lambeta, P.-W. Chou, S. Tian, B. Yang, B. Maloon, V. R. Most, D. Stroud, R. Santos, A. Byagowi, G. Kammerer, et al · 2020
Later among the works it cites.
Tacto: A fast, flexible and open-source simulator for high-resolution vision-based tactile sensors
S. Wang, M. Lambeta, P.-W. Chou, and R. Calandra · 2020
Later among the works it cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng · 2020
Later among the works it cites.
Object-centric neural scene rendering
M. Guo, A. Fathi, J. Wu, and T. Funkhouser · 2020
Later among the works it cites.
The open images dataset v4
A. Kuznetsova, H. Rom, N. Alldrin, J. Uijlings, I. Krasin, J. Pont-Tuset, S. Kamali, S. Popov, M. Malloci, A. Kolesnikov, et al · 2020
Later among the works it cites.
Convolutional occupancy networks
S. Peng, M. Niemeyer, L. Mescheder, M. Pollefeys, and A. Geiger · 2020
Later among the works it cites.
Swoosh! rattle! thump!–actions that sound
D. Gandhi, A. Gupta, and L. Pinto · 2020
Later among the works it cites.
Stressd: Sim-to-real from sound for stochastic dynamics
C. Matl, Y. Narang, D. Fox, R. Bajcsy, and F. Ramos · 2020
Later among the works it cites.
3d shape reconstruction from vision and touch
E. J. Smith, R. Calandra, A. Romero, G. Gkioxari, D. Meger, J. Malik, and M. Drozdzal · 2020
Later among the works it cites.
Deep-modal: real-time impact sound synthesis for arbitrary shapes
X. Jin, S. Li, T. Qu, D. Manocha, and G. Wang · 2020
Later among the works it cites.
Omnitact: A multi-directional high-resolution touch sensor
A. Padmanabha, F. Ebert, S. Tian, R. Calandra, C. Finn, and S. Levine · 2020
Later among the works it cites.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
T. Yu, D. Quillen, Z. He, R. Julian, K. Hausman, C. Finn, and S. Levine · 2020
Later among the works it cites.