Fetching the paper…
Reading the bibliography…
We present Vision in Action (ViA), an active perception system for bimanual robot manipulation.
Active perception
R. Bajcsy · 1988
Earlier work this paper cites.
Active vision
J. Aloimonos, I. Weiss, and A. Bandyopadhyay · 1988
Earlier work this paper cites.
Real time vergence control
T. Olson and R. Potter · 1989
Earlier work this paper cites.
Animate vision
D. H. Ballard · 1991
Earlier work this paper cites.
A head-eye system—analysis and design
K. Pahlavan and J.-O. Eklundh · 1992
Earlier work this paper cites.
Gaze control for a binocular camera head
J. L. Crowley, P. Bobet, and M. Mesrabi · 1992
Earlier work this paper cites.
Modeling visual attention via selective tuning
J. K. Tsotsos, S. M. Culhane, W. Y. K. Wai, Y. Lai, N. Davis, and F. Nuflo · 1995
Earlier work this paper cites.
A model of saliency-based visual attention for rapid scene analysis
L. Itti, C. Koch, and E. Niebur · 1998
Earlier work this paper cites.
A solution to the next best view problem for automated surface acquisition
R. Pito · 1999
Earlier work this paper cites.
Computational modelling of visual attention
L. Itti and C. Koch · 2001
Earlier work this paper cites.
Humanoid robot hrp-3
K. Kaneko, K. Harada, F. Kanehiro, G. Miyamori, and K. Akachi · 2008
Earlier work this paper cites.
The karlsruhe humanoid head
T. Asfour, K. Welke, P. Azad, A. Ude, and R. Dillmann · 2008
Earlier work this paper cites.
Online motion planning for hoap-2 humanoid robot navigation
M. Elmogy, C. Habel, and J. Zhang · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
An active vision system for detecting, fixating and manipulating objects in the real world
B. Rasolzadeh, M. Björkman, K. Huebner, and D. Kragic · 2010
Earlier work this paper cites.
Autonomous generation of complete 3d object models using next best view manipulation planning
M. Krainin, B. Curless, and D. Fox · 2011
Earlier work this paper cites.
An autonomous manipulation system based on force control and optimization
L. Righetti, M. Kalakrishnan, P. Pastor, J. Binney, J. Kelly, R. Voorhies, G. Sukhatme, and S. Schaal · 2014
Earlier work this paper cites.
R. Bajcsy, Y. Aloimonos, and J. K. Tsotsos · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
Autonomous view selection and gaze stabilization for humanoid robots
M. Grotz, T. Habra, R. Ronsse, and T. Asfour · 2017
Earlier work this paper cites.
Learning to look around: Intelligently exploring unseen environments for unknown tasks, 2017
D. Jayaraman and K. Grauman · 2017
Earlier work this paper cites.
Estimating the motion-to-photon latency in head mounted displays
J. Zhao, R. S. Allison, M. Vinnikov, and S. Jennings · 2017
Earlier work this paper cites.
Real-time perception meets reactive motion generation
D. Kappler, F. Meier, J. Issac, J. Mainprice, C. G. Cifuentes, M. Wüthrich, V. Berenz, S. Schaal, N. Ratliff, and J. Bohg · 2018
Cited alongside, same era.
Deep imitation learning for complex manipulation tasks from virtual reality teleoperation
T. Zhang, Z. McCarthy, O. Jow, D. Lee, X. Chen, K. Goldberg, and P. Abbeel · 2018
Cited alongside, same era.
ASIMO and Humanoid Robot Research at Honda , pages 55–90
S. Shigemi · 2019
Cited alongside, same era.
Reinforcement learning of active vision for manipulating objects under occlusions, 2019
R. Cheng, A. Agarwal, and K. Fragkiadaki · 2019
Cited alongside, same era.
Perceptual tolerance to motion-to-photon latency with head movement in virtual reality
M. Yang, J. Zhang, and L. Yu · 2019
Cited alongside, same era.
Open-television: Teleoperation with immersive active visual feedback
X. Cheng, J. Li, S. Yang, G. Yang, and X. Wang · 2024
Later among the works it cites.
Learning to look around: Enhancing teleoperation and learning with a human-like actuated neck, 2024
B. Sen, M. Wang, N. Thakur, A. Agarwal, and P. Agrawal · 2024
Later among the works it cites.
The interaction of top–down and bottom–up attention in visual working memory
W. Zheng, Y. Sun, H. Wu, H. Sun, and D. Zhang · 2024
Later among the works it cites.
Learning to look: Seeking information for decision making via policy factorization
S. Dass, J. Hu, B. Abbatematteo, P. Stone, and R. Martín-Martín · 2024
Later among the works it cites.
Spin: Simultaneous perception interaction and navigation
S. Uppal, A. Agarwal, H. Xiong, K. Shaw, and D. Pathak · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
U. A. Chattha, U. I. Janjua, F. Anwar, T. M. Madni, M. F. Cheema, and S. I. Janjua · 2020
Cited alongside, same era.
What matters in learning from offline human demonstrations for robot manipulation
A. Mandlekar, D. Xu, J. Wong, S. Nasiriany, C. Wang, R. Kulkarni, L. Fei-Fei, S. Savarese, Y. Zhu, and R. Martín-Martín · 2021
Cited alongside, same era.
Bimanual telemanipulation with force and haptic feedback through an anthropomorphic avatar system
C. Lenz and S. Behnke · 2022
Cited alongside, same era.
Learning fine-grained bimanual manipulation with low-cost hardware
T. Z. Zhao, V. Kumar, S. Levine, and C. Finn · 2023
Cited alongside, same era.
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, S. Feng, Y. Du, Z. Xu, E. Cousineau, B. Burchfiel, and S. Song · 2023
Cited alongside, same era.
SAM-RL: Sensing-Aware Model-Based Reinforcement Learning via Differentiable Physics-Based Simulation and Rendering
J. Lv, Y. Feng, C. Zhang, S. Zhao, L. Shao, and C. Lu · 2023
Cited alongside, same era.
Active vision reinforcement learning under limited visual observability, 2023
J. Shang and M. S. Ryoo · 2023
Cited alongside, same era.
Later among the works it cites.
Tidybot++: An open-source holonomic mobile manipulator for robot learning
J. Wu, W. Chong, R. Holmberg, A. Prasad, Y. Gao, O. Khatib, S. Song, S. Rusinkiewicz, and J. Bohg · 2024
Later among the works it cites.
Gello: A general, low-cost, and intuitive teleoperation framework for robot manipulators
P. Wu, Y. Shentu, Z. Yi, X. Lin, and P. Abbeel · 2024
Later among the works it cites.
Open teach: A versatile teleoperation system for robotic manipulation
A. Iyer, Z. Peng, Y. Dai, I. Guzey, S. Haldar, S. Chintala, and L. Pinto · 2024
Later among the works it cites.
Radiance fields for robotic teleoperation
M. Wilder-Smith, V. Patil, and M. Hutter · 2024
Later among the works it cites.
UMI on legs: Making manipulation policies mobile with manipulation-centric whole-body controllers
H. Ha, Y. Gao, Z. Fu, J. Tan, and S. Song · 2024
Later among the works it cites.
3d diffusion policy
Y. Ze, G. Zhang, K. Zhang, C. Hu, M. Wang, and H. Xu · 2024
Later among the works it cites.
4d gaussian splatting for real-time dynamic scene rendering
G. Wu, T. Yi, J. Fang, L. Xie, X. Zhang, W. Wei, W. Liu, Q. Tian, and X. Wang · 2024
Later among the works it cites.
Dynamics-guided diffusion model for robot manipulator design
X. Xu, H. Ha, and S. Song · 2024
Later among the works it cites.
Task-driven co-design of mobile manipulators
R. Schneider, D. Honerkamp, T. Welschehold, and A. Valada · 2024
Later among the works it cites.
Toddlerbot: Open-source ml-compatible humanoid platform for loco-manipulation, 2025
H. Shi, W. Wang, S. Song, and C. K. Liu · 2025
Closest in time.
Y. Jiang, R. Zhang, J. Wong, C. Wang, Y. Ze, H. Yin, C. Gokmen, S. Song, J. Wu, and L. Fei-Fei · 2025
Closest in time.
Robopanoptes: The all-seeing robot with whole-body dexterity
X. Xu, D. Bauer, and S. Song · 2025
Closest in time.
Y. Liu, S. Mu, X. Chao, Z. Li, Y. Mu, T. Chen, S. Li, C. Lyu, X.-p. Zhang, and W. Ding · 2025
Closest in time.
Active vision might be all you need: Exploring active vision in bimanual robotic manipulation, 2025
I. Chuang, A. Lee, D. Gao, M.-M. Naddaf-Sh, and I. Soltani · 2025
Closest in time.
Generalizable humanoid manipulation with 3d diffusion policies, 2025
Y. Ze, Z. Chen, W. Wang, T. Chen, X. He, Y. Yuan, X. B. Peng, and J. Wu · 2025
Closest in time.
Task-based grasp adaptation on a humanoid robot
J. Bohg, K. Welke, B. León, M. Do, D. Song, W. Wohlkinger, M. Madry, A. Aldóma, M. Przybylski, T. Asfour, H. Martí, D. Kragic, A. Morales, and M. Vincze · 2030
Closest in time.
View planning in robot active vision: A survey of systems, algorithms, and applications
R. Zeng, Y. Wen, W. Zhao, and Y.-J. Liu · 2096
Closest in time.