Fetching the paper…
Reading the bibliography…
We present Splat-MOVER, a modular robotics stack for open-vocabulary robotic manipulation, which leverages the editability of Gaussian Splatting (GSplat) scene representations to enable multi-stage manipulation tasks.
The ecological approach to the visual perception of pictures
J. J. Gibson · 1978
Earlier work this paper cites.
The psychology of everyday things
D. A. Norman · 1988
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Z. Wang, A. C. Bovik, H. R. Sheikh, and E. P. Simoncelli · 2004
Earlier work this paper cites.
Moveit![ros topics]
S. Chitta, I. Sucan, and S. Cousins · 2012
Earlier work this paper cites.
Reducing the barrier to entry of complex robotic software: a moveit! case study
D. Coleman, I. A. Şucan, S. Chitta, and N. Correll · 2014
Earlier work this paper cites.
Learning to detect visual grasp affordance
H. O. Song, M. Fritz, D. Goehring, and T. Darrell · 2015
Earlier work this paper cites.
Structure-from-motion revisited
J. L. Schönberger and J.-M. Frahm · 2016
Earlier work this paper cites.
Grasp pose detection in point clouds
A. Ten Pas, M. Gualtieri, K. Saenko, and R. Platt · 2017
Earlier work this paper cites.
A brief review of affordance in robotic manipulation research
N. Yamanobe, W. Wan, I. G. Ramirez-Alpizar, D. Petit, T. Tsuji, S. Akizuki, M. Hashimoto, K. Nagata, and K. Harada · 2017
Earlier work this paper cites.
Scaling egocentric vision: The epic-kitchens dataset
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, et al · 2018
Earlier work this paper cites.
Learning grasp affordance reasoning through semantic relations
P. Ardón, E. Pairet, R. P. Petrick, S. Ramamoorthy, and K. S. Lohan · 2019
Earlier work this paper cites.
Grounded human-object interaction hotspots from video
T. Nagarajan, C. Feichtenhofer, and K. Grauman · 2019
Earlier work this paper cites.
Graspnet-1billion: A large-scale benchmark for general object grasping
H.-S. Fang, C. Wang, M. Gou, and C. Lu · 2020
Earlier work this paper cites.
Dex-NeRF: Using a neural radiance field to grasp transparent objects
J. Ichnowski*, Y. Avigal*, J. Kerr, and K. Goldberg · 2020
Earlier work this paper cites.
End-to-end object detection with transformers
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko · 2020
Earlier work this paper cites.
Nerf: Representing scenes as neural radiance fields for view synthesis
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng · 2021
Earlier work this paper cites.
imap: Implicit mapping and positioning in real-time
E. Sucar, S. Liu, J. Ortiz, and A. J. Davison · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Earlier work this paper cites.
Emerging properties in self-supervised vision transformers
M. Caron, H. Touvron, I. Misra, H. Jégou, J. Mairal, P. Bojanowski, and A. Joulin · 2021
Earlier work this paper cites.
Dynamic head: Unifying object detection heads with attentions
X. Dai, Y. Chen, B. Xiao, D. Chen, M. Liu, L. Yuan, and L. Zhang · 2021
Earlier work this paper cites.
Extract free dense labels from CLIP
C. Zhou, C. C. Loy, and B. Dai · 2022
Cited alongside, same era.
Rt-1: Robotics transformer for real-world control at scale
A. Brohan, N. Brown, J. Carbajal, Y. Chebotar, J. Dabis, C. Finn, K. Gopalakrishnan, K. Hausman, A. Herzog, J. Hsu, J. Ibarz, B. Ichter, A. Irpan, T. Jackson, S. Jesmonth, N. Joshi, R. Julian, D. Kalashnikov, Y. Kuang, I. Leal, K.-H. Lee, S. Levine, Y. Lu, U. Malla, D. Manjunath, I. Mordatch, O. Nachum, C. Parada, J. Peralta, E. Perez, K. Pertsch, J. Quiambao, K. Rao, M. Ryoo, G. Salazar, P. Sanketi, K. Sayed, J. Singh, S. Sontakke, A. Stone, C. Tan, H. Tran, V. Vanhoucke, S. Vega, Q. Vuong, F. Xia, T. Xiao, P. Xu, S. Xu, T. Yu, and B. Zitkovich · 2022
Cited alongside, same era.
Vision-only robot navigation in a neural radiance world
M. Adamkiewicz, T. Chen, A. Caccavale, R. Gardner, P. Culbertson, J. Bohg, and M. Schwager · 2022
Cited alongside, same era.
Evo-nerf: Evolving nerf for sequential robot grasping of transparent objects
J. Kerr, L. Fu, H. Huang, Y. Avigal, M. Tancik, J. Ichnowski, A. Kanazawa, and K. Goldberg · 2022
Cited alongside, same era.
Feature 3dgs: Supercharging 3d gaussian splatting to enable distilled feature fields
S. Zhou, H. Chang, S. Jiang, Z. Fan, Z. Zhu, D. Xu, P. Chari, S. You, Z. Wang, and A. Kadambi · 2023
Later among the works it cites.
Anygrasp: Robust and efficient grasp perception in spatial and temporal domains
H.-S. Fang, C. Wang, H. Fang, M. Gou, J. Liu, H. Yan, W. Liu, Y. Xie, and C. Lu · 2023
Later among the works it cites.
Language Segment-Anything
L. Medeiros · 2023
Later among the works it cites.
Nerfstudio: A framework for neural radiance field development
M. Tancik, E. Weber, R. Li, B. Yi, T. Wang, A. Kristoffersen, J. Austin, K. Salahi, A. Ahuja, D. McAllister, A. Kanazawa, and E. Ng · 2023
Later among the works it cites.
ConceptFusion: Open-set multimodal 3d mapping
K. M. Jatavallabhula, A. Kuwajerwala, Q. Gu, M. Omama, T. Chen, S. Li, G. Iyer, S. Saryazdi, N. Keetha, A. Tewari, et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Kobayashi, E. Matsumoto, and V. Sitzmann · 2022
Cited alongside, same era.
CLIP-NeRF: Text-and-image driven manipulation of neural radiance fields
C. Wang, M. Chai, M. He, D. Chen, and J. Liao · 2022
Cited alongside, same era.
Language-driven semantic segmentation
B. Li, K. Q. Weinberger, S. Belongie, V. Koltun, and R. Ranftl · 2022
Cited alongside, same era.
Dino: Detr with improved denoising anchor boxes for end-to-end object detection
H. Zhang, F. Li, S. Liu, L. Zhang, H. Su, J. Zhu, L. M. Ni, and H.-Y. Shum · 2022
Cited alongside, same era.
Human hands as probes for interactive object understanding
M. Goyal, S. Modi, R. Goyal, and S. Gupta · 2022
Cited alongside, same era.
Language embedded radiance fields for zero-shot task-oriented grasping
A. Rashid, S. Sharma, C. M. Kim, J. Kerr, L. Y. Chen, A. Kanazawa, and K. Goldberg · 2023
Cited alongside, same era.
Distilled feature fields enable few-shot language-guided manipulation
W. Shen, G. Yang, A. Yu, J. Wong, L. P. Kaelbling, and P. Isola · 2023
Cited alongside, same era.
3D Gaussian splatting for real-time radiance field rendering
B. Kerbl, G. Kopanas, T. Leimkühler, and G. Drettakis · 2023
Cited alongside, same era.
Foundation models in robotics: Applications, challenges, and the future
R. Firoozi, J. Tucker, S. Tian, A. Majumdar, J. Sun, W. Liu, Y. Zhu, S. Song, A. Kapoor, K. Hausman, et al · 2023
Later among the works it cites.
Octo: An open-source generalist robot policy
O. M. Team, D. Ghosh, H. Walke, K. Pertsch, K. Black, O. Mees, S. Dasari, J. Hejna, T. Kreiman, C. Xu, et al · 2024
Closest in time.
Openvla: An open-source vision-language-action model
M. J. Kim, K. Pertsch, S. Karamcheti, T. Xiao, A. Balakrishna, S. Nair, R. Rafailov, E. Foster, G. Lam, P. Sanketi, et al · 2024
Closest in time.
Catnips: Collision avoidance through neural implicit probabilistic scenes
T. Chen, P. Culbertson, and M. Schwager · 2024
Closest in time.
Semantic anything in 3d gaussians
X. Hu, Y. Wang, L. Fan, J. Fan, J. Peng, Z. Lei, Q. Li, and Z. Zhang · 2024
Closest in time.
Fmgs: Foundation model embedded 3d gaussian splatting for holistic 3d scene understanding
X. Zuo, P. Samangouei, Y. Zhou, Y. Di, and M. Li · 2024
Closest in time.
G. Liao, J. Li, Z. Bao, X. Ye, J. Wang, Q. Li, and K. Liu · 2024
Closest in time.
Gaussiangrasper: 3d language gaussian splatting for open-vocabulary robotic grasping
Y. Zheng, X. Chen, Y. Zheng, S. Gu, R. Yang, B. Jin, P. Li, C. Zhong, Z. Wang, L. Liu, et al · 2024
Closest in time.
Manigaussian: Dynamic gaussian splatting for multi-task robotic manipulation
G. Lu, S. Zhang, Z. Wang, C. Liu, J. Lu, and Y. Tang · 2024
Closest in time.
Object-aware gaussian splatting for robotic manipulation
Y. Li and D. Pathak · 2024
Closest in time.
Graspsplats: Efficient manipulation with 3d feature splatting
M. Ji, R.-Z. Qiu, X. Zou, and X. Wang · 2024
Closest in time.
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, Z. Xu, S. Feng, E. Cousineau, Y. Du, B. Burchfiel, R. Tedrake, and S. Song · 2024
Closest in time.
Object-centric reconstruction and tracking of dynamic unknown objects using 3d gaussian splatting
K. R. Barad, A. Richard, J. Dentler, M. Olivares-Mendez, and C. Martinez · 2024
Closest in time.
Splat-nav: Safe real-time robot navigation in gaussian splatting maps
T. Chen, O. Shorinwa, W. Zeng, J. Bruno, P. Dames, and M. Schwager · 2024
Closest in time.