Fetching the paper…
Reading the bibliography…
In robotic manipulation, vision-language-action (VLA) models have emerged as a promising paradigm for learning generalizable and scalable robot policies.
A mathematical theory of communication
Claude E Shannon · 1948
Earlier work this paper cites.
On measures of entropy and information
Alfréd Rényi · 1961
Earlier work this paper cites.
Minimum entropy of error principle in estimation
Martin Janzura, Timo Koski, and Antonin Otáhal · 1994
Earlier work this paper cites.
Information-theoretic learning
Jose Principe and John Iii · 2000
Earlier work this paper cites.
An error-entropy minimization algorithm for supervised training of nonlinear adaptive systems
Deniz Erdogmus and Jose C Principe · 2002
Earlier work this paper cites.
Mean-square convergence analysis of adaline training with minimum error entropy criterion
Badong Chen, Yu Zhu, and Jinchun Hu · 2010
Earlier work this paper cites.
Kernel minimum error entropy algorithm
Badong Chen, Zejian Yuan, Nanning Zheng, and José C Príncipe · 2013
Earlier work this paper cites.
Insights into the robustness of minimum error entropy estimation
Badong Chen, Lei Xing, Bin Xu, Haiquan Zhao, and Jose C Principe · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Attention is all you need
A Vaswani · 2017
Earlier work this paper cites.
Alexei Botchkarev · 2018
Earlier work this paper cites.
Quantized minimum error entropy criterion
Badong Chen, Lei Xing, Nanning Zheng, and Jose C Principe · 2018
Earlier work this paper cites.
Minimum error entropy kalman filter
Badong Chen, Lujuan Dang, Yuantao Gu, Nanning Zheng, and José C Príncipe · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Robust power system state estimation with minimum error entropy unscented kalman filter
Lujuan Dang, Badong Chen, Shiyuan Wang, Wentao Ma, and Pengju Ren · 2020
Cited alongside, same era.
Fixed-point minimum error entropy with fiducial points
Yuqing Xie, Yingsong Li, Yuantao Gu, Jiuwen Cao, and Badong Chen · 2020
Cited alongside, same era.
Generalized minimum error entropy for robust learning
Jiacheng He, Gang Wang, Kui Cao, He Diao, Guotai Wang, and Bei Peng · 2023
Cited alongside, same era.
Libero: Benchmarking knowledge transfer for lifelong robot learning
Bo Liu, Yifeng Zhu, Chongkai Gao, Yihao Feng, Qiang Liu, Yuke Zhu, and Peter Stone · 2023
Cited alongside, same era.
Rt-2: Vision-language-action models transfer web knowledge to robotic control
Brianna Zitkovich, Tianhe Yu, Sichun Xu, Peng Xu, Ted Xiao, Fei Xia, Jialin Wu, Paul Wohlhart, Stefan Welker, Ayzaan Wahid, et al · 2023
Cited alongside, same era.
p i _ 0 pi\_0 : A vision-language-action flow model for general robot control
Diffusion policy: Visuomotor policy learning via action diffusion
Cheng Chi, Zhenjia Xu, Siyuan Feng, Eric Cousineau, Yilun Du, Benjamin Burchfiel, Russ Tedrake, and Shuran Song · 2025
Later among the works it cites.
Can Cui, Pengxiang Ding, Wenxuan Song, Shuanghao Bai, Xinyang Tong, Zirui Ge, Runze Suo, Wanqi Zhou, Yang Liu, Bofang Jia, et al · 2025
Later among the works it cites.
Long-vla: Unleashing long-horizon capability of vision language action model for robot manipulation
Yiguo Fan, Shuanghao Bai, Xinyang Tong, Pengxiang Ding, Yuyang Zhu, Hongchao Lu, Fengqi Dai, Wei Zhao, Yang Liu, Siteng Huang, et al · 2025
Later among the works it cites.
π 0.5 \pi_{0.5} : A vision-language-action model with open-world generalization
Physical Intelligence, Kevin Black, Noah Brown, James Darpinian, Karan Dhabalia, Danny Driess, Adnan Esmail, Michael Equi, Chelsea Finn, Niccolo Fusai, et al · 2025
Later among the works it cites.
Evaluating real-world robot manipulation policies in simulation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kevin Black, Noah Brown, Danny Driess, Adnan Esmail, Michael Equi, Chelsea Finn, Niccolo Fusai, Lachy Groom, Karol Hausman, Brian Ichter, et al · 2024
Cited alongside, same era.
Towards synergistic, generalized, and efficient dual-system for robotic manipulation
Qingwen Bu, Hongyang Li, Li Chen, Jisong Cai, Jia Zeng, Heming Cui, Maoqing Yao, and Yu Qiao · 2024
Cited alongside, same era.
Octo: An open-source generalist robot policy
Dibya Ghosh, Homer Rich Walke, Karl Pertsch, Kevin Black, Oier Mees, Sudeep Dasari, Joey Hejna, Tobias Kreiman, Charles Xu, Jianlan Luo, et al · 2024
Cited alongside, same era.
A comprehensive survey of regression-based loss functions for time series forecasting
Aryan Jadon, Avinash Patil, and Shruti Jadon · 2024
Cited alongside, same era.
Dinov2: Learning robust visual features without supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, et al · 2024
Cited alongside, same era.
Multimodal diffusion transformer: Learning versatile behavior from multimodal goals
Moritz Reuss, Ömer Erdinç Yağmurlu, Fabian Wenzel, and Rudolf Lioutikov · 2024
Cited alongside, same era.
Gr00t n1: An open foundation model for generalist humanoid robots
Johan Bjorck, Fernando Castañeda, Nikita Cherniadev, Xingye Da, Runyu Ding, Linxi Fan, Yu Fang, Dieter Fox, Fengyuan Hu, Spencer Huang, et al · 2025
Cited alongside, same era.
Xuanlin Li, Kyle Hsu, Jiayuan Gu, Oier Mees, Karl Pertsch, Homer Rich Walke, Chuyuan Fu, Ishikaa Lunawat, Isabel Sieh, Sean Kirmani, et al · 2025
Later among the works it cites.
Starvla: A lego-like codebase for vision-language-action model developing
starVLA Contributors · 2025
Later among the works it cites.
Gen-0: Embodied foundation models that scale with physical interaction
GA Team et al · 2025
Later among the works it cites.
Vq-vla: Improving vision-language-action models via scaling vector-quantized action tokenizers
Yating Wang, Haoyi Zhu, Mingyu Liu, Jiange Yang, Hao-Shu Fang, and Tong He · 2025
Later among the works it cites.
Align-then-steer: Adapting the vision-language action models through unified latent guidance
Yang Zhang, Chenwei Wang, Ouyang Lu, Yuan Zhao, Yunfei Ge, Zhenglong Sun, Xiu Li, Chi Zhang, Chenjia Bai, and Xuelong Li · 2025
Later among the works it cites.
Vlas: Vision-language-action model with speech instructions for customized robot manipulation
Wei Zhao, Pengxiang Ding, Zhang Min, Zhefei Gong, Shuanghao Bai, Han Zhao, and Donglin Wang · 2025
Later among the works it cites.
An information-theoretic approach for heterogeneous differentiable causal discovery
Wanqi Zhou, Shuanghao Bai, Yuqing Xie, Yicong He, Qibin Zhao, and Badong Chen · 2025
Later among the works it cites.
Cronusvla: Transferring latent motion across time for multi-frame prediction in manipulation
Hao Li, Shuai Yang, Yilun Chen, Yang Tian, Xiaoda Yang, Xinyi Chen, Hanqing Wang, Tai Wang, Feng Zhao, Dahua Lin, et al · 2026
Closest in time.
Vla-adapter: An effective paradigm for tiny-scale vision-language-action model
Yihao Wang, Pengxiang Ding, Lingxiao Li, Can Cui, Zirui Ge, Xinyang Tong, Wenxuan Song, Han Zhao, Wei Zhao, Pengxu Hou, et al · 2026
Closest in time.