Fetching the paper…
Reading the bibliography…
We present MoE-Loco, a Mixture of Experts (MoE) framework for multitask locomotion for legged robots.
R. A. Jacobs, M. I. Jordan, S. J. Nowlan, and G. E. Hinton, “Adaptive mixtures of local experts,” Neural computation , vol. 3, no. 1, pp. 79–87, 1991
1991
Earlier work this paper cites.
M. I. Jordan and R. A. Jacobs, “Hierarchical mixtures of experts and the em algorithm,” Neural computation , vol. 6, no. 2, pp. 181–214, 1994
1994
Earlier work this paper cites.
R. Caruana, “Multitask learning,” Machine learning , vol. 28, pp. 41–75, 1997
1997
Earlier work this paper cites.
R. Collobert and J. Weston, “A unified architecture for natural language processing: Deep neural networks with multitask learning,” in Proceedings of the 25th international conference on Machine learning , 2008, pp. 160–167
2008
Earlier work this paper cites.
M. Deisenroth and J. W. Ng, “Distributed gaussian processes,” in International conference on machine learning . PMLR, 2015, pp. 1481–1490
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
L. Pinto and A. Gupta, “Learning to push by grasping: Using multiple tasks for effective learning,” in 2017 IEEE international conference on robotics and automation (ICRA) . IEEE, 2017, pp. 2161–2168
2017
Earlier work this paper cites.
C. Finn, P. Abbeel, and S. Levine, “Model-agnostic meta-learning for fast adaptation of deep networks,” in International conference on machine learning . PMLR, 2017, pp. 1126–1135
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, and M. Hutter, “Learning agile and dynamic motor skills for legged robots,” Science Robotics , vol. 4, no. 26, p. eaau5872, 2019
2019
Earlier work this paper cites.
J. Lee, J. Hwangbo, L. Wellhausen, V. Koltun, and M. Hutter, “Learning quadrupedal locomotion over challenging terrain,” Science robotics , vol. 5, no. 47, p. eabc5986, 2020
2020
Earlier work this paper cites.
T. Yu, S. Kumar, A. Gupta, S. Levine, K. Hausman, and C. Finn, “Gradient surgery for multi-task learning,” Advances in Neural Information Processing Systems , vol. 33, pp. 5824–5836, 2020
2020
Earlier work this paper cites.
C. Yang, K. Yuan, Q. Zhu, W. Yu, and Z. Li, “Multi-expert learning of adaptive legged locomotion,” Science Robotics , vol. 5, no. 49, p. eabb2174, 2020
2020
Earlier work this paper cites.
C. Yang, K. Yuan, Q. Zhu, W. Yu, and Z. Li, “Multi-expert learning of adaptive legged locomotion,” Science Robotics , vol. 5, no. 49, p. eabb2174, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
S. Vandenhende, S. Georgoulis, W. Van Gansbeke, M. Proesmans, D. Dai, and L. Van Gool, “Multi-task learning for dense prediction tasks: A survey,” IEEE transactions on pattern analysis and machine intelligence , vol. 44, no. 7, pp. 3614–3633, 2021
2021
Cited alongside, same era.
B. Liu, X. Liu, X. Jin, P. Stone, and Q. Liu, “Conflict-averse gradient descent for multi-task learning,” Advances in Neural Information Processing Systems , vol. 34, pp. 18 878–18 890, 2021
2021
Cited alongside, same era.
S. Sodhani, A. Zhang, and J. Pineau, “Multi-task reinforcement learning with context-based representations,” in International Conference on Machine Learning . PMLR, 2021, pp. 9767–9779
2021
Cited alongside, same era.
G. Ji, J. Mun, H. Kim, and J. Hwangbo, “Concurrent training of a control policy and a state estimator for dynamic and robust legged locomotion,” IEEE Robotics and Automation Letters , vol. 7, no. 2, pp. 4630–4637, 2022
2022
Cited alongside, same era.
2024
Later among the works it cites.
X. Cheng, K. Shi, A. Agarwal, and D. Pathak, “Extreme parkour with legged robots,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 11 443–11 450
2024
Later among the works it cites.
2024
Later among the works it cites.
K. Li, M. Cucuringu, L. Sánchez-Betancourt, and T. Willi, “Mixtures of experts for scaling up neural networks in order execution,” in Proceedings of the 5th ACM International Conference on AI in Finance , 2024, pp. 669–676
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Zhou, W. Zhang, J. Jiang, W. Zhong, J. Gu, and W. Zhu, “On the convergence of stochastic multi-objective gradient manipulation and beyond,” Advances in Neural Information Processing Systems , vol. 35, pp. 38 103–38 115, 2022
2022
Cited alongside, same era.
K. N. Kumar, I. Essa, and S. Ha, “Cascaded compositional residual learning for complex interactive behaviors,” IEEE Robotics and Automation Letters , vol. 8, no. 8, pp. 4601–4608, 2023
2023
Cited alongside, same era.
A. Klipfel, N. Sontakke, R. Liu, and S. Ha, “Learning a single policy for diverse behaviors on a quadrupedal robot using scalable motion imitation,” in 2023 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2023, pp. 2768–2775
2023
Cited alongside, same era.
2023
Cited alongside, same era.
J. Wu, G. Xin, C. Qi, and Y. Xue, “Learning robust and agile legged locomotion using adversarial motion priors,” IEEE Robotics and Automation Letters , 2023
2023
Cited alongside, same era.
S. Liu, Z. Chen, Y. Liu, Y. Wang, D. Yang, Z. Zhao, Z. Zhou, X. Yi, W. Li, W. Zhang, et al. , “Improving generalization in visual reinforcement learning via conflict-aware gradient agreement augmentation,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 23 436–23 446
2023
Cited alongside, same era.
M. Mittal, C. Yu, Q. Yu, J. Liu, N. Rudin, D. Hoeller, J. L. Yuan, R. Singh, Y. Guo, H. Mazhar, A. Mandlekar, B. Babich, G. State, M. Hutter, and A. Garg, “Orbit: A unified simulation framework for interactive robot learning environments,” IEEE Robotics and Automation Letters , vol. 8, no. 6, pp. 3740–3747, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
G. B. Margolis, G. Yang, K. Paigwar, T. Chen, and P. Agrawal, “Rapid locomotion via reinforcement learning,” The International Journal of Robotics Research , vol. 43, no. 4, pp. 572–587, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
Y. Li, J. Li, W. Fu, and Y. Wu, “Learning agile bipedal motions on a quadrupedal robot,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 9735–9742
2024
Later among the works it cites.
2024
Later among the works it cites.
D. Hoeller, N. Rudin, D. Sako, and M. Hutter, “Anymal parkour: Learning agile navigation for quadrupedal robots,” Science Robotics , vol. 9, no. 88, p. eadi7566, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
M. Shafiee, G. Bellegarda, and A. Ijspeert, “Manyquadrupeds: Learning a single locomotion policy for diverse quadruped robots,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 3471–3477
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Zhang, S. Liu, J. Yu, Q. Cai, X. Zhao, C. Zhang, Z. Liu, Q. Liu, H. Zhao, L. Hu, et al. , “M3oe: Multi-domain multi-task mixture-of experts recommendation framework,” in Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval , 2024, pp. 893–902
2024
Later among the works it cites.
2024
Later among the works it cites.