Fetching the paper…
Reading the bibliography…
Deep Reinforcement Learning (RL) has been explored and verified to be effective in solving decision-making tasks in various domains, such as robotics, transportation, recommender systems, etc.
Eugene Ie, Vihan Jain, Jing Wang, Sanmit Narvekar, Ritesh Agarwal, Rui Wu, Heng-Tze Cheng, Morgane Lustman, Vince Gatto, Paul Covington, Jim McFadden, Tushar Chandra, and Craig Boutilier. 2019b · 1905
Earlier work this paper cites.
RLBench: The Robot Learning Benchmark and Learning Environment
Stephen James, Zicong Ma, David Rovick Arrojo, and Andrew J. Davison. 2019a · 1909
Earlier work this paper cites.
Solving Rubik’s Cube with a Robot Hand
OpenAI, Ilge Akkaya, Marcin Andrychowicz, Maciek Chociej, Mateusz Litwin, Bob McGrew, Arthur Petron, Alex Paino, Matthias Plappert, Glenn Powell, Raphael Ribas, Jonas Schneider, Nikolas Tezak, Jerry Tworek, Peter Welinder, Lilian Weng, Qiming Yuan, Wojciech Zaremba, and Lei Zhang. 2019 · 1910
Earlier work this paper cites.
SUMMIT: A Simulator for Urban Driving in Massive Mixed Traffic
Panpan Cai, Yiyuan Lee, Yuanfu Luo, and David Hsu. 2020 · 1911
Earlier work this paper cites.
Colight: Learning network-level cooperation for traffic signal control. In Proceedings of the 28th ACM international conference on information and knowledge management . 1913–1922
Hua Wei, Nan Xu, Huichu Zhang, Guanjie Zheng, Xinshi Zang, Chacha Chen, Weinan Zhang, Yanmin Zhu, Kai Xu, and Zhenhui Li. 2019b · 1922
Earlier work this paper cites.
Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning
Haokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta, Tenghao Huang, Mohit Bansal, and Colin A Raffel. 2022 · 1965
Earlier work this paper cites.
Metacognition and cognitive monitoring: A new area of cognitive–developmental inquiry
John H Flavell. 1979 · 1979
Earlier work this paper cites.
Effect of values on perception and decision making: A study of alternative work values measures
Elizabeth C Ravlin and Bruce M Meglino. 1987 · 1987
Earlier work this paper cites.
Temporal Probabilistic Logic Programs. In Logic Programming: The 1999 International Conference, Las Cruces, New Mexico, USA, November 29 - December 4, 1999 , Danny De Schreye (Ed.). MIT Press, 109–123
Alex Dekhtyar, Michael I. Dekhtyar, and V. S. Subrahmanian. 1999 · 1999
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S Sutton, David McAllester, Satinder Singh, and Yishay Mansour. 1999 · 1999
Earlier work this paper cites.
EnergyPlus: creating a new-generation building energy simulation program
Drury B Crawley, Linda K Lawrie, Frederick C Winkelmann, Walter F Buhl, Y Joe Huang, Curtis O Pedersen, Richard K Strand, Richard J Liesen, Daniel E Fisher, Michael J Witte, et al · 2001
Earlier work this paper cites.
Design and use paradigms for gazebo, an open-source multi-robot simulator. In 2004 IEEE/RSJ international conference on intelligent robots and systems (IROS)(IEEE Cat. No. 04CH37566) , Vol. 3. Ieee, 2149–2154
Nathan Koenig and Andrew Howard. 2004 · 2004
Earlier work this paper cites.
Execution monitoring in robotics: A survey
Ola Pettersson. 2005 · 2005
Earlier work this paper cites.
Reinforcement learning-based multi-agent system for network traffic signal control
Itamar Arel, Cong Liu, Tom Urbanik, and Airton G Kohls. 2010 · 2010
Earlier work this paper cites.
Urban traffic signal control using reinforcement learning agents
PG Balaji, X German, and Dipti Srinivasan. 2010 · 2010
Earlier work this paper cites.
States versus rewards: dissociable neural prediction error signals underlying model-based and model-free reinforcement learning
Jan Gläscher, Nathaniel Daw, Peter Dayan, and John P O’Doherty. 2010 · 2010
Earlier work this paper cites.
Control delay in reinforcement learning for real-time dynamic systems: A memoryless approach. In 2010 IEEE/RSJ international conference on intelligent robots and systems . IEEE, 3226–3231
Erik Schuitema, Lucian Buşoniu, Robert Babuška, and Pieter Jonker. 2010 · 2010
Earlier work this paper cites.
SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving
Ming Zhou, Jun Luo, Julian Villella, Yaodong Yang, David Rusu, Jiayu Miao, Weinan Zhang, Montgomery Alban, Iman Fadakar, Zheng Chen, Aurora Chongxi Huang, Ying Wen, Kimia Hassanzadeh, Daniel Graves, Dong Chen, Zhengbang Zhu, Nhat Nguyen, Mohamed Elsayed, Kun Shao, Sanjeevan Ahilan, Baokuan Zhang, Jiannan Wu, Zhengang Fu, Kasra Rezaee, Peyman Yadmellat, Mohsen Rohani, Nicolas Perez Nieves, Yihan Ni, Seyedershad Banijamali, Alexander Cowen Rivers, Zheng Tian, Daniel Palenicek, Haitham bou Ammar, Hongbo Zhang, Wulong Liu, Jianye Hao, and Jun Wang. 2020 · 2010
Earlier work this paper cites.
SoftGym: Benchmarking Deep Reinforcement Learning for Deformable Object Manipulation
Xingyu Lin, Yufei Wang, Jake Olkin, and David Held. 2021 · 2011
Earlier work this paper cites.
Annotated probabilistic temporal logic
Paulo Shakarian, Austin Parker, Gerardo I. Simari, and V. S. Subrahmanian. 2011 · 2011
Earlier work this paper cites.
Habits, action sequences and reinforcement learning
Amir Dezfouli and Bernard W Balleine. 2012 · 2012
Earlier work this paper cites.
Handbook of Markov decision processes: methods and applications . Vol. 40
Eugene A Feinberg and Adam Shwartz. 2012 · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control. In 2012 IEEE/RSJ international conference on intelligent robots and systems . IEEE, 5026–5033
Emanuel Todorov, Tom Erez, and Yuval Tassa. 2012 · 2012
Earlier work this paper cites.
The Arcade Learning Environment: An Evaluation Platform for General Agents
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling. 2013 · 2013
Earlier work this paper cites.
Reinforcement learning in robotics: A survey
Jens Kober, J Andrew Bagnell, and Jan Peters. 2013 · 2013
Earlier work this paper cites.
Reinforcement learning in robotics: Applications and real-world challenges
Petar Kormushev, Sylvain Calinon, and Darwin G Caldwell. 2013 · 2013
Earlier work this paper cites.
DTALite: A queue-based mesoscopic traffic simulator for fast model evaluation and calibration
Xuesong Zhou and Jeffrey Taylor. 2014 · 2014
Earlier work this paper cites.
Solving simultaneous route guidance and traffic signal optimization problem using space-phase-time hypernetwork
Pengfei Li, Pitu Mirchandani, and Xuesong Zhou. 2015 · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra. 2015 · 2015
Earlier work this paper cites.
Charles Beattie, Joel Z. Leibo, Denis Teplyashin, Tom Ward, Marcus Wainwright, Heinrich Küttler, Andrew Lefrancq, Simon Green, Víctor Valdés, Amir Sadik, Julian Schrittwieser, Keith Anderson, Sarah York, Max Cant, Adam Cain, Adrian Bolton, Stephen Gaffney, Helen King, Demis Hassabis, Shane Legg, and Stig Petersen. 2016 · 2016
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Earlier work this paper cites.
Introspective Perception: Learning to Predict Failures in Vision Systems
Shreyansh Daftry, Sam Zeng, J. Andrew Bagnell, and Martial Hebert. 2016 · 2016
Earlier work this paper cites.
Continuous adaptation via meta-learning in nonstationary and competitive environments
Maruan Al-Shedivat, Trapit Bansal, Yuri Burda, Ilya Sutskever, Igor Mordatch, and Pieter Abbeel. 2017 · 2017
Earlier work this paper cites.
Reinforcement learning for pivoting task
Rika Antonova, Silvia Cruciani, Christian Smith, and Danica Kragic. 2017 · 2017
Earlier work this paper cites.
Sensor fusion for robot control through deep reinforcement learning. In 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . Ieee, 2365–2370
Steven Bohez, Tim Verbelen, Elias De Coninck, Bert Vankeirsbilck, Pieter Simoens, and Bart Dhoedt. 2017 · 2017
Earlier work this paper cites.
Unsupervised pixel-level domain adaptation with generative adversarial networks. In Proceedings of the IEEE conference on computer vision and pattern recognition . 3722–3731
Konstantinos Bousmalis, Nathan Silberman, David Dohan, Dumitru Erhan, and Dilip Krishnan. 2017 · 2017
Earlier work this paper cites.
CARLA: An open urban driving simulator. In Conference on robot learning . PMLR, 1–16
Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun. 2017 · 2017
Earlier work this paper cites.
Grounded action transformation for robot learning in simulation. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 31
Josiah Hanna and Peter Stone. 2017 · 2017
Earlier work this paper cites.
Asymmetric Actor Critic for Image-Based Robot Learning
Pinto Lerrel, Andrychowicz Marcin, Welinder Peter, Zaremba Wojciech, and Abbeel Pieter. 2017 · 2017
Earlier work this paper cites.
Deep reinforcement learning for dynamic treatment regimes on medical registry data. In 2017 IEEE international conference on healthcare informatics (ICHI) . IEEE, 380–385
Ying Liu, Brent Logan, Ning Liu, Zhiyuan Xu, Jian Tang, and Yangzhi Wang. 2017 · 2017
Earlier work this paper cites.
Duckietown: an open, inexpensive and flexible platform for autonomy education and research. In 2017 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 1497–1504
Liam Paull, Jacopo Tani, Heejin Ahn, Javier Alonso-Mora, Luca Carlone, Michal Cap, Yu Fan Chen, Changhyun Choi, Jeff Dusek, Yajun Fang, et al · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Earlier work this paper cites.
Domain randomization for transferring deep neural networks from simulation to the real world. In 2017 IEEE/RSJ international conference on intelligent robots and systems (IROS) . IEEE, 23–30
Josh Tobin, Rachel Fong, Alex Ray, Jonas Schneider, Wojciech Zaremba, and Pieter Abbeel. 2017 · 2017
Earlier work this paper cites.
Adversarial discriminative domain adaptation. In Proceedings of the IEEE conference on computer vision and pattern recognition . 7167–7176
Eric Tzeng, Judy Hoffman, Kate Saenko, and Trevor Darrell. 2017 · 2017
Earlier work this paper cites.
Safe reinforcement learning via shielding. In Proceedings of the AAAI conference on artificial intelligence , Vol. 32
Mohammed Alshiekh, Roderick Bloem, Rüdiger Ehlers, Bettina Könighofer, Scott Niekum, and Ufuk Topcu. 2018 · 2018
Earlier work this paper cites.
Using simulation and domain adaptation to improve efficiency of deep robotic grasping. In 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 4243–4250
Konstantinos Bousmalis, Alex Irpan, Paul Wohlhart, Yunfei Bai, Matthew Kelcey, Mrinal Kalakrishnan, Laura Downs, Julian Ibarz, Peter Pastor, Kurt Konolige, et al · 2018
Earlier work this paper cites.
Multi-task domain adaptation for deep learning of instance grasping from simulation. In 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 3516–3523
Kuan Fang, Yunfei Bai, Stefan Hinterstoisser, Silvio Savarese, and Mrinal Kalakrishnan. 2018 · 2018
Earlier work this paper cites.
At human speed: Deep reinforcement learning with action delay
Vlad Firoiu, Tina Ju, and Josh Tenenbaum. 2018 · 2018
Earlier work this paper cites.
Deep q-learning from demonstrations. In Proceedings of the AAAI conference on artificial intelligence , Vol. 32
Todd Hester, Matej Vecerik, Olivier Pietquin, Marc Lanctot, Tom Schaul, Bilal Piot, Dan Horgan, John Quan, Andrew Sendonaris, Ian Osband, et al · 2018
Earlier work this paper cites.
An Environment for Autonomous Driving Decision-Making
Edouard Leurent. 2018 · 2018
Earlier work this paper cites.
Conditional adversarial domain adaptation
Mingsheng Long, Zhangjie Cao, Jianmin Wang, and Michael I Jordan. 2018 · 2018
Earlier work this paper cites.
Microscopic traffic simulation using SUMO. In 2018 21st international conference on intelligent transportation systems (ITSC) . IEEE, 2575–2582
Pablo Alvarez Lopez, Michael Behrisch, Laura Bieker-Walz, Jakob Erdmann, Yun-Pang Flötteröd, Robert Hilbrich, Leonhard Lücken, Johannes Rummel, Peter Wagner, and Evamarie Wießner. 2018 · 2018
Earlier work this paper cites.
Unsupervised reverse domain adaptation for synthetic medical images via adversarial training
Faisal Mahmood, Richard Chen, and Nicholas J Durr. 2018 · 2018
Earlier work this paper cites.
Multi-adversarial domain adaptation. In Proceedings of the AAAI conference on artificial intelligence , Vol. 32
Zhongyi Pei, Zhangjie Cao, Mingsheng Long, and Jianmin Wang. 2018 · 2018
Earlier work this paper cites.
Sim-to-real transfer of robotic control with dynamics randomization. In 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 3803–3810
Xue Bin Peng, Marcin Andrychowicz, Wojciech Zaremba, and Pieter Abbeel. 2018 · 2018
Earlier work this paper cites.
Failing to Learn: Autonomously Identifying Perception Failures for Self-driving Cars
Manikandasriram Srinivasan Ramanagopal, Cyrus Anderson, Ram Vasudevan, and Matthew Johnson-Roberson. 2018 · 2018
Earlier work this paper cites.
David Rohde, Stephen Bonner, Travis Dunlop, Flavian Vasile, and Alexandros Karatzoglou. 2018 · 2018
Earlier work this paper cites.
Sim2real viewpoint invariant visual servoing by recurrent control. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 4691–4699
Fereshteh Sadeghi, Alexander Toshev, Eric Jang, and Sergey Levine. 2018 · 2018
Earlier work this paper cites.
Virtual-Taobao: Virtualizing Real-world Online Retail Environment for Reinforcement Learning
Jing-Cheng Shi, Yang Yu, Qing Da, Shi-Yong Chen, and An-Xiang Zeng. 2018 · 2018
Earlier work this paper cites.
Intellilight: A reinforcement learning approach for intelligent traffic light control. In Proceedings of the 24th ACM SIGKDD international conference on knowledge discovery & data mining . 2496–2505
Hua Wei, Guanjie Zheng, Huaxiu Yao, and Zhenhui Li. 2018 · 2018
Earlier work this paper cites.
Hierarchical decision and control for continuous multitarget problem: Policy evaluation with action delay
Jiangcheng Zhu, Jun Zhu, Zhepei Wang, Shan Guo, and Chao Xu. 2018 · 2018
Earlier work this paper cites.
Sensor transfer: Learning optimal sensor effect image augmentation for sim-to-real domain adaptation
Alexandra Carlson, Katherine A Skinner, Ram Vasudevan, and Matthew Johnson-Roberson. 2019 · 2019
Earlier work this paper cites.
Challenges of real-world reinforcement learning
Gabriel Dulac-Arnold, Daniel Mankowitz, and Todd Hester. 2019 · 2019
Earlier work this paper cites.
Guidelines for reinforcement learning in healthcare
Omer Gottesman, Fredrik Johansson, Matthieu Komorowski, Aldo Faisal, David Sontag, Finale Doshi-Velez, and Leo Anthony Celi. 2019 · 2019
Earlier work this paper cites.
Recsim: A configurable simulation platform for recommender systems
Eugene Ie, Chih-wei Hsu, Martin Mladenov, Vihan Jain, Sanmit Narvekar, Jing Wang, Rui Wu, and Craig Boutilier. 2019a · 2019
Earlier work this paper cites.
Are we making real progress in simulated environments? measuring the sim2real gap in embodied visual navigation
Abhishek Kadian, Joanne Truong, Aaron Gokaslan, Alexander Clegg, Erik Wijmans, Stefan Lee, Manolis Savva, Sonia Chernova, and Dhruv Batra. 2019 · 2019
Earlier work this paper cites.
Reinforcement learning based VNF scheduling with end-to-end delay guarantee. In 2019 IEEE/CIC International Conference on Communications in China (ICCC) . IEEE, 572–577
Junling Li, Weisen Shi, Ning Zhang, and Xuemin Sherman Shen. 2019 · 2019
Earlier work this paper cites.
Safe reinforcement learning with model uncertainty estimates. In 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 8662–8668
Björn Lütjens, Michael Everett, and Jonathan P How. 2019 · 2019
Earlier work this paper cites.
Benchmarking Safe Exploration in Deep Reinforcement Learning
Alex Ray, Joshua Achiam, and Dario Amodei. 2019 · 2019
Earlier work this paper cites.
Mvfst-rl: An asynchronous rl framework for congestion control with delayed actions
Viswanath Sivakumar, Olivier Delalleau, Tim Rocktäschel, Alexander H Miller, Heinrich Küttler, Nantas Nardelli, Mike Rabbat, Joelle Pineau, and Sebastian Riedel. 2019 · 2019
Earlier work this paper cites.
Action robust reinforcement learning and applications in continuous control. In International Conference on Machine Learning . PMLR, 6215–6224
Chen Tessler, Yonathan Efroni, and Shie Mannor. 2019 · 2019
Earlier work this paper cites.
A survey on traffic signal control methods
Hua Wei, Guanjie Zheng, Vikash Gayah, and Zhenhui Li. 2019c · 2019
Earlier work this paper cites.
Universal domain adaptation. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 2720–2729
Kaichao You, Mingsheng Long, Zhangjie Cao, Jianmin Wang, and Michael I Jordan. 2019 · 2019
Earlier work this paper cites.
Vr-goggles for robots: Real-to-sim domain adaptation for visual control
Jingwei Zhang, Lei Tai, Peng Yun, Yufeng Xiong, Ming Liu, Joschka Boedecker, and Wolfram Burgard. 2019b · 2019
Earlier work this paper cites.
A study on challenges of testing robotic systems. In 2020 IEEE 13th International Conference on Software Testing, Validation and Verification (ICST) . IEEE, 96–107
Afsoon Afzal, Claire Le Goues, Michael Hilton, and Christopher Steven Timperley. 2020 · 2020
Earlier work this paper cites.
Learning precise 3d manipulation from multiple uncalibrated cameras. In 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 4616–4622
Iretiayo Akinola, Jacob Varley, and Dmitry Kalashnikov. 2020 · 2020
Earlier work this paper cites.
Reinforcement learning framework for delay sensitive energy harvesting wireless sensor networks
Hanan Al-Tous and Imad Barhumi. 2020 · 2020
Earlier work this paper cites.
Reinforcement learning with random delays. In International conference on learning representations
Yann Bouteiller, Simon Ramstedt, Giovanni Beltrame, Christopher Pal, and Jonathan Binas. 2020 · 2020
Earlier work this paper cites.
Uncertainty-aware action advising for deep reinforcement learning agents. In Proceedings of the AAAI conference on artificial intelligence , Vol. 34. 5792–5799
Felipe Leno Da Silva, Pablo Hernandez-Leal, Bilal Kartal, and Matthew E Taylor. 2020 · 2020
Earlier work this paper cites.
An imitation from observation approach to transfer learning with dynamics mismatch
Siddharth Desai, Ishan Durugkar, Haresh Karnan, Garrett Warnell, Josiah Hanna, and Peter Stone. 2020a · 2020
Earlier work this paper cites.
Stochastic grounded action transformation for robot learning in simulation. In 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 6106–6111
Siddharth Desai, Haresh Karnan, Josiah P Hanna, Garrett Warnell, and Peter Stone. 2020b · 2020
Earlier work this paper cites.
Longitudinal vehicle speed estimation for four-wheel-independently-actuated electric vehicles based on multi-sensor fusion
Xiaolin Ding, Zhenpo Wang, Lei Zhang, and Cong Wang. 2020 · 2020
Earlier work this paper cites.
Assistive gym: A physics simulation framework for assistive robotics. In 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 10169–10176
Zackory Erickson, Vamsee Gangaram, Ariel Kapusta, C Karen Liu, and Charles C Kemp. 2020 · 2020
Earlier work this paper cites.
Coupled real-synthetic domain adaptation for real-world deep depth enhancement
Xiao Gu, Yao Guo, Fani Deligianni, and Guang-Zhong Yang. 2020 · 2020
Earlier work this paper cites.
Utilizing reinforcement learning to autonomously mange buffers in a delay tolerant network node. In 2020 IEEE Aerospace Conference . IEEE, 1–8
Elizabeth Harkavy and Marc Sanchez Net. 2020 · 2020
Earlier work this paper cites.
Deep reinforcement learning for intelligent transportation systems: A survey
Ammar Haydari and Yasin Yılmaz. 2020 · 2020
Earlier work this paper cites.
LiDAR Object Detection and-Sensor Fusion in Simulation Environments Sensor modelling towards advancements in Real2Sim-Sim2Real
Christopher Höglind and Mahan Vahid Roudsari. 2020 · 2020
Earlier work this paper cites.
Self-supervised sim-to-real adaptation for visual robotic manipulation. In 2020 IEEE international conference on robotics and automation (ICRA) . IEEE, 2718–2724
Rae Jeong, Yusuf Aytar, David Khosid, Yuxiang Zhou, Jackie Kay, Thomas Lampe, Konstantinos Bousmalis, and Francesco Nori. 2020 · 2020
Earlier work this paper cites.
Sim2real predictivity: Does evaluation in simulation predict real-world performance?
Abhishek Kadian, Joanne Truong, Aaron Gokaslan, Alexander Clegg, Erik Wijmans, Stefan Lee, Manolis Savva, Sonia Chernova, and Dhruv Batra. 2020 · 2020
Cited alongside, same era.
Reinforced grounded action transformation for sim-to-real transfer. In 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 4397–4402
Haresh Karnan, Siddharth Desai, Josiah P Hanna, Garrett Warnell, and Peter Stone. 2020 · 2020
Cited alongside, same era.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Cited alongside, same era.
Delay-aware VNF scheduling: A reinforcement learning approach with variable action set
Junling Li, Weisen Shi, Ning Zhang, and Xuemin Shen. 2020 · 2020
Cited alongside, same era.
Off-policy learning in two-stage recommender systems. In Proceedings of The Web Conference 2020 . 463–473
Dinov2: Learning robust visual features without supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, et al · 2023
Later among the works it cites.
Adaptive Reinforcement Learning with LLM-augmented Reward Functions
Alex Place. 2023 · 2023
Later among the works it cites.
Bridging the Reality Gap Between Virtual and Physical Environments Through Reinforcement Learning
Mahesh Ranaweera and Qusay H. Mahmoud. 2023 · 2023
Later among the works it cites.
Sim-to-real via latent prediction: Transferring visual non-prehensile manipulation policies
Carlo Rizzardo, Fei Chen, and Darwin Caldwell. 2023 · 2023
Later among the works it cites.
Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning
Suraj Singireddy, Precious Nwaorgu, Andre Beckus, Aden McKinney, Chinwendu Enyioha, Sumit Kumar Jha, George K Atia, and Alvaro Velasquez. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jiaqi Ma, Zhe Zhao, Xinyang Yi, Ji Yang, Minmin Chen, Jiaxi Tang, Lichan Hong, and Ed H Chi. 2020 · 2020
Cited alongside, same era.
Active domain randomization. In Conference on Robot Learning . PMLR, 1162–1176
Bhairav Mehta, Manfred Diaz, Florian Golemo, Christopher J Pal, and Liam Paull. 2020 · 2020
Cited alongside, same era.
Intervention design for effective sim2real transfer
Melissa Mozifian, Amy Zhang, Joelle Pineau, and David Meger. 2020 · 2020
Cited alongside, same era.
Simulation-based reinforcement learning for real-world autonomous driving. In 2020 IEEE international conference on robotics and automation (ICRA) . IEEE, 6411–6418
Błażej Osiński, Adam Jakubowski, Paweł Zięcina, Piotr Miłoś, Christopher Galias, Silviu Homoceanu, and Henryk Michalewski. 2020 · 2020
Cited alongside, same era.
Sim-to-Real quadrotor landing via sequential deep Q-Networks and domain randomization
Riccardo Polvara, Massimiliano Patacchiola, Marc Hanheide, and Gerhard Neumann. 2020 · 2020
Cited alongside, same era.
Deepdrive Zero
Craig Quiter. 2020 · 2020
Cited alongside, same era.
Rl-cyclegan: Reinforcement learning aware simulation-to-real. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11157–11166
Kanishka Rao, Chris Harris, Alex Irpan, Sergey Levine, Julian Ibarz, and Mohi Khansari. 2020 · 2020
Cited alongside, same era.
A sim2real deep learning approach for the transformation of images from multiple vehicle-mounted cameras to a semantically segmented image in bird’s eye view. In 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) . IEEE, 1–7
Lennart Reiher, Bastian Lampe, and Lutz Eckstein. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
AutoVRL: A High Fidelity Autonomous Ground Vehicle Simulator for Sim-to-Real Deep Reinforcement Learning
Shathushan Sivashangaran, Apoorva Khairnar, and Azim Eskandarian. 2023 · 2023
Later among the works it cites.
Dynamic collaborative optimization of end-to-end delay and power consumption in wireless sensor networks for smart distribution grids
Wei Sun, Lei Zhang, Qiushuo Lv, Zhi Liu, Weitao Li, and Qiyue Li. 2023 · 2023
Later among the works it cites.
DROPO: Sim-to-real transfer with offline domain randomization
Gabriele Tiboni, Karol Arndt, and Ville Kyrki. 2023 · 2023
Later among the works it cites.
What truly matters in trajectory prediction for autonomous driving?
Phong Tran, Haoran Wu, Cunjun Yu, Panpan Cai, Sifa Zheng, and David Hsu. 2023 · 2023
Later among the works it cites.
Rethinking sim2real: Lower fidelity simulation leads to higher sim2real transfer in navigation. In Conference on Robot Learning . PMLR, 859–870
Joanne Truong, Max Rudolph, Naoki Harrison Yokoyama, Sonia Chernova, Dhruv Batra, and Akshara Rai. 2023 · 2023
Later among the works it cites.
Two-layer adaptive signal control framework for large-scale dynamically-congested networks: Combining efficient Max Pressure with Perimeter Control
Dimitrios Tsitsokas, Anastasios Kouvelas, and Nikolas Geroliminis. 2023 · 2023
Later among the works it cites.
Transfer from Imprecise and Abstract Models to Autonomous Technologies (TIAMAT)
Alvaro Velasquez. 2023 · 2023
Later among the works it cites.
Text2reward: Automated dense reward function generation for reinforcement learning
Tianbao Xie, Siheng Zhao, Chen Henry Wu, Yitao Liu, Qian Luo, Victor Zhong, Yanchao Yang, and Tao Yu. 2023 · 2023
Later among the works it cites.
Diffscene: Diffusion-based safety-critical scenario generation for autonomous vehicles. In The Second Workshop on New Frontiers in Adversarial Machine Learning
Chejian Xu, Ding Zhao, Alberto Sangiovanni-Vincentelli, and Bo Li. 2023 · 2023
Later among the works it cites.
A Real-World Reinforcement Learning Framework for Safe and Human-Like Tactical Decision-Making
Muharrem Ugur Yavas, Tufan Kumbasar, and Nazim Kemal Ure. 2023 · 2023
Later among the works it cites.
KuaiSim: A Comprehensive Simulator for Recommender Systems
Kesen Zhao, Shuchang Liu, Qingpeng Cai, Xiangyu Zhao, Ziru Liu, Dong Zheng, Peng Jiang, and Kun Gai. 2023 · 2023
Later among the works it cites.
Guided conditional diffusion for controllable traffic simulation. In 2023 IEEE international conference on robotics and automation (ICRA) . IEEE, 3560–3566
Ziyuan Zhong, Davis Rempe, Danfei Xu, Yuxiao Chen, Sushant Veer, Tong Che, Baishakhi Ray, and Marco Pavone. 2023 · 2023
Later among the works it cites.
Safety-Driven Deep Reinforcement Learning Framework for Cobots: A Sim2Real Approach. In 2024 10th International Conference on Control, Decision and Information Technologies (CoDIT) . IEEE, 2917–2923
Ammar N Abbas, Shakra Mehak, Georgios C Chasparis, John D Kelleher, Michael Guilfoyle, Maria Chiara Leva, and Aswin K Ramasubramanian. 2024 · 2024
Later among the works it cites.
Dynamic traffic signal control for heterogeneous traffic conditions using max pressure and reinforcement learning
Amit Agarwal, Deorishabh Sahu, Rishabh Mohata, Kuldeep Jeengar, Anuj Nautiyal, and Dhish Kumar Saxena. 2024 · 2024
Later among the works it cites.
Genesis: A Universal and Generative Physics Engine for Robotics and Beyond
Genesis Authors. 2024 · 2024
Later among the works it cites.
Donghoon Baek, Youngwoo Sim, Amartya Purushottam, Saurabh Gupta, and Joao Ramos. 2024 · 2024
Later among the works it cites.
Synthetic Vision: Training Vision-Language Models to Understand Physics
Vahid Balazadeh, Mohammadmehdi Ataei, Hyunmin Cheong, Amir Hosein Khasahmadi, and Rahul G Krishnan. 2024 · 2024
Later among the works it cites.
High-Fidelity Simulation of a Cartpole for Sim-to-Real Deep Reinforcement Learning. In 2024 4th Interdisciplinary Conference on Electrics and Computer (INTCEC) . IEEE, 1–6
Linus Bantel, Peter Domanski, and Dirk Pflüger. 2024 · 2024
Later among the works it cites.
EvoPrompting: language models for code-level neural architecture search
Angelica Chen, David Dohan, and David So. 2024a · 2024
Later among the works it cites.
Chen Chen, Ziyao Liu, Weifeng Jiang, Si Qi Goh, and KwoK-Yan Lam. 2024c · 2024
Later among the works it cites.
Dong Chen and Yanbo Huang. 2024 · 2024
Later among the works it cites.
RLingua: Improving Reinforcement Learning Sample Efficiency in Robotic Manipulations With Large Language Models
Liangliang Chen, Yutian Lei, Shiyu Jin, Ying Zhang, and Liangjun Zhang. 2024b · 2024
Later among the works it cites.
SynTraC: A Synthetic Dataset for Traffic Signal Control from Traffic Monitoring Cameras
Tiejin Chen, Prithvi Shirke, Bharatesh Chakravarthi, Arpitsinh Vaghela, Longchao Da, Duo Lu, Yezhou Yang, and Hua Wei. 2024d · 2024
Later among the works it cites.
PyBullet: Real-Time Physics Simulation
Erwin Coumans. 2024 · 2024
Later among the works it cites.
Longchao Da, Tiejin Chen, Lu Cheng, and Hua Wei. 2024a · 2024
Later among the works it cites.
Segment as You Wish–Free-Form Language-Based Segmentation for Medical Images
Longchao Da, Rui Wang, Xiaojian Xu, Parminder Bhatia, Taha Kass-Hout, Hua Wei, and Cao Xiao. 2024e · 2024
Later among the works it cites.
Local Policies Enable Zero-shot Long-horizon Manipulation
Murtaza Dalal, Min Liu, Walter Talbott, Chen Chen, Deepak Pathak, Jian Zhang, and Ruslan Salakhutdinov. 2024 · 2024
Later among the works it cites.
Understanding biases in chatgpt-based recommender systems: Provider fairness, temporal stability, and recency
Yashar Deldjoo. 2024 · 2024
Later among the works it cites.
Learning Vision-Based Bipedal Locomotion for Challenging Terrain. In 2024 IEEE International Conference on Robotics and Automation (ICRA) . 56–62
Helei Duan, Bikram Pandit, Mohitvishnu S. Gadde, Bart Van Marum, Jeremy Dao, Chanho Kim, and Alan Fern. 2024 · 2024
Later among the works it cites.
Domain adaption as auxiliary task for sim-to-real transfer in vision-based neuro-robotic control. In 2024 International Joint Conference on Neural Networks (IJCNN) . IEEE, 1–8
Connor Gäde, Jan-Gerrit Habekost, and Stefan Wermter. 2024 · 2024
Later among the works it cites.
Causal inference in recommender systems: A survey and future directions
Chen Gao, Yu Zheng, Wenjie Wang, Fuli Feng, Xiangnan He, and Yong Li. 2024 · 2024
Later among the works it cites.
Robust Gymnasium: A Unified Modular Benchmark for Robust Reinforcement Learning
Shangding Gu, Laixi Shi, Muning Wen, Ming Jin, Eric Mazumdar, Yuejie Chi, Adam Wierman, and Costas Spanos. 2024a · 2024
Later among the works it cites.
Sim2Real Rope Cutting With a Surgical Robot Using Vision-Based Reinforcement Learning
Mustafa Haiderbhai, Radian Gondokaryono, Andrew Wu, and Lueder A Kahrs. 2024 · 2024
Later among the works it cites.
Bridging the sim-to-real gap from the information bottleneck perspective. In 8th Annual Conference on Robot Learning
Haoran He, Peilin Wu, Chenjia Bai, Hang Lai, Lingxiao Wang, Ling Pan, Xiaolin Hu, and Weinan Zhang. 2024 · 2024
Later among the works it cites.
Gensim2: Scaling robot data generation with multi-modal and reasoning llms
Pu Hua, Minghuan Liu, Annabella Macaluso, Yunfeng Lin, Weinan Zhang, Huazhe Xu, and Lirui Wang. 2024 · 2024
Later among the works it cites.
DISCOVERSE: Efficient Robot Simulation in Complex High-Fidelity Environments
Yufei Jia, Guangyu Wang, Yuhang Dong, Junzhe Wu, Yupei Zeng, Haizhou Ge, Kairui Ding, Zike Yan, Weibin Gu, Chuxuan Li, Ziming Wang, Yunjie Cheng, Wei Sui, Ruqi Huang, and Guyue Zhou. 2024 · 2024
Later among the works it cites.
Gen2sim: Scaling up robot learning in simulation with generative models. In 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 6672–6679
Pushkal Katara, Zhou Xian, and Katerina Fragkiadaki. 2024 · 2024
Later among the works it cites.
A multimodal latent-features-based service recommendation system for the social Internet of Things
Amar Khelloufi, Huansheng Ning, Abdenacer Naouri, Abdelkarim Ben Sada, Attia Qammar, Abdelkader Khalil, Lingfeng Mao, and Sahraoui Dhelim. 2024 · 2024
Later among the works it cites.
Error Detection and Constraint Recovery in Hierarchical Multi-Label Classification without Prior Knowledge. In Proceedings of the 33rd ACM International Conference on Information and Knowledge Management, CIKM 2024, Boise, ID, USA, October 21-25, 2024 , Edoardo Serra and Francesca Spezzano (Eds.). ACM, 3842–3846
Joshua Shay Kricheli, Khoa Vo, Aniruddha Datta, Spencer Ozgur, and Paulo Shakarian. 2024 · 2024
Later among the works it cites.
Jonathan Wilder Lavington, Ke Zhang, Vasileios Lioutas, Matthew Niedoba, Yunpeng Liu, Dylan Green, Saeid Naderiparizi, Xiaoxuan Liang, Setareh Dabiri, Adam Ścibior, Berend Zwartsenberg, and Frank Wood. 2024 · 2024
Later among the works it cites.
NeuronsGym: A Hybrid Framework and Benchmark for Robot Navigation With Sim2Real Policy Learning
Haoran Li, Guangzheng Hu, Shasha Liu, Mingjun Ma, Yaran Chen, and Dongbin Zhao. 2024a · 2024
Later among the works it cites.
Optimizing automated picking systems in warehouse robots using machine learning
Keqin Li, Jin Wang, Xubo Wu, Xirui Peng, Runmian Chang, Xiaoyu Deng, Yiwen Kang, Yue Yang, Fanghao Ni, and Bo Hong. 2024b · 2024
Later among the works it cites.
Uncertainty Quantification for In-Context Learning of Large Language Models. In Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers) . 3357–3370
Chen Ling, Xujiang Zhao, Xuchao Zhang, Wei Cheng, Yanchi Liu, Yiyou Sun, Mika Oishi, Takao Osaki, Katsushi Matsuda, Jie Ji, et al · 2024
Later among the works it cites.
TD3 Based Collision Free Motion Planning for Robot Navigation
Hao Liu, Yi Shen, Chang Zhou, Yuelin Zou, Zijun Gao, and Qi Wang. 2024a · 2024
Later among the works it cites.
Upper and lower bounds for distributionally robust off-dynamics reinforcement learning
Zhishuai Liu, Weixin Wang, and Pan Xu. 2024b · 2024
Later among the works it cites.
Distributionally robust off-dynamics reinforcement learning: Provable efficiency with linear function approximation. In International Conference on Artificial Intelligence and Statistics . PMLR, 2719–2727
Zhishuai Liu and Pan Xu. 2024 · 2024
Later among the works it cites.
Cheap and quick: Efficient vision-language instruction tuning for large language models
Gen Luo, Yiyi Zhou, Tianhe Ren, Shengxin Chen, Xiaoshuai Sun, and Rongrong Ji. 2024 · 2024
Later among the works it cites.
Quantifying the sim2real gap for GPS and IMU sensors
Ishaan Mahajan, Huzaifa Unjhawala, Harry Zhang, Zhenhao Zhou, Aaron Young, Alexis Ruiz, Stefan Caldararu, Nevindu Batagoda, Sriram Ashokkumar, and Dan Negrut. 2024 · 2024
Later among the works it cites.
Libsignal: An open library for traffic signal control
Hao Mei, Xiaoliang Lei, Longchao Da, Bin Shi, and Hua Wei. 2024 · 2024
Later among the works it cites.
Wisdom from the crowd: Can recommender systems predict employee turnover and its destinations?
Hanyi Min, Baojiang Yang, David G Allen, Alicia A Grandey, and Mengqiao Liu. 2024 · 2024
Later among the works it cites.
Scalable Semantic Non-Markovian Simulation Proxy for Reinforcement Learning. In 18th IEEE International Conference on Semantic Computing, ICSC 2024, Laguna Hills, CA, USA, February 5-7, 2024 . IEEE, 183–190
Kaustuv Mukherji, Devendra Parkar, Lahari Pokala, Dyuman Aditya, Paulo Shakarian, and Clark Dorman. 2024 · 2024
Later among the works it cites.
Evolutionary reward design and optimization with multimodal large language models. In Proceedings of the 3rd Workshop on Advances in Language and Vision Research (ALVR) . 202–208
Ali Narin. 2024 · 2024
Later among the works it cites.
ChatGPT Label: Comparing the Quality of Human-Generated and LLM-Generated Annotations in Low-resource Language NLP Tasks
Arbi Haza Nasution and Aytug Onan. 2024 · 2024
Later among the works it cites.
OpenAI Gym Retro
OpenAI. 2024 · 2024
Later among the works it cites.
Scalable reinforcement learning framework for traffic signal control under communication delays
Aoyu Pang, Maonan Wang, Yirong Chen, Man-On Pun, and Michael Lepech. 2024 · 2024
Later among the works it cites.
A Survey on Sim-to-Real Transfer Methods for Robotic Manipulation. In 2024 IEEE 22nd Jubilee International Symposium on Intelligent Systems and Informatics (SISY) . IEEE, 000259–000266
Andrei Pitkevich and Ilya Makarov. 2024 · 2024
Later among the works it cites.
Saynav: Grounding large language models for dynamic planning to navigation in new environments. In Proceedings of the International Conference on Automated Planning and Scheduling , Vol. 34. 464–474
Abhinav Rajvanshi, Karan Sikka, Xiao Lin, Bhoram Lee, Han-Pang Chiu, and Alvaro Velasquez. 2024 · 2024
Later among the works it cites.
InfiniteWorld: A Unified Scalable Simulation Framework for General Visual-Language Robot Interaction
Pengzhen Ren, Min Li, Zhen Luo, Xinshuai Song, Ziwei Chen, Weijia Liufu, Yixuan Yang, Hao Zheng, Rongtao Xu, Zitong Huang, et al · 2024
Later among the works it cites.
Kanghyun Ryu, Qiayuan Liao, Zhongyu Li, Koushil Sreenath, and Negar Mehr. 2024 · 2024
Later among the works it cites.
Selected Issues, Methods, and Trends in the Energy Consumption of Industrial Robots
Agnieszka Sękala, Tomasz Blaszczyk, Krzysztof Foit, and Gabriel Kost. 2024 · 2024
Later among the works it cites.
Shengjie Sun, Runze Liu, Jiafei Lyu, Jing-Wen Yang, Liangpeng Zhang, and Xiu Li. 2024 · 2024
Later among the works it cites.
Robust Offline Reinforcement Learning with Linearly Structured f f -Divergence Regularization
Cheng Tang, Zhishuai Liu, and Pan Xu. 2024 · 2024
Later among the works it cites.
Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RL
Andrew Wagenmaker, Kevin Huang, Liyiming Ke, Byron Boots, Kevin Jamieson, and Abhishek Gupta. 2024 · 2024
Later among the works it cites.
Return augmented decision transformer for off-dynamics reinforcement learning
Ruhan Wang, Yu Yang, Zhishuai Liu, Dongruo Zhou, and Pan Xu. 2024 · 2024
Later among the works it cites.
Deep learning-based personalized learning recommendation system design for" T++" Guzheng Pedagogy
Xingyue Wang. 2024 · 2024
Later among the works it cites.
Metacognitive AI: Framework and the Case for a Neurosymbolic Approach. In Neural-Symbolic Learning and Reasoning - 18th International Conference, NeSy 2024, Barcelona, Spain, September 9-12, 2024, Proceedings, Part II (Lecture Notes in Computer Science, Vol. 14980) , Tarek R. Besold, Artur d’Avila Garcez, Ernesto Jiménez-Ruiz, Roberto Confalonieri, Pranava Madhyastha, and Benedikt Wagner (Eds.). Springer, 60–67
Hua Wei, Paulo Shakarian, Christian Lebiere, Bruce A. Draper, Nikhil Krishnaswamy, and Sergei Nirenburg. 2024 · 2024
Later among the works it cites.
Integrating AI for Enhanced Exploration of Video Recommendation Algorithm via Improved Collaborative Filtering
Yafei Xiang, Shuning Huo, Yichao Wu, Yulu Gong, and Mengran Zhu. 2024 · 2024
Later among the works it cites.
Kangming Xu, Huiming Zhou, Haotian Zheng, Mingwei Zhu, and Qi Xin. 2024c · 2024
Later among the works it cites.
Lanling Xu, Junjie Zhang, Bingqian Li, Jinpeng Wang, Mingchen Cai, Wayne Xin Zhao, and Ji-Rong Wen. 2024b · 2024
Later among the works it cites.
Ped-mp: A pedestrian-friendly max-pressure signal control policy for city networks
Te Xu, Yashveer Bika, and Michael W Levin. 2024a · 2024
Later among the works it cites.
CoMAL: Collaborative Multi-Agent Large Language Models for Mixed-Autonomy Traffic
Huaiyuan Yao, Longchao Da, Vishnu Nandam, Justin Turnau, Zhiwei Liu, Linsey Pang, and Hua Wei. 2024 · 2024
Later among the works it cites.
Natural Language Can Help Bridge the Sim2Real Gap
Albert Yu, Adeline Foote, Raymond Mooney, and Roberto Martín-Martín. 2024a · 2024
Later among the works it cites.
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning
Zhecheng Yuan, Tianming Wei, Shuiqi Cheng, Gu Zhang, Yuanpei Chen, and Huazhe Xu. 2024 · 2024
Later among the works it cites.
A simple framework for intrinsic reward-shaping for rl using llm feedback
Alex Zhang, Ananya Parashar, and Dwaipayan Saha. 2024b · 2024
Later among the works it cites.
Di Zhang, Xiaoshui Huang, Dongzhan Zhou, Yuqiang Li, and Wanli Ouyang. 2024a · 2024
Later among the works it cites.
LLM-Optic: Unveiling the Capabilities of Large Language Models for Universal Visual Grounding
Haoyu Zhao, Wenhang Ge, and Ying-cong Chen. 2024 · 2024
Later among the works it cites.
A survey on efficient inference for large language models
Zixuan Zhou, Xuefei Ning, Ke Hong, Tianyu Fu, Jiaming Xu, Shiyao Li, Yuming Lou, Luning Wang, Zhihang Yuan, Xiuhong Li, et al · 2024
Later among the works it cites.
RRLS: Robust Reinforcement Learning Suite
Adil Zouitine, David Bertoin, Pierre Clavier, Matthieu Geist, and Emmanuel Rachelson. 2024 · 2024
Later among the works it cites.
Robust Autonomy Emerges from Self-Play
Marco Cusumano-Towner, David Hafner, Alex Hertzberg, Brody Huval, Aleksei Petrenko, Eugene Vinitsky, Erik Wijmans, Taylor Killian, Stuart Bowers, Ozan Sener, et al · 2025
Closest in time.
Off-dynamics reinforcement learning via domain adaptation and reward augmented imitation
Yihong Guo, Yixuan Wang, Yuanyuan Shi, Pan Xu, and Anqi Liu. 2025 · 2025
Closest in time.
Minimax optimal and computationally efficient algorithms for distributionally robust offline reinforcement learning
Zhishuai Liu and Pan Xu. 2025 · 2025
Closest in time.
Energy-efficient robot configuration and motion planning using genetic algorithm and particle swarm optimization
Kazuki Nonoyama, Ziang Liu, Tomofumi Fujiwara, Md Moktadir Alam, and Tatsushi Nishi. 2022 · 2074
Closest in time.