Fetching the paper…
Reading the bibliography…
Autopilot systems are typically composed of an "inner loop" providing stability and control, while an "outer loop" is responsible for mission-level objectives, e.g.
1942
Earlier work this paper cites.
H. P. Whitaker, J. Yamron, and A. Kezer, Design of model-reference adaptive control systems for aircraft . Massachusetts Institute of Technology, Instrumentation Laboratory, 1958
1958
Earlier work this paper cites.
O. Miglino, H. H. Lund, and S. Nolfi, “Evolving mobile robots in simulated and real environments,” Artificial life , vol. 2, no. 4, pp. 417–434, 1995
1995
Earlier work this paper cites.
N. Jakobi, P. Husbands, and I. Harvey, “Noise and the reality gap: The use of simulation in evolutionary robotics,” Advances in artificial life , pp. 704–720, 1995
1995
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press Cambridge, 1998, vol. 1, no. 1
1998
Earlier work this paper cites.
D. J. Leith and W. E. Leithead, “Survey of gain-scheduling analysis and design,” International journal of control , vol. 73, no. 11, pp. 1001–1025, 2000
2000
Earlier work this paper cites.
S. Bouabdallah, P. Murrieri, and R. Siegwart, “Design and control of an indoor micro quadrotor,” in Robotics and Automation, 2004. Proceedings. ICRA’04. 2004 IEEE International Conference on , vol. 5. IEEE, 2004, pp. 4393–4398
2004
Earlier work this paper cites.
N. Koenig and A. Howard, “Design and use paradigms for gazebo, an open-source multi-robot simulator,” in Intelligent Robots and Systems, 2004.(IROS 2004). Proceedings. 2004 IEEE/RSJ International Conference on , vol. 3. IEEE, pp. 2149–2154
2004
Earlier work this paper cites.
P. S. Williams-Hayes, “Flight test implementation of a second generation intelligent flight control system,” infotech@ Aerospace, AIAA-2005-6995 , pp. 26–29, 2005
2005
Earlier work this paper cites.
S. L. Waslander, G. M. Hoffmann, J. S. Jang, and C. J. Tomlin, “Multi-agent quadrotor testbed control design: Integral sliding mode vs. reinforcement learning,” in Intelligent Robots and Systems, 2005.(IROS 2005). 2005 IEEE/RSJ International Conference on . IEEE, 2005, pp. 3712–3717
2005
Earlier work this paper cites.
T. Dierks and S. Jagannathan, “Output feedback control of a quadrotor uav using neural networks,” IEEE transactions on neural networks , vol. 21, no. 1, pp. 50–66, 2010
2010
Earlier work this paper cites.
J. F. Shepherd III and K. Tumer, “Robust neuro-control for a micro quadrotor,” in Proceedings of the 12th annual conference on Genetic and evolutionary computation . ACM, 2010, pp. 1131–1138
2010
Cited alongside, same era.
N. Hovakimyan, C. Cao, E. Kharisov, E. Xargay, and I. M. Gregory, “L1 adaptive control for safety-critical systems,” IEEE Control Systems , vol. 31, no. 5, pp. 54–104, 2011
2011
Cited alongside, same era.
2013
Cited alongside, same era.
M. Fatan, B. L. Sefidgari, and A. V. Barenji, “An adaptive neuro pid for controlling the altitude of quadcopter robot,” in Methods and models in automation and robotics (mmar), 2013 18th international conference on . IEEE, 2013, pp. 662–665
2013
Cited alongside, same era.
F. Santoso, M. A. Garratt, and S. G. Anavatti, “State-of-the-art intelligent flight control systems in unmanned aerial vehicles,” IEEE Transactions on Automation Science and Engineering , 2017
2017
Later among the works it cites.
J. Hwangbo, I. Sa, R. Siegwart, and M. Hutter, “Control of a quadrotor with reinforcement learning,” IEEE Robotics and Automation Letters , vol. 2, no. 4, pp. 2096–2103, 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Openai baselines,” https://github.com/openai/baselines , 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International Conference on Machine Learning , 2015, pp. 1889–1897
2015
Cited alongside, same era.
K. N. Maleki, K. Ashenayi, L. R. Hook, J. G. Fuller, and N. Hutchins, “A reliable system design for nondeterministic adaptive controllers in small uav autopilots,” in Digital Avionics Systems Conference (DASC), 2016 IEEE/AIAA 35th . IEEE, 2016, pp. 1–5
2016
Cited alongside, same era.
A. Bobtsov, A. Guirik, M. Budko, and M. Budko, “Hybrid parallel neuro-controller for multirotor unmanned aerial vehicle,” in Ultra Modern Telecommunications and Control Systems and Workshops (ICUMT), 2016 8th International Congress on . IEEE, 2016, pp. 1–4
2016
Cited alongside, same era.
2016
Cited alongside, same era.
T. Gabor, L. Belzner, M. Kiermeier, M. T. Beck, and A. Neitz, “A simulation-based architecture for smart cyber-physical systems,” in Autonomic Computing (ICAC), 2016 IEEE International Conference on . IEEE, 2016, pp. 374–379
2016
Cited alongside, same era.
A. Karpathy, “Deep Reinforcement Learning: Pong from Pixels,” 2018. [Online]. Available: http://karpathy.github.io/2016/05/31/rl/
2016
Cited alongside, same era.
2018
Closest in time.
W. Koch, “GymFC,” https://github.com/wil3/gymfc , 2018
2018
Closest in time.
“BetaFlight,” 2018. [Online]. Available: https://github.com/betaflight/betaflight
2018
Closest in time.
“Protocol Buffers,” 2018. [Online]. Available: https://developers.google.com/protocol-buffers/
2018
Closest in time.
“ArduPilot,” 2018. [Online]. Available: http://ardupilot.org/
2018
Closest in time.
“gzserver doesn’t close disconnected sockets,” 2018. [Online]. Available: https://bitbucket.org/osrf/gazebo/issues/2397/gzserver-doesnt-close-disconnected-sockets
2018
Closest in time.