Fetching the paper…
Reading the bibliography…
Rapid advancements in deep learning have led to many recent breakthroughs.
A. G. Barto, R. S. Sutton, C. W. Anderson, Neuronlike adaptive elements that can solve difficult learning control problems, IEEE transactions on systems, man, and cybernetics (5) (1983) 834–846
1983
Earlier work this paper cites.
A. W. Moore, Efficient memory-based learning for robot control (1990)
1990
Earlier work this paper cites.
R. A. Jacobs, M. I. Jordan, S. J. Nowlan, G. E. Hinton, et al., Adaptive mixtures of local experts., Neural computation 3 (1) (1991) 79–87
1991
Earlier work this paper cites.
M. I. Jordan, L. Xu, Convergence results for the EM approach to mixtures of experts architectures, Neural networks 8 (9) (1995) 1409–1431
1995
Earlier work this paper cites.
L. Breiman, N. Shang, Born again trees, University of California, Berkeley, Berkeley, CA, Technical Report 1 (1996) 2
1996
Earlier work this paper cites.
R. Kohavi, Scaling up the accuracy of naive-bayes classifiers: A decision-tree hybrid., in: Kdd, Vol. 96, 1996, pp. 202–207
1996
Earlier work this paper cites.
R. S. Sutton, Generalization in reinforcement learning: Successful examples using sparse coarse coding, in: Advances in neural information processing systems, 1996, pp. 1038–1044
1996
Earlier work this paper cites.
S. Schaal, Is imitation learning the route to humanoid robots?, Trends in cognitive sciences (1999)
1999
Earlier work this paper cites.
P. Abbeel, A. Y. Ng, Apprenticeship learning via inverse reinforcement learning, in: ICML, 2004
2004
Earlier work this paper cites.
C. Buciluǎ, R. Caruana, A. Niculescu-Mizil, Model compression, in: Proceedings of the 12th ACM SIGKDD international conference on Knowledge discovery and data mining, ACM, 2006, pp. 535–541
2006
Earlier work this paper cites.
L. De Moura, N. Bjørner, Z3: An efficient SMT solver, in: International conference on Tools and Algorithms for the Construction and Analysis of Systems, Springer, 2008, pp. 337–340
2008
Earlier work this paper cites.
A. Biere, M. Heule, H. van Maaren, T. Walsh (Eds.), Handbook of Satisfiability, Vol. 185 of Frontiers in Artificial Intelligence and Applications, IOS Press, 2009
2009
Earlier work this paper cites.
X. Niuniu, L. Yuxun, Notice of retraction: Review of decision trees, in: 2010 3rd international conference on computer science and information technology, Vol. 5, IEEE, 2010, pp. 105–109
2010
Earlier work this paper cites.
S. Ross, G. Gordon, D. Bagnell, A reduction of imitation learning and structured prediction to no-regret online learning, in: Proceedings of the fourteenth international conference on artificial intelligence and statistics, 2011, pp. 627–635
2011
Earlier work this paper cites.
A. Frank, A. Asuncion, et al., Uci machine learning repository, 2010, URL http://archive.ics.uci.edu/ml 15 (2011) 22
2011
Earlier work this paper cites.
S. E. Yuksel, J. N. Wilson, P. D. Gader, Twenty years of mixture of experts, IEEE transactions on neural networks and learning systems 23 (8) (2012) 1177–1193
2012
Earlier work this paper cites.
O. Irsoy, O. T. Yıldız, E. Alpaydın, Soft decision trees, in: Proceedings of the 21st International Conference on Pattern Recognition (ICPR2012), IEEE, 2012, pp. 1819–1822
2012
Earlier work this paper cites.
S. B. Kotsiantis, Decision trees: a recent overview, Artificial Intelligence Review 39 (4) (2013) 261–283
2013
Earlier work this paper cites.
T. Hester, P. Stone, Texplore: real-time sample-efficient reinforcement learning for robots, Machine learning 90 (3) (2013) 385–429
2013
Earlier work this paper cites.
2015
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al., Human-level control through deep reinforcement learning, Nature 518 (7540) (2015) 529
2015
Cited alongside, same era.
2015
Cited alongside, same era.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al., Mastering the game of Go with deep neural networks and tree search, Nature 529 (7587) (2016) 484
2016
Cited alongside, same era.
J.-Z. Cheng, D. Ni, Y.-H. Chou, J. Qin, C.-M. Tiu, Y.-C. Chang, C.-S. Huang, D. Shen, C.-M. Chen, Computer-aided diagnosis with deep learning architecture: applications to breast lesions in us images and pulmonary nodules in ct scans, Scientific reports 6 (1) (2016) 1–13
2019
Closest in time.
X. Li, Z. Serlin, G. Yang, C. Belta, A formal methods approach to interpretable reinforcement learning for robotic planning, Science Robotics 4 (37) (2019)
2019
Closest in time.
J. Wang, S. Pandit, Towards high-level, verifiable autonomous behaviors with temporal specifications, in: 2019 IEEE National Aerospace and Electronics Conference (NAECON), IEEE, 2019, pp. 92–99
2019
Closest in time.
2019
Closest in time.
A. Koul, A. Fern, S. Greydanus, Learning finite state representations of recurrent policy networks, in: ICLR, 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
M. T. Ribeiro, S. Singh, C. Guestrin, "Why Should I Trust You?": Explaining the Predictions of Any Classifier, in: KDD, 2016
2016
Cited alongside, same era.
A. A. Rusu, S. G. Colmenarejo, Ç. Gülçehre, G. Desjardins, J. Kirkpatrick, R. Pascanu, V. Mnih, K. Kavukcuoglu, R. Hadsell, Policy distillation, in: 4th International Conference on Learning Representations, ICLR 2016, San Juan, Puerto Rico, May 2-4, 2016, Conference Track Proceedings, 2016
2016
Cited alongside, same era.
Z. C. Lipton, The mythos of model interpretability, arXiv preprint arXiv:1606.03490 (2016)
2016
Cited alongside, same era.
M. Cicero, A. Bilbily, E. Colak, T. Dowdell, B. Gray, K. Perampaladas, J. Barfett, Training and validating a deep convolutional neural network for computer-aided detection and classification of abnormalities on frontal chest radiographs, Investigative radiology 52 (5) (2017) 281–287
2017
Cited alongside, same era.
T. Kooi, G. Litjens, B. Van Ginneken, A. Gubern-Mérida, C. I. Sánchez, R. Mann, A. den Heeten, N. Karssemeijer, Large scale deep learning for computer aided detection of mammographic lesions, Medical image analysis 35 (2017) 303–312
2017
Cited alongside, same era.
P. Van Wesel, A. E. Goodloe, Challenges in the verification of reinforcement learning algorithms, Tech. rep. (2017)
2017
Cited alongside, same era.
2017
Cited alongside, same era.
R. Miotto, F. Wang, S. Wang, X. Jiang, J. T. Dudley, Deep learning for healthcare: review, opportunities and challenges, Briefings in bioinformatics 19 (6) (2018) 1236–1246
2018
Cited alongside, same era.
2019
Closest in time.
A. Zhang, Z. C. Lipton, L. Pineda, K. Azizzadenesheli, A. Anandkumar, L. Itti, J. Pineau, T. Furlanello, Learning causal state representations of partially observable environments, CoRR (2019)
2019
Closest in time.
C. Niu, F. Wu, S. Tang, S. Ma, G. Chen, Toward verifiable and privacy preserving machine learning prediction, IEEE Transactions on Dependable and Secure Computing (2020)
2020
Closest in time.
J. Törnblom, S. Nadjm-Tehrani, Formal verification of input-output mappings of tree ensembles, Science of Computer Programming 194 (2020) 102450
2020
Closest in time.
R. Roscher, B. Bohn, M. F. Duarte, J. Garcke, Explainable machine learning for scientific insights and discoveries, Ieee Access 8 (2020) 42200–42216
2020
Closest in time.
E. Puiutta, E. Veith, Explainable reinforcement learning: A survey, in: International cross-domain conference for machine learning and knowledge extraction, Springer, 2020, pp. 77–95
2020
Closest in time.
G. Amir, M. Schapira, G. Katz, Towards scalable verification of deep reinforcement learning, in: 2021 Formal Methods in Computer Aided Design (FMCAD), IEEE, 2021, pp. 193–203
2021
Closest in time.
A. Heuillet, F. Couthouis, N. Díaz-Rodríguez, Explainability in deep reinforcement learning, Knowledge-Based Systems 214 (2021) 106685
2021
Closest in time.
L. Wells, T. Bednarz, Explainable ai and reinforcement learning—a systematic review of current approaches and trends, Frontiers in artificial intelligence 4 (2021) 48
2021
Closest in time.
S. Shi, J. Li, G. Li, P. Pan, K. Liu, Xpm: An explainable deep reinforcement learning framework for portfolio management, in: Proceedings of the 30th ACM International Conference on Information & Knowledge Management, 2021, pp. 1661–1670
2021
Closest in time.
2021
Closest in time.
J. Gou, B. Yu, S. J. Maybank, D. Tao, Knowledge distillation: A survey, International Journal of Computer Vision 129 (6) (2021) 1789–1819
2021
Closest in time.
Z. Wang, Y. Wei, F. Wu, Knowledge distillation based cooperative reinforcement learning for connectivity preservation in uav networks, in: 2021 International Conference on UK-China Emerging Technologies (UCET), IEEE, 2021, pp. 171–176
2021
Closest in time.
A. Tsantekidis, N. Passalis, A. Tefas, Diversity-driven knowledge distillation for financial trading using deep reinforcement learning, Neural Networks 140 (2021) 193–202
2021
Closest in time.
Z. Gao, K. Xu, B. Ding, H. Wang, Knowru: Knowledge reuse via knowledge distillation in multi-agent reinforcement learning, Entropy 23 (8) (2021) 1043
2021
Closest in time.