Fetching the paper…
Reading the bibliography…
Computational agents support humans in many areas of life and are therefore found in heterogeneous contexts.
G.D. Forney, The Viterbi algorithm, Proceedings of the IEEE
1973
Earlier work this paper cites.
L. Rabiner and B. Juang, An introduction to hidden Markov models, IEEE ASSP Magazine
1986
Earlier work this paper cites.
E. Wilson, Active vibration analysis of thin-walled beams, PhD thesis, University of Virginia, 1991
1991
Earlier work this paper cites.
P.S. Meltzer, A. Kallioniemi and J.M. Trent, Chromosome alterations in human solid tumors, in: The Genetic Basis of Human Cancer
2002
Earlier work this paper cites.
P.R. Murray, K.S. Rosenthal, G.S. Kobayashi and M.A. Pfaller, Medical Microbiology
2002
Earlier work this paper cites.
M. Stamp, A revealing introduction to hidden markov models, Science
2004
Earlier work this paper cites.
D. Osswald, J. Martin, C. Burghart, R. Mikut, H. Wörn and G. Bretthauer, Integrating a flexible anthropomorphic, robot hand into the control, system of a humanoid robot, Robotics and Autonomous Systems
2004
Earlier work this paper cites.
C. Burghart, R. Mikut, R. Stiefelhagen, T. Asfour, H. Holzapfel, P. Steinhaus and R. Dillmann, A cognitive architecture for a humanoid robot: a first approach, in: 5th IEEE-RAS International Conference on Humanoid Robots, 2005
2005
Earlier work this paper cites.
F. Fernández, J. García and M. Veloso, Probabilistic Policy Reuse for inter-task transfer learning, Robotics and Autonomous Systems
2010
Earlier work this paper cites.
J. Peters, Policy gradient methods, Scholarpedia
2010
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra and M.A. Riedmiller, Playing Atari with Deep Reinforcement Learning, CoRR
2013
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G.S. Corrado and J. Dean, Distributed Representations of Words and Phrases and their Compositionality, in: Advances in Neural Information Processing Systems
2013
Earlier work this paper cites.
T. Mikolov, K. Chen, G. Corrado and J. Dean, Efficient Estimation of Word Representations in Vector Space, 2013
2013
Earlier work this paper cites.
S.P. Chatzis and D. Kosmopoulos, A Partially-Observable Markov Decision Process for Dealing with Dynamically Changing Environments, in: Artificial Intelligence Applications and Innovations
2014
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A.A. Rusu, J. Veness, M.G. Bellemare, A. Graves, M. Riedmiller, A.K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg and D. Hassabis, Human-level control through deep reinforcement learning, Nature
2015
Cited alongside, same era.
J. Garcıa and F. Fernández, A comprehensive survey on safe reinforcement learning, Journal of Machine Learning Research
2015
Cited alongside, same era.
2015
Cited alongside, same era.
D. Silver, A. Huang, C.J. Maddison, A. Guez, L. Sifre, G. van den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, S. Dieleman, D. Grewe, J. Nham, N. Kalchbrenner, I. Sutskever, T. Lillicrap, M. Leach, K. Kavukcuoglu, T. Graepel and D. Hassabis, Mastering the Game of Go with Deep Neural Networks and Tree Search, Nature
M.K. Hanawal, H. Liu, H. Zhu and I.C. Paschalidis, Learning Policies for Markov Decision Processes From Data, IEEE Transactions on Automatic Control
2018
Later among the works it cites.
C. Allen and T. Hospedales, Analogies Explained: Towards Understanding Word Embeddings, in: Proceedings of the 36th International Conference on Machine Learning
2019
Later among the works it cites.
Y. Xian, Z. Fu, S. Muthukrishnan, G. de Melo and Y. Zhang, Reinforcement Knowledge Graph Reasoning for Explainable Recommendation, in: Proceedings of the 42nd International ACM SIGIR Conference on Research and Development in Information Retrieval
2019
Later among the works it cites.
B. Runck, S. Manson, E. Shook, M. Gini and N. Jordan, Using word embeddings to generate data-driven human agent decision-making from natural language, GeoInformatica
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
A. Pandey, G.A. Moreno, J. Cámara and D. Garlan, Hybrid Planning for Decision Making in Self-Adaptive Systems, in: 2016 IEEE 10th International Conference on Self-Adaptive and Self-Organizing Systems (SASO)
2016
Cited alongside, same era.
H. Wang, X. Wang, X. Hu, X. Zhang and M. Gu, A multi-agent reinforcement learning approach to dynamic service composition, Information Sciences
2016
Cited alongside, same era.
E. Lefever, A hybrid approach to domain-independent taxonomy learning, Applied Ontology
2016
Cited alongside, same era.
D. Poole and A. Mackworth, Artificial Intelligence: Foundations of Computational Agents
2017
Cited alongside, same era.
Y. Liu, B. Logan, N. Liu, Z. Xu, J. Tang and Y. Wang, Deep Reinforcement Learning for Dynamic Treatment Regimes on Medical Registry Data, in: 2017 IEEE International Conference on Healthcare Informatics (ICHI)
2017
Cited alongside, same era.
H. Hanke and D. Knees, A phase-field damage model based on evolving microstructure, Asymptotic Analysis
2017
Cited alongside, same era.
A. Verma and S. Kumar, Cognitive Robotics in Artificial Intelligence, in: 2018 8th International Conference on Cloud Computing, Data Science & Engineering (Confluence)
2018
Cited alongside, same era.
P. Kaiser and T. Asfour, Autonomous Detection and Experimental Validation of Affordances, IEEE Robotics and Automation Letters
2018
Cited alongside, same era.
2020
Later among the works it cites.
L. Miralles-Pechuán, F. Jiménez, H. Ponce and L. Martínez-Villaseñor, A Methodology Based on Deep Q-Learning/Genetic Algorithms for Optimizing COVID-19 Pandemic Government Actions, in: Proceedings of the 29th ACM International Conference on Information & Knowledge Management
2020
Later among the works it cites.
P. Black, DADS: The On-Line Dictionary of Algorithms and Data Structures, NIST Interagency/Internal Report (NISTIR), National Institute of Standards and Technology, Gaithersburg, MD, 2020. doi:https://doi.org/10.6028/NIST.IR.8318
2020
Later among the works it cites.
E. Dohmatob, G. Dumas and D. Bzdok, Dark control: The default mode network as a reinforcement learning agent, Human Brain Mapping
2020
Later among the works it cites.
F.L.D. Silva, G. Warnell, A.H.R. Costa and P. Stone, Agents Teaching Agents: A Survey on Inter-Agent Transfer Learning, in: Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems
2020
Later among the works it cites.
K. Zhao, X. Wang, Y. Zhang, L. Zhao, Z. Liu, C. Xing and X. Xie, Leveraging Demonstrations for Reinforcement Recommendation Reasoning over Knowledge Graphs, in: Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval
2020
Later among the works it cites.
A. Kanervisto, C. Scheller and V. Hautamäki, Action Space Shaping in Deep Reinforcement Learning, in: 2020 IEEE Conference on Games (CoG)
2020
Later among the works it cites.
Z. Weng, F. Paus, A. Varava, H. Yin, T. Asfour and D. Kragic, Graph-based Task-specific Prediction Models for Interactions between Deformable and Rigid Objects, in: 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2021
Later among the works it cites.
2021
Later among the works it cites.
R. Agarwal, M.C. Machado, P.S. Castro and M.G. Bellemare, Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement Learning, in: International Conference on Learning Representations
2021
Later among the works it cites.