Fetching the paper…
Reading the bibliography…
In this paper, we discuss the learning of generalised policies for probabilistic and classical planning problems using Action Schema Networks (ASNets).
International Joint Conferences on Artificial Intelligence (IJCAI)-ECAI keynote: Model-free, model-based, and general intelligence.
Geffner, H. (2018a) · 1906
Earlier work this paper cites.
Induction of decision trees
Quinlan, J. R. (1986) · 1986
Earlier work this paper cites.
Finding a shortest solution for the n × \times n extension of the 15-puzzle is intractable.
Ratner, D., and Warmuth, M. K. (1986) · 1986
Earlier work this paper cites.
Learning decision lists
Rivest, R. L. (1987) · 1987
Earlier work this paper cites.
Learning to act using real-time dynamic programming
Barto, A. G., Bradtke, S. J., and Singh, S. P. (1995) · 1995
Earlier work this paper cites.
Convolutional networks for images, speech, and time series
LeCun, Y., and Bengio, Y. (1995) · 1995
Earlier work this paper cites.
Neuro-Dynamic Programming
Bertsekas, D., and Tsitsiklis, J. N. (1996) · 1996
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Tibshirani, R. (1996) · 1996
Earlier work this paper cites.
Sokoban is PSPACE-complete
Culberson, J. (1997) · 1997
Earlier work this paper cites.
Long short-term memory
Hochreiter, S., and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Top-down induction of first-order logical decision trees
Blockeel, H., and de Raedt, L. (1998) · 1998
Earlier work this paper cites.
Learning action strategies for planning domains
Khardon, R. (1999) · 1999
Earlier work this paper cites.
Ensemble methods in machine learning
Dietterich, T. G. (2000) · 2000
Earlier work this paper cites.
Admissible heuristics for optimal planning
Haslum, P., and Geffner, H. (2000) · 2000
Earlier work this paper cites.
Learning generalized policies in planning using concept languages
Martin, M., and Geffner, H. (2000) · 2000
Earlier work this paper cites.
FF: The Fast-Forward planning system
Hoffmann, J. (2001) · 2001
Earlier work this paper cites.
Blocks world revisited
Slaney, J., and Thiébaux, S. (2001) · 2001
Earlier work this paper cites.
Inductive policy selection for first-order MDPs
Yoon, S., Fern, A., and Givan, R. (2002) · 2002
Earlier work this paper cites.
Labeled RTDP: Improving the convergence of real-time dynamic programming
Bonet, B., and Geffner, H. (2003) · 2003
Earlier work this paper cites.
Approximate policy iteration with a policy language bias
Fern, A., Yoon, S., and Givan, R. (2004) · 2004
Earlier work this paper cites.
Exploiting first-order regression in inductive policy selection
Gretton, C., and Thiébaux, S. (2004) · 2004
Earlier work this paper cites.
PPDDL1.0: an extension to PDDL for expressing planning domains with probabilistic effects
Younes, H. L., and Littman, M. L. (2004) · 2004
Earlier work this paper cites.
Classifying relational data with neural networks
Uwents, W., and Blockeel, H. (2005) · 2005
Earlier work this paper cites.
The first probabilistic track of the international planning competition
Younes, H. L., Littman, M. L., Weissman, D., and Asmuth, J. (2005) · 2005
Earlier work this paper cites.
The Fast Downward planning system
Helmert, M. (2006) · 2006
Earlier work this paper cites.
Discrepancy search with reactive policies for planning
Yoon, S., Fern, A., and Givan, R. (2006) · 2006
Earlier work this paper cites.
Model predictive control
Camacho, E. F., and Alba, C. B. (2007) · 2007
Earlier work this paper cites.
Probabilistic planning vs. replanning
Little, I., and Thiébaux, S. (2007) · 2007
Earlier work this paper cites.
Discriminative learning of beam-search heuristics for planning
Xu, Y., Fern, A., and Yoon, S. W. (2007) · 2007
Cited alongside, same era.
Using learned policies in heuristic-search planning
Yoon, S. W., Fern, A., and Givan, R. (2007) · 2007
Cited alongside, same era.
6th International Planning Competition: Uncertainty part
Bryce, D., and Buffet, O. (2008) · 2008
Cited alongside, same era.
The factored policy-gradient planner
Buffet, O., and Aberdeen, D. (2009) · 2009
Cited alongside, same era.
Landmarks, critical paths and abstractions: what’s the difference anyway?
Helmert, M., and Domshlak, C. (2009) · 2009
Cited alongside, same era.
Scaling up heuristic planning with relational decision trees
de la Rosa, T., Jiménez, S., Fuentetaja, R., and Borrajo, D. (2011) · 2011
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., and Macherey, K. (2016) · 2016
Later among the works it cites.
Geometric deep learning: going beyond Euclidean data
Bronstein, M. M., Bruna, J., LeCun, Y., Szlam, A., and Vandergheynst, P. (2017) · 2017
Later among the works it cites.
Bagging strategies for learning planning policies
de la Rosa, T., and Fuentetaja, R. (2017) · 2017
Later among the works it cites.
Schema networks: Zero-shot transfer with a generative causal model of intuitive physics
Kansky, K., Silver, T., Mély, D. A., Eldawy, M., Lázaro-Gredilla, M., Lou, X., Dorfman, N., Sidor, S., Phoenix, S., and George, D. (2017) · 2017
Later among the works it cites.
Reluplex: An efficient SMT solver for verifying deep neural networks
Katz, G., Barrett, C., Dill, D. L., Julian, K., and Kochenderfer, M. J. (2017) · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The first learning track of the international planning competition
Fern, A., Khardon, R., and Tadepalli, P. (2011) · 2011
Cited alongside, same era.
Generalized planning: Synthesizing plans that work for multiple environments
Hu, Y., and De Giacomo, G. (2011) · 2011
Cited alongside, same era.
LAMA 2008 and 2011
Richter, S., Westphal, M., and Helmert, M. (2011) · 2011
Cited alongside, same era.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D. (2011) · 2011
Cited alongside, same era.
Directed search for generalized plans using classical planners.
Srivastava, S., Immerman, N., Zilberstein, S., and Zhang, T. (2011) · 2011
Cited alongside, same era.
A survey of the seventh international planning competition
Coles, A., Coles, A., Olaya, A. G., Jiménez, S., López, C. L., Sanner, S., and Yoon, S. (2012) · 2012
Cited alongside, same era.
Generalized value iteration networks: Life beyond lattices
Niu, S., Chen, S., Guo, H., Targonski, C., Smith, M. C., and Kovačević, J. (2017) · 2017
Later among the works it cites.
Nonlinear hybrid planning with deep net learned transition models and mixed-integer linear programming
Say, B., Wu, G., Zhou, Y. Q., and Sanner, S. (2017) · 2017
Later among the works it cites.
Occupation measure heuristics for probabilistic planning
Trevizan, F., Thiébaux, S., and Haslum, P. (2017) · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I. (2017) · 2017
Later among the works it cites.
Classical planning in deep latent space: Bridging the subsymbolic-symbolic boundary
Asai, M., and Fukunaga, A. (2018) · 2018
Later among the works it cites.
Transfer of deep reactive policies for mdp planning
Bajpai, A. N., Garg, S., and Mausam (2018) · 2018
Later among the works it cites.
Features, projections, and representation change for generalized planning
Bonet, B., and Geffner, H. (2018) · 2018
Later among the works it cites.
Learning generalized reactive policies using deep neural networks
Groshev, E., Goldstein, M., Tamar, A., Srivastava, S., and Abbeel, P. (2018) · 2018
Later among the works it cites.
Training deep reactive policies for probabilistic planning problems
Issakkimuthu, M., Fern, A., and Tadepalli, P. (2018) · 2018
Later among the works it cites.
Tune: A research platform for distributed model selection and training
Liaw, R., Liang, E., Nishihara, R., Moritz, P., Gonzalez, J. E., and Stoica, I. (2018) · 2018
Later among the works it cites.
Lifted relational neural networks: Efficient learning of latent relational structures
Sourek, G., Aschenbrenner, V., Zelezny, F., Schockaert, S., and Kuzelka, O. (2018) · 2018
Later among the works it cites.
Action schema networks: Generalised policies with deep learning
Toyer, S., Trevizan, F., Thiébaux, S., and Xie, L. (2018) · 2018
Later among the works it cites.
Learning features and abstract actions for computing generalized plans
Bonet, B., Frances, G., and Geffner, H. (2019) · 2019
Later among the works it cites.
Generalized potential heuristics for classical planning
Francès, G., Corrêa, A. B., Geissmann, C., and Pommerening, F. (2019) · 2019
Later among the works it cites.
Size-independent neural transfer for rddl planning
Garg, S., Bajpai, A., and Mausam (2019) · 2019
Later among the works it cites.
Learning classical planning strategies with policy gradient
Gomoluch, P., Alrajeh, D., and Russo, A. (2019) · 2019
Later among the works it cites.
Guiding search with generalized policies for probabilistic planning
Shen, W., Trevizan, F. W., Toyer, S., Thiébaux, S., and Xie, L. (2019) · 2019
Later among the works it cites.
Deep learning for cost-optimal planning: Task-dependent planner selection
Sievers, S., Katz, M., Sohrabi, S., Samulowitz, H., and Ferber, P. (2019) · 2019
Later among the works it cites.
Evaluating robustness of neural networks with mixed integer programming
Tjeng, V., Xiao, K. Y., and Tedrake, R. (2019) · 2019
Later among the works it cites.
Deep reinforcement learning with relational inductive biases
Zambaldi, V., Raposo, D., Santoro, A., Bapst, V., Li, Y., Babuschkin, I., Tuyls, K., Reichert, D., Lillicrap, T., Lockhart, E., Shanahan, M., Langston, V., Pascanu, R., Botvinick, M., Vinyals, O., and Battaglia, P. (2019) · 2019
Later among the works it cites.
Neural network heuristics for classical planning: A study of hyperparameter space
Ferber, P., Helmert, M., and Hoffman, J. (2020) · 2020
Later among the works it cites.
Learning domain-independent planning heuristics with hypergraph networks
Shen, W., Trevizan, F., and Thiébaux, S. (2020) · 2020
Later among the works it cites.