Fetching the paper…
Reading the bibliography…
We introduce Compositional Imitation Learning and Execution (CompILE): a framework for learning reusable, variable-length segments of hierarchically-structured behavior from demonstration data.
Inquiries into Truth and Interpretation
Davidson, D · 1984
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J · 1997
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Sutton, R. S., Precup, D., and Singh, S · 1999
Earlier work this paper cites.
Topic segmentation with an aspect hidden markov model
Blei, D. M. and Moreno, P. J · 2001
Earlier work this paper cites.
Perceiving, remembering, and communicating structure in events
Zacks, J. M., Tversky, B., and Iyer, G · 2001
Earlier work this paper cites.
A bayesian framework for word segmentation: Exploring the effects of context
Goldwater, S., Griffiths, T. L., and Johnson, M · 2009
Earlier work this paper cites.
What constitutes an episode in episodic memory?
Ezzyat, Y. and Davachi, L · 2011
Earlier work this paper cites.
Supervised sequence labelling
Graves, A · 2012
Earlier work this paper cites.
Incremental semantically grounded learning from demonstration
Niekum, S., Chitta, S., Barto, A. G., Marthi, B., and Osentoski, S · 2013
Earlier work this paper cites.
Generative adversarial nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2014
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Rezende, D. J., Mohamed, S., and Wierstra, D · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q. V · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2015
Earlier work this paper cites.
Towards learning hierarchical skills for multi-phase manipulation tasks
Kroemer, O., Daniel, C., Neumann, G., Van Hoof, H., and Peters, J · 2015
Earlier work this paper cites.
Learning structured output representation using deep conditional generative models
Sohn, K., Lee, H., and Yan, X · 2015
Earlier work this paper cites.
Show and tell: A neural image caption generator
Vinyals, O., Toshev, A., Bengio, S., and Erhan, D · 2015
Earlier work this paper cites.
Ba, J. L., Kiros, J. R., and Hinton, G. E · 2016
Earlier work this paper cites.
Daps: Deep action proposals for action understanding
Escorcia, V., Heilbron, F. C., Niebles, J. C., and Ghanem, B · 2016
Earlier work this paper cites.
Composing graphical models with neural networks for structured representations and fast inference
Johnson, M., Duvenaud, D. K., Wiltschko, A., Adams, R. P., and Datta, S. R · 2016
Cited alongside, same era.
Hierarchical deep reinforcement learning: Integrating temporal abstraction and intrinsic motivation
Kulkarni, T. D., Narasimhan, K., Saeedi, A., and Tenenbaum, J · 2016
Cited alongside, same era.
A decomposable attention model for natural language inference
Parikh, A. P., Täckström, O., Das, D., and Uszkoreit, J · 2016
Cited alongside, same era.
Modular multitask reinforcement learning with policy sketches
Andreas, J., Klein, D., and Levine, S · 2017
Cited alongside, same era.
The option-critic architecture
Bacon, P.-L., Harb, J., and Precup, D · 2017
Cited alongside, same era.
Discovering event structure in continuous narrative perception and memory
Zero-shot task generalization with multi-task deep reinforcement learning
Oh, J., Singh, S., Lee, H., and Kohli, P · 2017
Later among the works it cites.
Event boundaries in memory and cognition
Radvansky, G. A. and Zacks, J. M · 2017
Later among the works it cites.
Constructing experience: event models from perception to action
Richmond, L. L. and Zacks, J. M · 2017
Later among the works it cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Later among the works it cites.
FeUdal networks for hierarchical reinforcement learning
Vezhnevets, A. S., Osindero, S., Schaul, T., Heess, N., Jaderberg, M., Silver, D., and Kavukcuoglu, K · 2017
Later among the works it cites.
Sequence modeling via segmentations
Wang, C., Wang, Y., Huang, P.-S., Mohamed, A., Zhou, D., and Deng, L · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Baldassano, C., Chen, J., Zadbood, A., Pillow, J. W., Hasson, U., and Norman, K. A · 2017
Cited alongside, same era.
Latent sequence decompositions
Chan, W., Zhang, Y., Le, Q., and Jaitly, N · 2017
Cited alongside, same era.
Recurrent hidden semi-markov model
Dai, H., Dai, B., Zhang, Y.-M., Li, S., and Song, L · 2017
Cited alongside, same era.
Denil, M., Colmenarejo, S. G., Cabi, S., Saxton, D., and de Freitas, N · 2017
Cited alongside, same era.
Stochastic neural networks for hierarchical reinforcement learning
Florensa, C., Duan, Y., and Abbeel, P · 2017
Cited alongside, same era.
Multi-level discovery of deep options
Fox, R., Krishnan, S., Stoica, I., and Goldberg, K · 2017
Cited alongside, same era.
Multi-modal imitation learning from unstructured demonstrations using generative adversarial nets
Hausman, K., Chebotar, Y., Schaal, S., Sukhatme, G., and Lim, J. J · 2017
Cited alongside, same era.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., et al · 2018
Closest in time.
Parametrized hierarchical procedures for neural programming
Fox, R., Shin, R., Krishnan, S., Goldberg, K., Song, D., and Stoica, I · 2018
Closest in time.
Time-agnostic prediction: Predicting predictable video frames
Jayaraman, D., Ebert, F., Efros, A. A., and Levine, S · 2018
Closest in time.
Adaptive skip intervals: Temporal abstraction for recurrent dynamical models
Neitz, A., Parascandolo, G., Bauer, S., and Schölkopf, B · 2018
Closest in time.
Learning abstract options
Riemer, M., Liu, M., and Tesauro, G · 2018
Closest in time.
Directed-info gail: Learning hierarchical policies from unsegmented demonstrations using directed information
Sharma, A., Sharma, M., Rhinehart, N., and Kitani, K. M · 2018
Closest in time.
TACO: Learning task decomposition via temporal alignment for control
Shiarlis, K., Wulfmeier, M., Salter, S., Whiteson, S., and Posner, I · 2018
Closest in time.
Neural program synthesis from diverse demonstration videos
Sun, S.-H., Noh, H., Somasundaram, S., and Lim, J · 2018
Closest in time.
Subgoal discovery for hierarchical dialogue policy learning
Tang, D., Li, X., Gao, J., Wang, C., Li, L., and Jebara, T · 2018
Closest in time.
Tassa, Y., Doron, Y., Muldal, A., Erez, T., Li, Y., Casas, D. d. L., Budden, D., Abdolmaleki, A., Merel, J., Lefrancq, A., et al · 2018
Closest in time.
Composing complex skills by learning transition policies with proximity reward induction
Lee, Y., Sun, S.-H., Somasundaram, S., Hu, E., and Lim, J. J · 2019
Closest in time.
Keyin: Discovering subgoal structure with keyframe-based video prediction
Pertsch, K., Rybkin, O., Yang, J., Derpanis, K., Lim, J., Daniilidis, K., and Jaegle, A · 2019
Closest in time.