mixup: Beyond empirical risk minimization
Original
Zhang, H., Cisse, M., Dauphin, Y. N., and Lopez-Paz, D. (2017) · 2017
Later among the works it cites.
Quantifying generalization in reinforcement learning
Original
Cobbe, K., Klimov, O., Hesse, C., Kim, T., and Schulman, J. (2018) · 2018
Later among the works it cites.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
Original
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., et al. (2018) · 2018
Later among the works it cites.
Generalization and regularization in dqn
Original
Farebrother, J., Machado, M. C., and Bowling, M. (2018) · 2018
Later among the works it cites.
Transfer learning for related reinforcement learning tasks via image-to-image translation
Original
Gamrian, S. and Goldberg, Y. (2018) · 2018
Later among the works it cites.
Rainbow: Combining improvements in deep reinforcement learning
Hessel, M., Modayil, J., Van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M., and Silver, D. (2018) · 2018
Later among the works it cites.
Assessing generalization in deep reinforcement learning
Original
Packer, C., Gao, K., Kos, J., Krähenbühl, P., Koltun, V., and Song, D. (2018) · 2018
Later among the works it cites.
Quantifying generalization in reinforcement learning
Original
Cobbe, K., Klimov, O., Hesse, C., Kim, T., and Schulman, J. (2018) · 2018
Later among the works it cites.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
Original
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., et al. (2018) · 2018
Later among the works it cites.
Rainbow: Combining improvements in deep reinforcement learning
Hessel, M., Modayil, J., Van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M., and Silver, D. (2018) · 2018
Later among the works it cites.
Generalization in reinforcement learning with selective noise injection and information bottleneck
Igl, M., Ciosek, K., Li, Y., Tschiatschek, S., Zhang, C., Devlin, S., and Hofmann, K. (2019) · 2019
Later among the works it cites.
On lipschitz bounds of general convolutional neural networks
Zou, D., Balan, R., and Singh, M. (2019) · 2019
Later among the works it cites.
Network randomization: A simple technique for generalization in deep reinforcement learning
Lee, K., Lee, K., Shin, J., and Lee, H. (2020) · 2020
Closest in time.
Network randomization: A simple technique for generalization in deep reinforcement learning
Lee, K., Lee, K., Shin, J., and Lee, H. (2020) · 2020
Closest in time.