Stable baselines
Hill, A., Raffin, A., Ernestus, M., Gleave, A., Kanervisto, A., Traore, R., Dhariwal, P., Hesse, C., Klimov, O., Nichol, A., Plappert, M., Radford, A., Schulman, J., Sidor, S., and Wu, Y · 2018
Later among the works it cites.
ROBEL: RObotics BEnchmarks for Learning with low-cost robots
Ahn, M., Zhu, H., Hartikainen, K., Ponte, H., Gupta, A., Levine, S., and Kumar, V · 2019
Later among the works it cites.
Conditioning by adaptive sampling for robust design
Original
Brookes, D. H., Park, H., and Listgarten, J · 2019
Later among the works it cites.
Bayesian optimization under heavy-tailed payoffs
Chowdhury, S. R. and Gopalan, A · 2019
Later among the works it cites.
Show your work: Improved reporting of experimental results
Dodge, J., Gururangan, S., Card, D., Schwartz, R., and Smith, N. A · 2019
Later among the works it cites.
Model inversion networks for model-based optimization
Original
Kumar, A. and Levine, S · 2019
Later among the works it cites.
Evaluating protein transfer learning with TAPE
Rao, R., Bhattacharya, N., Thomas, N., Duan, Y., Chen, P., Canny, J. F., Abbeel, P., and Song, Y. S · 2019
Later among the works it cites.
Human 5‘ utr design and variant effect prediction from a massively parallel translation assay
Sample, P. J., Wang, B., Reid, D. W., Presnyak, V., McFadyen, I. J., Morris, D. R., and Seelig, G · 2019
Later among the works it cites.
Model-based reinforcement learning for biological sequence design
Angermueller, C., Dohan, D., Belanger, D., Deshpande, R., Murphy, K., and Colwell, L · 2020
Later among the works it cites.
Population-based black-box optimization for biological sequence design
Angermüller, C., Belanger, D., Gane, A., Mariet, Z., Dohan, D., Murphy, K., Colwell, L., and Sculley, D · 2020
Later among the works it cites.
Autofocused oracles for model-based design
Original
Fannjiang, C. and Listgarten, J · 2020
Later among the works it cites.
D4rl: Datasets for deep data-driven reinforcement learning
Original
Fu, J., Kumar, A., Nachum, O., Tucker, G., and Levine, S · 2020
Later among the works it cites.
Firefly algorithm
Yang, X.-S. and Slowik, A · 2020
Later among the works it cites.
Deep reinforcement learning at the edge of the statistical precipice
Agarwal, R., Schwarzer, M., Castro, P. S., Courville, A., and Bellemare, M. G · 2021
Later among the works it cites.
Offline model-based optimization via normalized maximum likelihood estimation
Fu, J. and Levine, S · 2021
Later among the works it cites.
Conservative objective models for effective offline model-based optimization
Trabucco, B., Kumar, A., Geng, X., and Levine, S · 2021
Later among the works it cites.