Behavior priors for efficient reinforcement learning
Original
Tirumala, D., Galashov, A., Noh, H., Hasenclever, L., Pascanu, R., Schwarz, J., Desjardins, G., Czarnecki, W. M., Ahuja, A., Teh, Y. W., et al · 2020
Later among the works it cites.
Huggingface’s transformers: State-of-the-art natural language processing, 2020
Wolf, T., Debut, L., Sanh, V., Chaumond, J., Delangue, C., Moi, A., Cistac, P., Rault, T., Louf, R., Funtowicz, M., Davison, J., Shleifer, S., von Platen, P., Ma, C., Jernite, Y., Plu, J., Xu, C., Scao, T. L., Gugger, S., Drame, M., Lhoest, Q., and Rush, A. M · 2020
Later among the works it cites.
Learning to see before learning to act: Visual pre-training for manipulation
Yen-Chen, L., Zeng, A., Song, S., Isola, P., and Lin, T.-Y · 2020
Later among the works it cites.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Yu, T., Quillen, D., He, Z., Julian, R., Hausman, K., Finn, C., and Levine, S · 2020
Later among the works it cites.
Transporter networks: Rearranging the visual world for robotic manipulation
Original
Zeng, A., Florence, P., Tompson, J., Welker, S., Chien, J., Attarian, M., Armstrong, T., Krasin, I., Duong, D., Sindhwani, V., et al · 2020
Later among the works it cites.
robosuite: A modular simulation framework and benchmark for robot learning
Original
Zhu, Y., Wong, J., Mandlekar, A., and Martín-Martín, R · 2020
Later among the works it cites.
Deep reinforcement learning at the edge of the statistical precipice
Agarwal, R., Schwarzer, M., Castro, P. S., Courville, A. C., and Bellemare, M. G · 2021
Later among the works it cites.
Actionable models: Unsupervised offline reinforcement learning of robotic skills
Original
Chebotar, Y., Hausman, K., Lu, Y., Xiao, T., Kalashnikov, D., Varley, J., Irpan, A., Eysenbach, B., Julian, R., Finn, C., et al · 2021
Later among the works it cites.
Decision transformer: Reinforcement learning via sequence modeling
Chen, L., Lu, K., Rajeswaran, A., Lee, K., Grover, A., Laskin, M., Abbeel, P., Srinivas, A., and Mordatch, I · 2021
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale, 2021
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., Uszkoreit, J., and Houlsby, N · 2021
Later among the works it cites.
A minimalist approach to offline reinforcement learning
Original
Fujimoto, S. and Gu, S. S · 2021
Later among the works it cites.
Generalized decision transformer for offline hindsight information matching
Original
Furuta, H., Matsuo, Y., and Gu, S. S · 2021
Later among the works it cites.
Emaq: Expected-max q-learning operator for simple yet effective offline and online rl
Ghasemipour, S. K. S., Schuurmans, D., and Gu, S. S · 2021
Later among the works it cites.
Mastering atari with discrete world models, 2021
Hafner, D., Lillicrap, T., Norouzi, M., and Ba, J · 2021
Later among the works it cites.
Reinforcement learning as one big sequence modeling problem
Original
Janner, M., Li, Q., and Levine, S · 2021
Later among the works it cites.
Swin transformer: Hierarchical vision transformer using shifted windows, 2021
Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., and Guo, B · 2021
Later among the works it cites.
Pretrained transformers as universal computation engines, 2021
Lu, K., Grover, A., Abbeel, P., and Mordatch, I · 2021
Later among the works it cites.
Language conditioned imitation learning over unstructured data
Lynch, C. and Sermanet, P · 2021
Later among the works it cites.
Tool as embodiment for recursive manipulation
Original
Noguchi, Y., Matsushima, T., Matsuo, Y., and Gu, S. S · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I · 2021
Later among the works it cites.
Multimodal few-shot learning with frozen language models, 2021
Tsimpoukelli, M., Menick, J., Cabi, S., Eslami, S. M. A., Vinyals, O., and Hill, F · 2021
Later among the works it cites.
Cliport: What and where pathways for robotic manipulation
Shridhar, M., Manuelli, L., and Fox, D · 2022
Closest in time.