Fetching the paper…
Reading the bibliography…
Meta-Reinforcement Learning (Meta-RL) agents can struggle to operate across tasks with varying environmental features that require different optimal skills (i.e., different modes of behaviour).
Multivariate information transmission
William McGill · 1954
Earlier work this paper cites.
Causality: models, reasoning, and inference, by judea pearl, cambridge university press, 2000
Leland Gerson Neuberg · 2003
Earlier work this paper cites.
Mathematical statistics and data analysis , volume 371
John A Rice and John A Rice · 2007
Earlier work this paper cites.
Visualizing data using t-sne
Laurens Van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Contrastive learning with hard negative samples, 2021
Joshua Robinson, Ching-Yao Chuang, Suvrit Sra, and Stefanie Jegelka · 2010
Earlier work this paper cites.
A fast and simple algorithm for training neural probabilistic language models
Andriy Mnih and Yee Whye Teh · 2012
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Earlier work this paper cites.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman · 2014
Earlier work this paper cites.
Geometry: the language of space and form
John Tabak · 2014
Earlier work this paper cites.
Deep learning and the information bottleneck principle
Naftali Tishby and Noga Zaslavsky · 2015
Earlier work this paper cites.
Reinforcement learning with unsupervised auxiliary tasks, 2016
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z Leibo, David Silver, and Koray Kavukcuoglu · 2016
Earlier work this paper cites.
Learning invariances for policy generalization, 2018
Remi Tachet des Combes, Philip Bachman, and Harm van Seijen · 2018
Earlier work this paper cites.
Diversity is all you need: Learning skills without a reward function, 2018
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Earlier work this paper cites.
Learning an embedding space for transferable robot skills
Karol Hausman, Jost Tobias Springenberg, Ziyu Wang, Nicolas Heess, and Martin Riedmiller · 2018
Earlier work this paper cites.
Unsupervised feature learning via non-parametric instance discrimination
Zhirong Wu, Yuanjun Xiong, Stella X Yu, and Dahua Lin · 2018
Earlier work this paper cites.
A theoretical analysis of contrastive unsupervised representation learning
Sanjeev Arora, Hrishikesh Khandeparkar, Mikhail Khodak, Orestis Plevrakis, and Nikunj Saunshi · 2019
Earlier work this paper cites.
Learning representations by maximizing mutual information across views
Philip Bachman, R Devon Hjelm, and William Buchwalter · 2019
Cited alongside, same era.
Estimating information flow in deep neural networks, 2019
Ziv Goldfeld, Ewout van den Berg, Kristjan Greenewald, Igor Melnyk, Nam Nguyen, Brian Kingsbury, and Yury Polyanskiy · 2019
Cited alongside, same era.
Learning deep representations by mutual information estimation and maximization
R Devon Hjelm, Alex Fedorov, Samuel Lavoie-Marchildon, Karan Grewal, Phil Bachman, Adam Trischler, and Yoshua Bengio · 2019
Cited alongside, same era.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals · 2019
Cited alongside, same era.
On variational bounds of mutual information
Ben Poole, Sherjil Ozair, Aaron Van Den Oord, Alex Alemi, and George Tucker · 2019
Cited alongside, same era.
Towards effective context for meta-reinforcement learning: an approach based on contrastive learning
Haotian Fu, Hongyao Tang, Jianye Hao, Chen Chen, Xidong Feng, Dong Li, and Wulong Liu · 2021
Later among the works it cites.
panda-gym: Open-Source Goal-Conditioned Environments for Robotic Learning
Quentin Gallouédec, Nicolas Cazin, Emmanuel Dellandréa, and Liming Chen · 2021
Later among the works it cites.
Provably improved context-based offline meta-rl with attention and contrastive learning
Lanqing Li, Yuanhao Huang, Mingzhe Chen, Siteng Luo, Dijun Luo, and Junzhou Huang · 2021
Later among the works it cites.
Understanding negative samples in instance discriminative self-supervised representation learning
Kento Nozawa and Issei Sato · 2021
Later among the works it cites.
Stable-baselines3: Reliable reinforcement learning implementations
Antonin Raffin, Ashley Hill, Adam Gleave, Anssi Kanervisto, Maximilian Ernestus, and Noah Dormann · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Efficient off-policy meta-reinforcement learning via probabilistic context variables
Kate Rakelly, Aurick Zhou, Chelsea Finn, Sergey Levine, and Deirdre Quillen · 2019
Cited alongside, same era.
Environment probing interaction policies
Wenxuan Zhou, Lerrel Pinto, and Abhinav Gupta · 2019
Cited alongside, same era.
Momentum contrast for unsupervised visual representation learning
Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross Girshick · 2020
Cited alongside, same era.
Curl: Contrastive unsupervised representations for reinforcement learning
Michael Laskin, Aravind Srinivas, and Pieter Abbeel · 2020
Cited alongside, same era.
Context-aware dynamics model for generalization in model-based reinforcement learning
Kimin Lee, Younggyo Seo, Seunghyun Lee, Honglak Lee, and Jinwoo Shin · 2020
Cited alongside, same era.
Umap: Uniform manifold approximation and projection for dimension reduction, 2020
Leland McInnes, John Healy, and James Melville · 2020
Cited alongside, same era.
Trajectory-wise multiple choice learning for dynamics generalization in reinforcement learning
Younggyo Seo, Kimin Lee, Ignasi Clavera Gilaberte, Thanard Kurutach, Jinwoo Shin, and Pieter Abbeel · 2020
Cited alongside, same era.
Improving context-based meta-reinforcement learning with self-supervised trajectory contrastive learning, 2021
Bernie Wang, Simon Xu, Kurt Keutzer, Yang Gao, and Bichen Wu · 2021
Later among the works it cites.
seaborn: statistical data visualization
Michael L. Waskom · 2021
Later among the works it cites.
Contrastive learning as goal-conditioned reinforcement learning
Benjamin Eysenbach, Tianjun Zhang, Sergey Levine, and Russ R Salakhutdinov · 2022
Later among the works it cites.
Tight mutual information estimation with contrastive fenchel-legendre optimization
Qing Guo, Junya Chen, Dong Wang, Yuewei Yang, Xinwei Deng, Jing Huang, Larry Carin, Fan Li, and Chenyang Tao · 2022
Later among the works it cites.
Functional connectivity inference from fmri data using multivariate information measures
Qiang Li · 2022
Later among the works it cites.
DOMINO: Decomposed mutual information optimization for generalized context in meta-reinforcement learning
Yao Mu, Yuzheng Zhuang, Fei Ni, Bin Wang, Jianyu Chen, Jianye HAO, and Ping Luo · 2022
Later among the works it cites.
Pandr: Fast adaptation to new environments from offline experiences via decoupling policy and environment representations, 2022
Tong Sang, Hongyao Tang, Yi Ma, Jianye Hao, Yan Zheng, Zhaopeng Meng, Boyan Li, and Zhen Wang · 2022
Later among the works it cites.
Skill-based model-based reinforcement learning
Lucy Xiaoyang Shi, Joseph J. Lim, and Youngwoon Lee · 2022
Later among the works it cites.
A survey of zero-shot generalisation in deep reinforcement learning
Robert Kirk, Amy Zhang, Edward Grefenstette, and Tim Rocktäschel · 2023
Later among the works it cites.
Multi-view disentanglement for reinforcement learning with multiple cameras
Mhairi Dunion and Stefano V Albrecht · 2024
Closest in time.
DRED: Zero-shot transfer in reinforcement learning via data-regularised environment design
Samuel Garcin, James Doran, Shangmin Guo, Christopher G. Lucas, and Stefano V. Albrecht · 2024
Closest in time.
Multi-horizon representations with hierarchical forward models for reinforcement learning
Trevor McInroe, Lukas Schäfer, and Stefano V. Albrecht · 2024
Closest in time.