Hierarchical reinforcement learning for zero-shot generalization with subtask dependencies
Original
Sungryull Sohn, Junhyuk Oh, and Honglak Lee · 2018
Later among the works it cites.
Mask-guided contrastive attention model for person re-identification
Chunfeng Song, Yan Huang, Wanli Ouyang, and Liang Wang · 2018
Later among the works it cites.
Neural program synthesis from diverse demonstration videos
Shao-Hua Sun, Hyeonwoo Noh, Sriram Somasundaram, and Joseph Lim · 2018
Later among the works it cites.
Programmatically interpretable reinforcement learning
Abhinav Verma, Vijayaraghavan Murali, Rishabh Singh, Pushmeet Kohli, and Swarat Chaudhuri · 2018
Later among the works it cites.
Improving the universality and learnability of neural programmer-interpreters with combinator abstraction
Original
Da Xiao, Jo-Yu Liao, and Xingyuan Yuan · 2018
Later among the works it cites.
Neural-symbolic vqa: Disentangling reasoning from vision and language understanding
Original
Kexin Yi, Jiajun Wu, Chuang Gan, Antonio Torralba, Pushmeet Kohli, and Joshua B Tenenbaum · 2018
Later among the works it cites.
Neural symbolic reader: Scalable integration of distributed and symbolic representations for reading comprehension
Xinyun Chen, Chen Liang, Adams Wei Yu, Denny Zhou, Dawn Song, and Quoc V Le · 2019
Later among the works it cites.
Neural-symbolic computing: An effective methodology for principled integration of machine learning and reasoning
Original
Artur d’Avila Garcez, Marco Gori, Luis C Lamb, Luciano Serafini, Michael Spranger, and Son N Tran · 2019
Later among the works it cites.
Unicoder: A universal language encoder by pre-training with multiple cross-lingual tasks
Original
Haoyang Huang, Yaobo Liang, Nan Duan, Ming Gong, Linjun Shou, Daxin Jiang, and Ming Zhou · 2019
Later among the works it cites.
Perceptual visual reasoning with knowledge propagation
Guohao Li, Xin Wang, and Wenwu Zhu · 2019
Later among the works it cites.
Learning to describe scenes with programs
Yunchao Liu and Zheng Wu · 2019
Later among the works it cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Original
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Later among the works it cites.
Lxmert: Learning cross-modality encoder representations from transformers
Hao Tan and Mohit Bansal · 2019
Later among the works it cites.
Learning to infer and execute 3d shape programs
Original
Yonglong Tian, Andrew Luo, Xingyuan Sun, Kevin Ellis, William T Freeman, Joshua B Tenenbaum, and Jiajun Wu · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Original
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Ruslan Salakhutdinov, and Quoc V Le · 2019
Later among the works it cites.
Large batch optimization for deep learning: Training bert in 76 minutes
Original
Yang You, Jing Li, Sashank Reddi, Jonathan Hseu, Sanjiv Kumar, Srinadh Bhojanapalli, Xiaodan Song, James Demmel, Kurt Keutzer, and Cho-Jui Hsieh · 2019
Later among the works it cites.
End-to-end object detection with transformers
Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko · 2020
Later among the works it cites.
Compositional generalization via neural-symbolic stack machines
Original
Xinyun Chen, Chen Liang, Adams Wei Yu, Dawn Song, and Denny Zhou · 2020
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Original
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Later among the works it cites.
One policy to control them all: Shared modular policies for agent-agnostic control
Wenlong Huang, Igor Mordatch, and Deepak Pathak · 2020
Later among the works it cites.
Object-centric diagnosis of visual reasoning
Original
Jianwei Yang, Jiayuan Mao, Jiajun Wu, Devi Parikh, David D Cox, Joshua B Tenenbaum, and Chuang Gan · 2020
Later among the works it cites.
Augmenting policy learning with routines discovered from a demonstration
Original
Zelin Zhao, Chuang Gan, Jiajun Wu, Xiaoxiao Guo, and Joshua Tenenbaum · 2020
Later among the works it cites.
Meta module network for compositional visual reasoning
Wenhu Chen, Zhe Gan, Linjie Li, Yu Cheng, William Wang, and Jingjing Liu · 2021
Closest in time.
Mask attention networks: Rethinking and strengthen transformer
Original
Zhihao Fan, Yeyun Gong, Dayiheng Liu, Zhongyu Wei, Siyuan Wang, Jian Jiao, Nan Duan, Ruofei Zhang, and Xuanjing Huang · 2021
Closest in time.
Swin transformer: Hierarchical vision transformer using shifted windows
Original
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Closest in time.
Learning to synthesize programs as interpretable and generalizable policies, 2021
Dweep Trivedi, Jesse Zhang, Shao-Hua Sun, and Joseph J. Lim · 2021
Closest in time.
Rest: An efficient transformer for visual recognition
Original
QingLong Zhang and Yubin Yang · 2021
Closest in time.