Fetching the paper…
Reading the bibliography…
Deep learning has been the mainstream technique in natural language processing (NLP) area.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J Liu. 2019 · 1910
Earlier work this paper cites.
Catastrophic interference in connectionist networks: The sequential learning problem
Michael McCloskey and Neal J Cohen. 1989 · 1989
Earlier work this paper cites.
Dimensions of meaning
H Schütze. 1992 · 1992
Earlier work this paper cites.
Foundations of statistical natural language processing
Christopher Manning and Hinrich Schutze. 1999 · 1999
Earlier work this paper cites.
Introduction to the CoNLL-2003 shared task: Language-independent named entity recognition
Erik Tjong Kim Sang and Fien De Meulder. 2003 · 2003
Earlier work this paper cites.
Meta-learning for few-shot natural language processing: A survey
Wenpeng Yin. 2020 · 2007
Earlier work this paper cites.
Recurrent neural network based language model
Tomáš Mikolov, Martin Karafiát, Lukáš Burget, Jan Černockỳ, and Sanjeev Khudanpur. 2010 · 2010
Earlier work this paper cites.
Meta-KD: A meta knowledge distillation framework for language model compression across domains
Haojie Pan, Chengyu Wang, Minghui Qiu, Yichang Zhang, Yaliang Li, and Jun Huang. 2020 · 2012
Earlier work this paper cites.
Findings of the 2014 workshop on statistical machine translation
Ondrej Bojar, Christian Buck, Christian Federmann, Barry Haddow, Philipp Koehn, Johannes Leveling, Christof Monz, Pavel Pecina, Matt Post, Herve Saint-Amand, Radu Soricut, Lucia Specia, and Ales Tamchyna. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Earlier work this paper cites.
Learning to learn by gradient descent by gradient descent
Marcin Andrychowicz, Misha Denil, Sergio Gomez, Matthew W. Hoffman, David Pfau, Tom Schaul, Brendan Shillingford, and Nando de Freitas. 2016 · 2016
Earlier work this paper cites.
Matching networks for one shot learning
Oriol Vinyals, Charles Blundell, Timothy Lillicrap, Koray Kavukcuoglu, and Daan Wierstra. 2016 · 2016
Earlier work this paper cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Earlier work this paper cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A. Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al. 2017 · 2017
Earlier work this paper cites.
Learning without forgetting
Zhizhong Li and Derek Hoiem. 2017 · 2017
Earlier work this paper cites.
Gradient episodic memory for continual learning
David Lopez-Paz and Marc’Aurelio Ranzato. 2017 · 2017
Earlier work this paper cites.
Revisiting activation regularization for language rnns
Stephen Merity, Bryan McCann, and Richard Socher. 2017 · 2017
Earlier work this paper cites.
Optimization as a model for few-shot learning
Sachin Ravi and Hugo Larochelle. 2017 · 2017
Earlier work this paper cites.
Prototypical networks for few-shot learning
Jake Snell, Kevin Swersky, and Richard S. Zemel. 2017 · 2017
Earlier work this paper cites.
Continual learning through synaptic intelligence
Friedemann Zenke, Ben Poole, and Surya Ganguli. 2017 · 2017
Earlier work this paper cites.
Neural architecture search with reinforcement learning
Barret Zoph and Quoc V. Le. 2017 · 2017
Earlier work this paper cites.
Memory aware synapses: Learning what (not) to forget
Rahaf Aljundi, Francesca Babiloni, Mohamed Elhoseiny, Marcus Rohrbach, and Tinne Tuytelaars. 2018 · 2018
Earlier work this paper cites.
Lifelong machine learning
Zhiyuan Chen and Bing Liu. 2018 · 2018
Earlier work this paper cites.
Meta-learning for low-resource neural machine translation
Jiatao Gu, Yong Wang, Yun Chen, Victor O. K. Li, and Kyunghyun Cho. 2018 · 2018
Earlier work this paper cites.
Natural language to structured query generation via meta-learning
Po-Sen Huang, Chenglong Wang, Rishabh Singh, Wen tau Yih, and Xiaodong He. 2018 · 2018
Earlier work this paper cites.
Learning to adapt:a meta-learning approach for speaker adaptation
Ondřej Klejch, Joachim Fainberg, and Peter Bell. 2018 · 2018
Earlier work this paper cites.
Learning to generalize: Meta-learning for domain generalization
Da Li, Yongxin Yang, Yi-Zhe Song, and Timothy M. Hospedales. 2018 · 2018
Earlier work this paper cites.
A simple neural attentive meta-learner
Nikhil Mishra, Mostafa Rohaninejad, Xi Chen, and Pieter Abbeel. 2018 · 2018
Earlier work this paper cites.
On first-order meta-learning algorithms
Alex Nichol, Joshua Achiam, and John Schulman. 2018 · 2018
Earlier work this paper cites.
Progress & compress: A scalable framework for continual learning
Jonathan Schwarz, Wojciech Czarnecki, Jelena Luketina, Agnieszka Grabska-Barwinska, Yee Whye Teh, Razvan Pascanu, and Raia Hadsell. 2018 · 2018
Earlier work this paper cites.
Memory-based parameter adaptation
Pablo Sprechmann, Siddhant M. Jayakumar, Jack W. Rae, Alexander Pritzel, Adrià Puigdomènech Badia, Benigno Uria, Oriol Vinyals, Demis Hassabis, Razvan Pascanu, and Charles Blundell. 2018 · 2018
Earlier work this paper cites.
Memory, show the way: Memory based few shot word representation learning
Jingyuan Sun, Shaonan Wang, and Chengqing Zong. 2018 · 2018
Earlier work this paper cites.
Learning to compare: Relation network for few-shot learning
Flood Sung, Yongxin Yang, Li Zhang, Tao Xiang, Philip H.S. Torr, and Timothy M. Hospedales. 2018 · 2018
Earlier work this paper cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
One-shot relational learning for knowledge graphs
Wenhan Xiong, Mo Yu, Shiyu Chang, Xiaoxiao Guo, and William Yang Wang. 2018 · 2018
Earlier work this paper cites.
Diverse few-shot text classification with multiple metrics
Mo Yu, Xiaoxiao Guo, Jinfeng Yi, Shiyu Chang, Saloni Potdar, Yu Cheng, Gerald Tesauro, Haoyu Wang, and Bowen Zhou. 2018 · 2018
Earlier work this paper cites.
Learning transferable architectures for scalable image recognition
Barret Zoph, Vijay Vasudevan, Jonathon Shlens, and Quoc V. Le. 2018 · 2018
Earlier work this paper cites.
How to train your MAML
Antreas Antoniou, Harrison Edwards, and Amos Storkey. 2019 · 2019
Earlier work this paper cites.
Episodic memory in lifelong language learning
Cyprien de Masson d'Autume, Sebastian Ruder, Lingpeng Kong, and Dani Yogatama. 2019 · 2019
Earlier work this paper cites.
Leveraging end-to-end speech recognition with neural architecture search
Ahmed Baruwa, Mojeed Abisiga, Ibrahim Gbadegesin, and Afeez Fakunle. 2019 · 2019
Earlier work this paper cites.
Meta relational learning for few-shot link prediction in knowledge graphs
Mingyang Chen, Wen Zhang, Wei Zhang, Qiang Chen, and Huajun Chen. 2019a · 2019
Earlier work this paper cites.
Meta learning for hyperparameter optimization in dialogue system
Jen-Tzung Chien and Wei Xiang Lieow. 2019 · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Investigating meta-learning algorithms for low-resource natural language understanding tasks
Zi-Yi Dou, Keyi Yu, and Antonios Anastasopoulos. 2019 · 2019
Earlier work this paper cites.
Multimodal one-shot learning of speech and images
Ryan Eloff, Herman A. Engelbrecht, and Herman Kamper. 2019 · 2019
Earlier work this paper cites.
Online meta-learning
Chelsea Finn, Aravind Rajeswaran, Sham Kakade, and Sergey Levine. 2019 · 2019
Cited alongside, same era.
FewRel 2.0: Towards more challenging few-shot relation classification
Tianyu Gao, Xu Han, Hao Zhu, Zhiyuan Liu, Peng Li, Maosong Sun, and Jie Zhou. 2019b · 2019
Cited alongside, same era.
Induction networks for few-shot text classification
Ruiying Geng, Binhua Li, Yongbin Li, Xiaodan Zhu, Ping Jian, and Jian Sun. 2019 · 2019
Cited alongside, same era.
Coupling retrieval and meta-learning for context-dependent semantic parsing
Daya Guo, Duyu Tang, Nan Duan, Ming Zhou, and Jian Yin. 2019 · 2019
Cited alongside, same era.
Few-shot representation learning for out-of-vocabulary words
Ziniu Hu, Ting Chen, Kai-Wei Chang, and Yizhou Sun. 2019 · 2019
Cited alongside, same era.
Improved differentiable architecture search for language modeling and named entity recognition
Yufan Jiang, Chi Hu, Tong Xiao, Chunliang Zhang, and Jingbo Zhu. 2019 · 2019
MobileBERT: a compact task-agnostic BERT for resource-limited devices
Zhiqing Sun, Hongkun Yu, Xiaodan Song, Renjie Liu, Yiming Yang, and Denny Zhou. 2020 · 2020
Later among the works it cites.
Meta-Dataset: A dataset of datasets for learning to learn from few examples
Eleni Triantafillou, Tyler Zhu, Vincent Dumoulin, Pascal Lamblin, Utku Evci, Kelvin Xu, Ross Goroshin, Carles Gelada, Kevin Swersky, Pierre-Antoine Manzagol, and Hugo Larochelle. 2020 · 2020
Later among the works it cites.
Efficient meta lifelong-learning with limited memory
Zirui Wang, Sanket Vaibhav Mehta, Barnabas Poczos, and Jaime Carbonell. 2020c · 2020
Later among the works it cites.
Enhanced meta-learning for cross-lingual named entity recognition with minimal resources
Qianhui Wu, Zijia Lin, Guoxin Wang, Hui Chen, Börje F. Karlsson, Biqing Huang, and Chin-Yew Lin. 2020 · 2020
Later among the works it cites.
Multi-source meta transfer for low resource multiple-choice question answering
Ming Yan, Hao Zhang, Di Jin, and Joey Tianyi Zhou. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Speaker adaptive training using model agnostic meta-learning
Ondřej Klejch, Joachim Fainberg, Peter Bell, and Steve Renals. 2019 · 2019
Cited alongside, same era.
Adapting meta knowledge graph information for multi-hop reasoning over few-shot relations
Xin Lv, Yuxian Gu, Xu Han, Lei Hou, Juanzi Li, and Zhiyuan Liu. 2019 · 2019
Cited alongside, same era.
Personalizing dialogue agents via meta-learning
Andrea Madotto, Zhaojiang Lin, Chien-Sheng Wu, and Pascale Fung. 2019 · 2019
Cited alongside, same era.
Improving keyword spotting and language identification via neural architecture search at scale
Hanna Mazzawi, Xavi Gonzalvo, Aleks Kracun, Prashant Sridhar, Niranjan Subrahmanya, Ignacio Lopez Moreno, Hyun Jin Park, and Patrick Violette. 2019 · 2019
Cited alongside, same era.
Meta-learning for low-resource natural language generation in task-oriented dialogue systems
Fei Mi, Minlie Huang, Jiyong Zhang, and Boi Faltings. 2019 · 2019
Cited alongside, same era.
Meta-learning improves lifelong relation extraction
Abiola Obamuyide and Andreas Vlachos. 2019a · 2019
Cited alongside, same era.
Few-shot knowledge graph completion
Chuxu Zhang, Huaxiu Yao, Chao Huang, Meng Jiang, Zhenhui Li, and Nitesh V Chawla. 2020 · 2020
Later among the works it cites.
SpeechT5: Unified-modal encoder-decoder pre-training for spoken language processing
Junyi Ao, Rui Wang, Long Zhou, Shujie Liu, Shuo Ren, Yu Wu, Tom Ko, Qing Li, Yu Zhang, Zhihua Wei, Yao Qian, Jinyu Li, and Furu Wei. 2021 · 2021
Later among the works it cites.
Diverse distributions of self-supervised tasks for meta-learning in NLP
Trapit Bansal, Karthick Prasad Gunasekaran, Tong Wang, Tsendsuren Munkhdalai, and Andrew McCallum. 2021 · 2021
Later among the works it cites.
FLEX: Unifying evaluation for few-shot nlp
Jonathan Bragg, Arman Cohan, Kyle Lo, and Iz Beltagy. 2021 · 2021
Later among the works it cites.
SpeechNet: A universal modularized model for speech processing tasks
Yi-Chen Chen, Po-Han Chi, Shu wen Yang, Kai-Wei Chang, Jheng hao Lin, Sung-Feng Huang, Da-Rong Liu, Chi-Liang Liu, Cheng-Kuang Lee, and Hung yi Lee. 2021 · 2021
Later among the works it cites.
Meta-learning to compositionally generalize
Henry Conklin, Bailin Wang, Kenny Smith, and Ivan Titov. 2021 · 2021
Later among the works it cites.
Editing factual knowledge in language models
Nicola De Cao, Wilker Aziz, and Ivan Titov. 2021 · 2021
Later among the works it cites.
Few shot dialogue state tracking using meta-learning
Saket Dingliwal, Shuyang Gao, Sanchit Agarwal, Chien-Wei Lin, Tagyoung Chung, and Dilek Hakkani-Tur. 2021 · 2021
Later among the works it cites.
EfficientBERT: Progressively searching multilayer perceptron via warm-up knowledge distillation
Chenhe Dong, Guangrun Wang, Hang Xu, Jiefeng Peng, Xiaozhe Ren, and Xiaodan Liang. 2021 · 2021
Later among the works it cites.
Continual learning in recurrent neural networks
Benjamin Ehret, Christian Henning, Maria Cervera, Alexander Meulemans, Johannes Von Oswald, and Benjamin F Grewe. 2021 · 2021
Later among the works it cites.
Cross-lingual transfer with MAML on trees
Jezabel Garcia, Federica Freddi, Jamie McGowan, Tim Nieradzik, Feng-Ting Liao, Ye Tian, Da-shan Shiu, and Alberto Bernacchia. 2021 · 2021
Later among the works it cites.
Multilingual and cross-lingual document classification: A meta-learning approach
Niels van der Heijden, Helen Yannakoudakis, Pushkar Mishra, and Ekaterina Shutova. 2021 · 2021
Later among the works it cites.
Meta-learning in neural networks: A survey
Timothy M Hospedales, Antreas Antoniou, Paul Micaelli, and Amos J. Storkey. 2021 · 2021
Later among the works it cites.
Multi-accent speech separation with one shot learning
Kuan-Po Huang, Yuan-Kuei Wu, and Hung yi Lee. 2021 · 2021
Later among the works it cites.
Metric learning for keyword spotting
Jaesung Huh, Minjae Lee, Heesoo Heo, Seongkyu Mun, and Joon Son Chung. 2021 · 2021
Later among the works it cites.
A survey of deep meta-learning
Mike Huisman, Jan N. van Rijn, and Aske Plaat. 2021 · 2021
Later among the works it cites.
Pre-training with meta learning for Chinese word segmentation
Zhen Ke, Liang Shi, Songtao Sun, Erli Meng, Bin Wang, and Xipeng Qiu. 2021 · 2021
Later among the works it cites.
A review of domain adaptation without target labels
Wouter M. Kouw and Marco Loog. 2021 · 2021
Later among the works it cites.
Meta-learning for fast cross-lingual adaptation in dependency parsing
Anna Langedijk, Verna Dankers, Phillip Lippe, Sander Bos, Bryan Cardenas Guevara, Helen Yannakoudakis, and Ekaterina Shutova. 2021 · 2021
Later among the works it cites.
Meta learning and its applications to natural language processing
Hung-yi Lee, Ngoc Thang Vu, and Shang-Wen Li. 2021b · 2021
Later among the works it cites.
Meta-learning for improving rare word recognition in end-to-end asr
Florian Lux and Ngoc Thang Vu. 2021 · 2021
Later among the works it cites.
X-METRA-ADA: Cross-lingual meta-transfer learning adaptation to natural language understanding and question answering
Meryem M’hamdi, Doo Soon Kim, Franck Dernoncourt, Trung Bui, Xiang Ren, and Jonathan May. 2021 · 2021
Later among the works it cites.
DReCa: A general task augmentation strategy for few-shot natural language inference
Shikhar Murty, Tatsunori B. Hashimoto, and Christopher Manning. 2021 · 2021
Later among the works it cites.
Data augmentation for meta-learning
Renkun Ni, Micah Goldblum, Amr Sharaf, Kezhi Kong, and Tom Goldstein. 2021 · 2021
Later among the works it cites.
Few-shot learning for slot tagging with attentive relational network
Cennet Oguz and Ngoc Thang Vu. 2021 · 2021
Later among the works it cites.
Unsupervised neural machine translation for low-resource domains via meta-learning
Cheonbok Park, Yunwon Tae, TaeHee Kim, Soyoung Yang, Mohammad Azam Khan, Lucy Park, and Jaegul Choo. 2021 · 2021
Later among the works it cites.
Meta back-translation
Hieu Pham, Xinyi Wang, Yiming Yang, and Graham Neubig. 2021 · 2021
Later among the works it cites.
A student-teacher architecture for dialog domain adaptation under the meta-learning setting
Kun Qian, Wei Wei, and Zhou Yu. 2021 · 2021
Later among the works it cites.
Gradient projection memory for continual learning
Gobinda Saha and Kaushik Roy. 2021 · 2021
Later among the works it cites.
Meta-learning for effective multi-task and multilingual modelling
Ishan Tarunesh, Sushil Khyalia, Vishwajeet Kumar, Ganesh Ramakrishnan, and Preethi Jyothi. 2021 · 2021
Later among the works it cites.
Meta-learning for domain generalization in semantic parsing
Bailin Wang, Mirella Lapata, and Ivan Titov. 2021a · 2021
Later among the works it cites.
Variance-reduced first-order meta-learning for natural language processing tasks
Lingxiao Wang, Kevin Huang, Tengyu Ma, Quanquan Gu, and Jing Huang. 2021b · 2021
Later among the works it cites.
MetaXL: Meta representation transformation for low-resource cross-lingual learning
Mengzhou Xia, Guoqing Zheng, Subhabrata Mukherjee, Milad Shokouhi, Graham Neubig, and Ahmed Hassan Awadallah. 2021 · 2021
Later among the works it cites.
Adversarial meta sampling for multilingual low-resource speech recognition
Yubei Xiao, Ke Gong, Pan Zhou, Guolin Zheng, Xiaodan Liang, and Liang Lin. 2021 · 2021
Later among the works it cites.
Addressing catastrophic forgetting in few-shot problems
Pauching Yap, Hippolyt Ritter, and David Barber. 2021 · 2021
Later among the works it cites.
CrossFit: A few-shot learning challenge for cross-task generalization in NLP
Qinyuan Ye, Bill Yuchen Lin, and Xiang Ren. 2021 · 2021
Later among the works it cites.
Meta label correction for noisy label learning
Guoqing Zheng, Ahmed H. Awadallah, and Susan Dumais. 2021 · 2021
Later among the works it cites.
Meta-learning via language model in-context tuning
Yanda Chen, Ruiqi Zhong, Sheng Zha, George Karypis, and He He. 2022 · 2022
Closest in time.
Meta-TTS: Meta-learning for few-shot speaker adaptive text-to-speech
Sung-feng Huang, Chyi-Jiunn Lin, Da-rong Liu, Yi-chen Chen, and Hung-yi Lee. 2022 · 2022
Closest in time.
BERT learns to teach: Knowledge distillation with meta learning
Wangchunshu Zhou, Canwen Xu, and Julian McAuley. 2022 · 2022
Closest in time.