Fetching the paper…
Reading the bibliography…
In many applications of machine learning, certain categories of examples may be underrepresented in the training data, causing systems to underperform on such "few-shot" cases at test time.
Fewrel 2.0: Towards more challenging few-shot relation classification
Tianyu Gao, Xu Han, Hao Zhu, Zhiyuan Liu, Peng Li, Maosong Sun, and Jie Zhou. 2019 · 1910
Earlier work this paper cites.
Exploiting Cloze Questions for Few Shot Text Classification and Natural Language Inference
Timo Schick and Hinrich Schütze. 2020 · 2001
Earlier work this paper cites.
Oshin Agarwal, Yinfei Yang, Byron C. Wallace, and Ani Nenkova. 2020 · 2004
Earlier work this paper cites.
Language Models are Few-Shot Learners
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2005
Earlier work this paper cites.
Example-Based Named Entity Recognition
Morteza Ziyadi, Yuting Sun, Abhishek Goswami, Jade Huang, and Weizhu Chen. 2020 · 2008
Earlier work this paper cites.
Overcoming Conflicting Data for Model Updates
David Gaddy, Alex Kouzemtchenko, Pavan Kumar Reddy, Prateek Kolhar, and Rushin Shah. 2020 · 2010
Earlier work this paper cites.
Evaluation of output embeddings for fine-grained image classification
Zeynep Akata, Scott Reed, Daniel Walter, Honglak Lee, and Bernt Schiele. 2015 · 2015
Earlier work this paper cites.
Data recombination for neural semantic parsing
Robin Jia and Percy Liang. 2016 · 2016
Earlier work this paper cites.
Matching networks for one shot learning
Oriol Vinyals, Charles Blundell, Timothy Lillicrap, koray kavukcuoglu, and Daan Wierstra. 2016 · 2016
Earlier work this paper cites.
Towards zero-shot frame semantic parsing for domain scaling
Ankur Bapna, Gokhan Tür, Dilek Hakkani-Tür, and Larry Heck. 2017 · 2017
Earlier work this paper cites.
The Effectiveness of Data Augmentation in Image Classification using Deep Learning
Luis Perez and Jason Wang. 2017 · 2017
Earlier work this paper cites.
Prototypical networks for few-shot learning
Jake Snell, Kevin Swersky, and Richard Zemel. 2017 · 2017
Earlier work this paper cites.
Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results
Antti Tarvainen and H. Valpola. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Alice Coucke, Alaa Saade, Adrien Ball, Théodore Bluche, Alexandre Caulier, David Leroy, Clément Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, Maël Primet, and Joseph Dureau. 2018 · 2018
Cited alongside, same era.
Slot-gated modeling for joint slot filling and intent prediction
Chih-Wen Goo, Guang Gao, Yun-Kai Hsu, Chih-Li Huo, Tsung-Chieh Chen, Keng-Wei Hsu, and Yun-Nung Chen. 2018 · 2018
Cited alongside, same era.
FewRel: A large-scale supervised few-shot relation classification dataset with state-of-the-art evaluation
Xu Han, Hao Zhu, Pengfei Yu, Ziyun Wang, Yuan Yao, Zhiyuan Liu, and Maosong Sun. 2018 · 2018
Cited alongside, same era.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Hierarchical attention prototypical networks for few-shot text classification
Shengli Sun, Qingfeng Sun, Kevin Zhou, and Tengchao Lv. 2019 · 2019
Later among the works it cites.
Do not have enough data? deep learning to the rescue!
Ateret Anaby-Tavor, Boaz Carmeli, Esther Goldbraich, Amir Kantor, George Kour, Segev Shlomov, Naama Tepper, and Naama Zwerdling. 2020 · 2020
Later among the works it cites.
Good-enough compositional data augmentation
Jacob Andreas. 2020 · 2020
Later among the works it cites.
Few-shot slot tagging with collapsed dependency transfer and label-enhanced task-adaptive projection network
Yutai Hou, Wanxiang Che, Yongkui Lai, Zhihan Zhou, Yijia Liu, Han Liu, and Ting Liu. 2020 · 2020
Later among the works it cites.
Learning to classify intents and slot labels given a handful of examples
Jason Krone, Yi Zhang, and Mona Diab. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Evaluation of named entity coreference
Oshin Agarwal, Sanjay Subramanian, Ani Nenkova, and Dan Roth. 2019 · 2019
Cited alongside, same era.
Matching the blanks: Distributional similarity for relation learning
Livio Baldini Soares, Nicholas FitzGerald, Jeffrey Ling, and Tom Kwiatkowski. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
A closer look at feature space data augmentation for few-shot intent classification
Varun Kumar, Hadrien Glaude, Cyprien de Lichy, and Wlliam Campbell. 2019 · 2019
Cited alongside, same era.
An evaluation dataset for intent classification and out-of-scope prediction
Stefan Larson, Anish Mahendran, Joseph J. Peper, Christopher Clarke, Andrew Lee, Parker Hill, Jonathan K. Kummerfeld, Kevin Leach, Michael A. Laurenzano, Lingjia Tang, and Jason Mars. 2019 · 2019
Cited alongside, same era.
Zero-shot entity linking by reading entity descriptions
Lajanugen Logeswaran, Ming-Wei Chang, Kenton Lee, Kristina Toutanova, Jacob Devlin, and Honglak Lee. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Cited alongside, same era.
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
How much knowledge can you pack into the parameters of a language model?
Adam Roberts, Colin Raffel, and Noam Shazeer. 2020 · 2020
Later among the works it cites.
AutoPrompt: Eliciting Knowledge from Language Models with Automatically Generated Prompts
Taylor Shin, Yasaman Razeghi, Robert L. Logan IV, Eric Wallace, and Sameer Singh. 2020 · 2020
Later among the works it cites.
Simple and effective few-shot named entity recognition with structured nearest neighbor learning
Yi Yang and Arzoo Katiyar. 2020 · 2020
Later among the works it cites.
Meta-learning without memorization
Mingzhang Yin, George Tucker, Mingyuan Zhou, Sergey Levine, and Chelsea Finn. 2020 · 2020
Later among the works it cites.
Discriminative nearest neighbor few-shot intent detection by transferring natural language inference
Jianguo Zhang, Kazuma Hashimoto, Wenhao Liu, Chien-Sheng Wu, Yao Wan, Philip Yu, Richard Socher, and Caiming Xiong. 2020 · 2020
Later among the works it cites.
Learning to recombine and resample data for compositional generalization
Ekin Akyürek, Afra Feyza Akyürek, and Jacob Andreas. 2021 · 2021
Closest in time.