Fetching the paper…
Reading the bibliography…
Sequence labeling is an important technique employed for many Natural Language Processing (NLP) tasks, such as Named Entity Recognition (NER), slot tagging for dialog systems and semantic parsing.
Probability of error of some adaptive pattern-recognition machines
H. J. Scudder III · 1965
Earlier work this paper cites.
Representing text chunks
Erik F Tjong, Kim Sang, and Jorn Veenstra · 1999
Earlier work this paper cites.
Introduction to the conll-2003 shared task: Language-independent named entity recognition
Erik F. Tjong Kim Sang and Fien De Meulder · 2003
Earlier work this paper cites.
Name tagging with word clusters and discriminative training
Scott Miller, Jethran Guinness, and Alex Zamanian · 2004
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston · 2009
Earlier work this paper cites.
Semi-supervised learning
Olivier Chapelle, Bernhard Schlkopf, and Alexander Zien · 2010
Earlier work this paper cites.
Self-paced learning for latent variable models
M. P. Kumar, Benjamin Packer, and Daphne Koller · 2010
Earlier work this paper cites.
Overview of the 2012 shared task on parsing the web
Slav Petrov and Ryan McDonald · 2012
Earlier work this paper cites.
Asgard: A portable architecture for multilingual dialogue systems
J. Liu, Panupong Pasupat, D. Cyphers, and James R. Glass · 2013
Earlier work this paper cites.
Learning with pseudo-ensembles
Philip Bachman, Ouais Alsharif, and Doina Precup · 2014
Earlier work this paper cites.
Semi-supervised learning with deep generative models
Diederik P. Kingma, Shakir Mohamed, Danilo Jimenez Rezende, and Max Welling · 2014
Earlier work this paper cites.
Semi-supervised learning with ladder networks
Antti Rasmus, Mathias Berglund, Mikko Honkala, Harri Valpola, and Tapani Raiko · 2015
Earlier work this paper cites.
Language as a latent variable: Discrete generative models for sentence compression
Yishu Miao and Phil Blunsom · 2016
Earlier work this paper cites.
Improving neural machine translation models with monolingual data
Rico Sennrich, Barry Haddow, and Alexandra Birch · 2016
Earlier work this paper cites.
Training region-based object detectors with online hard example mining
Abhinav Shrivastava, Abhinav Gupta, and Ross Girshick · 2016
Earlier work this paper cites.
Active bias: Training more accurate neural networks by emphasizing high variance samples
Haw-Shiuan Chang, Erik G. Learned-Miller, and Andrew McCallum · 2017
Earlier work this paper cites.
Deep bayesian active learning with image data
Yarin Gal, Riashat Islam, and Zoubin Ghahramani · 2017
Earlier work this paper cites.
Understanding black-box predictions via influence functions
Pang Wei Koh and Percy Liang · 2017
Cited alongside, same era.
Learning active learning from data
Ksenia Konyushkova, Raphael Sznitman, and Pascal Fua · 2017
Cited alongside, same era.
Temporal ensembling for semi-supervised learning
Samuli Laine and Timo Aila · 2017
Cited alongside, same era.
Self-paced co-training
Fan Ma, Deyu Meng, Qi Xie, Zina Li, and Xuanyi Dong · 2017
Cited alongside, same era.
Cross-lingual name tagging and linking for 282 languages
Xiaoman Pan, Boliang Zhang, Jonathan May, Joel Nothman, Kevin Knight, and Heng Ji · 2017
Cited alongside, same era.
Semi-supervised sequence tagging with bidirectional language models
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Later among the works it cites.
Revisiting self-training for neural sequence generation, 2019
Junxian He, Jiatao Gu, Jiajun Shen, and Marc’Aurelio Ranzato · 2019
Later among the works it cites.
Learning to self-train for semi-supervised few-shot classification
Xinzhe Li, Qianru Sun, Yaoyao Liu, Qin Zhou, Shibao Zheng, Tat-Seng Chua, and Bernt Schiele · 2019
Later among the works it cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Later among the works it cites.
Enhancing deep active learning using selective self-training for image classification
Emmeleia Panagiota Mastoropoulou · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Matthew E Peters, Waleed Ammar, Chandra Bhagavatula, and Russell Power · 2017
Cited alongside, same era.
Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results
Antti Tarvainen and Harri Valpola · 2017
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
Cited alongside, same era.
Multi-space variational encoder-decoders for semi-supervised labeled sequence transduction
Chunting Zhou and Graham Neubig · 2017
Cited alongside, same era.
Semi-supervised sequence modeling with cross-view training
Kevin Clark, Minh-Thang Luong, Christopher D Manning, and Quoc V Le · 2018
Cited alongside, same era.
Snips voice platform: an embedded spoken language understanding system for private-by-design voice interfaces
Alice Coucke, Alaa Saade, Adrien Ball, Théodore Bluche, Alexandre Caulier, David Leroy, Clément Doumouro, Thibault Gisselbrecht, Francesco Caltagirone, Thibaut Lavril, Maël Primet, and Joseph Dureau · 2018
Cited alongside, same era.
Virtual adversarial training: a regularization method for supervised and semi-supervised learning
Takeru Miyato, Shin-ichi Maeda, Masanori Koyama, and Shin Ishii · 2018
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever · 2019
Later among the works it cites.
Meta-weight-net: Learning an explicit mapping for sample weighting
Jun Shu, Qi Xie, Lixuan Yi, Qian Zhao, Sanping Zhou, Zongben Xu, and Deyu Meng · 2019
Later among the works it cites.
Meta-transfer learning for few-shot learning
Q. Sun, Y. Liu, T. Chua, and B. Schiele · 2019
Later among the works it cites.
Unsupervised data augmentation for consistency training, 2019
Qizhe Xie, Zihang Dai, Eduard Hovy, Minh-Thang Luong, and Quoc V. Le · 2019
Later among the works it cites.
Learning to few-shot learn across diverse natural language classification tasks, 2020
Trapit Bansal, Rishikesh Jha, and Andrew McCallum · 2020
Closest in time.
Using error decay prediction to overcome practical issues of deep active learning for named entity recognition
Haw-Shiuan Chang, Shankar Vembu, Sunil Mohan, Rheeya Uppaal, and Andrew McCallum · 2020
Closest in time.
Seqvat: Virtual adversarial training for semi-supervised sequence labeling
Luoxin Chen, Weitong Ruan, Xinyue Liu, and Jianhua Lu · 2020
Closest in time.
Understanding self-training for gradual domain adaptation
Ananya Kumar, Tengyu Ma, and Percy Liang · 2020
Closest in time.
Bond: Bert-assisted open-domain named entity recognition with distant supervision
Chen Liang, Yue Yu, Haoming Jiang, Siawpeng Er, Ruijia Wang, Tuo Zhao, and Chao Zhang · 2020
Closest in time.
Self-training with noisy student improves imagenet classification
Qizhe Xie, Minh-Thang Luong, Eduard Hovy, and Quoc V. Le · 2020
Closest in time.
Rethinking pre-training and self-training
Barret Zoph, Golnaz Ghiasi, Tsung-Yi Lin, Yin Cui, Hanxiao Liu, Ekin Dogus Cubuk, and Quoc Le · 2020
Closest in time.