Fetching the paper…
Reading the bibliography…
Annotation studies often require annotators to familiarize themselves with the task, its annotation scheme, and the data domain.
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
Sanh, Victor, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Deep Bayesian active learning with image data
Gal, Yarin, Riashat Islam, and Zoubin Ghahramani. 2017 · 1932
Earlier work this paper cites.
Teoria statistica delle classi e calcolo delle probabilita
Bonferroni, Carlo. 1936 · 1936
Earlier work this paper cites.
On the Comparison of Several Mean Values: An Alternative Approach
Welch, Bernard Lewis. 1951 · 1951
Earlier work this paper cites.
Use of Ranks in One-Criterion Variance Analysis
Kruskal, William H. and W. Allen Wallis. 1952 · 1952
Earlier work this paper cites.
Cloze Procedure”: A New Tool for Measuring Readability
Taylor, Wilson L. 1953 · 1953
Earlier work this paper cites.
Derivation Of New Readability Formulas (Automated Readability Index, Fog Count And Flesch Reading Ease Formula) For Navy Enlisted Personnel
Kincaid, J Peter, Robert P Fishburne Jr, Richard L Rogers, and Brad S Chissom. 1975 · 1975
Earlier work this paper cites.
Mind in society: The development of higher psychological processes
Vygotsky, Lev. 1978 · 1978
Earlier work this paper cites.
Principles and Practice in Second Language Acquisition
Krashen, Stephen. 1982 · 1982
Earlier work this paper cites.
A Sequential Algorithm for Training Text Classifiers
Lewis, David D. and William A. Gale. 1994 · 1994
Earlier work this paper cites.
The basics of item response theory
Baker, Frank. 2001 · 2001
Earlier work this paper cites.
Common European Framework of Reference for Languages: learning, teaching, assessment
Council of Europe. 2001 · 2001
Earlier work this paper cites.
Toward Optimal Active Learning through Sampling Estimation of Error Reduction
Roy, Nicholas and Andrew McCallum. 2001 · 2001
Earlier work this paper cites.
Active Learning Using Pre-Clustering
Nguyen, Hieu T. and Arnold Smeulders. 2004 · 2004
Earlier work this paper cites.
GameFlow: A Model for Evaluating Player Enjoyment in Games
Sweetser, Penelope and Peta Wyeth. 2005 · 2005
Earlier work this paper cites.
Active Learning with Real Annotation Costs
Settles, Burr, Mark Craven, and Lewis Friedland. 2008 · 2008
Earlier work this paper cites.
Cheap and Fast – But is it Good? Evaluating Non-Expert Annotations for Natural Language Tasks
Snow, Rion, Brendan O’Connor, Daniel Jurafsky, and Andrew Ng. 2008 · 2008
Earlier work this paper cites.
Active Learning with Sampling by Uncertainty and Density for Word Sense Disambiguation and Text Classification
Zhu, Jingbo, Huizhen Wang, Tianshun Yao, and Benjamin K Tsou. 2008 · 2008
Earlier work this paper cites.
Curriculum learning
Bengio, Yoshua, Jérôme Louradour, Ronan Collobert, and Jason Weston. 2009 · 2009
Earlier work this paper cites.
The Curriculum : Theory and Practice
Kelly, A. V. 2009 · 2009
Earlier work this paper cites.
Timed Annotations - Enhancing MUC7 Metadata by the Time It Takes to Annotate Named Entities
Tomanek, Katrin and Udo Hahn. 2009 · 2009
Earlier work this paper cites.
Influence of Pre-Annotation on POS-Tagged Corpus Development
Fort, Karën and Benoît Sagot. 2010 · 2010
Earlier work this paper cites.
Active Learning by Querying Informative and Representative Examples
Huang, Sheng-Jun, Rong Jin, and Zhi-Hua Zhou. 2010 · 2010
Earlier work this paper cites.
Self-paced learning for latent variable models
Kumar, M., Benjamin Packer, and Daphne Koller. 2010 · 2010
Earlier work this paper cites.
Twitter as a corpus for sentiment analysis and opinion mining
Pak, Alexander and Patrick Paroubek. 2010 · 2010
Earlier work this paper cites.
Numbers rule : the vexing mathematics of democracy, from Plato to the present
Szpiro, George. 2010 · 2010
Earlier work this paper cites.
Automatic Gap-fill Question Generation from Text Books
Agarwal, Manish and Prashanth Mannem. 2011 · 2011
Earlier work this paper cites.
Active Learning with Amazon Mechanical Turk
Laws, Florian, Christian Scheible, and Hinrich Schütze. 2011 · 2011
Earlier work this paper cites.
Scikit-Learn: Machine Learning in Python
Pedregosa, Fabian, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, Jake Vanderplas, Alexandre Passos, David Cournapeau, Matthieu Brucher, Matthieu Perrot, and Édouard Duchesnay. 2011 · 2011
Earlier work this paper cites.
Self-taught active learning from crowds
Fang, Meng, Xingquan Zhu, Bin Li, Wei Ding, and Xindong Wu. 2012 · 2012
Cited alongside, same era.
Generating Diagnostic Multiple Choice Comprehension Cloze Questions
Mostow, Jack and Hyeju Jang. 2012 · 2012
Cited alongside, same era.
Active Learning
Settles, Burr. 2012 · 2012
Cited alongside, same era.
On the definition of a confounder
VanderWeele, Tyler J and Ilya Shpitser. 2013 · 2013
Cited alongside, same era.
Difficult cases: From data to learning, and back
Beigman Klebanov, Beata and Eyal Beigman. 2014 · 2014
Cited alongside, same era.
Predicting the Difficulty of Language Proficiency Tests
Beinborn, Lisa, Torsten Zesch, and Iryna Gurevych. 2014 · 2014
Cited alongside, same era.
Practical Obstacles to Deploying Active Learning
Lowell, David, Zachary C. Lipton, and Byron C. Wallace. 2019 · 2019
Later among the works it cites.
Progression in a Language Annotation Game with a Purpose
Madge, Chris, Juntao Yu, Jon Chamberlain, Udo Kruschwitz, Silviu Paun, and Massimo Poesio. 2019 · 2019
Later among the works it cites.
To Tune or Not to Tune? Adapting Pretrained Representations to Diverse Tasks
Peters, Matthew E., Sebastian Ruder, and Noah A. Smith. 2019 · 2019
Later among the works it cites.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reimers, Nils and Iryna Gurevych. 2019 · 2019
Later among the works it cites.
Analysis of Automatic Annotation Suggestions for Hard Discourse-Level Tasks in Expert Domains
Schulz, Claudia, Christian M. Meyer, Jan Kiesewetter, Michael Sailer, Elisabeth Bauer, Martin R. Fischer, Frank Fischer, and Iryna Gurevych. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fang, Meng, Jie Yin, and Dacheng Tao. 2014 · 2014
Cited alongside, same era.
Evaluating the impact of pre-annotation on annotation speed and potential bias: natural language processing gold standard development for clinical named entity recognition in clinical trial announcements
Lingren, Todd, Louise Deleger, Katalin Molnar, Haijun Zhai, Jareen Meinzen-Derr, Megan Kaiser, Laura Stoutenborough, Qi Li, and Imre Solti. 2014 · 2014
Cited alongside, same era.
Automatic Annotation Suggestions and Custom Annotation Layers in WebAnno
Yimam, Seid Muhie, Chris Biemann, Richard Eckart de Castilho, and Iryna Gurevych. 2014 · 2014
Cited alongside, same era.
A large annotated corpus for learning natural language inference
Bowman, Samuel R., Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Cited alongside, same era.
Active learning for sense annotation
Martínez Alonso, Héctor, Barbara Plank, Anders Johannsen, and Anders Søgaard. 2015 · 2015
Cited alongside, same era.
Active Learning from Weak and Strong Labelers
Zhang, Chicheng and Kamalika Chaudhuri. 2015 · 2015
Cited alongside, same era.
Predicting Annotation Difficulty to Improve Task Routing and Model Performance for Biomedical Information Extraction
Yang, Yinfei, Oshin Agarwal, Chris Tar, Byron C. Wallace, and Ani Nenkova. 2019 · 2019
Later among the works it cites.
Difficulty-aware Distractor Generation for Gap-Fill Items
Yeung, Chak Yan, John Lee, and Benjamin Tsou. 2019 · 2019
Later among the works it cites.
Deep Batch Active Learning by Diverse, Uncertain Gradient Lower Bounds
Ash, Jordan T., Chicheng Zhang, Akshay Krishnamurthy, John Langford, and Alekh Agarwal. 2020 · 2020
Later among the works it cites.
Understanding the Tradeoff between Cost and Quality of Expert Annotations for Keyphrase Extraction
Chau, Hung, Saeid Balaneshin, Kai Liu, and Ondrej Linda. 2020 · 2020
Later among the works it cites.
"linguistic features for readability assessment"
Deutsch, Tovly, Masoud Jasbi, and Stuart Shieber. 2020 · 2020
Later among the works it cites.
Distractor Analysis and Selection for Multiple-Choice Cloze Questions for Second-Language Learners
Gao, Lingyu, Kevin Gimpel, and Arnar Jensson. 2020 · 2020
Later among the works it cites.
Don’t Stop Pretraining: Adapt Language Models to Domains and Tasks
Gururangan, Suchin, Ana Marasović, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A. Smith. 2020 · 2020
Later among the works it cites.
COVIDLies: Detecting COVID-19 misinformation on social media
Hossain, Tamanna, Robert L. Logan IV, Arjuna Ugarte, Yoshitomo Matsubara, Sean Young, and Sameer Singh. 2020 · 2020
Later among the works it cites.
Teaching citizen scientists to categorize glitches using machine learning guided training
Jackson, Corey, Carsten Østerlund, Kevin Crowston, Mahboobeh Harandi, Sarah Allen, Sara Bahaadini, Scotty Coughlin, Vicky Kalogera, Aggelos Katsaggelos, Shane Larson, et al. 2020 · 2020
Later among the works it cites.
Aggregation Driven Progression System for GWAPs
Kicikoglu, Osman Doruk, Richard Bartle, Jon Chamberlain, Silviu Paun, and Massimo Poesio. 2020 · 2020
Later among the works it cites.
From Zero to Hero: Human-In-The-Loop Entity Linking in Low Resource Domains
Klie, Jan-Christoph, Richard Eckart de Castilho, and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
Empowering Active Learning to Jointly Optimize System and User Demands
Lee, Ji-Ung, Christian M. Meyer, and Iryna Gurevych. 2020 · 2020
Later among the works it cites.
Should You Fine-Tune BERT for Automated Essay Scoring?
Mayfield, Elijah and Alan W Black. 2020 · 2020
Later among the works it cites.
Adversarial NLI: A New Benchmark for Natural Language Understanding
Nie, Yixin, Adina Williams, Emily Dinan, Mohit Bansal, Jason Weston, and Douwe Kiela. 2020 · 2020
Later among the works it cites.
Bias in word embeddings
Papakyriakopoulos, Orestis, Simon Hegelich, Juan Carlos Medina Serrano, and Fabienne Marco. 2020 · 2020
Later among the works it cites.
Masked Language Model Scoring
Salazar, Julian, Davis Liang, Toan Q. Nguyen, and Katrin Kirchhoff. 2020 · 2020
Later among the works it cites.
The Influence of Input Data Complexity on Crowdsourcing Quality
Tauchmann, Christopher, Johannes Daxenberger, and Margot Mieskes. 2020 · 2020
Later among the works it cites.
Cold-start Active Learning through Self-supervised Language Modeling
Yuan, Michelle, Hsuan-Tien Lin, and Jordan Boyd-Graber. 2020 · 2020
Later among the works it cites.
Investigating label suggestions for opinion mining in German covid-19 social media
Beck, Tilman, Ji-Ung Lee, Christina Viehmann, Marcus Maurer, Oliver Quiring, and Iryna Gurevych. 2021 · 2021
Closest in time.
Datasheets for datasets
Gebru, Timnit, Jamie Morgenstern, Briana Vecchione, Jennifer Wortman Vaughan, Hanna Wallach, Hal Daumé III, and Kate Crawford. 2021 · 2021
Closest in time.
Changing the world by changing the data
Rogers, Anna. 2021 · 2021
Closest in time.
Winogrande: An adversarial winograd schema challenge at scale
Sakaguchi, Keisuke, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2021 · 2021
Closest in time.
“Everyone Wants to Do the Model Work, Not the Data Work”: Data Cascades in High-Stakes AI
Sambasivan, Nithya, Shivani Kapania, Hannah Highfill, Diana Akrong, Praveen Paritosh, and Lora M Aroyo. 2021 · 2021
Closest in time.