Fetching the paper…
Reading the bibliography…
Recent work on transformer-based neural networks has led to impressive advances on multiple-choice natural language understanding (NLU) problems, such as Question Answering (QA) and abductive reasoning.
Cross-lingual language model pretraining
Lample, Guillaume and Alexis Conneau. 2019 · 1901
Earlier work this paper cites.
Right for the wrong reasons: Diagnosing syntactic heuristics in natural language inference
McCoy, R Thomas, Ellie Pavlick, and Tal Linzen. 2019 · 1902
Earlier work this paper cites.
Scibert: A pretrained language model for scientific text
Beltagy, Iz, Kyle Lo, and Arman Cohan. 2019 · 1903
Earlier work this paper cites.
Linguistic knowledge and transferability of contextual representations
Liu, Nelson F, Matt Gardner, Yonatan Belinkov, Matthew E Peters, and Noah A Smith. 2019a · 1903
Earlier work this paper cites.
Docbert: Bert for document classification
Adhikari, Ashutosh, Achyudh Ram, Raphael Tang, and Jimmy Lin. 2019 · 1904
Earlier work this paper cites.
Publicly available clinical bert embeddings
Alsentzer, Emily, John R Murphy, Willie Boag, Wei-Hung Weng, Di Jin, Tristan Naumann, and Matthew McDermott. 2019 · 1904
Earlier work this paper cites.
Zhang, Xingxing, Furu Wei, and Ming Zhou. 2019 · 1905
Earlier work this paper cites.
Patentbert: Patent classification with fine-tuning a pre-trained bert model
Lee, Jieh-Sheng and Jieh Hsiang. 2019 · 1906
Earlier work this paper cites.
Eli5: Long form question answering
Fan, Angela, Yacine Jernite, Ethan Perez, David Grangier, Jason Weston, and Michael Auli. 2019 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized bert pretraining approach
Liu, Yinhan, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Text summarization with pretrained encoders
Liu, Yang and Mirella Lapata. 2019 · 1908
Earlier work this paper cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Lu, Jiasen, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 1908
Earlier work this paper cites.
" mask and infill": Applying masked language model to sentiment transfer
Wu, Xing, Tao Zhang, Liangjun Zang, Jizhong Han, and Songlin Hu. 2019 · 1908
Earlier work this paper cites.
Albert: A lite bert for self-supervised learning of language representations
Lan, Zhenzhong, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2019 · 1909
Earlier work this paper cites.
Do nlp models know numbers? probing numeracy in embeddings
Wallace, Eric, Yizhong Wang, Sujian Li, Sameer Singh, and Matt Gardner. 2019b · 1909
Earlier work this paper cites.
Mlqa: Evaluating cross-lingual extractive question answering
Lewis, Patrick, Barlas Oğuz, Ruty Rinott, Sebastian Riedel, and Holger Schwenk. 2019 · 1910
Earlier work this paper cites.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Sanh, Victor, Lysandre Debut, Julien Chaumond, and Thomas Wolf. 2019 · 1910
Earlier work this paper cites.
Multilingual question answering from formatted text applied to conversational agents
Siblini, Wissam, Charlotte Pasqual, Axel Lavielle, and Cyril Cauchois. 2019 · 1910
Earlier work this paper cites.
Dialogpt: Large-scale generative pre-training for conversational response generation
Zhang, Yizhe, Siqi Sun, Michel Galley, Yen-Chun Chen, Chris Brockett, Xiang Gao, Jianfeng Gao, Jingjing Liu, and Bill Dolan. 2019 · 1911
Earlier work this paper cites.
Paralysis by analysis: is your planning system becoming too rational?
Lenz, RT and Marjorie A Lyles. 1985 · 1985
Earlier work this paper cites.
Natural language question answering: the view from here
Hirschman, Lynette and Robert Gaizauskas. 2001 · 2001
Earlier work this paper cites.
Attention module is not only a weight: Analyzing transformers with vector norms
Kobayashi, Goro, Tatsuki Kuribayashi, Sho Yokoi, and Kentaro Inui. 2020 · 2004
Earlier work this paper cites.
The paradox of choice: Why more is less
Schwartz, Barry. 2004 · 2004
Cited alongside, same era.
Perturbed masking: Parameter-free probing for analyzing and interpreting bert
Wu, Zhiyong, Yun Chen, Ben Kao, and Qun Liu. 2020 · 2004
Cited alongside, same era.
The pascal recognising textual entailment challenge
Dagan, Ido, Oren Glickman, and Bernardo Magnini. 2005 · 2005
Cited alongside, same era.
Unifiedqa: Crossing format boundaries with a single qa system
Khashabi, Daniel, Sewon Min, Tushar Khot, Ashish Sabharwal, Oyvind Tafjord, Peter Clark, and Hannaneh Hajishirzi. 2020 · 2005
Cited alongside, same era.
Beyond accuracy: Behavioral testing of nlp models with checklist
Ribeiro, Marco Tulio, Tongshuang Wu, Carlos Guestrin, and Sameer Singh. 2020 · 2005
Cited alongside, same era.
Abductive natural language inference (anli)
Bhagavatula, Chandra, Ronan Le Bras, Chaitanya Malaviya, Keisuke Sakaguchi, et al. 2019 · 2019
Later among the works it cites.
Target-dependent sentiment classification with bert
Gao, Zhengjie, Ao Feng, Xinyu Song, and Xi Wu. 2019 · 2019
Later among the works it cites.
What does bert learn about the structure of language?
Jawahar, Ganesh, Benoît Sagot, and Djamé Seddah. 2019 · 2019
Later among the works it cites.
Fine-grained sentiment classification using bert
Munikar, Manish, Sushil Shakya, and Aakash Shrestha. 2019 · 2019
Later among the works it cites.
Coqa: A conversational question answering challenge
Reddy, Siva, Danqi Chen, and Christopher D Manning. 2019 · 2019
Later among the works it cites.
Social iqa: Commonsense reasoning about social interactions
Sap, Maarten, Hannah Rashkin, Derek Chen, Ronan Le Bras, and Yejin Choi. 2019a · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The second pascal recognising textual entailment challenge
Haim, R Bar, Ido Dagan, Bill Dolan, Lisa Ferro, Danilo Giampiccolo, Bernardo Magnini, and Idan Szpektor. 2006 · 2006
Cited alongside, same era.
Selective question answering under domain shift
Kamath, Amita, Robin Jia, and Percy Liang. 2020 · 2006
Cited alongside, same era.
What’s in a name? are bert named entity representations just as good for any other name?
Balasubramanian, Sriram, Naman Jain, Gaurav Jindal, Abhijeet Awasthi, and Sunita Sarawagi. 2020 · 2007
Cited alongside, same era.
The third pascal recognizing textual entailment challenge
Giampiccolo, Danilo, Bernardo Magnini, Ido Dagan, and William B Dolan. 2007 · 2007
Cited alongside, same era.
Can neural networks acquire a structural bias from raw linguistic data?
Warstadt, Alex and Samuel R Bowman. 2020 · 2007
Cited alongside, same era.
The fifth pascal recognizing textual entailment challenge
Bentivogli, Luisa, Peter Clark, Ido Dagan, and Danilo Giampiccolo. 2009 · 2009
Cited alongside, same era.
Thinking, fast and slow
Kahneman, Daniel. 2011 · 2011
Cited alongside, same era.
Effective sentence scoring method using bert for speech recognition
Shin, Joonbo, Yoonhyung Lee, and Kyomin Jung. 2019 · 2019
Later among the works it cites.
Videobert: A joint model for video and language representation learning
Sun, Chen, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid. 2019 · 2019
Later among the works it cites.
Sentiment classification using document embeddings trained with cosine similarity
Thongtan, Tan and Tanasanee Phienthrakul. 2019 · 2019
Later among the works it cites.
Universal Adversarial Triggers for Attacking and Analyzing NLP
Wallace, Eric, Shi Feng, Nikhil Kandpal, Matt Gardner, and Sameer Singh. 2019a · 2019
Later among the works it cites.
Augmenting dialogue response generation with unstructured textual knowledge
Wang, Yanmeng, Wenge Rong, Yuanxin Ouyang, and Zhang Xiong. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Zhilin, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
Abductive commonsense reasoning
Bhagavatula, Chandra, Ronan Le Bras, Chaitanya Malaviya, Keisuke Sakaguchi, Ari Holtzman, Hannah Rashkin, Doug Downey, Wen tau Yih, and Yejin Choi. 2020 · 2020
Later among the works it cites.
What bert is not: Lessons from a new suite of psycholinguistic diagnostics for language models
Ettinger, Allyson. 2020 · 2020
Later among the works it cites.
Biobert: a pre-trained biomedical language representation model for biomedical text mining
Lee, Jinhyuk, Wonjin Yoon, Sungdong Kim, Donghyeon Kim, Sunkyu Kim, Chan Ho So, and Jaewoo Kang. 2020 · 2020
Later among the works it cites.
K-bert: Enabling language representation with knowledge graph
Liu, Weijie, Peng Zhou, Zhe Zhao, Zhiruo Wang, Qi Ju, Haotang Deng, and Ping Wang. 2020 · 2020
Later among the works it cites.
Probing natural language inference models through semantic fragments
Richardson, Kyle, Hai Hu, Lawrence Moss, and Ashish Sabharwal. 2020 · 2020
Later among the works it cites.
A primer in bertology: What we know about how bert works
Rogers, Anna, Olga Kovaleva, and Anna Rumshisky. 2020 · 2020
Later among the works it cites.
Cord-19: The covid-19 open research dataset
Wang, Lucy Lu, Kyle Lo, Yoganand Chandrasekhar, Russell Reas, Jiangjiang Yang, Darrin Eide, Kathryn Funk, Rodney Kinney, Ziyang Liu, William Merrill, et al. 2020 · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Wolf, Thomas, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Later among the works it cites.
Knowing more about questions can help: Improving calibration in question answering
Zhang, Shujian, Chengyue Gong, and Eunsol Choi. 2021 · 2021
Later among the works it cites.