Fetching the paper…
Reading the bibliography…
Natural Language Inference (NLI) and Semantic Textual Similarity (STS) are widely used benchmark tasks for compositional evaluation of pre-trained language models.
RoBERTa: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Compound thoughts
Gottlob Frege. 1963 · 1963
Earlier work this paper cites.
The structure of a semantic theory
Jerrold Katz and Jerry Fodor. 1963 · 1963
Earlier work this paper cites.
Probing linguistic systematicity
Emily Goodwin, Koustuv Sinha, and Timothy J. O’Donnell. 2020 · 1969
Earlier work this paper cites.
The proper treatment of quantification in ordinary English
Richard Montague. 1973 · 1973
Earlier work this paper cites.
Definiteness and Indefiniteness. A Study in Reference and Grammaticality Prediction
John Hawkins. 1978 · 1978
Earlier work this paper cites.
The Semantics of Definite and Indefinite Noun Phrases
Irene Heim. 1982 · 1982
Earlier work this paper cites.
Logical form constraints and configurational structures in japanese
Hajime Hoji. 1985 · 1985
Earlier work this paper cites.
Some asymmetries in Japanese and their theoretical implications
Mamoru Saito. 1985 · 1985
Earlier work this paper cites.
Japanese: Descriptive Grammar
John Hinds. 1986 · 1986
Earlier work this paper cites.
The Languages of Japan
Masayoshi Shibatani. 1990 · 1990
Earlier work this paper cites.
FraCaS–a framework for computational semantics
Robin Cooper, Richard Crouch, Jan van Eijck, Chris Fox, Josef van Genabith, Jan Jaspers, Hans Kamp, Manfred Pinkal, Massimo Poesio, Stephen Pulman, et al. 1994 · 1994
Earlier work this paper cites.
Compositionality
Theo Janssen and Barbara Partee. 1997 · 1997
Earlier work this paper cites.
ipadic version 2.7.0 User’s Manual
Masayuki Asahara and Yuji Matsumoto. 2003 · 2003
Earlier work this paper cites.
Japanese plurals are exceptional
Kimiko Nakanishi and Satoshi Tomioka. 2004 · 2004
Earlier work this paper cites.
The pascal recognising textual entailment challenge
Ido Dagan, Oren Glickman, and Bernardo Magnini. 2006 · 2006
Earlier work this paper cites.
Tregex and tsurgeon: tools for querying and manipulating tree data structures
Roger Levy and Galen Andrew. 2006 · 2006
Earlier work this paper cites.
FarsTail: A Persian natural language inference dataset
Hossein Amirkhani, Mohammad Azari Jafari, Azadeh Amirak, Zohreh Pourjafari, Soroush Faridan Jahromi, and Zeinab Kouhkan. 2020 · 2009
Earlier work this paper cites.
Japanese and korean voice search
Mike Schuster and Kaisuke Nakajima. 2012 · 2012
Earlier work this paper cites.
A SICK cure for the evaluation of compositional distributional semantic models
Marco Marelli, Stefano Menini, Marco Baroni, Luisa Bentivogli, Raffaella Bernardi, and Roberto Zamparelli. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Morphological analysis for unsegmented languages using recurrent neural network language model
Hajime Morita, Daisuke Kawahara, and Sadao Kurohashi. 2015 · 2015
Earlier work this paper cites.
SemEval-2016 task 1: Semantic textual similarity, monolingual and cross-lingual evaluation
Eneko Agirre, Carmen Banea, Daniel Cer, Mona Diab, Aitor Gonzalez-Agirre, Rada Mihalcea, German Rigau, and Janyce Wiebe. 2016 · 2016
Cited alongside, same era.
SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation
Daniel Cer, Mona Diab, Eneko Agirre, Iñigo Lopez-Gazpio, and Lucia Specia. 2017 · 2017
Cited alongside, same era.
Textual inference: getting logic from humans
Aikaterini-Lida Kalouli, Livy Real, and Valeria de Paiva. 2017 · 2017
Cited alongside, same era.
An inference problem set for evaluating semantic theories and semantic processing systems for japanese
Ai Kawazoe, Ribeka Tanaka, Koji Mineshima, and Daisuke Bekki. 2017 · 2017
Cited alongside, same era.
A* CCG parsing with a supertag and dependency factored model
Masashi Yoshikawa, Hiroshi Noji, and Yuji Matsumoto. 2017 · 2017
Cited alongside, same era.
KorNLI and KorSTS: New benchmark datasets for Korean natural language understanding
Jiyeon Ham, Yo Joong Choe, Kyubyong Park, Ilji Choi, and Hyungjoon Soh. 2020 · 2020
Later among the works it cites.
Japanese realistic textual entailment corpus
Yuta Hayashibe. 2020 · 2020
Later among the works it cites.
OCNLI: Original Chinese Natural Language Inference
Hai Hu, Kyle Richardson, Liang Xu, Lu Li, Sandra Kübler, and Lawrence Moss. 2020 · 2020
Later among the works it cites.
The state and fate of linguistic diversity and inclusion in the NLP world
Pratik Joshi, Sebastin Santy, Amar Budhiraja, Kalika Bali, and Monojit Choudhury. 2020 · 2020
Later among the works it cites.
FlauBERT: Unsupervised language model pre-training for French
Hang Le, Loïc Vial, Jibril Frej, Vincent Segonne, Maximin Coavoux, Benjamin Lecouteux, Alexandre Allauzen, Benoit Crabbé, Laurent Besacier, and Didier Schwab. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
XNLI: Evaluating cross-lingual sentence representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Cited alongside, same era.
Breaking NLI systems with sentences that require simple lexical inferences
Max Glockner, Vered Shwartz, and Yoav Goldberg. 2018 · 2018
Cited alongside, same era.
Stress test evaluation for natural language inference
Aakanksha Naik, Abhilasha Ravichander, Norman Sadeh, Carolyn Rose, and Graham Neubig. 2018 · 2018
Cited alongside, same era.
SICK-BR: A portuguese corpus for inference
Livy Real et al. 2018 · 2018
Cited alongside, same era.
Juman++: A morphological analysis toolkit for scriptio continua
Arseny Tolmachev, Daisuke Kawahara, and Sadao Kurohashi. 2018 · 2018
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Yaobo Liang, Nan Duan, Yeyun Gong, Ning Wu, Fenfei Guo, Weizhen Qi, Ming Gong, Linjun Shou, Daxin Jiang, Guihong Cao, Xiaodong Fan, Ruofei Zhang, Rahul Agrawal, Edward Cui, Sining Wei, Taroon Bharti, Ying Qiao, Jiun-Hung Chen, Winnie Wu, Shuguang Liu, Fan Yang, Daniel Campos, Rangan Majumder, and Ming Zhou. 2020 · 2020
Later among the works it cites.
How can we accelerate progress towards human-like linguistic generalization?
Tal Linzen. 2020 · 2020
Later among the works it cites.
RussianSuperGLUE: A Russian language understanding evaluation benchmark
Tatiana Shavrina, Alena Fenogenova, Emelyanov Anton, Denis Shevelev, Ekaterina Artemova, Valentin Malykh, Vladislav Mikhailov, Maria Tikhonova, Andrey Chertok, and Andrey Evlampiev. 2020 · 2020
Later among the works it cites.
CLUE: A Chinese language understanding evaluation benchmark
Liang Xu, Hai Hu, Xuanwei Zhang, Lu Li, Chenjie Cao, Yudong Li, Yechen Xu, Kai Sun, Dian Yu, Cong Yu, Yin Tian, Qianqian Dong, Weitang Liu, Bo Shi, Yiming Cui, Junyi Li, Jun Zeng, Rongzhao Wang, Weijian Xie, Yanting Li, Yina Patterson, Zuoyu Tian, Yiwen Zhang, He Zhou, Shaoweihua Liu, Zhe Zhao, Qipeng Zhao, Cong Yue, Xinrui Zhang, Zhengliang Yang, Kyle Richardson, and Zhenzhong Lan. 2020 · 2020
Later among the works it cites.
Multilingualization of natural language inference datasets using machine translation (in Japanese)
Takumi Yoshikoshi, Daisuke Kawahara, and Sadao Kurohashi. 2020 · 2020
Later among the works it cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Later among the works it cites.
BERT & family eat word salad: Experiments with text understanding
Ashim Gupta, Giorgi Kvernadze, and Vivek Srikumar. 2021 · 2021
Later among the works it cites.
How effective is BERT without word ordering? implications for language understanding and data privacy
Jack Hessel and Alexandra Schofield. 2021 · 2021
Later among the works it cites.
Lower perplexity is not always human-like
Tatsuki Kuribayashi, Yohei Oseki, Takumi Ito, Ryo Yoshida, Masayuki Asahara, and Kentaro Inui. 2021 · 2021
Later among the works it cites.
KLUE: Korean language understanding evaluation
Sungjoon Park, Jihyung Moon, Sungdong Kim, Won Ik Cho, Jiyoon Han, Jangwon Park, Chisung Song, Junseong Kim, Yongsook Song, Taehwan Oh, Joohong Lee, Juhyun Oh, Sungwon Lyu, Younghoon Jeong, Inkwon Lee, Sangwoo Seo, Dongjun Lee, Hyunwoo Kim, Myeonghwa Lee, Seongbo Jang, Seungwon Do, Sunkyoung Kim, Kyungtae Lim, Jongwon Lee, Kyumin Park, Jamin Shin, Seonghyun Kim, Lucy Park, Alice Oh, Jung-Woo Ha, and Kyunghyun Cho. 2021 · 2021
Later among the works it cites.
Out of order: How important is the sequential order of words in a sentence in natural language understanding tasks?
Thang Pham, Trung Bui, Long Mai, and Anh Nguyen. 2021 · 2021
Later among the works it cites.
How good is your tokenizer? on the monolingual performance of multilingual language models
Phillip Rust, Jonas Pfeiffer, Ivan Vulić, Sebastian Ruder, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
ALUE: Arabic language understanding evaluation
Haitham Seelawi, Ibraheem Tuffaha, Mahmoud Gzawi, Wael Farhan, Bashar Talafha, Riham Badawi, Zyad Sober, Oday Al-Dweik, Abed Alhakim Freihat, and Hussein Al-Natsheh. 2021 · 2021
Later among the works it cites.
Examining the inductive bias of neural language models with artificial languages
Jennifer C. White and Ryan Cotterell. 2021 · 2021
Later among the works it cites.
SICK-NL: A dataset for Dutch natural language inference
Gijs Wijnholds and Michael Moortgat. 2021 · 2021
Later among the works it cites.
Exploring transitivity in neural NLI models through veridicality
Hitomi Yanaka, Koji Mineshima, and Kentaro Inui. 2021 · 2021
Later among the works it cites.