Fetching the paper…
Reading the bibliography…
Recent work has raised concerns about the inherent limitations of text-only pretraining.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2019 · 1910
Earlier work this paper cites.
Computational analysis of present-day American English
H. Kucera and W. N. Francis. 1967 · 1967
Earlier work this paper cites.
Basic Color Terms: Their Universality and Evolution
Brent Berlin and Paul Kay. 1969 · 1969
Earlier work this paper cites.
Logic and conversation
Herbert P Grice. 1975 · 1975
Earlier work this paper cites.
The open images dataset v4
Alina Kuznetsova, Hassan Rom, Neil Alldrin, Jasper Uijlings, Ivan Krasin, Jordi Pont-Tuset, Shahab Kamali, Stefan Popov, Matteo Malloci, Alexander Kolesnikov, and et al. 2020 · 1981
Earlier work this paper cites.
Sensory and cognitive contributions of color to the recognition of natural scenes
Karl R Gegenfurtner and Jochem Rieger. 2000 · 2000
Earlier work this paper cites.
The university of south florida free association, rhyme, and word fragment norms
Douglas L Nelson, Cathy L McEvoy, and Thomas A Schreiber. 2004 · 2004
Earlier work this paper cites.
The world color survey
Paul Kay, Brent Berlin, Luisa Maffi, William R Merrifield, and Richard Cook. 2009 · 2009
Earlier work this paper cites.
Unbiased look at dataset bias
Antonio Torralba and Alexei A. Efros. 2011 · 2011
Earlier work this paper cites.
Syntactic annotations for the Google Books NGram corpus
Yuri Lin, Jean-Baptiste Michel, Erez Aiden Lieberman, Jon Orwant, Will Brockman, and Slav Petrov. 2012 · 2012
Earlier work this paper cites.
The centre for speech, language and the brain (cslb) concept property norms
Barry Devereux, Lorraine Tyler, Jeroen Geertzen, and Billi Randall. 2013 · 2013
Earlier work this paper cites.
Reporting bias and knowledge acquisition
Jonathan Gordon and Benjamin Van Durme. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Analogy-based detection of morphological and semantic relations with word embeddings: what works and what doesn’t
Anna Gladkova, Aleksandr Drozd, and Satoshi Matsuoka. 2016 · 2016
Earlier work this paper cites.
Seeing through the human reporting bias: Visual classifiers from noisy human-centric labels
Ishan Misra, C. Lawrence Zitnick, Margaret Mitchell, and Ross B. Girshick. 2016 · 2016
Cited alongside, same era.
Stereotyping and bias in the flickr30k dataset
Emiel van Miltenburg. 2016 · 2016
Cited alongside, same era.
Making the V in VQA matter: Elevating the role of image understanding in visual question answering
Yash Goyal, Tejas Khot, Douglas Summers-Stay, Dhruv Batra, and Devi Parikh. 2017 · 2017
Cited alongside, same era.
Women also snowboard: Overcoming bias in captioning models
Kaylee Burns, Lisa Anne Hendricks, Trevor Darrell, and Anna Rohrbach. 2018 · 2018
Cited alongside, same era.
Color statistics of objects, and color tuning of object cortex in macaque monkey
Isabelle Rosenthal, Sivalogeswaran Ratnasingam, Theodros Haile, Serena Eastman, Josh Fuller-Deets, and Bevil R. Conway. 2018 · 2018
Inducing relational knowledge from BERT
Zied Bouraoui, José Camacho-Collados, and Steven Schockaert. 2020 · 2020
Later among the works it cites.
What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models
Allyson Ettinger. 2020 · 2020
Later among the works it cites.
How can we know what language models know?
Zhengbao Jiang, Frank F. Xu, Jun Araki, and Graham Neubig. 2020 · 2020
Later among the works it cites.
ALBERT: A lite BERT for self-supervised learning of language representations
Zhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel, Piyush Sharma, and Radu Soricut. 2020 · 2020
Later among the works it cites.
Birds have four legs?! NumerSense: Probing Numerical Commonsense Knowledge of Pre-Trained Language Models
Bill Yuchen Lin, Seyeon Lee, Rahul Khanna, and Xiang Ren. 2020 · 2020
Later among the works it cites.
Do neural language models overcome reporting bias?
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Kelly Zhang and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
Cracking the contextual commonsense code: Understanding commonsense reasoning aptitude of deep contextual representations
Jeff Da and Jungo Kasai. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
GQA: A new dataset for real-world visual reasoning and compositional question answering
Drew A. Hudson and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
Open sesame: Getting inside BERT’s linguistic knowledge
Yongjie Lin, Yi Chern Tan, and Robert Frank. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Cited alongside, same era.
Vered Shwartz and Yejin Choi. 2020 · 2020
Later among the works it cites.
The influence of object-color knowledge on emerging object representations in the brain
Lina Teichmann, Genevieve L. Quek, Amanda K. Robinson, Tijl Grootswagers, Thomas A. Carlson, and Anina N. Rich. 2020 · 2020
Later among the works it cites.
Information-theoretic probing with minimum description length
Elena Voita and Ivan Titov. 2020 · 2020
Later among the works it cites.
BLiMP: The benchmark of linguistic minimal pairs for English
Alex Warstadt, Alicia Parrish, Haokun Liu, Anhad Mohananey, Wei Peng, Sheng-Fu Wang, and Samuel R. Bowman. 2020 · 2020
Later among the works it cites.
Probing neural language models for human tacit assumptions
Nathaniel Weir, Adam Poliak, and Benjamin Van Durme. 2020 · 2020
Later among the works it cites.
PROST: Physical reasoning about objects through space and time
Stéphane Aroca-Ouellette, Cory Paik, Alessandro Roncone, and Katharina Kann. 2021 · 2021
Closest in time.
Scaling up visual and vision-language representation learning with noisy text supervision
Chao Jia, Yinfei Yang, Ye Xia, Yi-Ting Chen, Zarana Parekh, Hieu Pham, Quoc V. Le, Yunhsuan Sung, Zhen Li, and Tom Duerig. 2021 · 2021
Closest in time.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. 2021 · 2021
Closest in time.
Evaluating representations by the complexity of learning low-loss predictors
William F. Whitney, Min Jae Song, David Brandfonbrener, Jaan Altosaar, and Kyunghyun Cho. 2021 · 2021
Closest in time.
Vokenization: Improving language understanding with contextualized, visual-grounded supervision
Hao Tan and Mohit Bansal. 2020 · 2080
Closest in time.