Fetching the paper…
Reading the bibliography…
While contextualized word embeddings have been a de-facto standard, learning contextualized phrase embeddings is less explored and being hindered by the lack of a human-annotated benchmark that tests machine understanding of phrase semantics given a context sentence or paragraph (instead of phrases alone).
Learning words from context
William E Nagy, Patricia A Herman, and Richard C Anderson. 1985 · 1985
Earlier work this paper cites.
Learning words from context and dictionaries: An experimental comparison
Ute Fischer. 1994 · 1994
Earlier work this paper cites.
SENSEVAL-2: Overview
Philip Edmonds and Scott Cotton. 2001 · 2001
Earlier work this paper cites.
Placing search in context: the concept revisited
Lev Finkelstein, Evgeniy Gabrilovich, Yossi Matias, Ehud Rivlin, Zach Solan, Gadi Wolfman, and Eytan Ruppin. 2001 · 2001
Earlier work this paper cites.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan. 2020 · 2004
Earlier work this paper cites.
Improving statistical machine translation using word sense disambiguation
Marine Carpuat and Dekai Wu. 2007b · 2007
Earlier work this paper cites.
Natural Language Processing with Python: Analyzing Text with the Natural Language Toolkit
Steven Bird, Ewan Klein, and Edward Loper. 2009 · 2009
Earlier work this paper cites.
Domain and function: A dual-space model of semantic relations and compositions
Peter D Turney. 2012 · 2012
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
PPDB 2.0: Better paraphrase ranking, fine-grained entailment relations, word embeddings, and style classification
Ellie Pavlick, Pushpendre Rastogi, Juri Ganitkevitch, Benjamin Van Durme, and Chris Callison-Burch. 2015 · 2015
Earlier work this paper cites.
From paraphrase database to compositional paraphrase model and back
John Wieting, Mohit Bansal, Kevin Gimpel, and Karen Livescu. 2015 · 2015
Earlier work this paper cites.
SQuAD: 100,000+ questions for machine comprehension of text
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Earlier work this paper cites.
End-to-end neural coreference resolution
Kenton Lee, Luheng He, Mike Lewis, and Luke Zettlemoyer. 2017 · 2017
Earlier work this paper cites.
Daniel Cer, Yinfei Yang, Sheng-yi Kong, Nan Hua, Nicole Limtiaco, Rhomni St John, Noah Constant, Mario Guajardo-Cespedes, Steve Yuan, Chris Tar, et al. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018 · 2018
Earlier work this paper cites.
Big BiRD: A large, fine-grained, bigram relatedness dataset for examining semantic composition
Shima Asaadi, Saif Mohammad, and Svetlana Kiritchenko. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Natural questions: A benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Cited alongside, same era.
WiC: the word-in-context dataset for evaluating context-sensitive meaning representations
Mohammad Taher Pilehvar and Jose Camacho-Collados. 2019 · 2019
Cited alongside, same era.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Learning dense representations of phrases at scale
Jinhyuk Lee, Mujeen Sung, Jaewoo Kang, and Danqi Chen. 2021 · 2021
Later among the works it cites.
Wikimedia downloads
Wikimedia Team. 2021b · 2021
Later among the works it cites.
Phrase-BERT: Improved phrase embeddings from BERT with an application to corpus exploration
Shufan Wang, Laure Thompson, and Mohit Iyyer. 2021 · 2021
Later among the works it cites.
princeton-nlp/sup-simcse-roberta-large · hugging face
Princeton NLP Group. 2022 · 2022
Closest in time.
super_glue · datasets at hugging face
Huggingface. 2022a · 2022
Closest in time.
transformers/examples/pytorch/question-answering at main · huggingface/transformers
Huggingface. 2022b · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
PAWS-X: A cross-lingual adversarial dataset for paraphrase identification
Yinfei Yang, Yuan Zhang, Chris Tar, and Jason Baldridge. 2019 · 2019
Cited alongside, same era.
PAWS: Paraphrase adversaries from word scrambling
Yuan Zhang, Jason Baldridge, and Luheng He. 2019 · 2019
Cited alongside, same era.
Multi-span question answering using span-image network
Tarik Arici, Hayreddin Ceker, and Ismail Baha Tutar. 2020 · 2020
Cited alongside, same era.
spaCy: Industrial-strength Natural Language Processing in Python
Matthew Honnibal, Ines Montani, Sofie Van Landeghem, and Adriane Boyd. 2020 · 2020
Cited alongside, same era.
SpanBERT: Improving pre-training by representing and predicting spans
Mandar Joshi, Danqi Chen, Yinhan Liu, Daniel S. Weld, Luke Zettlemoyer, and Omer Levy. 2020 · 2020
Cited alongside, same era.
Assessing phrasal representation and composition in transformers
Lang Yu and Allyson Ettinger. 2020 · 2020
Cited alongside, same era.
Recent trends in word sense disambiguation: A survey
Michele Bevilacqua, Tommaso Pasini, Alessandro Raganato, Roberto Navigli, et al. 2021 · 2021
Cited alongside, same era.
Languagetool - online grammar, style & spell checker
LanguageTool. 2022 · 2022
Closest in time.
upwork_annotation_guidelines.pdf
PiC. 2021a · 2022
Closest in time.
upwork_samples.pdf
PiC. 2021b · 2022
Closest in time.
upwork_samples.pdf
PiC. 2022 · 2022
Closest in time.
Api:categories - mediawiki
Wikimedia Team. 2021a · 2022
Closest in time.
Universal sentence encoder | tensorflow hub
TensorFlow. 2022 · 2022
Closest in time.
Webscope | yahoo labs
Yahoo. 2022a · 2022
Closest in time.
Webscope | yahoo labs
Yahoo. 2022b · 2022
Closest in time.
arcface-pytorch/test.py at master · ronghuaiyang/arcface-pytorch
Ronghui Yang. 2022 · 2022
Closest in time.