Fetching the paper…
Reading the bibliography…
Phrase representations derived from BERT often do not exhibit complex phrasal compositionality, as the model relies instead on lexical similarity to determine semantic relatedness.
Transfertransfo: A transfer learning approach for neural network based conversational agents
Thomas Wolf, Victor Sanh, Julien Chaumond, and Clement Delangue. 2019 · 1901
Earlier work this paper cites.
Topic modeling in embedding spaces
Adji B Dieng, Francisco J R Ruiz, and David M Blei. 2019 · 1907
Earlier work this paper cites.
SpanBERT: Improving pre-training by representing and predicting spans
Mandar Joshi, Danqi Chen, Yinhan Liu, Daniel S. Weld, Luke Zettlemoyer, and Omer Levy. 2019 · 1907
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions and reversals
Vladimir I Levenshtein. 1966 · 1966
Earlier work this paper cites.
Mallet: A machine learning for language toolkit
Andrew Kachites McCallum. 2002 · 2002
Earlier work this paper cites.
Latent dirichlet allocation
David M. Blei, Andrew Y. Ng, and Michael I. Jordan. 2003 · 2003
Earlier work this paper cites.
Topics in semantic representation
Thomas L Griffiths, Mark Steyvers, and Joshua B Tenenbaum. 2007 · 2007
Earlier work this paper cites.
Topical n-grams: Phrase and topic discovery, with an application to information retrieval
Xuerui Wang, Andrew McCallum, and Xing Wei. 2007 · 2007
Earlier work this paper cites.
Reading tea leaves: How humans interpret topic models
Jonathan Chang, Sean Gerrish, Chong Wang, Jordan Boyd-Graber, and David Blei. 2009 · 2009
Earlier work this paper cites.
Parsing natural scenes and natural language with recursive neural networks
Richard Socher, Cliff Chiung-Yu Lin, Andrew Y. Ng, and Christopher D. Manning. 2011 · 2011
Earlier work this paper cites.
Domain and function: A dual-space model of semantic relations and compositions
Peter D. Turney. 2012 · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Earlier work this paper cites.
Introduction to the special issue on multiword expressions: From theory to practice and use
Carlos Ramisch, Aline Villavicencio, and Valia Kordoni. 2013 · 2013
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Scalable topical phrase mining from text corpora
Ahmed El-Kishky, Yanglei Song, Chi Wang, Clare R. Voss, and Jiawei Han. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
The Stanford CoreNLP natural language processing toolkit
Christopher D. Manning, Mihai Surdeanu, John Bauer, Jenny Finkel, Steven J. Bethard, and David McClosky. 2014 · 2014
Cited alongside, same era.
Abstractive multi-document summarization via phrase selection and merging
Lidong Bing, Piji Li, Yi Liao, Wai Lam, Weiwei Guo, and Rebecca Passonneau. 2015 · 2015
Cited alongside, same era.
Using phrases in mallet topic models
David Mimno. 2015 · 2015
Cited alongside, same era.
Ppdb 2.0: Better paraphrase ranking, fine-grained entailment relations, word embeddings, and style classification
Ellie Pavlick, Pushpendre Rastogi, Juri Ganitkevitch, Benjamin Van Durme, and Chris Callison-Burch. 2015 · 2015
Cited alongside, same era.
Big BiRD: A large, fine-grained, bigram relatedness dataset for examining semantic composition
Shima Asaadi, Saif Mohammad, and Svetlana Kiritchenko. 2019 · 2019
Later among the works it cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2019 · 2019
Later among the works it cites.
Investigating sports commentator bias within a large corpus of american football broadcasts
Jack Merullo, Luke Yeh, Abram Handler, Alvin Grissom II, Brendan O’Connor, and Mohit Iyyer. 2019 · 2019
Later among the works it cites.
Language models are unsupervised multitask learners
A. Radford, Jeffrey Wu, R. Child, David Luan, Dario Amodei, and Ilya Sutskever. 2019 · 2019
Later among the works it cites.
Sentence-bert: Sentence embeddings using siamese bert-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning composition models for phrase embeddings
Mo Yu and Mark Dredze. 2015 · 2015
Cited alongside, same era.
Fine-grained analysis of sentence embeddings using auxiliary prediction tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2016 · 2016
Cited alongside, same era.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering
Ruining He and Julian McAuley. 2016 · 2016
Cited alongside, same era.
Feuding families and former friends: Unsupervised learning for dynamic fictional relationships
Mohit Iyyer, Anupam Guha, Snigdha Chaturvedi, Jordan Boyd-Graber, and Hal Daumé III. 2016 · 2016
Cited alongside, same era.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and R. Socher. 2017 · 2017
Cited alongside, same era.
Pushing the limits of paraphrastic sentence embeddings with millions of machine translations
John Wieting and Kevin Gimpel. 2017 · 2017
Cited alongside, same era.
Learning phrase embeddings from paraphrases with GRUs
Zhihao Zhou, Lifu Huang, and Heng Ji. 2017 · 2017
Cited alongside, same era.
PAWS: Paraphrase Adversaries from Word Scrambling
Yuan Zhang, Jason Baldridge, and Luheng He. 2019 · 2019
Later among the works it cites.
STORIUM: A Dataset and Evaluation Platform for Machine-in-the-Loop Story Generation
Nader Akoury, Shufan Wang, Josh Whiting, Stephen Hood, Nanyun Peng, and Mohit Iyyer. 2020 · 2020
Later among the works it cites.
Interpreting Pretrained Contextualized Representations via Reductions to Static Embeddings
Rishi Bommasani, Kelly Davis, and Claire Cardie. 2020 · 2020
Later among the works it cites.
Analyzing gender bias within narrative tropes
Dhruvil Gala, Mohammad Omar Khursheed, Hannah Lerner, Brendan O’Connor, and Mohit Iyyer. 2020 · 2020
Later among the works it cites.
The Pile: An 800gb dataset of diverse text for language modeling
Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, Shawn Presser, and Connor Leahy. 2020 · 2020
Later among the works it cites.
spaCy: Industrial-strength Natural Language Processing in Python
Matthew Honnibal, Ines Montani, Sofie Van Landeghem, and Adriane Boyd. 2020 · 2020
Later among the works it cites.
Reformulating unsupervised style transfer as paraphrase generation
Kalpesh Krishna, John Wieting, and Mohit Iyyer. 2020 · 2020
Later among the works it cites.
On the sentence embeddings from pre-trained language models
Bohan Li, Hao Zhou, Junxian He, Mingxuan Wang, Yiming Yang, and Lei Li. 2020 · 2020
Later among the works it cites.
A cross-task analysis of text span representations
Shubham Toshniwal, Haoyue Shi, Bowen Shi, Lingyu Gao, Karen Livescu, and Kevin Gimpel. 2020 · 2020
Later among the works it cites.
Assessing phrasal representation and composition in transformers
Lang Yu and Allyson Ettinger. 2020 · 2020
Later among the works it cites.
Learning dense representations of phrases at scale
Jinhyuk Lee, Mujeen Sung, Jaewoo Kang, and Danqi Chen. 2021 · 2021
Closest in time.