Fetching the paper…
Reading the bibliography…
Linguistic style is an integral component of language.
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Language style as audience design
Allan Bell. 1984 · 1984
Earlier work this paper cites.
On the utility of content analysis in author attribution: "the federalist"
Colin Martindale and Dean McKenzie. 1995 · 1995
Earlier work this paper cites.
Author identification, idiolect, and linguistic uniqueness
Malcolm Coulthard. 2004 · 2004
Earlier work this paper cites.
The importance of suppressing domain style in authorship analysis
Sebastian Bischoff, Niklas Deckers, Marcel Schliebs, Ben Thies, Matthias Hagen, Efstathios Stamatatos, Benno Stein, and Martin Potthast. 2020 · 2005
Earlier work this paper cites.
Dimensionality reduction by learning an invariant mapping
Raia Hadsell, Sumit Chopra, and Yann LeCun. 2006 · 2006
Earlier work this paper cites.
Stylometric analysis of bloggers’ age and gender
Sumit Goswami, Sudeshna Sarkar, and Mayur Rustagi. 2009 · 2009
Earlier work this paper cites.
Classifying latent user attributes in Twitter
Delip Rao, David Yarowsky, Abhishek Shreevats, and Manaswi Gupta. 2010 · 2010
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay. 2011 · 2011
Earlier work this paper cites.
Authorship analysis studies: A survey
Sara El Manar El and Ismail Kassou. 2014 · 2014
Earlier work this paper cites.
A large annotated corpus for learning natural language inference
Samuel R. Bowman, Gabor Angeli, Christopher Potts, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
Demographic factors improve classification performance
Dirk Hovy. 2015 · 2015
Earlier work this paper cites.
Optimizing statistical machine translation for text simplification
Wei Xu, Courtney Napoles, Ellie Pavlick, Quanze Chen, and Chris Callison-Burch. 2016 · 2016
Earlier work this paper cites.
SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation
Daniel Cer, Mona Diab, Eneko Agirre, Iñigo Lopez-Gazpio, and Lucia Specia. 2017 · 2017
Earlier work this paper cites.
Supervised learning of universal sentence representations from natural language inference data
Alexis Conneau, Douwe Kiela, Holger Schwenk, Loïc Barrault, and Antoine Bordes. 2017 · 2017
Earlier work this paper cites.
Controlling linguistic style aspects in neural language generation
Jessica Ficler and Yoav Goldberg. 2017 · 2017
Earlier work this paper cites.
Surveying stylometry techniques and applications
Tempestt Neal, Kalaivani Sundararajan, Aneez Fatima, Yiming Yan, Yingfei Xiang, and Damon Woodard. 2017 · 2017
Earlier work this paper cites.
A study of style in machine translation: Controlling the formality of machine translation output
Xing Niu, Marianna Martindale, and Marine Carpuat. 2017 · 2017
Cited alongside, same era.
Personalized machine translation: Preserving original author traits
Ella Rabinovich, Raj Nath Patel, Shachar Mirkin, Lucia Specia, and Shuly Wintner. 2017 · 2017
Cited alongside, same era.
Convolutional neural networks for authorship attribution of short texts
Prasha Shrestha, Sebastian Sierra, Fabio González, Manuel Montes, Paolo Rosso, and Thamar Solorio. 2017 · 2017
Cited alongside, same era.
Masking topic-related information to enhance authorship attribution
Efstathios Stamatatos. 2017 · 2017
Cited alongside, same era.
SentEval: An evaluation toolkit for universal sentence representations
Alexis Conneau and Douwe Kiela. 2018 · 2018
Cited alongside, same era.
Hypothesis only baselines in natural language inference
Deep dive into authorship verification of email messages with convolutional neural network
Marina Litvak. 2019 · 2019
Later among the works it cites.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Later among the works it cites.
The Pushshift Reddit dataset
Jason Baumgartner, Savvas Zannettou, Brian Keegan, Megan Squire, and Jeremy Blackburn. 2020 · 2020
Later among the works it cites.
ConvoKit: A toolkit for the analysis of conversations
Jonathan P. Chang, Caleb Chiam, Liye Fu, Andrew Wang, Justine Zhang, and Cristian Danescu-Niculescu-Mizil. 2020 · 2020
Later among the works it cites.
Array programming with NumPy
Charles R. Harris, K. Jarrod Millman, Stéfan J. van der Walt, Ralf Gommers, Pauli Virtanen, David Cournapeau, Eric Wieser, Julian Taylor, Sebastian Berg, Nathaniel J. Smith, Robert Kern, Matti Picus, Stephan Hoyer, Marten H. van Kerkwijk, Matthew Brett, Allan Haldane, Jaime Fernández del Río, Mark Wiebe, Pearu Peterson, Pierre Gérard-Marchant, Kevin Sheppard, Tyler Reddy, Warren Weckesser, Hameer Abbasi, Christoph Gohlke, and Travis E. Oliphant. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adam Poliak, Jason Naradowsky, Aparajita Haldar, Rachel Rudinger, and Benjamin Van Durme. 2018 · 2018
Cited alongside, same era.
Intrinsic author verification using topic modeling
Nektaria Potha and Efstathios Stamatatos. 2018 · 2018
Cited alongside, same era.
Dear sir or madam, may I introduce the GYAFC dataset: Corpus, benchmarks and metrics for formality style transfer
Sudha Rao and Joel Tetreault. 2018 · 2018
Cited alongside, same era.
Topic or style? Exploring the most useful features for authorship attribution
Yunita Sari, Mark Stevenson, and Andreas Vlachos. 2018 · 2018
Cited alongside, same era.
What represents “style” in authorship attribution?
Kalaivani Sundararajan and Damon Woodard. 2018 · 2018
Cited alongside, same era.
A broad-coverage challenge corpus for sentence understanding through inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
Learning semantic textual similarity from conversations
Yinfei Yang, Steve Yuan, Daniel Cer, Sheng-yi Kong, Noah Constant, Petr Pilar, Heming Ge, Yun-Hsuan Sung, Brian Strope, and Ray Kurzweil. 2018 · 2018
Cited alongside, same era.
Representation learning of writing style
Julien Hay, Bich-Lien Doan, Fabrice Popineau, and Ouassim Ait Elhara. 2020 · 2020
Later among the works it cites.
Deepstyle: User style embedding for authorship attribution of short texts
Zhiqiang Hu, Roy Ka-Wei Lee, Lei Wang, Ee-peng Lim, and Bo Dai. 2020 · 2020
Later among the works it cites.
Stylometrics features under domain shift: Do they really “context-independent”?
Tatiana Litvinova. 2020 · 2020
Later among the works it cites.
SciPy 1.0: Fundamental Algorithms for Scientific Computing in Python
Pauli Virtanen, Ralf Gommers, Travis E. Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, Stéfan J. van der Walt, Matthew Brett, Joshua Wilson, K. Jarrod Millman, Nikolay Mayorov, Andrew R. J. Nelson, Eric Jones, Robert Kern, Eric Larson, C J Carey, İlhan Polat, Yu Feng, Eric W. Moore, Jake VanderPlas, Denis Laxalde, Josef Perktold, Robert Cimrman, Ian Henriksen, E. A. Quintero, Charles R. Harris, Anne M. Archibald, Antônio H. Ribeiro, Fabian Pedregosa, Paul van Mulbregt, and SciPy 1.0 Contributors. 2020 · 2020
Later among the works it cites.
SimCSE: Simple contrastive learning of sentence embeddings
Tianyu Gao, Xingcheng Yao, and Danqi Chen. 2021 · 2021
Later among the works it cites.
DeCLUTR: Deep contrastive learning for unsupervised textual representations
John Giorgi, Osvald Nitski, Bo Wang, and Gary Bader. 2021 · 2021
Later among the works it cites.
Overview of the cross-domain authorship verification task at PAN 2021
Mike Kestemont, Enrique Manjavacas, Ilia Markov, Janek Bevendorff, Matti Wiegmann, Efstathios Stamatatos, Benno Stein, and Martin Potthast. 2021 · 2021
Later among the works it cites.
DialogueCSE: Dialogue-based contrastive learning of sentence embeddings
Che Liu, Rui Wang, Jinghua Liu, Jian Sun, Fei Huang, and Luo Si. 2021 · 2021
Later among the works it cites.
On learning and representing social meaning in NLP: a sociolinguistic perspective
Dong Nguyen, Laura Rosseel, and Jack Grieve. 2021 · 2021
Later among the works it cites.
Siamese networks for large-scale author identification
Chakaveh Saedi and Mark Dras. 2021 · 2021
Later among the works it cites.
Does it capture STEL? A modular, similarity-based linguistic style evaluation framework
Anna Wegmann and Dong Nguyen. 2021 · 2021
Later among the works it cites.
Idiosyncratic but not arbitrary: Learning idiolects in online registers reveals distinctive yet consistent individual styles
Jian Zhu and David Jurgens. 2021 · 2021
Later among the works it cites.