Fetching the paper…
Reading the bibliography…
The dominant approaches to text representation in natural language rely on learning embeddings on massive corpora which have convenient properties such as compositionality and distance preservation.
An informational theory of the statistical structure of language
Benoit Mandelbrot · 1953
Earlier work this paper cites.
Statistical methods for multivariate extremes: an application to structural design
Stuart G Coles and Jonathan A Tawn · 1994
Earlier work this paper cites.
Poisson mixtures
Kenneth W Church and William A Gale · 1995
Earlier work this paper cites.
Wordnet: a lexical database for english
George A Miller · 1995
Earlier work this paper cites.
Novelty detection using extreme value statistics
S.J. Roberts · 1999
Earlier work this paper cites.
Extreme value statistics for novelty detection in biomedical data processing
S.J Roberts · 2000
Earlier work this paper cites.
Word frequency distributions , volume 18
R Harald Baayen · 2002
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Jauvin · 2003
Earlier work this paper cites.
Simulating multivariate extreme value distributions of logistic type
A. Stephenson · 2003
Earlier work this paper cites.
Modeling word burstiness using the dirichlet distribution
Rasmus E Madsen, David Kauchak, and Charles Elkan · 2005
Earlier work this paper cites.
A review of machine learning approaches to spam filtering
Thiago S Guzella and Walmir M Caminhas · 2009
Earlier work this paper cites.
Information-based models for ad hoc ir
Stéphane Clinchant and Eric Gaussier · 2010
Earlier work this paper cites.
Cohesion, coherence, and expert evaluations of writing proficiency
Scott Crossley and Danielle McNamara · 2010
Earlier work this paper cites.
Novelty detection with multivariate extreme value statistics
D. A. Clifton, S. Hugueny, and L. Tarassenko · 2011
Earlier work this paper cites.
Hidden factors and hidden topics: understanding rating dimensions with review text
Julian McAuley and Jure Leskovec · 2013
Earlier work this paper cites.
Extreme values, regular variation and point processes
Sidney I Resnick · 2013
Earlier work this paper cites.
On power law distributions in large-scale taxonomies
Rohit Babbar, Cornelia Metzig, Ioannis Partalas, Eric Gaussier, and Massih-Reza Amini · 2014
Earlier work this paper cites.
Extreme bandits
A. Carpentier and M. Valko · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Cited alongside, same era.
Personalized entity recommendation: A heterogeneous information network approach
Xiao Yu, Xiang Ren, Yizhou Sun, Quanquan Gu, Bradley Sturt, Urvashi Khandelwal, Brandon Norick, and Jiawei Han · 2014
Cited alongside, same era.
Learning the dependence structure of rare events: a non-asymptotic study
N. Goix, A. Sabourin, and S. Clémençon · 2015
Cited alongside, same era.
From group to individual labels using deep features
Dimitrios Kotzias, Misha Denil, Nando De Freitas, and Padhraic Smyth · 2015
Cited alongside, same era.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan · 2015
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Later among the works it cites.
The effectiveness of data augmentation in image classification using deep learning
Jason Wang and Luis Perez · 2017
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Later among the works it cites.
Deep k k -means: Jointly clustering with k k -means and learning representations
Maziar Moradi Fard, Thibaut Thonet, and Eric Gaussier · 2018
Later among the works it cites.
On binary classification in extreme regions
Hamid Jalalzai, Stephan Clémençon, and Anne Sabourin · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jialu Liu, Jingbo Shang, Chi Wang, Xiang Ren, and Jiawei Han · 2015
Cited alongside, same era.
Alireza Makhzani, Jonathon Shlens, Navdeep Jaitly, Ian Goodfellow, and Brendan Frey · 2015
Cited alongside, same era.
Feature clustering for extreme events analysis, with application to extreme stream-flow data
Maël Chiapino and Anne Sabourin · 2016
Cited alongside, same era.
Sparse representation of multivariate extremes with applications to anomaly ranking
N. Goix, A. Sabourin, and S. Clémençon · 2016
Cited alongside, same era.
Deep Learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville · 2016
Cited alongside, same era.
Bag of tricks for efficient text classification
Armand Joulin, Edouard Grave, Piotr Bojanowski, and Tomas Mikolov · 2016
Cited alongside, same era.
Max k-armed bandit: On the extremehunter algorithm and beyond
Mastane Achab, Stephan Clémençon, Aurélien Garivier, Anne Sabourin, and Claire Vernade · 2017
Cited alongside, same era.
Sosuke Kobayashi · 2018
Later among the works it cites.
Autoencoding any data through kernel autoencoders
Pierre Laforgue, Stephan Clémençon, and Florence d’Alché Buc · 2018
Later among the works it cites.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer · 2018
Later among the works it cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever · 2018
Later among the works it cites.
Disney at iest 2018: Predicting emotions using an ensemble
Wojciech Witon, Pierre Colombo, Ashutosh Modi, and Mubbasir Kapadia · 2018
Later among the works it cites.
Affect-driven dialog generation
Pierre Colombo, Wojciech Witon, Ashutosh Modi, James Kennedy, and Mubbasir Kapadia · 2019
Later among the works it cites.
From the token to the review: A hierarchical multimodal approach to opinion mining
Alexandre Garcia, Pierre Colombo, Slim Essid, Florence d’Alché Buc, and Chloé Clavel · 2019
Later among the works it cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Maxwell Forbes, and Yejin Choi · 2019
Later among the works it cites.
Eda: Easy data augmentation techniques for boosting performance on text classification tasks
Jason Wei and Kai Zou · 2019
Later among the works it cites.
Multimix: A robust data augmentation strategy for cross-lingual nlp
M Saiful Bari, Muhammad Tasnim Mohiuddin, and Shafiq Joty · 2020
Closest in time.
Hierarchical pre-training for sequence labelling in spoken dialog
Emile Chapuis, Pierre Colombo, Matteo Manica, Matthieu Labeau, and Chloe Clavel · 2020
Closest in time.
Guiding attention in sequence-to-sequence models for dialogue act prediction
Pierre Colombo, Emile Chapuis, Matteo Manica, Emmanuel Vignon, Giovanna Varni, and Chloe Clavel · 2020
Closest in time.
The importance of fillers for text representations of speech transcripts
Tanvi Dinkar, Pierre Colombo, Matthieu Labeau, and Chloé Clavel · 2020
Closest in time.
Informative clusters for multivariate extremes
Hamid Jalalzai and Rémi Leluc · 2020
Closest in time.