Fetching the paper…
Reading the bibliography…
When working with textual data, a natural application of disentangled representations is fair classification where the goal is to make predictions without being biased (or influenced) by sensitive attributes that may be present in the data (e.g., age, gender or race).
Affect-driven dialog generation
Pierre Colombo, Wojciech Witon, Ashutosh Modi, James Kennedy, and Mubbasir Kapadia. 2019 · 1904
Earlier work this paper cites.
Evaluating gender bias in machine translation
Gabriel Stanovsky, Noah A Smith, and Luke Zettlemoyer. 2019 · 1906
Earlier work this paper cites.
Understanding the limitations of variational mutual information estimators
Jiaming Song and Stefano Ermon. 2019 · 1910
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
Information and the accuracy attainable in the estimation of statistical parameters
C. Radhakrishna Rao. 1945 · 1945
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
A relationship between arbitrary positive matrices and doubly stochastic matrices
Richard Sinkhorn. 1964 · 1964
Earlier work this paper cites.
Hausdorff distances and interpolations
Jean Serra. 1998 · 1998
Earlier work this paper cites.
Rao’s distance measure
Colin Atkinson and Ann F. S. Mitchell. 1981 · 2002
Earlier work this paper cites.
The trimmed iterative closest point algorithm
Dmitry Chetverikov, Dmitry Svirko, Dmitry Stepanov, and Pavel Krsek. 2002 · 2002
Earlier work this paper cites.
On the estimation of information measures of continuous distributions
Georg Pichler, Pablo Piantanida, and Günther Koliander. 2020 · 2002
Earlier work this paper cites.
Heavy-tailed representations, text polarity classification & data augmentation
Hamid Jalalzai, Pierre Colombo, Chloé Clavel, Eric Gaussier, Giovanna Varni, Emmanuel Vignon, and Anne Sabourin. 2020 · 2003
Earlier work this paper cites.
Estimation of entropy and mutual information
Liam Paninski. 2003 · 2003
Earlier work this paper cites.
Topics in Optimal Transportation
Cedric Villani. 2003 · 2003
Earlier work this paper cites.
Null it out: Guarding protected attributes by iterative nullspace projection
Shauli Ravfogel, Yanai Elazar, Hila Gonen, Michael Twiton, and Yoav Goldberg. 2020 · 2004
Earlier work this paper cites.
Language (technology) is power: A critical survey of" bias" in nlp
Su Lin Blodgett, Solon Barocas, Hal Daumé III, and Hanna Wallach. 2020 · 2005
Earlier work this paper cites.
Improving disentangled text representation learning with information-theoretic guidance
Pengyu Cheng, Martin Renqiang Min, Dinghan Shen, Christopher Malon, Yizhe Zhang, Yitong Li, and Lawrence Carin. 2020b · 2006
Earlier work this paper cites.
A systematic review of natural language processing for knowledge management in healthcare
Ganga Prasad Basyal, Bhaskar P Rimal, and David Zeng. 2020 · 2007
Earlier work this paper cites.
A kernel method for the two-sample-problem
Arthur Gretton, Karsten Borgwardt, Malte Rasch, Bernhard Schölkopf, and Alex Smola. 2007 · 2007
Earlier work this paper cites.
Approximating the kullback leibler divergence between gaussian mixture models
John R Hershey and Peder A Olsen. 2007 · 2007
Earlier work this paper cites.
A survey on natural language processing (nlp) and applications in insurance
Antoine Ly, Benno Uthayasooriyar, and Tingting Wang. 2020 · 2010
Earlier work this paper cites.
Estimating divergence functionals and the likelihood ratio by convex risk minimization
XuanLong Nguyen, Martin J Wainwright, and Michael I Jordan. 2010 · 2010
Earlier work this paper cites.
Differential-geometrical methods in statistics , volume 28
Shun-ichi Amari. 2012 · 2012
Earlier work this paper cites.
Japanese and korean voice search
Mike Schuster and Kaisuke Nakajima. 2012 · 2012
Earlier work this paper cites.
Sinkhorn distances: Lightspeed coputation of optimal transportation
Marco Cuturi, Olivier Teboul, and Jean-Philippe Vert. 2013 · 2013
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Overview of the 2nd author profiling task at pan 2014
Francisco Rangel, Paolo Rosso, Martin Potthast, Martin Trenkmann, Benno Stein, Ben Verhoeven, Walter Daelemans, et al. 2014 · 2014
Cited alongside, same era.
Fisher information distance: A geometrical reading
Sueli IR Costa, Sandra A Santos, and Joao E Strapasson. 2015 · 2015
Cited alongside, same era.
Empirical evaluation of rectified activations in convolutional network
Bing Xu, Naiyan Wang, Tianqi Chen, and Mu Li. 2015 · 2015
Cited alongside, same era.
Adversarial removal of demographic attributes revisited
Maria Barrett, Yova Kementchedjhieva, Yanai Elazar, Desmond Elliott, and Anders Søgaard. 2019 · 2019
Later among the works it cites.
Law and word order: Nlp in legal tech
Robert Dale. 2019 · 2019
Later among the works it cites.
Interpolating between optimal transport and mmd using sinkhorn divergences
Jean Feydy, Thibault Séjourné, François-Xavier Vialard, Shun-ichi Amari, Alain Trouve, and Gabriel Peyré. 2019 · 2019
Later among the works it cites.
Sample complexity of sinkhorn divergences
Aude Genevay, Lénaic Chizat, Francis Bach, Marco Cuturi, and Gabriel Peyré. 2019 · 2019
Later among the works it cites.
Computational optimal transport
Gabriel Peyré and Marco Cuturi. 2019 · 2019
Later among the works it cites.
Practical and consistent estimation of f-divergences
Paul Rubenstein, Olivier Bousquet, Josip Djolonga, Carlos Riquelme, and Ilya O Tolstikhin. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Demographic dialectal variation in social media: A case study of African-American English
Su Lin Blodgett, Lisa Green, and Brendan O’Connor. 2016 · 2016
Cited alongside, same era.
The multiscale laplacian graph kernel
Risi Kondor and Horace Pan. 2016 · 2016
Cited alongside, same era.
Natural language processing for enhancing teaching and learning
Diane Litman. 2016 · 2016
Cited alongside, same era.
The problem with bias: Allocative versus representational harms in machine learning
Solon Barocas, Kate Crawford, Aaron Shapiro, and Hanna Wallach. 2017 · 2017
Cited alongside, same era.
Data decisions and theoretical implications when adversarially learning fair representations
Alex Beutel, Jilin Chen, Zhe Zhao, and Ed H Chi. 2017 · 2017
Cited alongside, same era.
Style transfer in text: Exploration and evaluation
Zhenxin Fu, Xiaoye Tan, Nanyun Peng, Dongyan Zhao, and Rui Yan. 2017 · 2017
Cited alongside, same era.
Ethical by design: Ethics best practices for natural language processing
Jochen L Leidner and Vassilis Plachouras. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Functional isolation forest
Guillaume Staerman, Pavlo Mozharovskyi, Stephan Clémençon, and Florence d’Alché Buc. 2019 · 2019
Later among the works it cites.
Hierarchical pre-training for sequence labelling in spoken dialog
Emile Chapuis, Pierre Colombo, Matteo Manica, Matthieu Labeau, and Chloé Clavel. 2020 · 2020
Later among the works it cites.
Guiding attention in sequence-to-sequence models for dialogue act prediction
Pierre Colombo, Emile Chapuis, Matteo Manica, Emmanuel Vignon, Giovanna Varni, and Chloe Clavel. 2020 · 2020
Later among the works it cites.
The importance of fillers for text representations of speech transcripts
Tanvi Dinkar, Pierre Colombo, Matthieu Labeau, and Chloé Clavel. 2020 · 2020
Later among the works it cites.
Formal limitations on the measurement of mutual information
David McAllester and Karl Stratos. 2020 · 2020
Later among the works it cites.
The fisher–rao distance between multivariate normal distributions: Special cases, bounds and applications
Julianna Pinele, João E. Strapasson, and Sueli I. R. Costa. 2020 · 2020
Later among the works it cites.
The area of the convex hull of sampled curves: a robust functional statistical depth measure
Guillaume Staerman, Pavlo Mozharovskyi, and Stéphan Clémençon. 2020 · 2020
Later among the works it cites.
Mutual-information based few-shot classification
Malik Boudiaf, Ziko Imtiaz Masud, Jérôme Rony, Jose Dolz, Ismail Ben Ayed, and Pablo Piantanida. 2021 · 2021
Later among the works it cites.
Code-switched inspired losses for spoken dialog representations
Pierre Colombo, Emile Chapuis, Matthieu Labeau, and Chloé Clavel. 2021a · 2021
Later among the works it cites.
Improving multimodal fusion via mutual dependency maximisation
Pierre Colombo, Emile Chapuis, Matthieu Labeau, and Chloé Clavel. 2021b · 2021
Later among the works it cites.
Beam search with bidirectional strategies for neural response generation
Pierre Colombo, Chloé Clavel, Chouchang Yack, and Giovanna Varni. 2021e · 2021
Later among the works it cites.
Automatic text evaluation through the lens of wasserstein barycenters
Pierre Colombo, Guillaume Staerman, Chloé Clavel, and Pablo Piantanida. 2021f · 2021
Later among the works it cites.
Nl-augmenter: A framework for task-sensitive natural language augmentation
Kaustubh D. Dhole, Varun Gangal, Sebastian Gehrmann, Aadesh Gupta, Zhenhao Li, Saad Mahamood, Abinaya Mahendiran, Simon Mille, Ashish Srivastava, Samson Tan, Tongshuang Wu, Jascha Sohl-Dickstein, Jinho D. Choi, Eduard H. Hovy, Ondrej Dusek, Sebastian Ruder, Sajant Anand, Nagender Aneja, Rabin Banjade, Lisa Barthe, Hanna Behnke, Ian Berlot-Attwell, Connor Boyle, Caroline Brun, Marco Antonio Sobrevilla Cabezudo, Samuel Cahyawijaya, Emile Chapuis, Wanxiang Che, Mukund Choudhary, Christian Clauss, Pierre Colombo, Filip Cornell, Gautier Dagan, Mayukh Das, Tanay Dixit, Thomas Dopierre, Paul-Alexis Dray, Suchitra Dubey, Tatiana Ekeinhor, Marco Di Giovanni, Rishabh Gupta, Rishabh Gupta, Louanes Hamla, Sang Han, Fabrice Harel-Canada, Antoine Honore, Ishan Jindal, Przemyslaw K. Joniak, Denis Kleyko, Venelin Kovatchev, and et al. 2021 · 2021
Later among the works it cites.
Few-shot learning of an interleaved text summarization model by pretraining with synthetic data
Sanjeev Kumar Karn, Francine Chen, Yan-Ying Chen, Ulli Waltinger, and Hinrich Schütze. 2021 · 2021
Later among the works it cites.
Towards understanding and mitigating social biases in language models
Paul Pu Liang, Chiyu Wu, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2021 · 2021
Later among the works it cites.
What are the best systems? new perspectives on NLP benchmarking
Pierre Colombo, Nathan Noiry, Ekhine Irurozki, and Stéphan Clémençon. 2022 · 2022
Closest in time.
Knife: Kernelized-neural differential entropy estimation
Georg Pichler, Pierre Colombo, Malik Boudiaf, Gunther Koliander, and Pablo Piantanida. 2022 · 2022
Closest in time.
Functional anomaly detection: a benchmark study
Guillaume Staerman, Eric Adjakossa, Pavlo Mozharovskyi, Vera Hofer, Jayant Sen Gupta, and Stephan Clémençon. 2022 · 2022
Closest in time.