Fetching the paper…
Reading the bibliography…
Recent advancements in high-quality, large-scale English resources have pushed the frontier of English Automatic Text Simplification (ATS) research.
Subjective assessment of text complexity: A dataset for german language
Babak Naderi, Salar Mohtaj, Kaspar Ensikat, and Sebastian Möller. 2019 · 1904
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions, and reversals
Vladimir I. Levenshtein. 1965 · 1965
Earlier work this paper cites.
Compilation of a multilingual parallel corpus
Yasuhito Tanaka. 2001 · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Sentence alignment for monolingual comparable corpora
Regina Barzilay and Noemie Elhadad. 2003 · 2003
Earlier work this paper cites.
Syntactic simplification and text cohesion
Advaith Siddharthan. 2004 · 2004
Earlier work this paper cites.
Xcopa: A multilingual dataset for causal commonsense reasoning
Edoardo Maria Ponti, Goran Glavaš, Olga Majewska, Qianchu Liu, Ivan Vulić, and Anna Korhonen. 2020 · 2005
Earlier work this paper cites.
Towards robust context-sensitive sentence alignment for monolingual corpora
Rani Nelken and Stuart M. Shieber. 2006 · 2006
Earlier work this paper cites.
Text simplification for language learners: a corpus analysis
Sarah E Petersen and Mari Ostendorf. 2007 · 2007
Earlier work this paper cites.
A proper approach to Japanese morphological analysis: Dictionary, model, and evaluation
Yasuharu Den, Junpei Nakamura, Toshinobu Ogiso, and Hideki Ogura. 2008 · 2008
Earlier work this paper cites.
Mtop: A comprehensive multilingual task-oriented semantic parsing benchmark
Haoran Li, Abhinav Arora, Shuohui Chen, Anchit Gupta, Sonal Gupta, and Yashar Mehdad. 2020 · 2008
Earlier work this paper cites.
Manual de simplificação sintática para o português
Lucia Specia, Sandra Maria Aluísio, and Thiago A Salgueiro Pardo. 2008 · 2008
Earlier work this paper cites.
Natural language processing with Python: analyzing text with the natural language toolkit
Steven Bird, Ewan Klein, and Edward Loper. 2009 · 2009
Earlier work this paper cites.
Building a brazilian portuguese parallel corpus of original and simplified texts
Helena M Caseli, Tiago F Pereira, Lucia Specia, Thiago AS Pardo, Caroline Gasperin, and Sandra Maria Aluísio. 2009 · 2009
Earlier work this paper cites.
Fostering digital inclusion and accessibility: The PorSimples project for simplification of Portuguese texts
Sandra Aluísio and Caroline Gasperin. 2010 · 2010
Earlier work this paper cites.
Xl-wic: A multilingual benchmark for evaluating semantic contextualization
Alessandro Raganato, Tommaso Pasini, Jose Camacho-Collados, and Mohammad Taher Pilehvar. 2020 · 2010
Earlier work this paper cites.
MT-based sentence alignment for OCR-generated parallel texts
Rico Sennrich and Martin Volk. 2010 · 2010
Earlier work this paper cites.
Translating from complex to simplified sentences
Lucia Specia. 2010 · 2010
Earlier work this paper cites.
Pautas básicas de simplificación textual y diseño del corpus simplext
Alberto Anula. 2011 · 2011
Earlier work this paper cites.
Computing krippendorff’s alpha-reliability
Klaus Krippendorff. 2011 · 2011
Earlier work this paper cites.
DSim, a Danish parallel corpus for text simplification
Sigrid Klerke and Anders Søgaard. 2012 · 2012
Earlier work this paper cites.
Multifarm: A benchmark for multilingual ontology matching
Christian Meilicke, Raúl García-Castro, Fred Freitas, Willem Robert van Hage, Elena Montiel-Ponsoda, Ryan Ribeiro de Azevedo, Heiner Stuckenschmidt, Ondřej Šváb Zamazal, Vojtěch Svátek, Andrei Tamilin, Cássia Trojahn, and Shenghui Wang. 2012 · 2012
Earlier work this paper cites.
Building a German/simple German parallel corpus for automatic text simplification
David Klaper, Sarah Ebling, and Martin Volk. 2013 · 2013
Earlier work this paper cites.
Simple, readable sub-sentences
Sigrid Klerke and Anders Søgaard. 2013 · 2013
Earlier work this paper cites.
Text simplification for people with autistic spectrum disorders
C Orasan, R Evans, and I Dornescu. 2013 · 2013
Earlier work this paper cites.
A Neurophysiologically-Inspired Statistical Language Model
Jon Dehdari. 2014 · 2014
Earlier work this paper cites.
Euronews: a multilingual speech corpus for ASR
Roberto Gretter. 2014 · 2014
Earlier work this paper cites.
The fewer, the better? a contrastive study about ways to simplify
Ruslan Mitkov and Sanja Štajner. 2014 · 2014
Earlier work this paper cites.
Translating sentences from ’original’ to ’simplified’ spanish
Sanja Stajner. 2014 · 2014
Earlier work this paper cites.
Translating sentences from ’original’ to ’simplified’ spanish
Sanja Štajner. 2014 · 2014
Earlier work this paper cites.
Design and annotation of the first Italian corpus for text simplification
Dominique Brunato, Felice Dell’Orletta, Giulia Venturi, and Simonetta Montemagni. 2015 · 2015
Earlier work this paper cites.
MultiLing 2015: Multilingual summarization of single and multi-documents, on-line fora, and call-center conversations
George Giannakopoulos, Jeff Kubina, John Conroy, Josef Steinberger, Benoit Favre, Mijail Kabadjov, Udo Kruschwitz, and Massimo Poesio. 2015 · 2015
Earlier work this paper cites.
Japanese news simplification: tak design, data set construction, and analysis of simplified text
Isao Goto, Hideki Tanaka, and Tadashi Kumano. 2015 · 2015
Earlier work this paper cites.
Making it simplext: Implementation and evaluation of a text simplification system for spanish
Horacio Saggion, Sanja Štajner, Stefan Bott, Simon Mille, Luz Rello, and Biljana Drndarevic. 2015 · 2015
Earlier work this paper cites.
Automatic text simplification for Spanish: Comparative evaluation of various simplification strategies
Sanja Štajner, Iacer Calixto, and Horacio Saggion. 2015 · 2015
Earlier work this paper cites.
Problems in current text simplification research: New data can help
Wei Xu, Chris Callison-Burch, and Courtney Napoles. 2015 · 2015
Earlier work this paper cites.
Saud al-Sanousi’s Saaq al-Bambuu: The Authorized Abridged Edition for Students of Arabic
Saud Al-Sanousi. 2016 · 2016
Cited alongside, same era.
PaCCSS-IT: A parallel corpus of complex-simple sentences for automatic text simplification
Dominique Brunato, Andrea Cimino, Felice Dell’Orletta, and Giulia Venturi. 2016 · 2016
Cited alongside, same era.
Simpitiki: a simplification corpus for italian
Sara Tonelli, Alessio Palmero Aprosio, and Francesca Saltori. 2016 · 2016
Cited alongside, same era.
Optimizing statistical machine translation for text simplification
Wei Xu, Courtney Napoles, Ellie Pavlick, Quanze Chen, and Chris Callison-Burch. 2016 · 2016
Cited alongside, same era.
Fast Krippendorff: Fast computation of Krippendorff’s alpha agreement measure
Santiago Castro. 2017 · 2017
Cited alongside, same era.
OpenNMT: Open-source toolkit for neural machine translation
Controllable sentence simplification
Louis Martin, Éric de la Clergerie, Benoît Sagot, and Antoine Bordes. 2020 · 2020
Later among the works it cites.
fugashi, a tool for tokenizing Japanese in python
Paul McCann. 2020 · 2020
Later among the works it cites.
SimplifyUR: Unsupervised lexical text simplification for Urdu
Namoos Hayat Qasmi, Haris Bin Zia, Awais Athar, and Agha Ali Raza. 2020 · 2020
Later among the works it cites.
Benchmarking data-driven automatic text simplification for German
Andreas Säuberli, Sarah Ebling, and Martin Volk. 2020 · 2020
Later among the works it cites.
MLSUM: The multilingual summarization corpus
Thomas Scialom, Paul-Alexis Dray, Sylvain Lamprier, Benjamin Piwowarski, and Jacopo Staiano. 2020 · 2020
Later among the works it cites.
Multi-SimLex: A Large-Scale Evaluation of Multilingual and Crosslingual Lexical Semantic Similarity
Ivan Vulić, Simon Baker, Edoardo Maria Ponti, Ulla Petti, Ira Leviant, Kelly Wing, Olga Majewska, Eden Bar, Matt Malone, Thierry Poibeau, Roi Reichart, and Anna Korhonen. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Guillaume Klein, Yoon Kim, Yuntian Deng, Jean Senellart, and Alexander Rush. 2017 · 2017
Cited alongside, same era.
Sentence simplification with core vocabulary
Takumi Maruyama and Kazuhide Yamamoto. 2017 · 2017
Cited alongside, same era.
MASSAlign: Alignment and annotation of comparable documents
Gustavo Paetzold, Fernando Alva-Manchego, and Lucia Specia. 2017 · 2017
Cited alongside, same era.
Learning joint multilingual sentence representations with neural machine translation
Holger Schwenk and Matthijs Douze. 2017 · 2017
Cited alongside, same era.
Sentence alignment methods for improving text simplification systems
Sanja Štajner, Marc Franco-Salvador, Simone Paolo Ponzetto, Paolo Rosso, and Heiner Stuckenschmidt. 2017 · 2017
Cited alongside, same era.
Sentence simplification with deep reinforcement learning
Xingxing Zhang and Mirella Lapata. 2017 · 2017
Cited alongside, same era.
XNLI: Evaluating cross-lingual sentence representations
Alexis Conneau, Ruty Rinott, Guillaume Lample, Adina Williams, Samuel Bowman, Holger Schwenk, and Veselin Stoyanov. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
CCNet: Extracting high quality monolingual datasets from web crawl data
Guillaume Wenzek, Marie-Anne Lachaux, Alexis Conneau, Vishrav Chaudhary, Francisco Guzmán, Armand Joulin, and Edouard Grave. 2020 · 2020
Later among the works it cites.
Automated text simplification: A survey
Suha S. Al-Thanyyan and Aqil M. Azmi. 2021 · 2021
Later among the works it cites.
Olá, bonjour, salve! XFORMAL: A benchmark for multilingual formality style transfer
Eleftheria Briakou, Di Lu, Ke Zhang, and Joel Tetreault. 2021 · 2021
Later among the works it cites.
A quantitative study of simplification strategies in adapted texts for l2 learners of russian
Anna Dmitrieva, Antonina Laposhina, and Maria Lebedeva. 2021 · 2021
Later among the works it cites.
Creating an aligned Russian text simplification dataset from language learner data
Anna Dmitrieva and Jörg Tiedemann. 2021 · 2021
Later among the works it cites.
Text simplification with autoregressive models
Alena Fenogenova. 2021 · 2021
Later among the works it cites.
X-fact: A new benchmark dataset for multilingual fact checking
Ashim Gupta and Vivek Srikumar. 2021 · 2021
Later among the works it cites.
MKQA: A Linguistically Diverse Benchmark for Multilingual Open Domain Question Answering
Shayne Longpre, Yi Lu, and Joachim Daiber. 2021 · 2021
Later among the works it cites.
Controllable text simplification with explicit paraphrasing
Mounica Maddela, Fernando Alva-Manchego, and Wei Xu. 2021 · 2021
Later among the works it cites.
Text Simplification by Tagging
Kostiantyn Omelianchuk, Vipul Raheja, and Oleksandr Skurzhanskyi. 2021 · 2021
Later among the works it cites.
RuSimpleSentEval-2021 shared task: evaluating sentence simplification for russian
Andrey Sakhovskiy, Alexandra Izhevskaya, Alena Pestova, Elena Tutubalina, Valentin Malykh, Ivana Smurov, and Ekaterina Artemova. 2021 · 2021
Later among the works it cites.
Sentence simplification with rugpt3
AA Shatilov and AI Rey. 2021 · 2021
Later among the works it cites.
Automatic text simplification for social good: Progress and challenges
Sanja Stajner. 2021 · 2021
Later among the works it cites.
Revisiting the primacy of english in zero-shot cross-lingual transfer
Iulia Turc, Kenton Lee, Jacob Eisenstein, Ming-Wei Chang, and Kristina Toutanova. 2021 · 2021
Later among the works it cites.
Investigating text simplification evaluation
Laura Vásquez-Rodríguez, Matthew Shardlow, Piotr Przybyła, and Sophia Ananiadou. 2021 · 2021
Later among the works it cites.
mT5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
Later among the works it cites.
Klexikon: A German dataset for joint summarization and simplification
Dennis Aumiller and Michael Gertz. 2022 · 2022
Later among the works it cites.
FairLex: A multilingual benchmark for evaluating fairness in legal text processing
Ilias Chalkidis, Tommaso Pasini, Sheng Zhang, Letizia Tomada, Sebastian Schwemer, and Anders Søgaard. 2022 · 2022
Later among the works it cites.
Slovene text simplification dataset SloTS
Sabina Gorenc and Marko Robnik-Šikonja. 2022 · 2022
Later among the works it cites.
Bitext mining using distilled sentence representations for low-resource languages
Kevin Heffernan, Onur Çelebi, and Holger Schwenk. 2022 · 2022
Later among the works it cites.
Towards arabic sentence simplification via classification and generative approaches
Nouran Khallaf and Serge Sharoff. 2022 · 2022
Later among the works it cites.
MUSS: Multilingual unsupervised sentence simplification by mining paraphrases
Louis Martin, Angela Fan, Éric de la Clergerie, Antoine Bordes, and Benoît Sagot. 2022 · 2022
Later among the works it cites.
Neural readability pairwise ranking for sentences in Italian administrative language
Martina Miliani, Serena Auriemma, Fernando Alva-Manchego, and Alessandro Lenci. 2022 · 2022
Later among the works it cites.
Bloom: A 176b-parameter open-access multilingual language model
Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, Matthias Gallé, et al. 2022 · 2022
Later among the works it cites.
Sequence to sequence pretraining for a less-resourced slovenian language
Matej Ulčar and Marko Robnik-Šikonja. 2022 · 2022
Later among the works it cites.
Stanceosaurus: Classifying stance towards multilingual misinformation
Jonathan Zheng, Ashutosh Baheti, Tarek Naous, Wei Xu, and Alan Ritter. 2022 · 2022
Later among the works it cites.
Towards massively multi-domain multilingual readability assessment
Tarek Naous, Michael J. Ryan, Mohit Chandra, and Wei Xu. 2023 · 2023
Closest in time.
Xtreme-up: A user-centric scarce-data benchmark for under-represented languages
Sebastian Ruder, Jonathan H Clark, Alexander Gutkin, Mihir Kale, Min Ma, Massimo Nicosia, Shruti Rijhwani, Parker Riley, Jean-Michel A Sarr, Xinyi Wang, et al. 2023 · 2023
Closest in time.
Patient-friendly clinical notes: Towards a new text simplification dataset
Jan Trienes, Jörg Schlötterer, Hans-Ulrich Schildhaus, and Christin Seifert. 2023 · 2023
Closest in time.