Fetching the paper…
Reading the bibliography…
Recent research has focused on literary machine translation (MT) as a new challenge in MT.
Literary freedom: Project gutenberg
Bryan Stroube. 2003 · 2003
Earlier work this paper cites.
Kendall’s Tau , pages 713–715
Llukan Puka. 2011 · 2011
Earlier work this paper cites.
Interrater reliability: the kappa statistic
Mary L McHugh. 2012 · 2012
Earlier work this paper cites.
Towards a literary machine translation: The role of referential cohesion
Rob Voigt and Dan Jurafsky. 2012 · 2012
Earlier work this paper cites.
Continuous measurement scales in human evaluation of machine translation
Yvette Graham, Timothy Baldwin, Alistair Moffat, and Justin Zobel. 2013 · 2013
Earlier work this paper cites.
Using a new analytic measure for the annotation and analysis of mt errors on real data
Arle Lommel, Aljoscha Burchardt, Maja Popović, Kim Harris, Eleftherios Avramidis, and Hans Uszkoreit. 2014 · 2014
Earlier work this paper cites.
Normalization: A preprocessing stage
S Patro. 2015 · 2015
Earlier work this paper cites.
Best-worst scaling more reliable than rating scales: A case study on sentiment intensity annotation
Svetlana Kiritchenko and Saif Mohammad. 2017 · 2017
Earlier work this paper cites.
A call for clarity in reporting BLEU scores
Matt Post. 2018 · 2018
Earlier work this paper cites.
Attaining the unattainable? reassessing claims of human parity in neural machine translation
Antonio Toral, Sheila Castilho, Ke Hu, and Andy Way. 2018 · 2018
Earlier work this paper cites.
The challenges of using neural machine translation for literature
Evgeny Matusov. 2019 · 2019
Earlier work this paper cites.
Multilingual denoising pre-training for neural machine translation
Yinhan Liu, Jiatao Gu, Naman Goyal, Xian Li, Sergey Edunov, Marjan Ghazvininejad, Mike Lewis, and Luke Zettlemoyer. 2020 · 2020
Earlier work this paper cites.
BLEURT: Learning robust metrics for text generation
Thibault Sellam, Dipanjan Das, and Ankur Parikh. 2020 · 2020
Earlier work this paper cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Earlier work this paper cites.
Translation between creativity and reproducing an equivalent original text
Nabil Al-Awawdeh. 2021 · 2021
Earlier work this paper cites.
Beyond english-centric multilingual machine translation
Angela Fan, Shruti Bhosale, Holger Schwenk, Zhiyi Ma, Ahmed El-Kishky, Siddharth Goyal, Mandeep Baines, Onur Celebi, Guillaume Wenzek, Vishrav Chaudhary, Naman Goyal, Tom Birch, Vitaliy Liptchinsky, Sergey Edunov, Edouard Grave, Michael Auli, and Armand Joulin. 2021 · 2021
Earlier work this paper cites.
Experts, errors, and context: A large-scale study of human evaluation for machine translation
Markus Freitag, George Foster, David Grangier, Viresh Ratnakar, Qijun Tan, and Wolfgang Macherey. 2021 · 2021
Earlier work this paper cites.
No language left behind: Scaling human-centered machine translation
Marta R Costa-jussà, James Cross, Onur Çelebi, Maha Elbayad, Kenneth Heafield, Kevin Heffernan, Elahe Kalbassi, Janice Lam, Daniel Licht, Jean Maillard, et al. 2022 · 2022
Earlier work this paper cites.
Creativity in translation: Machine translation as a constraint for literary texts
Ana Guerberof-Arenas and Antonio Toral. 2022 · 2022
Earlier work this paper cites.
Human-adapted mt for literary texts: Reality or fantasy?
Damien Hansen and Emmanuelle Esperança-Rodier. 2022 · 2022
Earlier work this paper cites.
BlonDe: An automatic evaluation metric for document-level machine translation
Yuchen Jiang, Tianyu Liu, Shuming Ma, Dongdong Zhang, Jian Yang, Haoyang Huang, Rico Sennrich, Ryan Cotterell, Mrinmaya Sachan, and Ming Zhou. 2022 · 2022
Earlier work this paper cites.
Exploring document-level literary machine translation with parallel paragraphs from world literature
Katherine Thai, Marzena Karpinska, Kalpesh Krishna, Bill Ray, Moira Inghilleri, John Wieting, and Mohit Iyyer. 2022 · 2022
Cited alongside, same era.
Label Studio: Data labeling software
Maxim Tkachenko, Mikhail Malyuk, Andrey Holmanyuk, and Nikolai Liubimov. 2020-2022 · 2022
Cited alongside, same era.
GuoFeng: A benchmark for zero pronoun recovery and translation
Mingzhou Xu, Longyue Wang, Derek F. Wong, Hongye Liu, Linfeng Song, Lidia S. Chao, Shuming Shi, and Zhaopeng Tu. 2022 · 2022
Cited alongside, same era.
ByGPT5: End-to-end style-conditioned poetry generation with token-free language models
Jonas Belouadi and Steffen Eger. 2023 · 2023
Cited alongside, same era.
Missing information, unresponsive authors, experimental flaws: The impossibility of assessing the reproducibility of previous human evaluations in NLP
Anya Belz, Craig Thomson, Ehud Reiter, Gavin Abercrombie, Jose M. Alonso-Moral, Mohammad Arvan, Anouck Braggaar, Mark Cieliebak, Elizabeth Clark, Kees van Deemter, Tanvi Dinkar, Ondřej Dušek, Steffen Eger, Qixiang Fang, Mingqi Gao, Albert Gatt, Dimitra Gkatzia, Javier González-Corbelle, Dirk Hovy, Manuela Hürlimann, Takumi Ito, John D. Kelleher, Filip Klubicka, Emiel Krahmer, Huiyuan Lai, Chris van der Lee, Yiru Li, Saad Mahamood, Margot Mieskes, Emiel van Miltenburg, Pablo Mosteiro, Malvina Nissim, Natalie Parde, Ondřej Plátek, Verena Rieser, Jie Ruan, Joel Tetreault, Antonio Toral, Xiaojun Wan, Leo Wanner, Lewis Watson, and Diyi Yang. 2023 · 2023
Impact of translation workflows with and without MT on textual characteristics in literary translation
Joke Daems, Paola Ruffo, and Lieve Macken. 2024 · 2024
Closest in time.
xcomet: Transparent Machine Translation Evaluation through Fine-grained Error Detection
Nuno M. Guerreiro, Ricardo Rei, Daan van Stigt, Luisa Coheur, Pierre Colombo, and André F. T. Martins. 2024 · 2024
Closest in time.
Prometheus 2: An open source language model specialized in evaluating other language models
Seungone Kim, Juyoung Suk, Shayne Longpre, Bill Yuchen Lin, Jamin Shin, Sean Welleck, Graham Neubig, Moontae Lee, Kyungjae Lee, and Minjoon Seo. 2024 · 2024
Closest in time.
Error span annotation: A balanced approach for human evaluation of machine translation
Tom Kocmi, Vilém Zouhar, Eleftherios Avramidis, Roman Grundkiewicz, Marzena Karpinska, Maja Popović, Mrinmaya Sachan, and Mariya Shmatova. 2024 · 2024
Closest in time.
Ki – aber wie? Übersetzertag 2024
Kollektive-Intelligenz. 2024 · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Findings of the WMT 2023 shared task on quality estimation
Frederic Blain, Chrysoula Zerva, Ricardo Rei, Nuno M. Guerreiro, Diptesh Kanojia, José G. C. de Souza, Beatriz Silva, Tânia Vaz, Yan Jingxuan, Fatemeh Azadi, Constantin Orasan, and André Martins. 2023 · 2023
Cited alongside, same era.
FastKASSIM: A fast tree kernel-based syntactic similarity metric
Maximillian Chen, Caitlyn Chen, Xiao Yu, and Zhou Yu. 2023 · 2023
Cited alongside, same era.
Menli: Robust evaluation metrics from natural language inference
Yanran Chen and Steffen Eger. 2023 · 2023
Cited alongside, same era.
Training and meta-evaluating machine translation evaluation metrics at the paragraph level
Daniel Deutsch, Juraj Juraska, Mara Finkelstein, and Markus Freitag. 2023 · 2023
Cited alongside, same era.
Large language models effectively leverage document-level context for literary translation, but critical errors persist
Marzena Karpinska and Mohit Iyyer. 2023 · 2023
Cited alongside, same era.
GEMBA-MQM: Detecting translation quality error spans with GPT-4
Tom Kocmi and Christian Federmann. 2023 · 2023
Cited alongside, same era.
‘i am a bit surprised’: Literary translation and post-editing processes compared
Waltraud Kolb. 2023 · 2023
Cited alongside, same era.
Christoph Leiter and Steffen Eger. 2024 · 2024
Closest in time.
LLMs as narcissistic evaluators: When ego inflates evaluation scores
Yiqi Liu, Nafise Moosavi, and Chenghua Lin. 2024 · 2024
Closest in time.
Machine translation meets large language models: Evaluating ChatGPT’s ability to automatically post-edit literary texts
Lieve Macken. 2024 · 2024
Closest in time.
State of what art? a call for multi-prompt llm evaluation
Moran Mizrahi, Guy Kaplan, Dan Malkin, Rotem Dror, Dafna Shahaf, and Gabriel Stanovsky. 2024 · 2024
Closest in time.
Künstliche intelligenz – gefahr oder hilfe für literarische Übersetzung?
Magdalena Nizioł. 2024 · 2024
Closest in time.
Salute the classic: Revisiting challenges of machine translation in the age of large language models
Jianhui Pang, Fanghua Ye, Longyue Wang, Dian Yu, Derek F Wong, Shuming Shi, and Zhaopeng Tu. 2024 · 2024
Closest in time.
The translator’s canvas: Using LLMs to enhance poetry translation
Natália Resende and James Hadley. 2024 · 2024
Closest in time.
Whence the 3 percent?: How far have we come toward decentering america’s literary preference?
Markella B Rutherford, Peggy Levitt, and Erika Zhang. 2024 · 2024
Closest in time.
Gemma: Open models based on gemini research and technology
Gemma Team, Thomas Mesnard, Cassidy Hardin, Robert Dadashi, Surya Bhupatiraju, Shreya Pathak, Laurent Sifre, Morgane Rivière, Mihir Sanjay Kale, Juliette Love, et al. 2024 · 2024
Closest in time.
Common flaws in running human evaluation experiments in NLP
Craig Thomson, Ehud Reiter, and Anya Belz. 2024 · 2024
Closest in time.
Proceedings of the 1st Workshop on Creative-text Translation and Technology . European Association for Machine Translation, Sheffield, United Kingdom
Bram Vanroy, Marie-Aude Lefer, Lieve Macken, and Paola Ruffo, editors. 2024 · 2024
Closest in time.
Minghao Wu, Jiahao Xu, Yulin Yuan, Gholamreza Haffari, and Longyue Wang. 2024 · 2024
Closest in time.
Jianhao Yan, Pingchuan Yan, Yulong Chen, Judy Li, Xianchao Zhu, and Yue Zhang. 2024 · 2024
Closest in time.
An Yang, Baosong Yang, Binyuan Hui, Bo Zheng, Bowen Yu, Chang Zhou, Chengpeng Li, Chengyuan Li, Dayiheng Liu, Fei Huang, et al. 2024 · 2024
Closest in time.
Llm-based multi-agent poetry generation in non-cooperative environments
Ran Zhang and Steffen Eger. 2024 · 2024
Closest in time.
MQM-APE: Toward high-quality error annotation predictors with automatic post-editing in LLM translation evaluators
Qingyu Lu, Liang Ding, Kanjian Zhang, Jinxia Zhang, and Dacheng Tao. 2025 · 2025
Closest in time.