Fetching the paper…
Reading the bibliography…
We analyze the behaviors of open large language models (LLMs) on the task of data-to-text (D2T) generation, i.e., generating coherent and relevant text from structured data.
WikiDataSets : Standardized sub-graphs from WikiData
Armand Boschin. 2019 · 1906
Earlier work this paper cites.
Building applied natural language generation systems
Ehud Reiter and Robert Dale. 1997 · 1997
Earlier work this paper cites.
Sumtime-meteo: Parallel corpus of naturally occurring forecast texts and weather data
Somayajulu Sripada, Ehud Reiter, Jim Hunter, and Jin Yu. 2002 · 2002
Earlier work this paper cites.
Corpus-driven generation of weather forecasts
Anja Belz. 2005 · 2005
Earlier work this paper cites.
Evaluation of text generation: A survey
Asli Celikyilmaz, Elizabeth Clark, and Jianfeng Gao. 2020 · 2006
Earlier work this paper cites.
Investigating Pretrained Language Models for Graph-to-Text Generation
Leonardo F. R. Ribeiro, Martin Schmitt, Hinrich Schütze, and Iryna Gurevych. 2020 · 2007
Earlier work this paper cites.
Automatic generation of weather forecast texts using comprehensive probabilistic generation-space models
Anja Belz. 2008 · 2008
Earlier work this paper cites.
Generating Textual Summaries of Bar Charts
Seniz Demir, Sandra Carberry, and Kathleen F. McCoy. 2008 · 2008
Earlier work this paper cites.
Learning Semantic Correspondences with Less Supervision
Percy Liang, Michael I. Jordan, and Dan Klein. 2009 · 2009
Earlier work this paper cites.
A Simple Domain-Independent Probabilistic Approach to Generation
Gabor Angeli, Percy Liang, and Dan Klein. 2010 · 2010
Earlier work this paper cites.
Summarizing Information Graphics Textually
Seniz Demir, Sandra Carberry, and Kathleen F. McCoy. 2012 · 2012
Earlier work this paper cites.
Toward multi-domain language generation using recurrent neural networks
Tsung-Hsien Wen, Milica Gašic, Nikola Mrkšic, Lina M Rojas-Barahona, Pei-Hao Su, David Vandyke, and Steve Young. 2015 · 2015
Earlier work this paper cites.
Neural Text Generation from Structured Data with Application to the Biography Domain
Rémi Lebret, David Grangier, and Michael Auli. 2016 · 2016
Earlier work this paper cites.
Multi-domain Neural Network Language Generation for Spoken Dialogue Systems
Tsung-Hsien Wen, Milica Gasic, Nikola Mrksic, Lina Maria Rojas-Barahona, Pei-Hao Su, David Vandyke, and Steve J. Young. 2016 · 2016
Earlier work this paper cites.
The WebNLG Challenge: Generating Text from RDF Data
Claire Gardent, Anastasia Shimorina, Shashi Narayan, and Laura Perez-Beltrachini. 2017 · 2017
Earlier work this paper cites.
Why We Need New Evaluation Metrics for NLG
Jekaterina Novikova, Ondřej Dušek, Amanda Cercas Curry, and Verena Rieser. 2017 · 2017
Earlier work this paper cites.
A Statistical Framework for Product Description Generation
Jinpeng Wang, Yutai Hou, Jing Liu, Yunbo Cao, and Chin-Yew Lin. 2017 · 2017
Earlier work this paper cites.
Challenges in Data-to-Document Generation
Sam Wiseman, Stuart M. Shieber, and Alexander M. Rush. 2017 · 2017
Earlier work this paper cites.
Survey of the State of the Art in Natural Language Generation: Core tasks, applications and evaluation
Albert Gatt and Emiel Krahmer. 2018 · 2018
Earlier work this paper cites.
Operation-guided Neural Networks for High Fidelity Data-To-Text Generation
Feng Nie, Jinpeng Wang, Jin-Ge Yao, Rong Pan, and Chin-Yew Lin. 2018 · 2018
Earlier work this paper cites.
Constrained Decoding for Neural NLG from Compositional Representations in Task-Oriented Dialogue
Anusha Balakrishnan, Jinfeng Rao, Kartikeya Upasani, Michael White, and Rajen Subba. 2019 · 2019
Earlier work this paper cites.
Semantic noise matters for neural natural language generation
Ondrej Dušek, David M. Howcroft, and Verena Rieser. 2019 · 2019
Earlier work this paper cites.
Text Generation from Knowledge Graphs with Graph Transformers
Rik Koncel-Kedziorski, Dhanush Bekal, Yi Luan, Mirella Lapata, and Hannaneh Hajishirzi. 2019 · 2019
Earlier work this paper cites.
Data-to-Text Generation with Content Selection and Planning
Ratish Puduppully, Li Dong, and Mirella Lapata. 2019a · 2019
Earlier work this paper cites.
Best practices for the human evaluation of automatically generated text
Chris van der Lee, Albert Gatt, Emiel van Miltenburg, Sander Wubben, and Emiel Krahmer. 2019 · 2019
Earlier work this paper cites.
Revisiting Challenges in Data-to-Text Generation with Fact Grounding
Hongmin Wang. 2019 · 2019
Earlier work this paper cites.
The 2020 bilingual, bi-directional WebNLG+ shared task: Overview and evaluation results (WebNLG+ 2020)
Thiago Castro Ferreira, Claire Gardent, Nikolai Ilinykh, Chris van der Lee, Simon Mille, Diego Moussallem, and Anastasia Shimorina. 2020 · 2020
Earlier work this paper cites.
KGPT: Knowledge-Grounded Pre-Training for Data-to-Text Generation
Wenhu Chen, Yu Su, Xifeng Yan, and William Yang Wang. 2020 · 2020
Earlier work this paper cites.
Evaluating the state-of-the-art of end-to-end natural language generation: The E2E NLG challenge
Ondrej Dušek, Jekaterina Novikova, and Verena Rieser. 2020 · 2020
Cited alongside, same era.
Chart-to-Text: Generating Natural Language Descriptions for Charts by Adapting the Transformer Model
Jason Obeid and Enamul Hoque. 2020 · 2020
Cited alongside, same era.
A Hierarchical Model for Data-to-Text Generation
Clément Rebuffel, Laure Soulier, Geoffrey Scoutheeten, and Patrick Gallinari. 2020 · 2020
Cited alongside, same era.
A Gold Standard Methodology for Evaluating Accuracy in Data-To-Text Systems
Craig Thomson and Ehud Reiter. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-Art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Cited alongside, same era.
GPTScore: Evaluate as You Desire
Jinlan Fu, See-Kiong Ng, Zhengbao Jiang, and Pengfei Liu. 2023 · 2023
Later among the works it cites.
Repairing the Cracked Foundation: A Survey of Obstacles in Evaluation Practices for Generated Text
Sebastian Gehrmann, Elizabeth Clark, and Thibault Sellam. 2023 · 2023
Later among the works it cites.
Time Travel in LLMs: Tracing Data Contamination in Large Language Models
Shahriar Golchin and Mihai Surdeanu. 2023 · 2023
Later among the works it cites.
Ari Holtzman, Peter West, and Luke Zettlemoyer. 2023 · 2023
Later among the works it cites.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Yejin Bang, Andrea Madotto, and Pascale Fung. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Knowledge Graph Based Synthetic Corpus Generation for Knowledge-Enhanced Language Model Pre-training
Oshin Agarwal, Heming Ge, Siamak Shakeri, and Rami Al-Rfou. 2021 · 2021
Cited alongside, same era.
The GEM benchmark: Natural language generation, its evaluation and metrics
Sebastian Gehrmann, Tosin P. Adewumi, Karmanya Aggarwal, Pawan Sasanka Ammanamanchi, Aremu Anuoluwapo, Antoine Bosselut, Khyathi Raghavi Chandu, Miruna-Adriana Clinciu, Dipanjan Das, Kaustubh D. Dhole, Wanyu Du, Esin Durmus, Ondrej Dušek, Chris Emezue, Varun Gangal, Cristina Garbacea, Tatsunori Hashimoto, Yufang Hou, Yacine Jernite, Harsh Jhamtani, Yangfeng Ji, Shailza Jolly, Dhruv Kumar, Faisal Ladhak, Aman Madaan, Mounica Maddela, Khyati Mahajan, Saad Mahamood, Bodhisattwa Prasad Majumder, Pedro Henrique Martins, Angelina McMillan-Major, Simon Mille, Emiel van Miltenburg, Moin Nadeem, Shashi Narayan, Vitaly Nikolaev, Rubungo Andre Niyongabo, Salomey Osei, Ankur P. Parikh, Laura Perez-Beltrachini, Niranjan Ramesh Rao, Vikas Raunak, Juan Diego Rodriguez, Sashank Santhanam, João Sedoc, Thibault Sellam, Samira Shaikh, Anastasia Shimorina, Marco Antonio Sobrevilla Cabezudo, Hendrik Strobelt, Nishant Subramani, Wei Xu, Diyi Yang, Akhila Yerukola, and Jiawei Zhou. 2021 · 2021
Cited alongside, same era.
Data-to-text Generation with Macro Planning
Ratish Puduppully and Mirella Lapata. 2021 · 2021
Cited alongside, same era.
Controllable and Diverse Text Generation in E-commerce
Huajie Shao, Jun Wang, Haohong Lin, Xuezhou Zhang, Aston Zhang, Heng Ji, and Tarek F. Abdelzaher. 2021 · 2021
Cited alongside, same era.
TCube: Domain-Agnostic Neural Time-series Narration
Mandar Sharma, John S. Brownstein, and Naren Ramakrishnan. 2021 · 2021
Cited alongside, same era.
SportSett:Basketball - A Robust and Maintainable Dataset for Natural Language Generation
Craig Thomson, Ehud Reiter, and Somayajulu Sripada. 2021 · 2021
Cited alongside, same era.
Human evaluation of automatically generated text: Current trends and best practice guidelines
Chris van der Lee, Albert Gatt, Emiel van Miltenburg, and Emiel Krahmer. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de Las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, Lélio Renard Lavaud, Marie-Anne Lachaux, Pierre Stock, Teven Le Scao, Thibaut Lavril, Thomas Wang, Timothée Lacroix, and William El Sayed. 2023 · 2023
Later among the works it cites.
Mind the Labels: Describing Relations in Knowledge Graphs With Pretrained Models
Zdeněk Kasner, Ioannis Konstas, and Ondřej Dušek. 2023 · 2023
Later among the works it cites.
GEMBA-MQM: Detecting Translation Quality Error Spans with GPT-4
Tom Kocmi and Christian Federmann. 2023a · 2023
Later among the works it cites.
Large Language Models Are State-of-the-Art Evaluators of Translation Quality
Tom Kocmi and Christian Federmann. 2023b · 2023
Later among the works it cites.
Benchmarking cognitive biases in large language models as evaluators
Ryan Koo, Minhwa Lee, Vipul Raheja, Jong Inn Park, Zae Myung Kim, and Dongyeop Kang. 2023 · 2023
Later among the works it cites.
G-Eval: NLG Evaluation using Gpt-4 with Better Human Alignment
Yang Liu, Dan Iter, Yichong Xu, Shuohang Wang, Ruochen Xu, and Chenguang Zhu. 2023 · 2023
Later among the works it cites.
Data-to-text Generation for Severely Under-Resourced Languages with GPT-3.5: A Bit of Help Needed from Google Translate (WebNLG 2023)
Michela Lorandi and Anja Belz. 2023 · 2023
Later among the works it cites.
Evaluation Metrics in the Era of GPT-4: Reliably Evaluating Large Language Models on Sequence to Sequence Tasks
Andrea Sottana, Bin Liang, Kai Zou, and Zheng Yuan. 2023 · 2023
Later among the works it cites.
Evaluating factual accuracy in complex data-to-text
Craig Thomson, Ehud Reiter, and Barkavi Sundararajan. 2023 · 2023
Later among the works it cites.
Zephyr: Direct Distillation of LM Alignment
Lewis Tunstall, Edward Beeching, Nathan Lambert, Nazneen Rajani, Kashif Rasul, Younes Belkada, Shengyi Huang, Leandro von Werra, Clémentine Fourrier, Nathan Habib, Nathan Sarrazin, Omar Sanseviero, Alexander M. Rush, and Thomas Wolf. 2023 · 2023
Later among the works it cites.
Barriers and enabling factors for error analysis in NLG research
Emiel Van Miltenburg, Miruna Clinciu, Ondřej Dušek, Dimitra Gkatzia, Stephanie Inglis, Leo Leppänen, Saad Mahamood, Stephanie Schoch, Craig Thomson, and Luou Wen. 2023 · 2023
Later among the works it cites.
INSTRUCTSCORE: Explainable Text Generation Evaluation with Finegrained Feedback
Wenda Xu, Danqing Wang, Liangming Pan, Zhenqiao Song, Markus Freitag, William Yang Wang, and Lei Li. 2023 · 2023
Later among the works it cites.
Evaluating Generative Models for Graph-to-Text Generation
Shuzhou Yuan and Michael Färber. 2023 · 2023
Later among the works it cites.
Yilun Zhao, Haowei Zhang, Shengyun Si, Linyong Nan, Xiangru Tang, and Arman Cohan. 2023 · 2023
Later among the works it cites.
Judging LLM-as-a-judge with MT-Bench and Chatbot Arena
Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric P. Xing, Hao Zhang, Joseph E. Gonzalez, and Ion Stoica. 2023 · 2023
Later among the works it cites.
Leak, cheat, repeat: Data contamination and evaluation malpractices in closed-source LLMs
Simone Balloccu, Patrícia Schmidtová, Mateusz Lango, and Ondrej Dušek. 2024 · 2024
Closest in time.
Leave no context behind: Efficient infinite context transformers with infini-attention
Tsendsuren Munkhdalai, Manaal Faruqui, and Siddharth Gopal. 2024 · 2024
Closest in time.
Introducing ChatGPT
OpenAI. 2023b · 2024
Closest in time.
We should evaluate real-world impact!
Ehud Reiter. 2023 · 2024
Closest in time.
Closed AI Models Make Bad Baselines
Anna Rogers. 2023 · 2024
Closest in time.
Large Language Models are Inconsistent and Biased Evaluators
Rickard Stureborg, Dimitris Alikaniotis, and Yoshi Suhara. 2024 · 2024
Closest in time.
Preparing for the era of 32K context: Early learnings and explorations
TogetherAI. 2023 · 2024
Closest in time.
Data-to-text Generation with Entity Modeling
Ratish Puduppully, Li Dong, and Mirella Lapata. 2019b · 2035
Closest in time.