Fetching the paper…
Reading the bibliography…
Opinion summarization sets itself apart from other types of summarization tasks due to its distinctive focus on aspects and sentiments.
Unsupervised opinion summarization with noising and denoising
Reinald Kim Amplayo and Mirella Lapata. 2020 · 1945
Earlier work this paper cites.
A coefficient of agreement for nominal scales
Jacob Cohen. 1960 · 1960
Earlier work this paper cites.
An analysis of variance test for normality (complete samples)
S. S. Shapiro and M. B. Wilk. 1965 · 1965
Earlier work this paper cites.
Measuring nominal scale agreement among many raters
Joseph L Fleiss. 1971 · 1971
Earlier work this paper cites.
A solution to plato’s problem: The latent semantic analysis theory of acquisition, induction, and representation of knowledge
Thomas K Landauer and Susan T. Dumais. 1997 · 1997
Earlier work this paper cites.
Bleu: A method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Lexrank: Graph-based lexical centrality as salience in text summarization
Günes Erkan and Dragomir R. Radev. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Evaluating content selection in summarization: The pyramid method
Ani Nenkova and Rebecca Passonneau. 2004 · 2004
Earlier work this paper cites.
Centroid-based summarization of multiple documents
Dragomir R. Radev, Hongyan Jing, Małgorzata Styś, and Daniel Tam. 2004 · 2004
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
A study of translation edit rate with targeted human annotation
Matthew Snover, Bonnie Dorr, Rich Schwartz, Linnea Micciulla, and John Makhoul. 2006 · 2006
Earlier work this paper cites.
Computing inter-rater reliability and its variance in the presence of high agreement
Kilem Li Gwet. 2008 · 2008
Earlier work this paper cites.
Agile corpus annotation in practice: An overview of manual and automatic annotation of CVs
Bea Alex, Claire Grover, Rongzhou Shen, and Mijail Kabadjov. 2010 · 2010
Earlier work this paper cites.
Opinosis: A graph based approach to abstractive summarization of highly redundant opinions
Kavita Ganesan, ChengXiang Zhai, and Jiawei Han. 2010 · 2010
Earlier work this paper cites.
Computing krippendorff’s alpha-reliability
Klaus Krippendorff. 2011 · 2011
Earlier work this paper cites.
A comparison of greedy and optimal assessment of natural language student input using word-to-word similarity metrics
Vasile Rus and Mihai Lintean. 2012 · 2012
Earlier work this paper cites.
Bootstrapping dialog systems with word embeddings
Gabriel Forgues, Joelle Pineau, Jean-Marie Larchevêque, and Réal Tremblay. 2014 · 2014
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
From word embeddings to document distances
Matt Kusner, Yu Sun, Nicholas Kolkin, and Kilian Weinberger. 2015 · 2015
Earlier work this paper cites.
chrF: character n-gram F-score for automatic MT evaluation
Maja Popović. 2015 · 2015
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering
Ruining He and Julian McAuley. 2016 · 2016
Earlier work this paper cites.
Abstractive text summarization using sequence-to-sequence RNNs and beyond
Ramesh Nallapati, Bowen Zhou, Cicero dos Santos, Caglar Gulcehre, and Bing Xiang. 2016 · 2016
Earlier work this paper cites.
Learning to score system summaries for better content selection evaluation
Maxime Peyrard, Teresa Botschen, and Iryna Gurevych. 2017 · 2017
Earlier work this paper cites.
Summarizing opinions: Aspect extraction meets sentiment prediction and they are both weakly supervised
Stefanos Angelidis and Mirella Lapata. 2018 · 2018
Earlier work this paper cites.
Deep contextualized word representations
Matthew E. Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Earlier work this paper cites.
FLAIR: An easy-to-use framework for state-of-the-art NLP
Alan Akbik, Tanja Bergmann, Duncan Blythe, Kashif Rasul, Stefan Schweter, and Roland Vollgraf. 2019 · 2019
Cited alongside, same era.
MeanSum: A neural model for unsupervised multi-document abstractive summarization
Eric Chu and Peter Liu. 2019 · 2019
Cited alongside, same era.
Sentence mover’s similarity: Automatic evaluation for multi-sentence texts
Elizabeth Clark, Asli Celikyilmaz, and Noah A. Smith. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Unsupervised neural single-document summarization of reviews via learning latent discourse structure and its ranking
Masaru Isonuma, Junichiro Mori, and Ichiro Sakata. 2019 · 2019
Cited alongside, same era.
Automatic text evaluation through the lens of Wasserstein barycenters
Pierre Colombo, Guillaume Staerman, Chloé Clavel, and Pablo Piantanida. 2021 · 2021
Later among the works it cites.
Self-supervised and controlled multi-document opinion summarization
Hady Elsahar, Maximin Coavoux, Jos Rozen, and Matthias Gallé. 2021 · 2021
Later among the works it cites.
SummEval: Re-evaluating Summarization Evaluation
Alexander R. Fabbri, Wojciech Kryściński, Bryan McCann, Caiming Xiong, Richard Socher, and Dragomir Radev. 2021 · 2021
Later among the works it cites.
Convex Aggregation for Opinion Summarization
Hayate Iso, Xiaolan Wang, Yoshihiko Suhara, Stefanos Angelidis, and Wang-Chiew Tan. 2021 · 2021
Later among the works it cites.
Unsupervised abstractive opinion summarization by generating sentences with tree-structured topic guidance
Masaru Isonuma, Junichiro Mori, Danushka Bollegala, and Ichiro Sakata. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Towards personalized review summarization via user-aware sequence network
Junjie Li, Haoran Li, and Chengqing Zong. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Cited alongside, same era.
Answers unite! unsupervised metrics for reinforced summarization models
Thomas Scialom, Sylvain Lamprier, Benjamin Piwowarski, and Jacopo Staiano. 2019 · 2019
Cited alongside, same era.
Red-faced ROUGE: Examining the suitability of ROUGE for opinion summary evaluation
Wenyi Tay, Aditya Joshi, Xiuzhen Zhang, Sarvnaz Karimi, and Stephen Wan. 2019 · 2019
Cited alongside, same era.
Re-evaluating evaluation in text summarization
Manik Bhandari, Pranav Narayan Gour, Atabak Ashfaq, Pengfei Liu, and Graham Neubig. 2020 · 2020
Cited alongside, same era.
Few-shot learning for opinion summarization
Arthur Bražinskas, Mirella Lapata, and Ivan Titov. 2020a · 2020
Cited alongside, same era.
SUPERT: Towards new frontiers in unsupervised evaluation metrics for multi-document summarization
Yang Gao, Wei Zhao, and Steffen Eger. 2020 · 2020
Cited alongside, same era.
Nadav Oved and Ran Levy. 2021 · 2021
Later among the works it cites.
QuestEval: Summarization asks for fact-based evaluation
Thomas Scialom, Paul-Alexis Dray, Sylvain Lamprier, Benjamin Piwowarski, Jacopo Staiano, Alex Wang, and Patrick Gallinari. 2021 · 2021
Later among the works it cites.
TransSum: Translating aspect and sentiment embeddings for self-supervised opinion summarization
Ke Wang and Xiaojun Wan. 2021 · 2021
Later among the works it cites.
Bartscore: Evaluating generated text as text generation
Weizhe Yuan, Graham Neubig, and Pengfei Liu. 2021 · 2021
Later among the works it cites.
Efficient few-shot fine-tuning for opinion summarization
Arthur Brazinskas, Ramesh Nallapati, Mohit Bansal, and Markus Dreyer. 2022 · 2022
Later among the works it cites.
Infolm: A new metric to evaluate summarization & data2text generation
Pierre Jean A. Colombo, Chloé Clavel, and Pablo Piantanida. 2022 · 2022
Later among the works it cites.
QAFactEval: Improved QA-based factual consistency evaluation for summarization
Alexander Fabbri, Chien-Sheng Wu, Wenhao Liu, and Caiming Xiong. 2022 · 2022
Later among the works it cites.
DialSummEval: Revisiting summarization evaluation for dialogues
Mingqi Gao and Xiaojun Wan. 2022 · 2022
Later among the works it cites.
SummaC: Re-Visiting NLI-based Models for Inconsistency Detection in Summarization
Philippe Laban, Tobias Schnabel, Paul N. Bennett, and Marti A. Hearst. 2022 · 2022
Later among the works it cites.
Prompted opinion summarization with gpt-3.5
Adithya Bhaskar, Alexander R. Fabbri, and Greg Durrett. 2023 · 2023
Closest in time.
Gptscore: Evaluate as you desire
Jinlan Fu, See-Kiong Ng, Zhengbao Jiang, and Pengfei Liu. 2023 · 2023
Closest in time.
Human-like summarization evaluation with chatgpt
Mingqi Gao, Jie Ruan, Renliang Sun, Xunjian Yin, Shiping Yang, and Xiaojun Wan. 2023 · 2023
Closest in time.
Attributable and scalable opinion summarization
Tom Hosking, Hao Tang, and Mirella Lapata. 2023 · 2023
Closest in time.
G-eval: Nlg evaluation using gpt-4 with better human alignment
Yang Liu, Dan Iter, Yichong Xu, Shuohang Wang, Ruochen Xu, and Chenguang Zhu. 2023 · 2023
Closest in time.
Chatgpt as a factual inconsistency evaluator for text summarization
Zheheng Luo, Qianqian Xie, and Sophia Ananiadou. 2023 · 2023
Closest in time.
Automatically evaluating opinion prevalence in opinion summarization
Christopher Malon. 2023 · 2023
Closest in time.
Is chatgpt a good nlg evaluator? a preliminary study
Jiaan Wang, Yunlong Liang, Fandong Meng, Zengkui Sun, Haoxiang Shi, Zhixu Li, Jinan Xu, Jianfeng Qu, and Jie Zhou. 2023 · 2023
Closest in time.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Closest in time.
Revisiting automatic question summarization evaluation in the biomedical domain
Hongyi Yuan, Yaoyun Zhang, Fei Huang, and Songfang Huang. 2023 · 2023
Closest in time.
MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance
Wei Zhao, Maxime Peyrard, Fei Liu, Yang Gao, Christian M. Meyer, and Steffen Eger. 2019 · 2023
Closest in time.