Fetching the paper…
Reading the bibliography…
We propose a method to perform automatic document summarisation without using reference summaries.
The proof and measurement of association between two things
Charles Spearman. 1904 · 1904
Earlier work this paper cites.
A Law of Comparative Judgement
Louis Leon Thurstone. 1927 · 1927
Earlier work this paper cites.
Rank correlation methods
M.G. Kendall. 1948 · 1948
Earlier work this paper cites.
Rank analysis of incomplete block designs: I. The method of paired comparisons
Ralph Allan Bradley and Milton E Terry. 1952 · 1952
Earlier work this paper cites.
Temporal Credit Assignment in Reinforcement Learning
R. S. Sutton. 1984 · 1984
Earlier work this paper cites.
Least-squares temporal difference learning
Justin A. Boyan. 1999 · 1999
Earlier work this paper cites.
Least-squares policy iteration
Michail G Lagoudakis and Ronald Parr. 2003 · 2003
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Preference learning with Gaussian processes
Wei Chu and Zoubin Ghahramani. 2005 · 2005
Earlier work this paper cites.
Natural Language Processing with Python
Steven Bird, Ewan Klein, and Edward Loper. 2009 · 2009
Earlier work this paper cites.
Preference learning: An introduction
Johannes Fürnkranz and Eyke Hüllermeier. 2010 · 2010
Earlier work this paper cites.
Preference uncertainty, preference refinement and paired comparison choice experiments
David C Kingsley and Thomas C Brown. 2010 · 2010
Cited alongside, same era.
Optimal Bayesian recommendation sets and myopically optimal choice query sets
Paolo Viappiani and Craig Boutilier. 2010 · 2010
Cited alongside, same era.
Active ranking using pairwise comparisons
Kevin G. Jamieson and Robert D. Nowak. 2011 · 2011
Cited alongside, same era.
A survey of text summarization techniques
Ani Nenkova and Kathleen McKeown. 2012 · 2012
Cited alongside, same era.
Framework of automatic text summarization using reinforcement learning
Seonggi Ryang and Takeshi Abekawa. 2012 · 2012
Cited alongside, same era.
Fear the REAPER: A system for automatic multi-document summarization with reinforcement learning
Cody Rioux, Sadid A. Hasan, and Yllias Chali. 2014 · 2014
Model-free preference-based reinforcement learning
Christian Wirth, Johannes Fürnkranz, and Gerhard Neumann. 2016 · 2016
Later among the works it cites.
Deep Reinforcement Learning from Human Preferences
Paul F. Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Later among the works it cites.
Bandit structured prediction for neural sequence-to-sequence learning
Julia Kreutzer, Artem Sokolov, and Stefan Riezler. 2017 · 2017
Later among the works it cites.
Just sort it! A simple and effective approach to active preference learning
Lucas Maystre and Matthias Grossglauser. 2017 · 2017
Later among the works it cites.
A deep reinforced model for abstractive summarization
Romain Paulus, Caiming Xiong, and Richard Socher. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A reinforcement learning approach for adaptive single- and multi-document summarization
Stefan Henß, Margot Mieskes, and Iryna Gurevych. 2015 · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al. 2015 · 2015
Cited alongside, same era.
Sequence level training with recurrent neural networks
Marc’Aurelio Ranzato, Sumit Chopra, Michael Auli, and Wojciech Zaremba. 2015 · 2015
Cited alongside, same era.
Learning structured predictors from bandit feedback for interactive NLP
Artem Sokolov, Julia Kreutzer, Christopher Lo, and Stefan Riezler. 2016a · 2016
Cited alongside, same era.
Stochastic structured prediction under bandit feedback
Artem Sokolov, Julia Kreutzer, Stefan Riezler, and Christopher Lo. 2016b · 2016
Cited alongside, same era.
Avinesh P.V.S. and Christian M. Meyer. 2017 · 2017
Later among the works it cites.
A survey of preference-based reinforcement learning methods
Christian Wirth, Riad Akrour, Gerhard Neumann, and Johannes Fürnkranz. 2017 · 2017
Later among the works it cites.
Ranking sentences for extractive summarization with reinforcement learning
Shashi Narayan, Shay B. Cohen, and Mirella Lapata. 2018 · 2018
Closest in time.
Multi-reward reinforced summarization with saliency and entailment
Ramakanth Pasunuru and Mohit Bansal. 2018 · 2018
Closest in time.
Finding convincing arguments using scalable bayesian preference learning
Edwin D. Simpson and Iryna Gurevych. 2018 · 2018
Closest in time.
Estimating summary quality with pairwise preferences
Markus Zopf. 2018 · 2018
Closest in time.