“Sample Efficient Text Summarization Using a Single Pre-Trained Transformer”, 2019
Original
Urvashi Khandelwal, Kevin Clark, Dan Jurafsky and Lukasz Kaiser · 1905
Earlier work this paper cites.
“Discriminative Active Learning”, 2019
Original
Daniel Gissin and Shai Shalev-Shwartz · 1907
Earlier work this paper cites.
“Fine-Tuning Language Models from Human Preferences”, 2019
Original
Daniel Ziegler et al · 1909
Earlier work this paper cites.
“On the likelihood that one unknown probability exceeds another in view of the evidence of two samples”
William Thompson · 1933
Earlier work this paper cites.
“Query by Committee”
H.. Seung, M. Opper and H. Sompolinsky · 1992
Earlier work this paper cites.
“A Sequential Algorithm for Training Text Classifiers”
David. Lewis and William. Gale · 1994
Earlier work this paper cites.
“A Probability Analysis on the Value of Unlabeled Data for Classification Problems”
Tony Zhang and Frank. Oles · 2000
Earlier work this paper cites.
“An Analysis of Active Learning Strategies for Sequence Labeling Tasks”
Burr Settles and Mark Craven · 2008
Earlier work this paper cites.
“Multiple-Instance Active Learning”
Burr Settles, Mark Craven and Soumya Ray · 2008
Earlier work this paper cites.
“Active Learning Literature Survey”, 2009
Burr Settles · 2009
Earlier work this paper cites.
“Learning to summarize from human feedback”, 2020
Original
Nisan Stiennon et al · 2009
Earlier work this paper cites.
“APRIL: Active Preference Learning-Based Reinforcement Learning”
Riad Akrour, Marc Schoenauer and Michèle Sebag · 2012
Earlier work this paper cites.
“Weight Uncertainty in Neural Networks”
Charles Blundell, Julien Cornebise, Koray Kavukcuoglu and Daan Wierstra · 2015
Earlier work this paper cites.
“Teaching Machines to Read and Comprehend”
Karl Hermann et al · 2015
Earlier work this paper cites.
“Layer Normalization”, 2016
Original
Jimmy Ba, Jamie Kiros and Geoffrey. Hinton · 2016
Earlier work this paper cites.
“Dropout as a Bayesian Approximation: Representing Model Uncertainty in Deep Learning”
Yarin Gal and Zoubin Ghahramani · 2016
Earlier work this paper cites.
“HyperNetworks”
David Ha, Andrew Dai and Quoc Le · 2016
Earlier work this paper cites.
“Pointer Sentinel Mixture Models”, 2016
Original
Stephen Merity, Caiming Xiong, James Bradbury and Richard Socher · 2016
Earlier work this paper cites.