Fetching the paper…
Reading the bibliography…
Data generation and labeling are usually an expensive part of learning for robotics.
Individual comparisons by ranking methods
F. Wilcoxon · 1945
Earlier work this paper cites.
Least squares quantization in pcm
S. Lloyd · 1982
Earlier work this paper cites.
Clustering by means of medoids
L. Kaufman and P. Rousseeuw · 1987
Earlier work this paper cites.
An adaptive metropolis algorithm
H. Haario, E. Saksman, J. Tamminen, et al · 2001
Earlier work this paper cites.
Scalable training of l 1-regularized log-linear models
G. Andrew and J. Gao · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey · 2008
Earlier work this paper cites.
Discriminative batch mode active learning
Y. Guo and D. Schuurmans · 2008
Earlier work this paper cites.
Maximizing global entropy reduction for active learning in speech recognition
B. Varadarajan, D. Yu, L. Deng, and A. Acero · 2009
Earlier work this paper cites.
Preference learning in recommender systems
M. De Gemmis, L. Iaquinta, P. Lops, C. Musto, F. Narducci, and G. Semeraro · 2009
Earlier work this paper cites.
An axiomatic approach for result diversification
S. Gollapudi and A. Sharma · 2009
Earlier work this paper cites.
Preference-based policy learning
R. Akrour, M. Schoenauer, and M. Sebag · 2011
Earlier work this paper cites.
Preference-learning based inverse reinforcement learning for dialog control
H. Sugiyama, T. Meguro, and Y. Minami · 2012
Earlier work this paper cites.
April: Active preference learning-based reinforcement learning
R. Akrour, M. Schoenauer, and M. Sebag · 2012
Earlier work this paper cites.
Keyframe-based learning from demonstration
B. Akgun, M. Cakmak, K. Jiang, and A. L. Thomaz · 2012
Earlier work this paper cites.
Continuous inverse optimal control with locally optimal examples
S. Levine and V. Koltun · 2012
Cited alongside, same era.
Preference-based reinforcement learning: a formal framework and a policy iteration algorithm
J. Fürnkranz, E. Hüllermeier, W. Cheng, and S.-H. Park · 2012
Cited alongside, same era.
A bayesian approach for policy learning from trajectory preference queries
A. Wilson, A. Fern, and P. Tadepalli · 2012
Cited alongside, same era.
An active learning algorithm for ranking from pairwise preferences with an almost optimal query complexity
N. Ailon · 2012
Cited alongside, same era.
Max-sum diversification, monotone submodular functions and dynamic updates
A. Borodin, H. C. Lee, and Y. Ye · 2012
Cited alongside, same era.
Dissimilarity-based sparse subset selection
E. Elhamifar, G. Sapiro, and S. S. Sastry · 2016
Later among the works it cites.
Active comparison based learning incorporating user uncertainty and noise
R. Holladay, S. Javdani, A. Dragan, and S. Srinivasa · 2016
Later among the works it cites.
Planning for cars that coordinate with people: leveraging effects on human actions for planning and active information gathering over human internal state
D. Sadigh, N. Landolfi, S. S. Sastry, S. A. Seshia, and A. D. Dragan · 2016
Later among the works it cites.
Active reinforcement learning: Observing rewards at a cost
D. Krueger, J. Leike, O. Evans, and J. Salvatier · 2016
Later among the works it cites.
Planning for autonomous cars that leverage effects on human actions
D. Sadigh, S. Sastry, S. A. Seshia, and A. D. Dragan · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
E. Todorov, T. Erez, and Y. Tassa · 2012
Cited alongside, same era.
Determinantal point processes for machine learning
A. Kulesza, B. Taskar, et al · 2012
Cited alongside, same era.
Active learning for probabilistic hypotheses using the maximum gibbs error criterion
N. V. Cuong, W. S. Lee, N. Ye, K. M. A. Chai, and H. L. Chieu · 2013
Cited alongside, same era.
Near-optimal batch mode active learning and adaptive submodular optimization
Y. Chen and A. Krause · 2013
Cited alongside, same era.
Spatial variation , volume 36
B. Matérn · 2013
Cited alongside, same era.
Submodular function maximization., 2014
A. Krause and D. Golovin · 2014
Cited alongside, same era.
Learning preferences for manipulation tasks from online coactive feedback
A. Jain, S. Sharma, T. Joachims, and A. Saxena · 2015
Cited alongside, same era.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Later among the works it cites.
A geometric approach to active learning for convolutional neural networks
O. Sener and S. Savarese · 2017
Later among the works it cites.
Active preference-based learning of reward functions
D. Sadigh, A. D. Dragan, S. S. Sastry, and S. A. Seshia · 2017
Later among the works it cites.
Do you want your autonomous car to drive like you?
C. Basu, Q. Yang, D. Hungerman, M. Singhal, and A. D. Dragan · 2017
Later among the works it cites.
Deep reinforcement learning from human preferences
P. F. Christiano, J. Leike, T. Brown, M. Martic, S. Legg, and D. Amodei · 2017
Later among the works it cites.
Safe and Interactive Autonomy: Control, Learning, and Verification
D. Sadigh · 2017
Later among the works it cites.
Do you want your autonomous car to drive like you?
C. Basu, Q. Yang, D. Hungerman, A. Dragan, and M. Singhal · 2017
Later among the works it cites.
Determinantal point processes for mini-batch diversification
C. Zhang, H. Kjellstrom, and S. Mandt · 2017
Later among the works it cites.
Single shot active learning using pseudo annotators
Y. Yang and M. Loog · 2018
Closest in time.