Fetching the paper…
Reading the bibliography…
We propose a general-purpose approach to discovering active learning (AL) strategies from data.
C. J. C. H. Watkins and P. Dayan, “Q-learning,” Machine Learning , 1992
1992
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine Learning , 1992
1992
Earlier work this paper cites.
D. D. Lewis and W. A. Gale, “A sequential algorithm for training text classifiers,” in ACM SIGIR proceedings on Research and Development in Information Retrieval , 1994
1994
Earlier work this paper cites.
M. Asada, S. Noda, S. Tawaratsumida, and K. Hosoda, “Purposive behavior acquisition for a real robot by vision-based reinforcement learning,” Machine Learning , 1996
1996
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement Learning . MIT Press, 1998
1998
Earlier work this paper cites.
1999
Earlier work this paper cites.
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in Advances in Neural Information Processing Systems , 2000
2000
Earlier work this paper cites.
S. Tong and D. Koller, “Support vector machine active learning with applications to text classification,” Journal of Machine Learning Research , 2001
2001
Earlier work this paper cites.
Y. Baram, R. El-Yaniv, and K. Luz, “Online choice of active learning algorithms,” Journal of Machine Learning Research , 2004
2004
Earlier work this paper cites.
J. Michels, A. Saxena, and A. Y. Ng, “High speed obstacle avoidance using monocular vision and reinforcement learning,” in International Conference on Machine Learning , 2005
2005
Earlier work this paper cites.
R. Gilad-Bachrach, A. Navot, and N. Tishby, “Query by committee made real,” in Advances in Neural Information Processing Systems , 2005
2005
Earlier work this paper cites.
T. Osugi, D. Kun, and S. Scott, “Balancing exploration and exploitation: A new algorithm for active machine learning,” in International Conference on Data Mining , 2005
2005
Earlier work this paper cites.
H. van Hasselt, “Double Q-learning,” in Advances in Neural Information Processing Systems , 2010
2010
Earlier work this paper cites.
S.-J. Huang, R. Jin, and Z.-H. Zhou, “Active learning by querying informative and representative examples,” in Advances in Neural Information Processing Systems , 2010
2010
Earlier work this paper cites.
B. Settles, Active Learning . Morgan & Claypool Publishers, 2012
2012
Earlier work this paper cites.
S. Ebert, M. Fritz, and B. Schiele, “RALF: A reinforced active learning formulation for object class recognition,” in Conference on Computer Vision and Pattern Recognition , 2012
2012
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller, “Playing atari with deep reinforcement learning,” in Deep Learning Workshop at Advances in Neural Information Processing Systems , 2013
2013
Earlier work this paper cites.
W. Luo, A. G. Schwing, and R. Urtasun, “Latent structured active learning,” in Advances in Neural Information Processing Systems , 2013
2013
Cited alongside, same era.
A. Freytag, E. Rodner, and J. Denzler, “Selecting influential examples: Active learning with expected model output changes,” in European Conference on Computer Vision , 2014
2014
Cited alongside, same era.
J. C. Caicedo and S. Lazebnik, “Active object localization with deep reinforcement learning,” in International Conference on Computer Vision , 2015
2015
Cited alongside, same era.
O. Russakovsky, L.-J. Li, and L. Fei-Fei, “Best of both worlds: Human-machine collaboration for object annotation,” in Conference on Computer Vision and Pattern Recognition , 2015
2015
Cited alongside, same era.
W.-N. Hsu and H.-T. Lin, “Active learning by learning,” in American Association for Artificial Intelligence Conference , 2015
M. Fang, Y. Li, and T. Cohn, “Learning how to active learn: A deep reinforcement learning approach,” in Conference on Empirical Methods in Natural Language Processing , 2017
2017
Later among the works it cites.
G. Contardo, L. Denoyer, and T. Artières, “A meta-learning approach to one-step active-learning,” in CEUR International Workshop on Automatic Selection, Configuration and Composition of Machine Learning Algorithms , 2017
2017
Later among the works it cites.
B. Zoph and Q. V. Le, “Neural architecture search with reinforcement learning,” in International Conference on Learning Representations , 2017
2017
Later among the works it cites.
B. Baker, O. Gupta, N. Naik, and R. Raskar, “Designing neural network architectures using reinforcement learning,” in International Conference on Learning Representations , 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
Q. Sun, A. Laddha, and D. Batra, “Active learning for structured probabilistic models with histogram approximation,” in Conference on Computer Vision and Pattern Recognition , 2015
2015
Cited alongside, same era.
C. Käding, A. Freytag, E. Rodner, P. Bodesheim, and J. Denzler, “Active learning and discovery of object categories in the presence of unnameable instances,” in Conference on Computer Vision and Pattern Recognition , 2015
2015
Cited alongside, same era.
C. Long and G. Hua, “Multi-class multi-annotator active learning with robust Gaussian process for visual recognition,” in International Conference on Computer Vision , 2015
2015
Cited alongside, same era.
C. Zhang and K. Chaudhuri, “Active learning from weak and strong labelers,” in Advances in Neural Information Processing Systems , 2015
2015
Cited alongside, same era.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,” Journal of Machine Learning Research , 2016
2016
Cited alongside, same era.
M. Bellver, X. Giró-i Nieto, F. Marqués, and J. Torres, “Hierarchical object detection with deep reinforcement learning,” in Workshop on Deep Reinforcement Learning at Advances in Neural Information Processing Systems , 2016
2016
Cited alongside, same era.
H.-M. Chu and H.-T. Lin, “Can active learning experience be transferred?” in International Conference on Data Mining , 2016
2016
Cited alongside, same era.
R. Hu, J. Andreas, M. Rohrbach, T. Darrell, and K. Saenko, “Learning to reason: End-to-end module networks for visual question answering,” in International Conference on Computer Vision , 2017
2017
Later among the works it cites.
J. Supanc̆ic̆, III and D. Ramanan, “Tracking as online decision-making: Learning a policy from streaming videos with reinforcement learning,” in International Conference on Computer Vision , 2017
2017
Later among the works it cites.
D. Dheeru and E. Karra Taniskidou, “UCI machine learning repository,” 2017. [Online]. Available: http://archive.ics.uci.edu/ml
2017
Later among the works it cites.
M. Liu, W. Buntine, and G. Haffari, “Learning how to actively learn: A deep imitation learning approach,” in Annual Meeting of the Association for Computational Linguistics , 2018
2018
Closest in time.
S. Ravi and H. Larochelle, “Meta-learning for batch mode active learning,” in International Conference on Learning Representations Workshop Track , 2018
2018
Closest in time.
K. Pang, M. Dong, Y. Wu, and T. Hospedales, “Meta-learning transferable active learning policies by deep reinforcement learning,” in AutoML workshop at International Conference on Machine Learning , 2018
2018
Closest in time.
D. Jayaraman and K. Grauman, “Learning to look around: Intelligently exploring unseen environments for unknown tasks,” in Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied question answering,” in Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.
D. Acuna, H. Ling, A. Kar, and S. Fidler, “Efficient interactive annotation of segmentation datasets with polygon-RNN++,” in Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.
K. Konyushkova, J. R. R. Uijlings, C. H. Lampert, and V. Ferrari, “Learning intelligent dialogs for bounding box annotation,” in Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.
S. Hao, J. Lu, P. Zhao, C. Zhang, S. C. Hoi, and C. Miao, “Second-order online active learning and its applications,” IEEE Transactions on Knowledge and Data Engineering , 2018
2018
Closest in time.
W. H. Beluch, T. Genewein, A. Nürnberger, and J. M. Köhler, “The power of ensembles for active learning in image classification,” in Conference on Computer Vision and Pattern Recognition , 2018
2018
Closest in time.