Fetching the paper…
Reading the bibliography…
Deep reinforcement learning augments the reinforcement learning framework and utilizes the powerful representation of deep neural networks.
The long and the short of memory
Wayne A Wickelgren · 1973
Earlier work this paper cites.
Image segmentation techniques
Robert M Haralick and Linda G Shapiro · 1985
Earlier work this paper cites.
A theoretical framework for back-propagation
Yann LeCun, D Touresky, G Hinton, and T Sejnowski · 1988
Earlier work this paper cites.
Long-term adjuvant tamoxifen therapy for breast cancer
V Craig Jordan · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Learning long-term dependencies with gradient descent is difficult
Yoshua Bengio, Patrice Simard, and Paolo Frasconi · 1994
Earlier work this paper cites.
Active motion detection and object tracking
Joachim Denzler and Dietrich WR Paulus · 1994
Earlier work this paper cites.
Motion tracking with an active camera
Don Murray and Anup Basu · 1994
Earlier work this paper cites.
Tracking and recognizing rigid and non-rigid facial motions using local parametric models of image motion
Michael J Black and Yaser Yacoob · 1995
Earlier work this paper cites.
Bagging predictors
Leo Breiman · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
The mnist database of handwritten digits
Yann LeCun · 1998
Earlier work this paper cites.
Efficient backprop
Yann LeCun, Léon Bottou, Genevieve B Orr, and Klaus-Robert Müller · 1998
Earlier work this paper cites.
The architecture of mind: A connectionist approach
David E Rumelhart · 1998
Earlier work this paper cites.
Automatic retinal image registration scheme using global optimization techniques
George K Matsopoulos, Nicolaos A Mouravliansky, Konstantinos K Delibasis, and Konstantina S Nikita · 1999
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S. Sutton, David McAllester, Satinder Singh, and Yishay Mansour · 1999
Earlier work this paper cites.
Real-time tracking of non-rigid objects using mean shift
Dorin Comaniciu, Visvanathan Ramesh, and Peter Meer · 2000
Earlier work this paper cites.
Actor-critic algorithms
Vijay R Konda and John N Tsitsiklis · 2000
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y. Ng and Stuart J. Russell · 2000
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng, Stuart J Russell, et al · 2000
Earlier work this paper cites.
Genetic algorithms for a robust 3-d mr-ct registration
J-M Rouet, J-J Jacq, and Christian Roux · 2000
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S Sutton, David A. McAllester, Satinder P. Singh, and Yishay Mansour · 2000
Earlier work this paper cites.
Optimization of mutual information for multiresolution image registration
Philippe Thévenaz and Michael Unser · 2000
Earlier work this paper cites.
Autonomous helicopter control using reinforcement learning policy search methods
J. A. Bagnell and J. G. Schneider · 2001
Earlier work this paper cites.
The distribution of target registration error in rigid-body point-based registration
J Michael Fitzpatrick and Jay B West · 2001
Earlier work this paper cites.
Optimal denominations for coins and bank notes: in defense of the principle of least effort
Leo Van Hove · 2001
Earlier work this paper cites.
Minimax differential dynamic programming: application to a biped walking robot
J. Morimoto, G. Zeglin, and C. G. Atkeson · 2003
Earlier work this paper cites.
Meta-learning in reinforcement learning
Nicolas Schweighofer and Kenji Doya · 2003
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y. Ng · 2004
Earlier work this paper cites.
Unifying temporal and structural credit assignment problems
Adrian K. Agogino and Kagan Tumer · 2004
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
David G Lowe · 2004
Earlier work this paper cites.
A boosted particle filter: Multitarget detection and tracking
Kenji Okuma, Ali Taleghani, Nando De Freitas, James J Little, and David G Lowe · 2004
Earlier work this paper cites.
” grabcut” interactive foreground extraction using iterated graph cuts
Carsten Rother, Vladimir Kolmogorov, and Andrew Blake · 2004
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Zhou Wang, Alan C Bovik, Hamid R Sheikh, and Eero P Simoncelli · 2004
Earlier work this paper cites.
The itk software guide: updated for itk version 2.4, 2005
Luis Ibanez, Will Schroeder, Lydia Ng, and Josh Cates · 2005
Earlier work this paper cites.
Detecting and tracking moving object using an active camera
Kye Kyung Kim, Soo Hyun Cho, Hae Jin Kim, and Jae Yeon Lee · 2005
Earlier work this paper cites.
Fast reinforcement learning for vision-guided mobile robots
T. Martinez-Marin and T. Duckett · 2005
Earlier work this paper cites.
Efficient selectivity and backup operators in monte-carlo tree search
Rémi Coulom · 2006
Earlier work this paper cites.
Random walks for image segmentation
Leo Grady · 2006
Earlier work this paper cites.
A reinforcement learning framework for medical image segmentation
Farhang Sahba, Hamid R Tizhoosh, and Magdy MA Salama · 2006
Earlier work this paper cites.
The pascal visual object classes challenge 2007 (voc2007) results
Mark Everingham, Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman · 2007
Earlier work this paper cites.
Recovering surface layout from an image
Derek Hoiem, Alexei A Efros, and Martial Hebert · 2007
Earlier work this paper cites.
A multiagent approach to q-learning for daily stock trading
Jae Won Lee, Jonghun Park, Jangmin O, Jongwoo Lee, and Euyseok Hong · 2007
Earlier work this paper cites.
Application of opposition-based reinforcement learning in image segmentation
Farhang Sahba, Hamid R Tizhoosh, and Magdy MMA Salama · 2007
Earlier work this paper cites.
Minimal-bracketing sets for high-dynamic-range image capture
Neil Barakat, A Nicholas Hone, and Thomas E Darcie · 2008
Earlier work this paper cites.
Evaluating multiple object tracking performance: the clear mot metrics
Keni Bernardin and Rainer Stiefelhagen · 2008
Earlier work this paper cites.
A comprehensive survey of multiagent reinforcement learning
L. Busoniu, R. Babuska, and B. De Schutter · 2008
Earlier work this paper cites.
Policy gradient based reinforcement learning for real autonomous underwater cable tracking
A. El-Fakdi and M. Carreras · 2008
Earlier work this paper cites.
The alzheimer’s disease neuroimaging initiative (adni): Mri methods
Clifford R Jack Jr, Matt A Bernstein, Nick C Fox, Paul Thompson, Gene Alexander, Danielle Harvey, Bret Borowski, Paula J Britson, Jennifer L. Whitwell, Chadwick Ward, et al · 2008
Earlier work this paper cites.
A consolidated actor-critic model with function approximation for high-dimensional pomdps
C. Niedzwiedz, I. Elhanany, Zhenzhen Liu, and S. Livingston · 2008
Earlier work this paper cites.
Reinforcement learning of motor skills with policy gradients
Jan Peters and Stefan Schaal · 2008
Earlier work this paper cites.
Visual tracking with online multiple instance learning
Boris Babenko, Ming-Hsuan Yang, and Serge Belongie · 2009
Earlier work this paper cites.
Natural actorâ-critic algorithms
Shalabh Bhatnagar, Richard S. Sutton, Mohammad Ghavamzadeh, and Mark Lee · 2009
Earlier work this paper cites.
Apprenticeship learning for helicopter control
Adam Coates, Pieter Abbeel, and Andrew Y. Ng · 2009
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Pedestrian detection: A benchmark
Piotr Dollár, Christian Wojek, Bernt Schiele, and Pietro Perona · 2009
Earlier work this paper cites.
Elastix: a toolbox for intensity-based medical image registration
Stefan Klein, Marius Staring, Keelin Murphy, Max A Viergever, and Josien PW Pluim · 2009
Earlier work this paper cites.
Nonparametric representation of an approximated poincaré map for learning biped locomotion
Jun Morimoto and Christopher G. Atkeson · 2009
Earlier work this paper cites.
An actor-critic method using least squares temporal difference learning
I. C. Paschalidis, K. Li, and R. Moazzez Estanjini · 2009
Earlier work this paper cites.
Vision-based reinforcement learning using approximate policy iteration
M. R. Shaker, Shigang Yue, and T. Duckett · 2009
Earlier work this paper cites.
Autonomous helicopter aerobatics through apprenticeship learning
Pieter Abbeel, Adam Coates, and Andrew Y. Ng · 2010
Earlier work this paper cites.
An actor–critic algorithm with function approximation for discounted cost constrained markov decision processes
Shalabh Bhatnagar · 2010
Earlier work this paper cites.
Regression forests for efficient anatomy detection and localization in ct studies
Antonio Criminisi, Jamie Shotton, Duncan Robertson, and Ender Konukoglu · 2010
Earlier work this paper cites.
Human tracking using convolutional neural networks
Jialue Fan, Wei Xu, Ying Wu, and Yihong Gong · 2010
Earlier work this paper cites.
Double q-learning
Hado V Hasselt · 2010
Earlier work this paper cites.
Nhs fetal anomaly screening programme
Donna Kirwan · 2010
Earlier work this paper cites.
A contextual-bandit approach to personalized news article recommendation
Lihong Li, Wei Chu, John Langford, and Robert E Schapire · 2010
Earlier work this paper cites.
Multi-video summarization based on video-mmr
Yingbo Li and Bernard Merialdo · 2010
Earlier work this paper cites.
V3d enables real-time 3d visualization and quantitative analysis of large-scale biological image data sets
Hanchuan Peng, Zongcai Ruan, Fuhui Long, Julie H Simpson, and Eugene W Myers · 2010
Earlier work this paper cites.
Vsumm: A mechanism designed to produce static video summaries and a novel evaluation method
Sandra Eliza Fontes De Avila, Ana Paula Brandão Lopes, Antonio da Luz Jr, and Arnaldo de Albuquerque Araújo · 2011
Earlier work this paper cites.
The pascal visual object classes challenge 2012 (voc2012) development kit
Mark Everingham and John Winn · 2011
Earlier work this paper cites.
A real-time model-based reinforcement learning architecture for robot control
Todd Hester, Michael Quinlan, and Peter Stone · 2011
Earlier work this paper cites.
Tracking-learning-detection
Zdenek Kalal, Krystian Mikolajczyk, and Jiri Matas · 2011
Earlier work this paper cites.
Efficient inference in fully connected crfs with gaussian edge potentials
Philipp Krähenbühl and Vladlen Koltun · 2011
Earlier work this paper cites.
Extensions of recurrent neural network language model
Tomas Mikolov, Stefan Kombrink, Lukás Burget, Jan Cernocký, and Sanjeev Khudanpur · 2011
Earlier work this paper cites.
Reading digits in natural images with unsupervised feature learning
Yuval Netzer, Tao Wang, Adam Coates, Alessandro Bissacco, Bo Wu, and Andrew Y Ng · 2011
Earlier work this paper cites.
Globally-optimal greedy algorithms for tracking a variable number of objects
Hamed Pirsiavash, Deva Ramanan, and Charless C Fowlkes · 2011
Earlier work this paper cites.
Learning decision: Robustness, uncertainty, and approximation
J. Bagnell · 2012
Earlier work this paper cites.
See all by looking at a few: Sparse modeling for finding representative objects
Ehsan Elhamifar, Guillermo Sapiro, and Rene Vidal · 2012
Earlier work this paper cites.
Are we ready for autonomous driving? the kitti vision benchmark suite
A. Geiger, P. Lenz, and R. Urtasun · 2012
Earlier work this paper cites.
Single and multiple object tracking using log-euclidean riemannian subspace and block-division appearance model
Weiming Hu, Xi Li, Wenhan Luo, Xiaoqin Zhang, Stephen Maybank, and Zhongfei Zhang · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
Learning object class detectors from weakly annotated video
Alessandro Prest, Christian Leistner, Javier Civera, Cordelia Schmid, and Vittorio Ferrari · 2012
Earlier work this paper cites.
Imitating play from game trajectories: Temporal difference learning versus preference learning
T. P. Runarsson and S. M. Lucas · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus · 2012
Earlier work this paper cites.
View invariant human action recognition using histograms of 3d joints
Lu Xia, Chia-Chih Chen, and Jake K Aggarwal · 2012
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, and Geoffrey E. Hinton · 2013
Earlier work this paper cites.
Perceptual organization and recognition of indoor scenes from rgb-d images
Saurabh Gupta, Pablo Arbelaez, and Jitendra Malik · 2013
Earlier work this paper cites.
Multi-source multi-scale counting in extremely dense crowd images
Haroon Idrees, Imran Saleemi, Cody Seibert, and Mubarak Shah · 2013
Earlier work this paper cites.
Recurrent continuous translation models
Nal Kalchbrenner and Phil Blunsom · 2013
Earlier work this paper cites.
Reinforcement learning in robotics: A survey
Jens Kober, J. Andrew Bagnell, and Jan Peters · 2013
Earlier work this paper cites.
Data-efficient generalization of robot skills with contextual policy search
A. Kupcsik, M. Deisenroth, Jan Peters, and G. Neumann · 2013
Earlier work this paper cites.
Scale-space theory in computer vision
Tony Lindeberg · 2013
Earlier work this paper cites.
Lcc-demons: a robust and accurate symmetric diffeomorphic registration algorithm
Marco Lorenzi, Nicholas Ayache, Giovanni B Frisoni, Xavier Pennec, Alzheimer’s Disease Neuroimaging Initiative (ADNI, et al · 2013
Earlier work this paper cites.
Improving probabilistic image registration via reinforcement learning and uncertainty evaluation
Tayebeh Lotfi, Lisa Tang, Shawn Andrews, and Ghassan Hamarneh · 2013
Earlier work this paper cites.
System and method for 3-d/3-d registration between non-contrast-enhanced cbct and contrast-enhanced ct for abdominal aortic aneurysm stenting
Shun Miao, Rui Liao, Marcus Pfister, Li Zhang, and Vincent Ordy · 2013
Earlier work this paper cites.
Online feature selection for model-based reinforcement learning
Trung Thanh Nguyen, Zhuoru Li, Tomi Silander, and Tze-Yun Leong · 2013
Earlier work this paper cites.
Fast object segmentation in unconstrained video
Anestis Papazoglou and Vittorio Ferrari · 2013
Earlier work this paper cites.
Mit technology review
David Rotman · 2013
Earlier work this paper cites.
Deep neural networks for object detection
Christian Szegedy, Alexander Toshev, and Dumitru Erhan · 2013
Earlier work this paper cites.
Selective search for object recognition
Jasper RR Uijlings, Koen EA Van De Sande, Theo Gevers, and Arnold WM Smeulders · 2013
Earlier work this paper cites.
Learning a deep compact image representation for visual tracking
Naiyan Wang and Dit-Yan Yeung · 2013
Earlier work this paper cites.
Online object tracking: A benchmark
Yi Wu, Jongwoo Lim, and Ming-Hsuan Yang · 2013
Earlier work this paper cites.
App2: automatic tracing of 3d neuron morphology based on hierarchical pruning of a gray-weighted image distance-tree
Hang Xiao and Hanchuan Peng · 2013
Earlier work this paper cites.
Seamseg: Video object segmentation using patch seams
S Avinash Ramakanth and R Venkatesh Babu · 2014
Earlier work this paper cites.
Robust online multi-object tracking based on tracklet confidence and online discriminative appearance learning
Seung-Hwan Bae and Kuk-Jin Yoon · 2014
Earlier work this paper cites.
Approximate real-time optimal control based on sparse gaussian process models
J. Boedecker, J. T. Springenberg, J. Wülfing, and M. Riedmiller · 2014
Earlier work this paper cites.
Global contrast based salient region detection
Ming-Ming Cheng, Niloy J Mitra, Xiaolei Huang, Philip HS Torr, and Shi-Min Hu · 2014
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart van Merrienboer, Çaglar Gülçehre, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Population-based studies of myocardial hypertrophy: high resolution cardiovascular magnetic resonance atlases improve statistical power
Antonio de Marvao, Timothy JW Dawes, Wenzhe Shi, Christopher Minas, Niall G Keenan, Tamara Diamond, Giuliana Durighel, Giovanni Montana, Daniel Rueckert, Stuart A Cook, et al · 2014
Earlier work this paper cites.
Multi-task policy search for robotics
M. P. Deisenroth, P. Englert, J. Peters, and D. Fox · 2014
Earlier work this paper cites.
Comparison of model-free and model-based methods for time optimal hit control of a badminton robot
B. Depraetere, M. Liu, G. Pinte, I. Grondman, and R. BabuÅ¡ka · 2014
Earlier work this paper cites.
Long-term recurrent convolutional networks for visual recognition and description
Jeff Donahue, Lisa Anne Hendricks, Sergio Guadarrama, Marcus Rohrbach, Subhashini Venugopalan, Kate Saenko, and Trevor Darrell · 2014
Earlier work this paper cites.
Scalable object detection using deep neural networks
Dumitru Erhan, Christian Szegedy, Alexander Toshev, and Dragomir Anguelov · 2014
Earlier work this paper cites.
Jhu-isi gesture and skill assessment working set (jigsaws): A surgical activity dataset for human motion modeling
Yixin Gao, S Swaroop Vedula, Carol E Reiley, Narges Ahmidi, Balakrishnan Varadarajan, Henry C Lin, Lingling Tao, Luca Zappella, Benjamın Béjar, David D Yuh, et al · 2014
Earlier work this paper cites.
Multi-organ localization combining global-to-local regression and confidence maps
Romane Gauriau, Rémi Cuingnet, David Lesage, and Isabelle Bloch · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
Ross Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Learning rich features from rgb-d images for object detection and segmentation
Saurabh Gupta, Ross Girshick, Pablo Arbeláez, and Jitendra Malik · 2014
Earlier work this paper cites.
Creating summaries from user videos
Michael Gygli, Helmut Grabner, Hayko Riemenschneider, and Luc Van Gool · 2014
Earlier work this paper cites.
Simultaneous detection and segmentation
Bharath Hariharan, Pablo Arbeláez, Ross Girshick, and Jitendra Malik · 2014
Earlier work this paper cites.
High-speed tracking with kernelized correlation filters
João F Henriques, Rui Caseiro, Pedro Martins, and Jorge Batista · 2014
Earlier work this paper cites.
Deep features for text spotting
Max Jaderberg, Andrea Vedaldi, and Andrew Zisserman · 2014
Earlier work this paper cites.
Supervoxel-consistent foreground propagation in video
Suyog Dutt Jain and Kristen Grauman · 2014
Earlier work this paper cites.
Data fusion of radar and image measurements for multi-object tracking via kalman filtering
Du Yong Kim and Moongu Jeon · 2014
Earlier work this paper cites.
Learning an image-based motion context for multiple people tracking
Laura Leal-Taixé, Michele Fenzi, Alina Kuznetsova, Bodo Rosenhahn, and Silvio Savarese · 2014
Earlier work this paper cites.
Learning complex neural network policies with trajectory optimization
Sergey Levine and Vladlen Koltun · 2014
Earlier work this paper cites.
Evaluation of prostate segmentation algorithms for mri: the promise12 challenge
Geert Litjens, Robert Toth, Wendy van de Ven, Caroline Hoeks, Sjoerd Kerkstra, Bram van Ginneken, Graham Vincent, Gwenael Guillard, Neil Birbeck, Jindang Zhang, et al · 2014
Earlier work this paper cites.
Addressing the rare word problem in neural machine translation
Thang Luong, Ilya Sutskever, Quoc V. Le, Oriol Vinyals, and Wojciech Zaremba · 2014
Earlier work this paper cites.
Deep captioning with multimodal recurrent neural networks (m-rnn)
Junhua Mao, Wei Xu, Yi Yang, Jiang Wang, and Alan L. Yuille · 2014
Earlier work this paper cites.
Fully automatic lesion segmentation in breast mri using mean-shift and graph-cuts on a region adjacency graph
Darryl McClymont, Andrew Mehnert, Adnan Trakic, Dominic Kennedy, and Stuart Crozier · 2014
Earlier work this paper cites.
The multimodal brain tumor image segmentation benchmark (brats)
Bjoern H Menze, Andras Jakab, Stefan Bauer, Jayashree Kalpathy-Cramer, Keyvan Farahani, Justin Kirby, Yuliya Burren, Nicole Porz, Johannes Slotboom, Roland Wiest, et al · 2014
Earlier work this paper cites.
Probabilistic sparse matching for robust 3d/3d fusion in minimally invasive surgery
Dominik Neumann, Saša Grbić, Matthias John, Nassir Navab, Joachim Hornegger, and Razvan Ionasec · 2014
Earlier work this paper cites.
Category-specific video summarization
Danila Potapov, Matthijs Douze, Zaid Harchaoui, and Cordelia Schmid · 2014
Earlier work this paper cites.
Deterministic policy gradient algorithms
David Silver, Guy Lever, Nicolas Heess, Thomas Degris, Daan Wierstra, and Martin Riedmiller · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Deeppose: Human pose estimation via deep neural networks
Alexander Toshev and Christian Szegedy · 2014
Earlier work this paper cites.
Using trajectory data to improve bayesian optimization for reinforcement learning
Aaron Wilson, Alan Fern, and Prasad Tadepalli · 2014
Earlier work this paper cites.
Generic object detection with dense neural patterns and regionlets
Will Y Zou, Xiaoyu Wang, Miao Sun, and Yuanqing Lin · 2014
Earlier work this paper cites.
Variational autoencoder based anomaly detection using reconstruction probability
Jinwon An and Sungzoon Cho · 2015
Earlier work this paper cites.
Model-based reinforcement learning in continuous environments using real-time constrained optimization
O. Andersson, F. Heintz, and P. Doherty · 2015
Earlier work this paper cites.
Nci-isbi 2013 challenge: automated segmentation of prostate structures
N Bloch, A Madabhushi, H Huisman, J Freymann, J Kirby, M Grauer, A Enquobahrie, C Jaffe, L Clarke, and K Farahani · 2015
Earlier work this paper cites.
Multi-modality vertebra recognition in arbitrary views using 3d deformable hierarchical model
Yunliang Cai, Said Osman, Manas Sharma, Mark Landis, and Shuo Li · 2015
Earlier work this paper cites.
Active object localization with deep reinforcement learning
Juan C Caicedo and Svetlana Lazebnik · 2015
Cited alongside, same era.
Spectral–spatial classification of hyperspectral data based on deep belief network
Yushi Chen, Xing Zhao, and Xiuping Jia · 2015
Cited alongside, same era.
Near-online multi-target tracking with aggregated local flow descriptor
Wongun Choi · 2015
Cited alongside, same era.
Attention-based models for speech recognition
Jan Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, KyungHyun Cho, and Yoshua Bengio · 2015
Cited alongside, same era.
Video co-summarization: Video summarization by visual co-occurrence
Wen-Sheng Chu, Yale Song, and Alejandro Jaimes · 2015
Cited alongside, same era.
Learning spatially regularized correlation filters for visual tracking
Learning video object segmentation from static images
Federico Perazzi, Anna Khoreva, Rodrigo Benenson, Bernt Schiele, and Alexander Sorkine-Hornung · 2017
Later among the works it cites.
Tracking the untrackable: Learning to track multiple cues with long-term dependencies
Amir Sadeghian, Alexandre Alahi, and Silvio Savarese · 2017
Later among the works it cites.
Proximal Policy Optimization Algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Fast yolo: A fast you only look once system for real-time embedded object detection in video
Mohammad Javad Shafiee, Brendan Chywl, Francis Li, and Alexander Wong · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Martin Danelljan, Gustav Hager, Fahad Shahbaz Khan, and Michael Felsberg · 2015
Cited alongside, same era.
Flownet: Learning optical flow with convolutional networks
Alexey Dosovitskiy, Philipp Fischer, Eddy Ilg, Philip Hausser, Caner Hazirbas, Vladimir Golkov, Patrick Van Der Smagt, Daniel Cremers, and Thomas Brox · 2015
Cited alongside, same era.
Hierarchical recurrent neural network for skeleton based action recognition
Yong Du, Wei Wang, and Liang Wang · 2015
Cited alongside, same era.
Concurrent markov decision processes for robot team learning
Justin Girard and M Reza Emami · 2015
Cited alongside, same era.
Fast r-cnn
Ross Girshick · 2015
Cited alongside, same era.
Contextual action recognition with r* cnn
Georgia Gkioxari, Ross Girshick, and Jitendra Malik · 2015
Cited alongside, same era.
An active search strategy for efficient object class detection
Abel Gonzalez-Garcia, Alexander Vezhnevets, and Vittorio Ferrari · 2015
Cited alongside, same era.
Crest: Convolutional residual learning for visual tracking
Yibing Song, Chao Ma, Lijun Gong, Jiawei Zhang, Rynson WH Lau, and Ming-Hsuan Yang · 2017
Later among the works it cites.
Third-person imitation learning
Bradly C. Stadie, Pieter Abbeel, and Ilya Sutskever · 2017
Later among the works it cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke, and Alexander A Alemi · 2017
Later among the works it cites.
End-to-end representation learning for correlation filter based tracking
Jack Valmadre, Luca Bertinetto, Joao Henriques, Andrea Vedaldi, and Philip HS Torr · 2017
Later among the works it cites.
Sfm-net: Learning of structure and motion from video
Sudheendra Vijayanarasimhan, Susanna Ricco, Cordelia Schmid, Rahul Sukthankar, and Katerina Fragkiadaki · 2017
Later among the works it cites.
On the effects of batch and weight normalization in generative adversarial networks
Sitao Xiang and Hao Li · 2017
Later among the works it cites.
Action-decision networks for visual tracking with deep reinforcement learning
Sangdoo Yun, Jongwon Choi, Youngjoon Yoo, Kimin Yun, and Jin Young Choi · 2017
Later among the works it cites.
Deep reinforcement learning for visual object tracking in videos
Da Zhang, Hamid Maei, Xin Wang, and Yuan-Fang Wang · 2017
Later among the works it cites.
Spftn: A self-paced fine-tuning network for segmenting objects in weakly labelled videos
Dingwen Zhang, Le Yang, Deyu Meng, Dong Xu, and Junwei Han · 2017
Later among the works it cites.
Pyramid scene parsing network
Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, and Jiaya Jia · 2017
Later among the works it cites.
Scene parsing through ade20k dataset
Bolei Zhou, Hang Zhao, Xavier Puig, Sanja Fidler, Adela Barriuso, and Antonio Torralba · 2017
Later among the works it cites.
Deep feature flow for video recognition
Xizhou Zhu, Yuwen Xiong, Jifeng Dai, Lu Yuan, and Yichen Wei · 2017
Later among the works it cites.
Deep reinforcement learning with model learning and monte carlo tree search in minecraft
Stephan Alaniz · 2018
Later among the works it cites.
Iterative interaction training for segmentation editing networks
Gustav Bredell, Christine Tanner, and Ender Konukoglu · 2018
Later among the works it cites.
Real-time’actor-critic’tracking
Boyu Chen, Dong Wang, Peixia Li, Shuang Wang, and Huchuan Lu · 2018
Later among the works it cites.
Model-based reinforcement learning via meta-policy optimization
Ignasi Clavera, Jonas Rothfuss, John Schulman, Yasuhiro Fujita, Tamim Asfour, and Pieter Abbeel · 2018
Later among the works it cites.
Multi-step reinforcement learning: A unifying algorithm
Kristopher De Asis, J Fernando Hernandez-Garcia, G Zacharias Holland, and Richard S Sutton · 2018
Later among the works it cites.
An introduction to deep reinforcement learning
Vincent François-Lavet, Peter Henderson, Riashat Islam, Marc G Bellemare, and Joelle Pineau · 2018
Later among the works it cites.
Synthesizing programs for images using reinforced adversarial learning
Yaroslav Ganin, Tejas Kulkarni, Igor Babuschkin, SM Eslami, and Oriol Vinyals · 2018
Later among the works it cites.
Dynamic zoom-in network for fast object detection in large images
Mingfei Gao, Ruichi Yu, Ang Li, Vlad I Morariu, and Larry S Davis · 2018
Later among the works it cites.
Unsupervised video object segmentation for deep reinforcement learning
Vikash Goel, Jameson Weng, and Pascal Poupart · 2018
Later among the works it cites.
Dual-agent deep reinforcement learning for deformable face tracking
Minghao Guo, Jiwen Lu, and Jie Zhou · 2018
Later among the works it cites.
Meta-reinforcement learning of structured exploration strategies
Abhishek Gupta, Russell Mendonca, YuXuan Liu, Pieter Abbeel, and Sergey Levine · 2018
Later among the works it cites.
Reinforcement cutting-agent learning for video object segmentation
Junwei Han, Le Yang, Dingwen Zhang, Xiaojun Chang, and Xiaodan Liang · 2018
Later among the works it cites.
Virtual-to-real: Learning to control in visual semantic segmentation
Zhang-Wei Hong, Chen Yu-Ming, Shih-Yang Su, Tzu-Yun Shann, Yi-Hsiang Chang, Hsuan-Kung Yang, Brian Hsi-Lin Ho, Chih-Chieh Tu, Yueh-Chuan Chang, Tsu-Ching Hsiao, et al · 2018
Later among the works it cites.
Squeeze-and-excitation networks
J. Hu, L. Shen, and G. Sun · 2018
Later among the works it cites.
Efficient organ localization using multi-label convolutional neural networks in thorax-abdomen ct scans
Gabriel Efrain Humpire-Mamani, Arnaud Arindra Adiyoso Setio, Bram van Ginneken, and Colin Jacobs · 2018
Later among the works it cites.
Composition loss for counting, density map estimation and localization in dense crowds
Haroon Idrees, Muhmmad Tayyab, Kishan Athrey, Dong Zhang, Somaya Al-Maadeed, Nasir Rajpoot, and Mubarak Shah · 2018
Later among the works it cites.
Multiobject tracking in videos based on lstm and deep reinforcement learning
Ming-xin Jiang, Chao Deng, Zhi-geng Pan, Lan-fang Wang, and Xing Sun · 2018
Later among the works it cites.
The sixth visual object tracking vot2018 challenge results
Matej Kristan, Ales Leonardis, Jiri Matas, Michael Felsberg, Roman Pflugfelder, Luka Cehovin Zajc, Tomas Vojir, Goutam Bhat, Alan Lukezic, Abdelrahman Eldesokey, et al · 2018
Later among the works it cites.
Model-ensemble trust-region policy optimization
Thanard Kurutach, Ignasi Clavera, Yan Duan, Aviv Tamar, and Pieter Abbeel · 2018
Later among the works it cites.
Deep contextual recurrent residual networks for scene labeling
T Hoang Ngan Le, Chi Nhan Duong, Ligong Han, Khoa Luu, Kha Gia Quach, and Marios Savvides · 2018
Later among the works it cites.
Reformulating level sets as deep recurrent neural network approach to semantic segmentation
T Hoang Ngan Le, Kha Gia Quach, Khoa Luu, Chi Nhan Duong, and Marios Savvides · 2018
Later among the works it cites.
High performance visual tracking with siamese region proposal network
Bo Li, Junjie Yan, Wei Wu, Zheng Zhu, and Xiaolin Hu · 2018
Later among the works it cites.
Chao Li, Qiaoyong Zhong, Di Xie, and Shiliang Pu · 2018
Later among the works it cites.
Inverse reinforcement learning via function approximation for clinical motion analysis
K. Li, M. Rath, and J. W. Burdick · 2018
Later among the works it cites.
Deep visual tracking: Review and experimental comparison
Peixia Li, Dong Wang, Lijun Wang, and Huchuan Lu · 2018
Later among the works it cites.
Fast multiple landmark localisation using a patch-based iterative network
Yuanwei Li, Amir Alansary, Juan J Cerrolaza, Bishesh Khanal, Matthew Sinclair, Jacqueline Matthew, Chandni Gupta, Caroline Knight, Bernhard Kainz, and Daniel Rueckert · 2018
Later among the works it cites.
Deep reinforcement learning for surgical gesture segmentation and classification
Daochang Liu and Tingting Jiang · 2018
Later among the works it cites.
Crowd counting using deep recurrent spatial-aware network
Lingbo Liu, Hongjun Wang, Guanbin Li, Wanli Ouyang, and Liang Lin · 2018
Later among the works it cites.
Sim-to-real reinforcement learning for deformable object manipulation
Jan Matas, Stephen James, and Andrew J Davison · 2018
Later among the works it cites.
Learning to adapt in dynamic, real-world environments through meta-reinforcement learning
Anusha Nagabandi, Ignasi Clavera, Simin Liu, Ronald S Fearing, Pieter Abbeel, Sergey Levine, and Chelsea Finn · 2018
Later among the works it cites.
Overcoming exploration in reinforcement learning with demonstrations
A. Nair, B. McGrew, M. Andrychowicz, W. Zaremba, and P. Abbeel · 2018
Later among the works it cites.
Deep reinforcement learning of region proposal networks for object detection
Aleksis Pirinen and Cristian Sminchisescu · 2018
Later among the works it cites.
Yolov3: An incremental improvement
Joseph Redmon and Ali Farhadi · 2018
Later among the works it cites.
Collaborative deep reinforcement learning for multi-object tracking
Liangliang Ren, Jiwen Lu, Zifeng Wang, Qi Tian, and Jie Zhou · 2018
Later among the works it cites.
Deep reinforcement learning with iterative shift for visual tracking
Liangliang Ren, Xin Yuan, Jiwen Lu, Ming Yang, and Jie Zhou · 2018
Later among the works it cites.
Video summarization using fully convolutional sequence networks
Mrigank Rochan, Linwei Ye, and Yang Wang · 2018
Later among the works it cites.
Meta reinforcement learning with latent variable gaussian processes
Steindór Sæmundsson, Katja Hofmann, and Marc Peter Deisenroth · 2018
Later among the works it cites.
Seednet: Automatic seed generation with deep reinforcement learning for robust interactive segmentation
Gwangmo Song, Heesoo Myeong, and Kyoung Mu Lee · 2018
Later among the works it cites.
Robust multimodal image registration using deep recurrent reinforcement learning
Shanhui Sun, Jing Hu, Mingqing Yao, Jinrong Hu, Xiaodong Yang, Qi Song, and Xi Wu · 2018
Later among the works it cites.
Deep learning for biometrics: A survey
Kalaivani Sundararajan and Damon L. Woodard · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Deep progressive reinforcement learning for skeleton-based action recognition
Yansong Tang, Yi Tian, Jiwen Lu, Peiyang Li, and Jie Zhou · 2018
Later among the works it cites.
Improved image selection for stack-based hdr imaging
Peter van Beek · 2018
Later among the works it cites.
Deepigeos: a deep interactive geodesic framework for medical image segmentation
Guotai Wang, Maria A Zuluaga, Wenqi Li, Rosalind Pratt, Premal A Patel, Michael Aertsen, Tom Doel, Anna L David, Jan Deprest, Sébastien Ourselin, et al · 2018
Later among the works it cites.
Cosface: Large margin cosine loss for deep face recognition
Hao Wang, Yitong Wang, Zheng Zhou, Xing Ji, Dihong Gong, Jingchao Zhou, Zhifeng Li, and Wei Liu · 2018
Later among the works it cites.
Multitask learning for object localization with deep reinforcement learning
Yan Wang, Lei Zhang, Lituan Wang, and Zizhou Wang · 2018
Later among the works it cites.
Cbam: Convolutional block attention module
Sanghyun Woo, Jongchan Park, Joon-Young Lee, and In So Kweon · 2018
Later among the works it cites.
Dynamic video segmentation network
Yu-Syuan Xu, Tsu-Jui Fu, Hsuan-Kung Yang, and Chun-Yi Lee · 2018
Later among the works it cites.
Icnet for real-time semantic segmentation on high-resolution images
Hengshuang Zhao, Xiaojuan Qi, Xiaoyong Shen, Jianping Shi, and Jiaya Jia · 2018
Later among the works it cites.
Deep reinforcement learning for unsupervised video summarization with diversity-representativeness reward
Kaiyang Zhou, Yu Qiao, and Tao Xiang · 2018
Later among the works it cites.
Video summarisation by classification with deep reinforcement learning
Kaiyang Zhou, Tao Xiang, and Andrea Cavallaro · 2018
Later among the works it cites.
Robotics and Autonomous Systems
Advanced planning for autonomous vehicles using reinforcement learning and deep inverse reinforcement learning · 2019
Later among the works it cites.
Partial policy-based reinforcement learning for anatomical landmark localization in 3d medical images
Walid Abdullah Al and Il Dong Yun · 2019
Later among the works it cites.
Reinforcement learning-based automatic diagnosis of acute appendicitis in abdominal ct
Walid Abdullah Al, Il Dong Yun, and Kyong Joon Lee · 2019
Later among the works it cites.
Evaluating reinforcement learning agents for anatomical landmark detection
Amir Alansary, Ozan Oktay, Yuanwei Li, Loic Le Folgoc, Benjamin Hou, Ghislain Vaillant, Konstantinos Kamnitsas, Athanasios Vlontzos, Ben Glocker, Bernhard Kainz, et al · 2019
Later among the works it cites.
A general and adaptive robust loss function
Jonathan T Barron · 2019
Later among the works it cites.
Mvtec ad — a comprehensive real-world dataset for unsupervised anomaly detection
P. Bergmann, M. Fauser, D. Sattlegger, and C. Steger · 2019
Later among the works it cites.
M3d-rpn: Monocular 3d region proposal network for object detection
Garrick Brazil and Xiaoming Liu · 2019
Later among the works it cites.
Superhuman ai for multiplayer poker
Noam Brown and Tuomas Sandholm · 2019
Later among the works it cites.
Deep neural networks outperform human expert’s capacity in characterizing bioleaching bacterial biofilm composition
Antoine Buetti-Dinh, Vanni Galli, Sören Bellenberg, Olga Ilie, Malte Herold, Stephan Christel, Mariia Boretska, Igor V. Pivkin, Paul Wilmes, Wolfgang Sand, Mario Vera, and Mark Dopson · 2019
Later among the works it cites.
Deep reinforcement learning for subpixel neural tracking
Tianhong Dai, Magda Dubois, Kai Arulkumaran, Jonathan Campbell, Cher Bass, Benjamin Billot, Fatmatulzehra Uslu, Vincenzo De Paola, Claudia Clopath, and Anil Anthony Bharath · 2019
Later among the works it cites.
Arcface: Additive angular margin loss for deep face recognition
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou · 2019
Later among the works it cites.
Visual tracking by means of deep reinforcement learning and an expert demonstrator
Matteo Dunnhofer, Niki Martinel, Gian Luca Foresti, and Christian Micheloni · 2019
Later among the works it cites.
Mobiface: A lightweight deep learning face recognition on mobile devices
Chi Nhan Duong, Kha Gia Quach, Ibsa Jalata, Ngan Le, and Khoa Luu · 2019
Later among the works it cites.
Learning from longitudinal face demonstration–where tractable deep modeling meets inverse reinforcement learning
Chi Nhan Duong, Kha Gia Quach, Khoa Luu, T. Hoang Le, Marios Savvides, and Tien D. Bui · 2019
Later among the works it cites.
Lasot: A high-quality benchmark for large-scale single object tracking
Heng Fan, Liting Lin, Fan Yang, Peng Chu, Ge Deng, Sijia Yu, Hexin Bai, Yong Xu, Chunyuan Liao, and Haibin Ling · 2019
Later among the works it cites.
Got-10k: A large high-diversity benchmark for generic object tracking in the wild
Lianghua Huang, Xin Zhao, and Kaiqi Huang · 2019
Later among the works it cites.
Learning to paint with model-based deep reinforcement learning
Zhewei Huang, Wen Heng, and Shuchang Zhou · 2019
Later among the works it cites.
Multi-agent deep reinforcement learning for multi-object tracker
Mingxin Jiang, Tao Hai, Zhigeng Pan, Haiyan Wang, Yinjie Jia, and Chao Deng · 2019
Later among the works it cites.
Srm: A style-based recalibration module for convolutional neural networks
Hyunjae Lee, Hyo-Eun Kim, and Hyeonseob Nam · 2019
Later among the works it cites.
GS3D: an efficient 3d object detection framework for autonomous driving
Buyu Li, Wanli Ouyang, Lu Sheng, Xingyu Zeng, and Xiaogang Wang · 2019
Later among the works it cites.
Taming maml: Efficient unbiased meta-reinforcement learning
Hao Liu, Richard Socher, and Caiming Xiong · 2019
Later among the works it cites.
Context-aware crowd counting
Weizhe Liu, Mathieu Salzmann, and Pascal Fua · 2019
Later among the works it cites.
Biometric recognition using deep learning: A survey
Shervin Minaee, AmirAli Abdolrashidi, Hang Su, Mohammed Bennamoun, and David Zhang · 2019
Later among the works it cites.
Speech recognition using deep neural networks: A systematic review
Ali Bou Nassif, Ismail Shahin, Imtinan Attili, Mohammad Azzeh, and Khaled Shaalan · 2019
Later among the works it cites.
A review of deep learning based speech synthesis
Yishuang Ning, Sheng He, Zhiyong Wu, Chunxiao Xing, and Liang-Jie Zhang · 2019
Later among the works it cites.
Monogrnet: A geometric reasoning network for monocular 3d object localization
Zengyi Qin, Jinglu Wang, and Yan Lu · 2019
Later among the works it cites.
Efficient off-policy meta-reinforcement learning via probabilistic context variables
Kate Rakelly, Aurick Zhou, Chelsea Finn, Sergey Levine, and Deirdre Quillen · 2019
Later among the works it cites.
The carla autonomous driving challenge, 2019
German Ros, Vladfen Koltun, Felipe Codevilla, and Antonio Lopez · 2019
Later among the works it cites.
Multi-level bottom-top and top-bottom feature fusion for crowd counting
Vishwanath A Sindagi and Vishal M Patel · 2019
Later among the works it cites.
Reinforcement learning in stationary mean-field games
Jayakumar Subramanian and Aditya Mahajan · 2019
Later among the works it cites.
Generative adversarial imitation learning with deep p-network for robotic cloth manipulation
Y. Tsurumine, Y. Cui, K. Yamazaki, and T. Matsubara · 2019
Later among the works it cites.
Deep reinforcement learning with smooth policy update: Application to robotic cloth manipulation
Yoshihisa Tsurumine, Yunduan Cui, Eiji Uchibe, and Takamitsu Matsubara · 2019
Later among the works it cites.
AlphaStar: Mastering the Real-Time Strategy Game StarCraft II
Oriol Vinyals, Igor Babuschkin, Junyoung Chung, Michael Mathieu, Max Jaderberg, Wojtek Czarnecki, Andrew Dudzik, Aja Huang, Petko Georgiev, Richard Powell, Timo Ewalds, Dan Horgan, Manuel Kroiss, Ivo Danihelka, John Agapiou, Junhyuk Oh, Valentin Dalibard, David Choi, Laurent Sifre, Yury Sulsky, Sasha Vezhnevets, James Molloy, Trevor Cai, David Budden, Tom Paine, Caglar Gulcehre, Ziyu Wang, Tobias Pfaff, Toby Pohlen, Dani Yogatama, Julia Cohen, Katrina McKinney, Oliver Smith, Tom Schaul, Timothy Lillicrap, Chris Apps, Koray Kavukcuoglu, Demis Hassabis, and David Silver · 2019
Later among the works it cites.
Multiple landmark detection using multi-agent reinforcement learning
Athanasios Vlontzos, Amir Alansary, Konstantinos Kamnitsas, Daniel Rueckert, and Bernhard Kainz · 2019
Later among the works it cites.
Racial faces in the wild: Reducing racial bias by information maximization adaptation network
Mei Wang, Weihong Deng, Jiani Hu, Xunqiang Tao, and Yaohai Huang · 2019
Later among the works it cites.
Benchmarking model-based reinforcement learning
Tingwu Wang, Xuchan Bao, Ignasi Clavera, Jerrick Hoang, Yeming Wen, Eric Langlois, Shunshi Zhang, Guodong Zhang, Pieter Abbeel, and Jimmy Ba · 2019
Later among the works it cites.
From open set to closed set: Counting objects by spatial divide-and-conquer
Haipeng Xiong, Hao Lu, Chengxin Liu, Liang Liu, Zhiguo Cao, and Chunhua Shen · 2019
Later among the works it cites.
Learning adaptive discriminative correlation filters via temporal consistency preserving spatial feature selection for robust visual object tracking
Tianyang Xu, Zhen-Hua Feng, Xiao-Jun Wu, and Josef Kittler · 2019
Later among the works it cites.
Efficient multiple organ localization in ct image using 3d region proposal network
Xuanang Xu, Fugen Zhou, Bo Liu, Dongshan Fu, and Xiangzhi Bai · 2019
Later among the works it cites.
Perspective-guided convolution networks for crowd counting
Zhaoyi Yan, Yuchen Yuan, Wangmeng Zuo, Xiao Tan, Yezhen Wang, Shilei Wen, and Errui Ding · 2019
Later among the works it cites.
Reinforcement learning in healthcare: a survey
Chao Yu, Jiming Liu, and Shamim Nemati · 2019
Later among the works it cites.
Experience replay optimization
Daochen Zha, Kwei-Herng Lai, Kaixiong Zhou, and Xia Hu · 2019
Later among the works it cites.
A review of image set classification
Zhong-Qiu Zhao, Shou-Tao Xu, Dian Liu, Wei-Dong Tian, and Zhi-Da Jiang · 2019
Later among the works it cites.
Neural batch sampling with reinforcement learning for semi-supervised anomaly detection
Wen-Hsuan Chu and Kris M. Kitani · 2020
Later among the works it cites.
An empirical investigation of the challenges of real-world reinforcement learning, 2020
Gabriel Dulac-Arnold, Nir Levine, Daniel J. Mankowitz, Jerry Li, Cosmin Paduraru, Sven Gowal, and Todd Hester · 2020
Later among the works it cites.
An empirical investigation of the challenges of real-world reinforcement learning
Gabriel Dulac-Arnold, Nir Levine, Daniel J. Mankowitz, Jerry Li, Cosmin Paduraru, Sven Gowal, and Todd Hester · 2020
Later among the works it cites.
Ultrasound-guided robotic navigation with deep reinforcement learning
Hannes Hase, Mohammad Farid Azampour, Maria Tirindelli, Magdalini Paschali, Walter Simson, Emad Fatemizadeh, and Nassir Navab · 2020
Later among the works it cites.
Follow then forage exploration: Improving asynchronous advantage actor critic
James B. Holliday and Ngan T.H. Le · 2020
Later among the works it cites.
Robust automatic multiple landmark detection
Arjit Jain, Alexander Powers, and Hans J Johnson · 2020
Later among the works it cites.
Model-based reinforcement learning with value-targeted regression
Zeyu Jia, Lin Yang, Csaba Szepesvari, and Mengdi Wang · 2020
Later among the works it cites.
Offset curves loss for imbalanced problem in medical segmentation
Ngan Le, Trung Le, Kashu Yamazaki, Toan Duc Bui, Khoa Luu, and Marios Savides · 2020
Later among the works it cites.
A multi-task contextual atrous residual network for brain tumor detection & segmentation
Ngan Le, Kashu Yamazaki, Dat Truong, Kha Gia Quach, and Marios Savvides · 2020
Later among the works it cites.
Deep reinforced attention learning for quality-aware visual recognition
Duo Li and Qifeng Chen · 2020
Later among the works it cites.
Iteratively-refined interactive 3d medical image segmentation with multi-agent reinforcement learning
Xuan Liao, Wenhao Li, Qisen Xu, Xiangfeng Wang, Bo Jin, Xiaoyun Zhang, Yanfeng Wang, and Ya Zhang · 2020
Later among the works it cites.
Weighing counts: Sequential crowd counting by reinforcement learning
Liang Liu, Hao Lu, Hongwei Zou, Haipeng Xiong, Zhiguo Cao, and Chunhua Shen · 2020
Later among the works it cites.
Reinforced axial refinement network for monocular 3d object detection
Lijie Liu, Chufan Wu, Jiwen Lu, Lingxi Xie, Jie Zhou, and Qi Tian · 2020
Later among the works it cites.
Ultrasound video summarization using deep reinforcement learning
Tianrui Liu, Qingjie Meng, Athanasios Vlontzos, Jeremy Tan, Daniel Rueckert, and Bernhard Kainz · 2020
Later among the works it cites.
Deep reinforcement learning for organ localization in ct
Fernando Navarro, Anjany Sekuboyina, Diana Waldmannstetter, Jan C Peeken, Stephanie E Combs, and Bjoern H Menze · 2020
Later among the works it cites.
Refuge challenge: A unified framework for evaluating automated methods for glaucoma assessment from fundus photographs
José Ignacio Orlando, Huazhu Fu, João Barbosa Breda, Karel van Keer, Deepti R Bathula, Andrés Diaz-Pinto, Ruogu Fang, Pheng-Ann Heng, Jeyoung Kim, JoonHo Lee, et al · 2020
Later among the works it cites.
Deep model-based reinforcement learning for high-dimensional problems, a survey, 2020
Aske Plaat, Walter Kosters, and Mike Preuss · 2020
Later among the works it cites.
3d deep learning on medical images: A review, 2020
Satya P. Singh, Lipo Wang, Sukrit Gupta, Haveesh Goli, Parasuraman Padmanabhan, and Balázs Gulyás · 2020
Later among the works it cites.
Multi-step medical image segmentation based on reinforcement learning
Zhiqiang Tian, Xiangyu Si, Yaoyue Zheng, Zhang Chen, and Xiaojian Li · 2020
Later among the works it cites.
End-to-end model-free reinforcement learning for urban driving using implicit affordances
Marin Toromanoff, Emilie Wirbel, and Fabien Moutarde · 2020
Later among the works it cites.
Efficient object detection in large images using deep reinforcement learning
Burak Uzkent, Christopher Yeh, and Stefano Ermon · 2020
Later among the works it cites.
Mask-rl: Multiagent video object segmentation framework through reinforcement learning
Giuseppe Vecchio, Simone Palazzo, Daniela Giordano, Francesco Rundo, and Concetto Spampinato · 2020
Later among the works it cites.
Mitigating bias in face recognition using skewness-aware reinforcement learning
Mei Wang and Weihong Deng · 2020
Later among the works it cites.
Dynamic face video segmentation via reinforcement learning
Yujiang Wang, Mingzhi Dong, Jie Shen, Yang Wu, Shiyang Cheng, and Maja Pantic · 2020
Later among the works it cites.
Learning a reinforced agent for flexible exposure bracketing selection
Zhouxia Wang, Jiawei Zhang, Mude Lin, Jiong Wang, Ping Luo, and Jimmy Ren · 2020
Later among the works it cites.
Self-training with noisy student improves imagenet classification
Qizhe Xie, Minh-Thang Luong, Eduard Hovy, and Quoc V Le · 2020
Later among the works it cites.
Predicting goal-directed human attention using inverse reinforcement learning
Zhibo Yang, Lihan Huang, Yupei Chen, Zijun Wei, Seoyoung Ahn, Gregory Zelinsky, Dimitris Samaras, and Minh Hoai · 2020
Later among the works it cites.
Meta-world: A benchmark and evaluation for multi-task and meta reinforcement learning
Tianhe Yu, Deirdre Quillen, Zhanpeng He, Ryan Julian, Karol Hausman, Chelsea Finn, and Sergey Levine · 2020
Later among the works it cites.
Multi-modal visual tracking: Review and experimental comparison, 2020
Pengyu Zhang, Dong Wang, and Huchuan Lu · 2020
Later among the works it cites.
Roughness index and roughness distance for benchmarking medical segmentation
Vidhiwar Singh Rathour, Kashu Yamakazi, and T Le · 2021
Closest in time.
Agent-environment network for temporal action proposal generation
Kashu Yamakazi Akihiro Sugimoto Viet-Khoa Vo-Ho, Ngan T.H. Le and Triet Tran · 2021
Closest in time.
Invertible residual network with regularization for effective medical image segmentation
Kashu Yamazaki, Vidhiwar Singh Rathour, and T Le · 2021
Closest in time.