Fetching the paper…
Reading the bibliography…
This paper is about the problem of learning a stochastic policy for generating an object (like a molecular graph) from a sequence of actions, such that the probability of generating an object is proportional to a given positive reward for that object.
Gaussian processes for regression
C. K. Williams and C. Rasmussen · 1995
Earlier work this paper cites.
The properties of known drugs. 1. molecular frameworks
Guy W Bemis and Mark A Murcko · 1996
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng, Stuart J Russell, et al · 2000
Earlier work this paper cites.
A graph-based genetic algorithm and its application to the multiobjective evolution of median molecules
Nathan Brown, Ben McKay, François Gilardoni, and Johann Gasteiger · 2004
Earlier work this paper cites.
Transpositions and move groups in monte carlo tree search
Benjamin E Childs, James H Brodeur, and Levente Kocsis · 2008
Earlier work this paper cites.
Gaussian process optimization in the bandit setting: No regret and experimental design
Niranjan Srinivas, Andreas Krause, S. Kakade, and M. Seeger · 2010
Earlier work this paper cites.
Autodock vina: improving the speed and accuracy of docking with a new scoring function, efficient optimization, and multithreading
Oleg Trott and Arthur J Olson · 2010
Earlier work this paper cites.
The knowledge-gradient algorithm for sequencing experiments in drug discovery
Diana M. Negoescu, Peter I. Frazier, and Warren B. Powell · 2011
Earlier work this paper cites.
Quantifying the chemical beauty of drugs
G Richard Bickerton, Gaia V Paolini, Jérémy Besnard, Sorel Muresan, and Andrew L Hopkins · 2012
Earlier work this paper cites.
Fragment based drug design: from experimental to computational approaches
Ashutosh Kumar, A Voet, and KYJ Zhang · 2012
Earlier work this paper cites.
A capable neural network model for solving the maximum flow problem
Alireza Nazemi and Farahnaz Omidi · 2012
Earlier work this paper cites.
Data generation as sequential decision making
Philip Bachman and Doina Precup · 2015
Earlier work this paper cites.
Zinc 15–ligand discovery for everyone
Teague Sterling and John J Irwin · 2015
Earlier work this paper cites.
Variational inference with normalizing flows, 2016
Danilo Jimenez Rezende and Shakir Mohamed · 2016
Earlier work this paper cites.
Calibrating energy-based generative adversarial networks
Zihang Dai, Amjad Almahairi, Philip Bachman, Eduard Hovy, and Aaron Courville · 2017
Earlier work this paper cites.
Neural message passing for quantum chemistry, 2017
Justin Gilmer, Samuel S. Schoenholz, Patrick F. Riley, Oriol Vinyals, and George E. Dahl · 2017
Cited alongside, same era.
Reinforcement learning with deep energy-based policies
Tuomas Haarnoja, Haoran Tang, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Deep learning for network flow analysis and malware classification
RK Rahul, T Anjali, Vijay Krishna Menon, and KP Soman · 2017
Cited alongside, same era.
Proximal policy optimization algorithms, 2017
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Generating focussed molecule libraries for drug discovery with recurrent neural networks, 2017
Marwin H. S. Segler, Thierry Kogej, Christian Tyrchan, and Mark P. Waller · 2017
Cited alongside, same era.
Molgan: An implicit generative model for small molecular graphs, 2018
Discrete object generation with reversible inductive construction
Ari Seff, Wenda Zhou, Farhan Damani, Abigail Doyle, and Ryan P Adams · 2019
Later among the works it cites.
{MARS}: Markov molecular sampling for multi-objective drug discovery
Yutong Xie, Chence Shi, Hao Zhou, Yuwei Yang, Weinan Zhang, Yong Yu, and Lei Li · 2019
Later among the works it cites.
Model-based reinforcement learning for biological sequence design
Christof Angermueller, David Dohan, David Belanger, Ramya Deshpande, Kevin Murphy, and Lucy Colwell · 2020
Later among the works it cites.
Interference and generalization in temporal difference learning
Emmanuel Bengio, Joelle Pineau, and Doina Precup · 2020
Later among the works it cites.
Learning to solve network flow problems via neural decoding
Yize Chen and Baosen Zhang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nicola De Cao and Thomas Kipf · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
Approximate inference in discrete distributions with monte carlo tree search and value functions, 2019
Lars Buesing, Nicolas Heess, and Theophane Weber · 2019
Cited alongside, same era.
A graph-based genetic algorithm and generative model/monte carlo tree search for the exploration of chemical space
Jan H Jensen · 2019
Cited alongside, same era.
Batchbald: Efficient and diverse batch acquisition for deep bayesian active learning, 2019
Andreas Kirsch, Joost van Amersfoort, and Yarin Gal · 2019
Cited alongside, same era.
Autoregressive energy machines, 2019
Charlie Nash and Conor Durkan · 2019
Cited alongside, same era.
Learning discrete energy-based models via auxiliary-variable local exploration
Hanjun Dai, Rishabh Singh, Bo Dai, Charles Sutton, and Dale Schuurmans · 2020
Later among the works it cites.
Learning to navigate the synthetically accessible chemical space using reinforcement learning, 2020
Sai Krishna Gottipati, Boris Sattarov, Sufeng Niu, Yashaswi Pathak, Haoran Wei, Shengchao Liu, Karam M. J. Thomas, Simon Blackburn, Connor W. Coley, Jian Tang, Sarath Chandar, and Yoshua Bengio · 2020
Later among the works it cites.
Implicit under-parameterization inhibits data-efficient deep reinforcement learning, 2020
Aviral Kumar, Rishabh Agarwal, Dibya Ghosh, and Sergey Levine · 2020
Later among the works it cites.
Graphaf: a flow-based autoregressive model for molecular graph generation, 2020
Chence Shi, Minkai Xu, Zhaocheng Zhu, Weinan Zhang, Ming Zhang, and Jian Tang · 2020
Later among the works it cites.
Amortized bayesian optimization over discrete spaces
Kevin Swersky, Yulia Rubanova, David Dohan, and Kevin Murphy · 2020
Later among the works it cites.
Oops i took a gradient: Scalable sampling for discrete distributions, 2021
Will Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud, and Chris J. Maddison · 2021
Closest in time.
Graphdf: A discrete flow model for molecular graph generation, 2021
Youzhi Luo, Keqiang Yan, and Shuiwang Ji · 2021
Closest in time.
Masked label prediction: Unified message passing model for semi-supervised classification, 2021
Yunsheng Shi, Zhengjie Huang, Shikun Feng, Hui Zhong, Wenjin Wang, and Yu Sun · 2021
Closest in time.
Chapter 11. junction tree variational autoencoder for molecular graph generation
Wengong Jin, Regina Barzilay, and Tommi Jaakkola · 2041
Closest in time.