Fetching the paper…
Reading the bibliography…
Reinforcement learning (RL) over text representations can be effective for finding high-value policies that can search over graphs.
Learning Internal Representations by Error Propagation , pages 318–362
David E. Rumelhart and James L. McClelland · 1987
Earlier work this paper cites.
Smiles, a chemical language and information system. 1. introduction to methodology and encoding rules
David Weininger · 1988
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S Sutton, David McAllester, Satinder Singh, and Yishay Mansour · 1999
Earlier work this paper cites.
A neural probabilistic language model
Yoshua Bengio, Réjean Ducharme, Pascal Vincent, and Christian Janvin · 2003
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning, 2011
Stephane Ross, Geoffrey J. Gordon, and J. Andrew Bagnell · 2011
Earlier work this paper cites.
Chembl: A large-scale bioactivity database for drug discovery
Anna Gaulton, Louisa J. Bellis, A. Patricia Bento, Jon Chambers, Mark Davies, Anne Hersey, Yvonne Light, Shaun McGlinchey, David Michalovich, Bissan Al-Lazikani, and John P. Overington · 2012
Earlier work this paper cites.
Zinc: a free tool to discover chemistry for biology
John J Irwin, Teague Sterling, Michael M Mysinger, Erin S Bolstad, and Ryan G Coleman · 2012
Earlier work this paper cites.
Inchi-the worldwide chemical structure identifier standard
Stephen Heller, Alan McNaught, Stephen Stein, Dmitrii Tchekhovskoi, and Igor Pletnev · 2013
Earlier work this paper cites.
Rdkit: A software suite for cheminformatics, computational chemistry, and predictive modeling
Greg Landrum et al · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning, 2013
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Earlier work this paper cites.
Fast, accurate, and reliable molecular docking with QuickVina 2
Amr Alhossary, Stephanus Daniel Handoko, Yuguang Mu, and Chee-Keong Kwoh · 2015
Earlier work this paper cites.
Concrete problems in ai safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané · 2016
Earlier work this paper cites.
Generative adversarial imitation learning, 2016
Jonathan Ho and Stefano Ermon · 2016
Earlier work this paper cites.
Chemgan challenge for drug discovery: can ai reproduce natural chemical diversity?, 2017
Mostapha Benhenda · 2017
Earlier work this paper cites.
Molecular generation with recurrent neural networks (rnns)
Esben Jannik Bjerrum and Richard Threlfall · 2017
Earlier work this paper cites.
Sgdr: Stochastic gradient descent with warm restarts, 2017
Ilya Loshchilov and Frank Hutter · 2017
Earlier work this paper cites.
Molecular de novo design through deep reinforcement learning, 2017
Marcus Olivecrona, Thomas Blaschke, Ola Engkvist, and Hongming Chen · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms, 2017
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Automatic chemical design using a data-driven continuous representation of molecules
Rafael Gó mez-Bombarelli, Jennifer N. Wei, David Duvenaud, José Miguel Hernández-Lobato, Benjamín Sánchez-Lengeling, Dennis Sheberla, Jorge Aguilera-Iparraguirre, Timothy D. Hirzel, Ryan P. Adams, and Alán Aspuru-Guzik · 2018
Earlier work this paper cites.
Automatic chemical design using a data-driven continuous representation of molecules
Rafael Gómez-Bombarelli, Jennifer N Wei, David Duvenaud, José Miguel Hernández-Lobato, Benjamín Sánchez-Lengeling, Dennis Sheberla, Jorge Aguilera-Iparraguirre, Timothy D Hirzel, Ryan P Adams, and Alán Aspuru-Guzik · 2018
Earlier work this paper cites.
Generative recurrent networks for de novo drug design
Anvita Gupta, Alex T Müller, Berend JH Huisman, Jens A Fuchs, Petra Schneider, and Gisbert Schneider · 2018
Earlier work this paper cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor, 2018
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Earlier work this paper cites.
Junction tree variational autoencoder for molecular graph generation, 2018
Wengong Jin, Regina Barzilay, and Tommi Jaakkola · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever · 2018
Earlier work this paper cites.
Graph convolutional policy network for goal-directed molecular graph generation, 2018
Jiaxuan You, Bowen Liu, Rex Ying, Vijay Pande, and Jure Leskovec · 2018
Earlier work this paper cites.
Policy transfer with strategy optimization, 2018
Wenhao Yu, C. Karen Liu, and Greg Turk · 2018
Cited alongside, same era.
All smiles variational autoencoder
Zaccary Alperstein, Artem Cherkasov, and Jason Tyler Rolfe · 2019
Cited alongside, same era.
GuacaMol: Benchmarking models for de novo molecular design
Nathan Brown, Marco Fiscato, Marwin H.S. Segler, and Alain C. Vaucher · 2019
Cited alongside, same era.
Towards safe artificial general intelligence
Tom Everitt · 2019
Cited alongside, same era.
Reinforcement learning for improving agent design
David Ha · 2019
Cited alongside, same era.
Chembl: towards direct deposition of bioassay data
David Mendez, Anna Gaulton, A Patrícia Bento, Jon Chambers, Marleen De Veij, Eloy Félix, María Paula Magariños, Juan F Mosquera, Prudence Mutowo, Michał Nowotka, et al · 2019
Janus: Parallel tempered genetic algorithm guided by deep neural networks for inverse molecular design, 2021
AkshatKumar Nigam, Robert Pollice, and Alan Aspuru-Guzik · 2021
Later among the works it cites.
Deep molecular dreaming: inverse machine learning for de-novo molecular design and interpretability with surjective representations
Cynthia Shen, Mario Krenn, Sagi Eppel, and Alá n Aspuru-Guzik · 2021
Later among the works it cites.
Deep reinforcement learning for transportation network combinatorial optimization: A survey
Qi Wang and Chunlei Tang · 2021
Later among the works it cites.
Hit and lead discovery with explorative rl and fragment-based molecule generation, 2021
Soojung Yang, Doyeong Hwang, Seul Lee, Seongok Ryu, and Sung Ju Hwang · 2021
Later among the works it cites.
Mastering visual continuous control: Improved data-augmented reinforcement learning, 2021
Denis Yarats, Rob Fergus, Alessandro Lazaric, and Lerrel Pinto · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Stabilizing transformers for reinforcement learning, 2019
Emilio Parisotto, H. Francis Song, Jack W. Rae, Razvan Pascanu, Caglar Gulcehre, Siddhant M. Jayakumar, Max Jaderberg, Raphael Lopez Kaufman, Aidan Clark, Seb Noury, Matthew M. Botvinick, Nicolas Heess, and Raia Hadsell · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library, 2019
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Köpf, Edward Yang, Zach DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala · 2019
Cited alongside, same era.
Jointly learning to construct and control agents using deep reinforcement learning
Charles Schaff, David Yunis, Ayan Chakrabarti, and Matthew R Walter · 2019
Cited alongside, same era.
Recurrent neural networks (rnns): A gentle introduction and overview
Robin M Schmidt · 2019
Cited alongside, same era.
Deep representation learning for social network analysis
Qiaoyu Tan, Ninghao Liu, and Xia Hu · 2019
Cited alongside, same era.
Smiles-bert: large scale unsupervised pre-training for molecular property prediction
Sheng Wang, Yuzhi Guo, Yuhong Wang, Hongmao Sun, and Junzhou Huang · 2019
Cited alongside, same era.
Bag of tricks for training deeper graph neural networks: A comprehensive benchmark study
Tianlong Chen, Kaixiong Zhou, Keyu Duan, Wenqing Zheng, Peihao Wang, Xia Hu, and Zhangyang Wang · 2022
Later among the works it cites.
Benchmarking graph neural networks, 2022
Vijay Prakash Dwivedi, Chaitanya K. Joshi, Anh Tuan Luu, Thomas Laurent, Yoshua Bengio, and Xavier Bresson · 2022
Later among the works it cites.
Translation between molecules and natural language
Carl Edwards, Tuan Lai, Kevin Ros, Garrett Honke, Kyunghyun Cho, and Heng Ji · 2022
Later among the works it cites.
Language models can learn complex molecular distributions
Daniel Flam-Shepherd, Kevin Zhu, and Alán Aspuru-Guzik · 2022
Later among the works it cites.
Sample efficiency matters: A benchmark for practical molecular optimization, 2022
Wenhao Gao, Tianfan Fu, Jimeng Sun, and Connor W. Coley · 2022
Later among the works it cites.
Dockstring: easy molecular docking yields better benchmarks for ligand design
Miguel García-Ortegón, Gregor NC Simm, Austin J Tripp, José Miguel Hernández-Lobato, Andreas Bender, and Sergio Bacallado · 2022
Later among the works it cites.
Artificial intelligence foundation for therapeutic science
Kexin Huang, Tianfan Fu, Wenhao Gao, Yue Zhao, Yusuf Roohani, Jure Leskovec, Connor W Coley, Cao Xiao, Jimeng Sun, and Marinka Zitnik · 2022
Later among the works it cites.
Chemformer: a pre-trained transformer for computational chemistry
Ross Irwin, Spyridon Dimitriadis, Jiazhen He, and Esben Jannik Bjerrum · 2022
Later among the works it cites.
Selfies and the future of molecular string representations
Mario Krenn, Qianxiang Ai, Senja Barthel, Nessa Carson, Angelo Frei, Nathan C Frey, Pascal Friederich, Théophile Gaudin, Alberto Alexander Gayle, Kevin Maik Jablonka, et al · 2022
Later among the works it cites.
Data-driven offline optimization for architecting hardware accelerators, 2022
Aviral Kumar, Amir Yazdanbakhsh, Milad Hashemi, Kevin Swersky, and Sergey Levine · 2022
Later among the works it cites.
Neuroevolution-enhanced multi-objective optimization for mixed-precision quantization
Santiago Miret, Vui Seng Chua, Mattias Marder, Mariano Phiellip, Nilesh Jain, and Somdeb Majumdar · 2022
Later among the works it cites.
Defining and characterizing reward hacking
Joar Skalse, Nikolaus HR Howe, Dmitrii Krasheninnikov, and David Krueger · 2022
Later among the works it cites.
Galactica: A large language model for science
Ross Taylor, Marcin Kardas, Guillem Cucurull, Thomas Scialom, Anthony Hartshorn, Elvis Saravia, Andrew Poulton, Viktor Kerkez, and Robert Stojnic · 2022
Later among the works it cites.
Augmented hill-climb increases reinforcement learning efficiency for language-based de novo molecule generation
Morgan Thomas, Noel O’Boyle, Andreas Bender, and Chris Graaf · 2022
Later among the works it cites.
An evaluation framework for the objective functions of de novo drug design benchmarks
Austin Tripp, Wenlin Chen, and José Miguel Hernández-Lobato · 2022
Later among the works it cites.
A deep-learning system bridging molecule structure and biomedical text with comprehension comparable to human professionals
Zhenni Zeng, Yuan Yao, Zhiyuan Liu, and Maosong Sun · 2022
Later among the works it cites.
Faster and more diverse de novo molecular optimization with double-loop reinforcement learning using augmented smiles, 2023
Esben Jannik Bjerrum, Christian Margreitter, Thomas Blaschke, and Raquel López-Ríos de Castro · 2023
Closest in time.
Group selfies: a robust fragment-based molecular string representation
Austin H Cheng, Andy Cai, Santiago Miret, Gustavo Malkomes, Mariano Phielipp, and Alán Aspuru-Guzik · 2023
Closest in time.
Proto-value networks: Scaling representation learning with auxiliary tasks, 2023
Jesse Farebrother, Joshua Greaves, Rishabh Agarwal, Charline Le Lan, Ross Goroshin, Pablo Samuel Castro, and Marc G. Bellemare · 2023
Closest in time.
Robustness of graph neural networks at scale, 2023
Simon Geisler, Tobias Schmidt, Hakan Şirin, Daniel Zügner, Aleksandar Bojchevski, and Stephan Günnemann · 2023
Closest in time.
Offline q-learning on diverse multi-task data both scales and generalizes, 2023
Aviral Kumar, Rishabh Agarwal, Xinyang Geng, George Tucker, and Sergey Levine · 2023
Closest in time.
Exploring chemical space with score-based out-of-distribution generation, 2023
Seul Lee, Jaehyeong Jo, and Sung Ju Hwang · 2023
Closest in time.