Fetching the paper…
Reading the bibliography…
Graph mining tasks arise from many different application domains, ranging from social networks, transportation to E-commerce, etc., which have been receiving great attention from the theoretical and algorithmic design communities in recent years, and there has been some pioneering work employing the research-rich Reinforcement Learning (RL) techniques to address graph data mining tasks.
G. Wan, S. Pan, C. Gong, C. Zhou, and G. Haffari, “Reasoning like human: Hierarchical reinforcement learning for knowledge graph reasoning,” in Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence , Yokohama, Yokohama, Japan, 2021, pp. 1926–1932
1932
Earlier work this paper cites.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Harley, T. P. Lillicrap, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in Proceedings of the 33rd International Conference on International Conference on Machine Learning , vol. 48, 2016, pp. 1928–1937
1937
Earlier work this paper cites.
A. K. Debnath, R. L. Lopez de Compadre, G. Debnath, A. J. Shusterman, and C. Hansch, “Structure-activity relationship of mutagenic aromatic and heteroaromatic nitro compounds. correlation with molecular orbital energies and hydrophobicity,” Journal of medicinal chemistry , vol. 34, no. 2, pp. 786–797, 1991
1991
Earlier work this paper cites.
C. J. Watkins and P. Dayan, “Q-learning,” Machine learning , vol. 8, no. 3, pp. 279–292, 1992
1992
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3, pp. 229–256, 1992
1992
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, no. 8, pp. 1735–1780, 1997
1997
Earlier work this paper cites.
V. R. Konda and J. N. Tsitsiklis, “Actor-critic algorithms,” in Advances in neural information processing systems , vol. 12, Denver, Colorado, USA, 1999, pp. 1008–1014
1999
Earlier work this paper cites.
R. S. Sutton, D. McAllester, S. Singh, and Y. Mansour, “Policy gradient methods for reinforcement learning with function approximation,” in Proceedings of the 12th International Conference on Neural Information Processing Systems , Denver, CO, 1999, pp. 1057–1063
1999
Earlier work this paper cites.
A. L. Goldberger, L. A. Amaral, L. Glass, J. M. Hausdorff, P. C. Ivanov, R. G. Mark, J. E. Mietus, G. B. Moody, C.-K. Peng, and H. E. Stanley, “Physiobank, physiotoolkit, and physionet: components of a new research resource for complex physiologic signals,” circulation , vol. 101, no. 23, pp. e215–e220, 2000
2000
Earlier work this paper cites.
Z. Wang, T. Schaul, M. Hessel, H. Hasselt, M. Lanctot, and N. Freitas, “Dueling network architectures for deep reinforcement learning,” in Proceedings of The 33rd International Conference on Machine Learning , vol. 48, New York, New York, USA, 20–22 Jun 2016, pp. 1995–2003
2003
Earlier work this paper cites.
H. Toivonen, A. Srinivasan, R. D. King, S. Kramer, and C. Helma, “Statistical evaluation of the predictive toxicology challenge 2000-2001,” Bioinformatics , vol. 19, no. 10, pp. 1183–1193, 2003
2003
Earlier work this paper cites.
P. D. Dobson and A. J. Doig, “Distinguishing enzyme structures from non-enzymes without alignments,” Journal of molecular biology , vol. 330, no. 4, pp. 771–783, 2003
2003
Earlier work this paper cites.
K. M. Borgwardt, C. S. Ong, S. Schönauer, S. Vishwanathan, A. J. Smola, and H.-P. Kriegel, “Protein function prediction via graph kernels,” Bioinformatics , vol. 21, pp. i47–i56, 2005
2005
Earlier work this paper cites.
S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z. Ives, “Dbpedia: A nucleus for a web of open data,” in The semantic web , 2007, pp. 722–735
2007
Earlier work this paper cites.
A. Mislove, M. Marcon, K. P. Gummadi, P. Druschel, and B. Bhattacharjee, “Measurement and analysis of online social networks,” in Proceedings of the 7th ACM SIGCOMM Conference on Internet Measurement , San Diego, California, USA, 2007, pp. 29–42
2007
Earlier work this paper cites.
J. Leskovec, J. Kleinberg, and C. Faloutsos, “Graph evolution: Densification and shrinking diameters,” ACM transactions on Knowledge Discovery from Data (TKDD) , vol. 1, no. 1, pp. 2–es, 2007
2007
Earlier work this paper cites.
K. Bollacker, C. Evans, P. Paritosh, T. Sturge, and J. Taylor, “Freebase: A collaboratively created graph database for structuring human knowledge,” in Proceedings of the 2008 ACM SIGMOD International Conference on Management of Data , New York, NY, USA, 2008, pp. 1247–1250
2008
Earlier work this paper cites.
P. Sen, G. Namata, M. Bilgic, L. Getoor, B. Galligher, and T. Eliassi-Rad, “Collective classification in network data,” AI magazine , vol. 29, no. 3, pp. 93–93, 2008
2008
Earlier work this paper cites.
N. Wale, I. A. Watson, and G. Karypis, “Comparison of descriptor spaces for chemical compound retrieval and classification,” Knowledge and Information Systems , vol. 14, no. 3, pp. 347–375, 2008
2008
Earlier work this paper cites.
N. Shervashidze, P. Schweitzer, E. J. van Leeuwen, K. Mehlhorn, and K. M. Borgwardt, “Weisfeiler-lehman graph kernels,” The Journal of Machine Learning Research , vol. 12, pp. 2539–2561, 2011
2011
Earlier work this paper cites.
W.-t. Yih, K. Toutanova, J. C. Platt, and C. Meek, “Learning discriminative projections for text similarity measures,” in Proceedings of the Fifteenth Conference on Computational Natural Language Learning , Portland, Oregon, 2011, pp. 247–256
2011
Earlier work this paper cites.
J. Leskovec and J. Mcauley, “Learning to discover social circles in ego networks,” in Advances in Neural Information Processing Systems , F. Pereira, C. Burges, L. Bottou, and K. Weinberger, Eds., vol. 25, 2012
2012
Earlier work this paper cites.
M. Gardner, P. Talukdar, B. Kisiel, and T. Mitchell, “Improving learning and inference in a large knowledge-base using latent syntactic cues,” in Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing , Seattle, Washington, USA, 2013, pp. 833–838
2013
Earlier work this paper cites.
A. Bordes, N. Usunier, A. Garcia-Durán, J. Weston, and O. Yakhnenko, “Translating embeddings for modeling multi-relational data,” in Proceedings of the 26th International Conference on Neural Information Processing Systems , vol. 2, Lake Tahoe, Nevada, 2013, pp. 2787–2795
2013
Earlier work this paper cites.
A. Mukherjee, V. Venkataraman, B. Liu, and N. Glance, “What yelp fake review filter might be doing?” in Proceedings of the International AAAI Conference on Web and Social Media , vol. 7, 2013, pp. 409–418
2013
Earlier work this paper cites.
J. J. McAuley and J. Leskovec, “From amateurs to connoisseurs: modeling the evolution of user expertise through online reviews,” in Proceedings of the 22nd international conference on World Wide Web , 2013, pp. 897–908
2013
Earlier work this paper cites.
H. Kawano, “Hierarchical sub-task decomposition for reinforcement learning of multi-robot delivery mission,” in Proceedings of the IEEE International Conference on Robotics and Automation , 2013, pp. 828–835
2013
Earlier work this paper cites.
Z. Wang, J. Zhang, J. Feng, and Z. Chen, “Knowledge graph embedding by translating on hyperplanes,” in Proceedings of the Twenty-Eighth AAAI Conference on Artificial Intelligence , Québec City, Québec, Canada, 2014, pp. 1112–1119
2014
Earlier work this paper cites.
D. Silver, G. Lever, N. Heess, T. Degris, D. Wierstra, and M. Riedmiller, “Deterministic policy gradient algorithms,” in Proceedings of the 31st International Conference on International Conference on Machine Learning , vol. 32, 2014, pp. I–387–I–395
2014
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
F. Mahdisoltani, J. Biega, and F. Suchanek, “Yago3: A knowledge base from multilingual wikipedias,” in 7th biennial conference on innovative data systems research , Asilomar, California, USA, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
K. Toutanova, D. Chen, P. Pantel, H. Poon, P. Choudhury, and M. Gamon, “Representing text for joint embedding of text and knowledge bases,” in Proceedings of the 2015 conference on empirical methods in natural language processing , 2015, pp. 1499–1509
2015
Earlier work this paper cites.
Y. Lin, Z. Liu, M. Sun, Y. Liu, and X. Zhu, “Learning entity and relation embeddings for knowledge graph completion,” in Proceedings of the Twenty-Ninth AAAI Conference on Artificial Intelligence , Austin, Texas, 2015, pp. 2181–2187
2015
Earlier work this paper cites.
Y. Duan, X. Chen, R. Houthooft, J. Schulman, and P. Abbeel, “Benchmarking deep reinforcement learning for continuous control,” in International conference on machine learning , 2016, pp. 1329–1338
2016
Earlier work this paper cites.
Y. Deng, F. Bao, Y. Kong, Z. Ren, and Q. Dai, “Deep direct reinforcement learning for financial signal representation and trading,” IEEE transactions on neural networks and learning systems , vol. 28, no. 3, pp. 653–664, 2016
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, and M. Lanctot, “Mastering the game of go with deep neural networks and tree search,” nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
J. Foerster, I. A. Assael, N. De Freitas, and S. Whiteson, “Learning to communicate with deep multi-agent reinforcement learning,” in Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
S. Sukhbaatar and R. Fergus, “Learning multiagent communication with backpropagation,” Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
W. Xiong, T. Hoang, and W. Y. Wang, “Deeppath: A reinforcement learning method for knowledge graph reasoning,” in Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing , Copenhagen, Denmark, 2017, pp. 564–573
2017
Earlier work this paper cites.
J. B. Heaton, N. G. Polson, and J. H. Witte, “Deep learning for finance: deep portfolios,” Applied Stochastic Models in Business and Industry , vol. 33, no. 1, pp. 3–12, 2017
2017
Earlier work this paper cites.
W. L. Hamilton, R. Ying, and J. Leskovec, “Inductive representation learning on large graphs,” in Proceedings of the 31st International Conference on Neural Information Processing Systems , Long Beach, California, USA, 2017, pp. 1025–1035
2017
Earlier work this paper cites.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. J. Rennie, E. Marcheret, Y. Mroueh, J. Ross, and V. Goel, “Self-critical sequence training for image captioning,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2017, pp. 1179–1195
2017
Earlier work this paper cites.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, Y. Chen, T. Lillicrap, F. Hui, L. Sifre, G. van den Driessche, T. Graepel, and D. Hassabis, “Mastering the game of go without human knowledge,” Nature , vol. 550, no. 7676, pp. 354–359, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
C. Tessler, S. Givony, T. Zahavy, D. Mankowitz, and S. Mannor, “A deep hierarchical approach to lifelong learning in minecraft,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 31, 2017
2017
Earlier work this paper cites.
A. S. Vezhnevets, S. Osindero, T. Schaul, N. Heess, M. Jaderberg, D. Silver, and K. Kavukcuoglu, “Feudal networks for hierarchical reinforcement learning,” in Proceedings of the 34th International Conference on Machine Learning , vol. 70, 2017, pp. 3540–3549
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
H. Van Seijen, M. Fatemi, J. Romoff, R. Laroche, T. Barnes, and J. Tsang, “Hybrid reward architecture for reinforcement learning,” in Proceedings of the Advances in Neural Information Processing Systems , vol. 30, 2017
2017
Earlier work this paper cites.
S. Levine, P. Pastor, A. Krizhevsky, J. Ibarz, and D. Quillen, “Learning hand-eye coordination for robotic grasping with deep learning and large-scale data collection,” The International journal of robotics research , vol. 37, no. 4-5, pp. 421–436, 2018
2018
Earlier work this paper cites.
M. Obara, T. Kashiyama, and Y. Sekimoto, “Deep reinforcement learning approach for train rescheduling utilizing graph theory,” in 2018 IEEE International Conference on Big Data (Big Data) , 2018, pp. 4525–4533
2018
Earlier work this paper cites.
M. Zhang and Y. Chen, “Link prediction based on graph neural networks,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
H. Dai, H. Li, T. Tian, X. Huang, L. Wang, J. Zhu, and L. Song, “Adversarial attack on graph structured data,” in Proceedings of the 35th International Conference on Machine Learning , vol. 80, 2018, pp. 1115–1124
2018
Earlier work this paper cites.
K. Zhang, Z. Yang, H. Liu, T. Zhang, and T. Basar, “Fully decentralized multi-agent reinforcement learning with networked agents,” in International Conference on Machine Learning , vol. 80, 2018, pp. 5872–5881
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
T. Wang, R. Liao, J. Ba, and S. Fidler, “Nervenet: Learning structured policy with graph neural networks,” in International conference on learning representations , 2018
2018
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press, 2018
2018
Earlier work this paper cites.
T. Dettmers, P. Minervini, P. Stenetorp, and S. Riedel, “Convolutional 2d knowledge graph embeddings,” in Proceedings of the Thirty-Second AAAI Conference on Artificial Intelligence and Thirtieth Innovative Applications of Artificial Intelligence Conference and Eighth AAAI Symposium on Educational Advances in Artificial Intelligence , New Orleans, Louisiana, USA, 2018
2018
Earlier work this paper cites.
Z. Li, X. Jin, S. Guan, Y. Wang, and X. Cheng, “Path reasoning over knowledge graph: A multi-agent and reinforcement learning based method,” in 2018 IEEE International Conference on Data Mining Workshops (ICDMW) , 2018, pp. 929–936
2018
Earlier work this paper cites.
X. V. Lin, R. Socher, and C. Xiong, “Multi-hop knowledge graph reasoning with reward shaping,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing , Brussels, Belgium, 2018, pp. 3243–3253
2018
Earlier work this paper cites.
L. Cai and W. Y. Wang, “Kbgan: Adversarial learning for knowledge graph embeddings,” in Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) , 2018, pp. 1470–1480
2018
Earlier work this paper cites.
Z. Wu, B. Ramsundar, E. N. Feinberg, J. Gomes, C. Geniesse, A. S. Pappu, K. Leswing, and V. Pande, “Moleculenet: a benchmark for molecular machine learning,” Chemical science , vol. 9, no. 2, pp. 513–530, 2018
2018
Cited alongside, same era.
H. Gao, Z. Wang, and S. Ji, “Large-scale learnable graph convolutional networks,” in Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , London, United Kingdom, 2018, pp. 1416–1424
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Y. Khemchandani, S. O’Hagan, S. Samanta, N. Swainston, T. J. Roberts, D. Bollegala, and D. B. Kell, “Deepgraphmolgen, a multi-objective, computational strategy for generating molecules with desirable properties: a graph convolution and reinforcement learning approach,” Journal of cheminformatics , vol. 12, no. 1, pp. 1–17, 2020
2020
Later among the works it cites.
W. Jin, R. Barzilay, and T. Jaakkola, “Multi-objective molecule generation using interpretable substructures,” in International conference on machine learning , vol. 119, 2020, pp. 4849–4859
2020
Later among the works it cites.
H. Wang, K. Wang, J. Yang, L. Shen, N. Sun, H. S. Lee, and S. Han, “Gcn-rl circuit designer: Transferable transistor sizing with graph neural networks and reinforcement learning,” in Proceedings of the 57th ACM/IEEE Design Automation Conference (DAC) , 2020, pp. 1–6
2020
Later among the works it cites.
C. Seshadhri, A. Sharma, A. Stolman, and A. Goel, “The impossibility of low-rank representations for triangle-rich complex networks,” Proceedings of the National Academy of Sciences , vol. 117, no. 11, pp. 5631–5637, 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. B. Lee, R. Rossi, and X. Kong, “Graph classification using structural attention,” in Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , London, United Kingdom, 2018, pp. 1666–1674
2018
Cited alongside, same era.
D. Seyler, P. Chandar, and M. Davis, “An information retrieval framework for contextual suggestion based on heterogeneous information network embeddings,” in The 41st International ACM SIGIR Conference on Research & Development in Information Retrieval , Ann Arbor, MI, USA, 2018, pp. 953–956
2018
Cited alongside, same era.
C. Yang, Y. Feng, P. Li, Y. Shi, and J. Han, “Meta-graph based hin spectral embedding: Methods, analyses, and insights,” in 2018 IEEE International Conference on Data Mining (ICDM) , 2018, pp. 657–666
2018
Cited alongside, same era.
D. Zügner, A. Akbarnejad, and S. Günnemann, “Adversarial attacks on neural networks for graph data,” in Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , London, United Kingdom, 2018, pp. 2847–2856
2018
Cited alongside, same era.
L. Chen, B. Tan, S. Long, and K. Yu, “Structured dialogue policy with graph neural networks,” in Proceedings of the 27th International Conference on Computational Linguistics , Santa Fe, New Mexico, USA, 2018, pp. 1257–1268
2018
Cited alongside, same era.
X. Zhao, L. Zhang, Z. Ding, L. Xia, J. Tang, and D. Yin, “Recommendations with negative feedback via pairwise deep reinforcement learning,” in Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining , London, United Kingdom, 2018, pp. 1040–1048
2018
Cited alongside, same era.
M. Yousefi, N. Mtetwa, Y. Zhang, and H. Tianfield, “A reinforcement learning approach for attack graph analysis,” in 2018 17th IEEE International Conference On Trust, Security And Privacy In Computing And Communications/ 12th IEEE International Conference On Big Data Science And Engineering (TrustCom/BigDataSE) , 2018, pp. 212–217
2018
Cited alongside, same era.
T. Nishi, K. Otaki, K. Hayakawa, and T. Yoshimura, “Traffic signal control based on reinforcement learning with graph convolutional neural nets,” in Proceedings of the 2018 21st International Conference on Intelligent Transportation Systems (ITSC) , 2018, pp. 877–883
2018
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
F. Soleymani and E. Paquet, “Deep graph convolutional reinforcement learning for financial portfolio management-deeppocket,” Expert Systems with Applications , vol. 182, pp. 115–127, 2021
2021
Later among the works it cites.
H. Peng, B. Du, M. Liu, M. Liu, S. Ji, S. Wang, X. Zhang, and L. He, “Dynamic graph convolutional network for long-term traffic flow prediction with reinforcement learning,” Information Sciences , vol. 578, pp. 401–416, 2021
2021
Later among the works it cites.
E. Meirom, H. Maron, S. Mannor, and G. Chechik, “Controlling graph dynamics with reinforcement learning and graph neural networks,” in International Conference on Machine Learning , vol. 139, 2021, pp. 7565–7577
2021
Later among the works it cites.
2021
Later among the works it cites.
S. Yu, A. Mazaheri, and A. Jannesari, “Auto graph encoder-decoder for neural network pruning,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 6362–6372
2021
Later among the works it cites.
H. Yuan, H. Yu, J. Wang, K. Li, and S. Ji, “On explainability of graph neural networks via subgraph explorations,” in Proceedings of the 38th International Conference on Machine Learning , vol. 139, 2021, pp. 12 241–12 252
2021
Later among the works it cites.
2021
Later among the works it cites.
Z. Xi, R. Pang, S. Ji, and T. Wang, “Graph backdoor,” in 30th USENIX Security Symposium (USENIX Security 21) , 2021, pp. 1523–1540
2021
Later among the works it cites.
Z. Zhang, J. Jia, B. Wang, and N. Z. Gong, “Backdoor attacks to graph neural networks,” in Proceedings of the 26th ACM Symposium on Access Control Models and Technologies , New York, NY, USA, 2021, pp. 15–26
2021
Later among the works it cites.
X. Zhao, L. Chen, and H. Chen, “A weighted heterogeneous graph-based dialog system,” IEEE Transactions on Neural Networks and Learning Systems , pp. 1–6, 2021
2021
Later among the works it cites.
H. Peng, R. Zhang, Y. Dou, R. Yang, J. Zhang, and P. S. Yu, “Reinforced neighborhood selection guided multi-relational graph neural networks,” ACM Transactions on Information Systems (TOIS) , vol. 40, no. 4, pp. 1–46, 2021
2021
Later among the works it cites.
Y. Deng, Y. Li, F. Sun, B. Ding, and W. Lam, “Unified conversational recommendation policy learning via graph-based reinforcement learning,” in Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval , 2021, pp. 1431–1441
2021
Later among the works it cites.
C. Wang, Y. Liu, X. Gao, and G. Chen, “A reinforcement learning model for influence maximization in social networks,” in International Conference on Database Systems for Advanced Applications , Cham, 2021, pp. 701–709
2021
Later among the works it cites.
X. Zhou, P. Wang, Q. Luo, and Z. Pan, “Multi-hop knowledge graph reasoning based on hyperbolic knowledge graph embedding and reinforcement learning,” in The 10th International Joint Conference on Knowledge Graphs , Virtual Event, Thailand, 2021, pp. 1–9
2021
Later among the works it cites.
M. Zheng, Y. Zhou, and Q. Cui, “Hierarchical policy network with multi-agent for knowledge graph reasoning based on reinforcement learning,” in International Conference on Knowledge Science, Engineering and Management , Cham, 2021, pp. 445–457
2021
Later among the works it cites.
P. Tiwari, H. Zhu, and H. M. Pandey, “Dapath: Distance-aware knowledge graph reasoning based on deep reinforcement learning,” Neural Networks , vol. 135, pp. 1–12, 2021
2021
Later among the works it cites.
S. Li, H. Wang, R. Pan, and M. Mao, “Memorypath: A deep reinforcement learning framework for incorporating memory component into knowledge graph reasoning,” Neurocomputing , vol. 419, pp. 273–286, 2021. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S0925231220312959
2021
Later among the works it cites.
X. Huang, D. Chen, T. Ren, and D. Wang, “A survey of community detection methods in multilayer networks,” Data Mining and Knowledge Discovery , vol. 35, no. 1, pp. 1–45, 2021
2021
Later among the works it cites.
Z. Li, Y. Sun, S. Tang, C. Zhang, and H. Ma, “Reinforcement learning with dual attention guided graph convolution for relation extraction,” in 2020 25th International Conference on Pattern Recognition (ICPR) . IEEE, 2021, pp. 946–953
2021
Later among the works it cites.
X. Fu, J. Li, J. Wu, Q. Sun, C. Ji, S. Wang, J. Tan, H. Peng, and P. S. Yu, “Ace-hgnn: Adaptive curvature exploration hyperbolic graph neural network,” in 2021 IEEE International Conference on Data Mining (ICDM) , 2021, pp. 111–120
2021
Later among the works it cites.
J. Xu, Y. Yang, S. Pu, Y. Fu, J. Feng, W. Jiang, J. Lu, and C. Wang, “Netrl: Task-aware network denoising via deep reinforcement learning,” IEEE Transactions on Knowledge and Data Engineering , pp. 1–1, 2021
2021
Later among the works it cites.
Y. Qin, X. Wang, P. Cui, and W. Zhu, “Gqnas: Graph q network for neural architecture search,” in Proceedings of the IEEE International Conference on Data Mining (ICDM) , 2021, pp. 1288–1293
2021
Later among the works it cites.
H. Peng, J. Li, Z. Wang, R. Yang, M. Liu, M. Zhang, P. Yu, and L. He, “Lifelong property price prediction: A case study for the toronto real estate market,” IEEE Transactions on Knowledge and Data Engineering , pp. 1–1, 2021
2021
Later among the works it cites.
J. Dineen, A. Haque, and M. Bielskas, “Reinforcement learning for data poisoning on graph neural networks,” in International Conference on Social Computing, Behavioral-Cultural Modeling and Prediction and Behavior Representation in Modeling and Simulation , 2021, pp. 141–150
2021
Later among the works it cites.
2021
Later among the works it cites.
Y. Sun, K. Zhang, and C. Sun, “Model-based transfer reinforcement learning based on graphical model representations,” IEEE Transactions on Neural Networks and Learning Systems , pp. 1–14, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
F.-X. Devailly, D. Larocque, and L. Charlin, “Ig-rl: Inductive graph reinforcement learning for massive-scale traffic signal control,” IEEE Transactions on Intelligent Transportation Systems , pp. 1–12, 2021
2021
Later among the works it cites.
Z. Yu and M. Hu, “Deep reinforcement learning with graph representation for vehicle repositioning,” IEEE Transactions on Intelligent Transportation Systems , pp. 1–14, 2021
2021
Later among the works it cites.
S. Chen, J. Dong, P. Ha, Y. Li, and S. Labi, “Graph neural network and reinforcement learning for multi-agent cooperative control of connected autonomous vehicles,” Computer‐Aided Civil and Infrastructure Engineering , vol. 36, no. 7, pp. 838–857, 2021
2021
Later among the works it cites.
Z. Zeng, “Graphlight: Graph-based reinforcement learning for traffic signal control,” in 2021 IEEE 6th International Conference on Computer and Communication Systems (ICCCS) , 2021, pp. 645–650
2021
Later among the works it cites.
C. Yang, H. Wang, J. Tang, C. Shi, M. Sun, G. Cui, and Z. Liu, “Full-scale information diffusion prediction with reinforced recurrent networks,” IEEE Transactions on Neural Networks and Learning Systems , pp. 1–13, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
J. Yan, S. Yang, and E. Hancock, “Learning for graph matching and related combinatorial optimization problems,” in Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence , Yokohama, Yokohama, Japan, 2021
2021
Later among the works it cites.
C. K. Joshi, Q. Cappart, L.-M. Rousseau, and T. Laurent, “Learning TSP Requires Rethinking Generalization,” in 27th International Conference on Principles and Practice of Constraint Programming (CP 2021) , vol. 210, Dagstuhl, Germany, 2021, pp. 33:1–33:21
2021
Later among the works it cites.
2021
Later among the works it cites.
G. S. Ramachandran, I. Brugere, L. R. Varshney, and C. Xiong, “Gaea: Graph augmentation for equitable access via reinforcement learning,” in Proceedings of the 2021 AAAI/ACM Conference on AI, Ethics, and Society , 2021, pp. 884–894
2021
Later among the works it cites.
P. Ammanabrolu and M. Riedl, “Playing text-adventure games with graph-based deep reinforcement learning,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) , 2019, pp. 3557–3565
2021
Later among the works it cites.
2022
Closest in time.
Y. Wu, X. Xi, and J. He, “Afgsl: Automatic feature generation based on graph structure learning,” Knowledge-Based Systems , vol. 238, p. 107835, 2022
2022
Closest in time.
M. Zhang, X. Yu, J. Rong, and L. Ou, “Graph pruning for model compression,” Applied Intelligence , pp. 1–13, 2022
2022
Closest in time.
J. Jin, S. Zhou, W. Zhang, T. He, Y. Yu, and R. Fakoor, “Graph-enhanced exploration for goal-oriented reinforcement learning,” in Proceedings of the Tenth International Conference on Learning Representations , 2022
2022
Closest in time.
A. Jain, N. Kosaka, K.-M. Kim, and J. J. Lim, “Know your action set: Learning action relations for reinforcement learning,” in Proceedings of the Tenth International Conference on Learning Representations , 2022
2022
Closest in time.
Z. Zhang, Z. Yang, H. Liu, P. Tokekar, and F. Huang, “Reinforcement learning under a multi-agent predictive state representation model: Method and theory,” in Proceedings of the Tenth International Conference on Learning Representations , 2022
2022
Closest in time.
S. Hong, D. Yoon, and K.-E. Kim, “Structure-aware transformer policy for inhomogeneous multi-task reinforcement learning,” in International Conference on Learning Representations , 2022
2022
Closest in time.
H. Liu, S. Zhou, C. Chen, T. Gao, J. Xu, and M. Shu, “Dynamic knowledge graph reasoning based on deep reinforcement learning,” Knowledge-Based Systems , vol. 241, p. 108235, 2022
2022
Closest in time.
2022
Closest in time.
R. Wang, F. Fang, J. Cui, and W. Zheng, “Learning self-driven collective dynamics with graph networks,” Scientific reports , vol. 12, no. 1, pp. 1–11, 2022
2022
Closest in time.
L. Chen, J. Cui, X. Tang, Y. Qian, Y. Li, and Y. Zhang, “Rlpath: a knowledge graph link prediction method using reinforcement learning based attentive relation path searching and representation learning,” Applied Intelligence , vol. 52, no. 4, pp. 4715–4726, 2022
2022
Closest in time.
2022
Closest in time.
D. Chen, M. Nie, H. Zhang, Z. Wang, and D. Wang, “Network embedding algorithm taking in variational graph autoencoder,” Mathematics , vol. 10, no. 3, p. 485, 2022
2022
Closest in time.
X. Zhao, Q. Dai, J. Wu, H. Peng, M. Liu, X. Bai, J. Tan, S. Wang, and P. Yu, “Multi-view tensor graph neural networks through reinforced aggregation,” IEEE Transactions on Knowledge and Data Engineering , pp. 1–1, 2022
2022
Closest in time.
H. Peng, R. Zhang, S. Li, Y. Cao, S. Pan, and P. Yu, “Reinforced, incremental and cross-lingual event detection from social messages,” IEEE Transactions on Pattern Analysis and Machine Intelligence , pp. 1–1, 2022
2022
Closest in time.
D. Bacciu and D. Numeroso, “Explaining deep graph networks via input perturbation,” IEEE Transactions on Neural Networks and Learning Systems , pp. 1–12, 2022
2022
Closest in time.
P. Shang, X. Liu, C. Yu, G. Yan, Q. Xiang, and X. Mi, “A new ensemble deep graph reinforcement learning network for spatio-temporal traffic volume forecasting in a freeway network,” Digital Signal Processing , vol. 123, p. 103419, 2022
2022
Closest in time.
H. Zhu, T. Shou, R. Guo, Z. Jiang, Z. Wang, Z. Wang, Z. Yu, W. Zhang, C. Wang, and L. Chen, “Redpacketbike: A graph-based demand modeling and crowd-driven station rebalancing framework for bike sharing systems,” IEEE Transactions on Mobile Computing , pp. 1–1, 2022
2022
Closest in time.
L. Ma, Z. Shao, X. Li, Q. Lin, J. Li, V. C. M. Leung, and A. K. Nandi, “Influence maximization in complex networks by using evolutionary deep reinforcement learning,” IEEE Transactions on Emerging Topics in Computational Intelligence , pp. 1–15, 2022
2022
Closest in time.
2022
Closest in time.
D. Chen, M. Nie, J. Yan, D. Wang, and Q. Gan, “Network representation learning algorithm based on complete subgraph folding,” Mathematics , vol. 10, no. 4, p. 581, 2022
2022
Closest in time.
W. Zhao, Y. Li, T. Fan, and F. Wu, “A novel embedding learning framework for relation completion and recommendation based on graph neural network and multi-task learning,” Soft Computing , pp. 1–13, 2022
2022
Closest in time.
X. Wang, H. Ji, C. Shi, B. Wang, Y. Ye, P. Cui, and P. S. Yu, “Heterogeneous graph attention network,” in The World Wide Web Conference , San Francisco, CA, USA, 2019, pp. 2022–2032
2032
Closest in time.
T. Trouillon, J. Welbl, S. Riedel, E. Gaussier, and G. Bouchard, “Complex embeddings for simple link prediction,” in Proceedings of The 33rd International Conference on Machine Learning , M. F. Balcan and K. Q. Weinberger, Eds., vol. 48, 2016, pp. 2071–2080
2080
Closest in time.
Q. Sun, J. Li, H. Peng, J. Wu, Y. Ning, P. S. Yu, and L. He, “Sugar: Subgraph neural network with reinforcement pooling and self-supervised mutual information mechanism,” in Proceedings of the Web Conference 2021 , Ljubljana, Slovenia, 2021, pp. 2081–2091
2091
Closest in time.
H. v. Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in Proceedings of the Thirtieth AAAI Conference on Artificial Intelligence , Phoenix, Arizona, 2016, pp. 2094–2100
2094
Closest in time.