Fetching the paper…
Reading the bibliography…
Recently, Graph Neural Networks (GNNs) have become state-of-the-art algorithms for analyzing non-euclidean graph data.
Measurement and analysis of online social networks. In Proceedings of the 7th ACM SIGCOMM conference on Internet measurement . 29–42
Alan Mislove, Massimiliano Marcon, Krishna P Gummadi, Peter Druschel, and Bobby Bhattacharjee. 2007 · 2007
Earlier work this paper cites.
Rubik: A Hierarchical Architecture for Efficient Graph Learning
Xiaobing Chen, Yuke Wang, Xinfeng Xie, Xing Hu, Abanti Basak, Ling Liang, Mingyu Yan, Lei Deng, Yufei Ding, Zidong Du, Yunji Chen, and Yuan Xie. 2020a · 2009
Earlier work this paper cites.
Graph embedding in vector spaces by node attribute statistics
Jaume Gibert, Ernest Valveny, and Horst Bunke. 2012 · 2012
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks. In Advances in Neural Information Processing Systems 25: 26th Annual Conference on Neural Information Processing Systems,NIPS , Peter L. Bartlett, Fernando C. N. Pereira, Christopher J. C. Burges, Léon Bottou, and Kilian Q. Weinberger (Eds.). 1106–1114
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012 · 2012
Earlier work this paper cites.
Diannao: A small-footprint high-throughput accelerator for ubiquitous machine-learning
Tianshi Chen, Zidong Du, Ninghui Sun, Jia Wang, Chengyong Wu, Yunji Chen, and Olivier Temam. 2014a · 2014
Earlier work this paper cites.
TOP-PIM: throughput-oriented programmable processing in memory. In The 23rd International Symposium on High-Performance Parallel and Distributed Computing, HPDC’14, Vancouver, BC, Canada - June 23 - 27, 2014 , Beth Plale, Matei Ripeanu, Franck Cappello, and Dongyan Xu (Eds.). ACM, 85–98
Dong Ping Zhang, Nuwan Jayasena, Alexander Lyashevsky, Joseph L. Greathouse, Lifan Xu, and Michael Ignatowski. 2014 · 2014
Earlier work this paper cites.
A scalable processing-in-memory accelerator for parallel graph processing. In Proceedings of the 42nd Annual International Symposium on Computer Architecture . 105–117
Junwhan Ahn, Sungpack Hong, Sungjoo Yoo, Onur Mutlu, and Kiyoung Choi. 2015 · 2015
Earlier work this paper cites.
NDA: Near-DRAM acceleration architecture leveraging commodity DRAM devices and standard memory modules. In 21st IEEE International Symposium on High Performance Computer Architecture, HPCA 2015, Burlingame, CA, USA, February 7-11, 2015 . IEEE Computer Society, 283–295
Amin Farmahini Farahani, Jung Ho Ahn, Katherine Morrow, and Nam Sung Kim. 2015 · 2015
Earlier work this paper cites.
Practical Near-Data Processing for In-Memory Analytics Frameworks. In 2015 International Conference on Parallel Architectures and Compilation, PACT 2015, San Francisco, CA, USA, October 18-21, 2015 . IEEE Computer Society, 113–124
Mingyu Gao, Grant Ayers, and Christos Kozyrakis. 2015 · 2015
Earlier work this paper cites.
Fast r-cnn. In Proceedings of the IEEE international conference on computer vision . 1440–1448
Ross Girshick. 2015 · 2015
Earlier work this paper cites.
The IBM z13 memory subsystem for big data
P. J. Meaney, L. D. Curley, G. D. Gilda, M. R. Hodges, D. J. Buerkle, R. D. Siegl, and R. K. Dong. 2015 · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun. 2015 · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation. In International Conference on Medical image computing and computer-assisted intervention . Springer, 234–241
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015 · 2015
Earlier work this paper cites.
Very Deep Convolutional Networks for Large-Scale Image Recognition. In 3rd International Conference on Learning Representations, ICLR , Yoshua Bengio and Yann LeCun (Eds.)
Karen Simonyan and Andrew Zisserman. 2015 · 2015
Earlier work this paper cites.
Optimizing fpga-based accelerator design for deep convolutional neural networks. In Proceedings of the 2015 ACM/SIGDA international symposium on field-programmable gate arrays . 161–170
Chen Zhang, Peng Li, Guangyu Sun, Yijin Guan, Bingjun Xiao, and Jason Cong. 2015 · 2015
Earlier work this paper cites.
Using GPUs to speed-up FCM-based community detection in Social Networks. In 2016 7th International Conference on Computer Science and Information Technology (CSIT) . 1–6
Mohammed Alandoli, Mohammed Shehab, Mahmoud Al-Ayyoub, Yaser Jararweh, and Mohammad Al-Smadi. 2016 · 2016
Earlier work this paper cites.
Chameleon: Versatile and practical near-DRAM acceleration architecture for large memory systems. In 2016 49th annual IEEE/ACM international symposium on Microarchitecture (MICRO) . IEEE, 1–13
Hadi Asghari-Moghaddam, Young Hoon Son, Jung Ho Ahn, and Nam Sung Kim. 2016 · 2016
Earlier work this paper cites.
Eyeriss: A Spatial Architecture for Energy-Efficient Dataflow for Convolutional Neural Networks. In 43rd ACM/IEEE Annual International Symposium on Computer Architecture, ISCA 2016, Seoul, South Korea, June 18-22, 2016 . IEEE Computer Society, 367–379
Yu-Hsin Chen, Joel S. Emer, and Vivienne Sze. 2016 · 2016
Earlier work this paper cites.
Nxgraph: An efficient graph processing system on a single machine. In 2016 IEEE 32nd International Conference on Data Engineering (ICDE) . IEEE, 409–420
Yuze Chi, Guohao Dai, Yu Wang, Guangyu Sun, Guoliang Li, and Huazhong Yang. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition. In 2016 IEEE Conference on Computer Vision and Pattern Recognition, CVPR . IEEE Computer Society, 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Transparent offloading and mapping (TOM) enabling programmer-transparent near-data processing in GPU systems
Kevin Hsieh, Eiman Ebrahimi, Gwangsun Kim, Niladrish Chatterjee, Mike O’Connor, Nandita Vijaykumar, Onur Mutlu, and Stephen W Keckler. 2016 · 2016
Earlier work this paper cites.
Ssd: Single shot multibox detector. In European conference on computer vision . Springer, 21–37
Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott Reed, Cheng-Yang Fu, and Alexander C Berg. 2016 · 2016
Earlier work this paper cites.
Deep gaussian embedding of graphs: Unsupervised inductive learning via ranking
Aleksandar Bojchevski and Stephan Günnemann. 2017 · 2017
Earlier work this paper cites.
Tetris: Scalable and efficient neural network acceleration with 3d memory. In Proceedings of the Twenty-Second International Conference on Architectural Support for Programming Languages and Operating Systems . 751–764
Mingyu Gao, Jing Pu, Xuan Yang, Mark Horowitz, and Christos Kozyrakis. 2017 · 2017
Earlier work this paper cites.
Learning Graph Representations with Embedding Propagation. In Advances in Neural Information Processing Systems 30: Annual Conference on Neural Information Processing Systems 2017, 4-9 December 2017, Long Beach, CA, USA , Isabelle Guyon, Ulrike von Luxburg, Samy Bengio, Hanna M. Wallach, Rob Fergus, S. V. N. Vishwanathan, and Roman Garnett (Eds.). 5119–5130
Alberto García-Durán and Mathias Niepert. 2017 · 2017
Earlier work this paper cites.
Inductive representation learning on large graphs. In Advances in neural information processing systems . 1024–1034
Will Hamilton, Zhitao Ying, and Jure Leskovec. 2017 · 2017
Earlier work this paper cites.
Neural message passing for jet physics
Isaac Henrion, Johann Brehmer, Joan Bruna, Kyunghyun Cho, Kyle Cranmer, Gilles Louppe, and Gaspar Rochette. 2017 · 2017
Earlier work this paper cites.
Mobilenets: Efficient convolutional neural networks for mobile vision applications
Andrew G Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam. 2017 · 2017
Earlier work this paper cites.
Densely Connected Convolutional Networks. In 2017 IEEE Conference on Computer Vision and Pattern Recognition, CVPR . IEEE Computer Society, 2261–2269
Gao Huang, Zhuang Liu, Laurens van der Maaten, and Kilian Q. Weinberger. 2017 · 2017
Earlier work this paper cites.
In-datacenter performance analysis of a tensor processing unit. In Proceedings of the 44th annual international symposium on computer architecture . 1–12
Norman P Jouppi, Cliff Young, Nishant Patil, David Patterson, Gaurav Agrawal, Raminder Bajwa, Sarah Bates, Suresh Bhatia, Nan Boden, Al Borchers, et al · 2017
Earlier work this paper cites.
Semi-Supervised Classification with Graph Convolutional Networks. In 5th International Conference on Learning Representations, ICLR 2017, Toulon, France, April 24-26, 2017, Conference Track Proceedings . OpenReview.net
Thomas N. Kipf and Max Welling. 2017 · 2017
Earlier work this paper cites.
GraphPIM: Enabling Instruction-Level PIM Offloading in Graph Computing Frameworks. In 2017 IEEE International Symposium on High Performance Computer Architecture, HPCA 2017, Austin, TX, USA, February 4-8, 2017 . IEEE Computer Society, 457–468
Lifeng Nai, Ramyad Hadidi, Jaewoong Sim, Hyojong Kim, Pranith Kumar, and Hyesoon Kim. 2017 · 2017
Earlier work this paper cites.
YOLO9000: better, faster, stronger. In Proceedings of the IEEE conference on computer vision and pattern recognition . 7263–7271
Joseph Redmon and Ali Farhadi. 2017 · 2017
Earlier work this paper cites.
Fastgcn: fast learning with graph convolutional networks via importance sampling
Jie Chen, Tengfei Ma, and Cao Xiao. 2018 · 2018
Earlier work this paper cites.
Graphh: A processing-in-memory architecture for large-scale graph processing
Guohao Dai, Tianhao Huang, Yuze Chi, Jishen Zhao, Guangyu Sun, Yongpan Liu, Yu Wang, Yuan Xie, and Huazhong Yang. 2018 · 2018
Earlier work this paper cites.
Multi-dimensional Parallel Training of Winograd Layer on Memory-Centric Architecture. In 51st Annual IEEE/ACM International Symposium on Microarchitecture, MICRO 2018, Fukuoka, Japan, October 20-24, 2018 . IEEE Computer Society, 682–695
Byungchul Hong, Yeonju Ro, and John Kim. 2018 · 2018
Earlier work this paper cites.
DeepTrain: A Programmable Embedded Platform for Training Deep Neural Networks
Duckhwan Kim, Taesik Na, Sudhakar Yalamanchili, and Saibal Mukhopadhyay. 2018 · 2018
Earlier work this paper cites.
Processing-in-memory for energy-efficient neural network training: A heterogeneous approach. In 2018 51st Annual IEEE/ACM International Symposium on Microarchitecture (MICRO) . IEEE, 655–668
Jiawen Liu, Hengyu Zhao, Matheus A Ogleari, Dong Li, and Jishen Zhao. 2018 · 2018
Earlier work this paper cites.
Machine learning in chemoinformatics and drug discovery
Yu-Chen Lo, Stefano E Rensi, Wen Torng, and Russ B Altman. 2018 · 2018
Cited alongside, same era.
A Scalable Near-Memory Architecture for Training Deep Neural Networks on Large In-Memory Datasets
Fabian Schuiki, Michael Schaffner, Frank K. Gürkaynak, and Luca Benini. 2019 · 2018
Cited alongside, same era.
Pitfalls of Graph Neural Network Evaluation
Oleksandr Shchur, Maximilian Mumme, Aleksandar Bojchevski, and Stephan Günnemann. 2018 · 2018
Cited alongside, same era.
Towards Memory-Efficient Allocation of CNNs on Processing-in-Memory Architecture
Yi Wang, Weixuan Chen, Jing Yang, and Tao Li. 2018 · 2018
Cited alongside, same era.
Parana: A Parallel Neural Architecture Considering Thermal Problem of 3D Stacked Memory
Shouyi Yin, Shibin Tang, Xinhan Lin, Peng Ouyang, Fengbin Tu, Leibo Liu, Jishen Zhao, Cong Xu, Shuangchen Li, Yuan Xie, and Shaojun Wei. 2019 · 2018
Cited alongside, same era.
EnGN: A High-Throughput and Energy-Efficient Accelerator for Large Graph Neural Networks
Shengwen Liang, Ying Wang, Cheng Liu, Lei He, LI Huawei, Dawen Xu, and Xiaowei Li. 2020b · 2020
Later among the works it cites.
Chip Placement with Deep Reinforcement Learning
Azalia Mirhoseini, Anna Goldie, Mustafa Yazgan, Joe Jiang, Ebrahim Songhori, Shen Wang, Young-Joon Lee, Eric Johnson, Omkar Pathak, Sungmin Bae, et al · 2020
Later among the works it cites.
A deep learning approach to antibiotic discovery
Jonathan M Stokes, Kevin Yang, Kyle Swanson, Wengong Jin, Andres Cubillos-Ruiz, Nina M Donghia, Craig R MacNair, Shawn French, Lindsey A Carfrae, Zohar Bloom-Ackermann, et al · 2020
Later among the works it cites.
Reducing communication in graph neural network training. In SC20: International Conference for High Performance Computing, Networking, Storage and Analysis . IEEE, 1–14
Alok Tripathy, Katherine Yelick, and Aydın Buluç. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Graph convolutional neural networks for web-scale recommender systems. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 974–983
Rex Ying, Ruining He, Kaifeng Chen, Pong Eksombatchai, William L Hamilton, and Jure Leskovec. 2018 · 2018
Cited alongside, same era.
Bisenet: Bilateral segmentation network for real-time semantic segmentation. In Proceedings of the European conference on computer vision (ECCV) . 325–341
Changqian Yu, Jingbo Wang, Chao Peng, Changxin Gao, Gang Yu, and Nong Sang. 2018 · 2018
Cited alongside, same era.
GraphP: Reducing communication for PIM-based graph processing with efficient data partition. In 2018 IEEE International Symposium on High Performance Computer Architecture (HPCA) . IEEE, 544–557
Mingxing Zhang, Youwei Zhuo, Chao Wang, Mingyu Gao, Yongwei Wu, Kang Chen, Christos Kozyrakis, and Xuehai Qian. 2018 · 2018
Cited alongside, same era.
Unet++: A nested u-net architecture for medical image segmentation
Zongwei Zhou, Md Mahfuzur Rahman Siddiquee, Nima Tajbakhsh, and Jianming Liang. 2018 · 2018
Cited alongside, same era.
Kgat: Knowledge graph attention network for recommendation. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 950–958
2019 · 2019
Cited alongside, same era.
Cluster-gcn: An efficient algorithm for training deep and large graph convolutional networks. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 257–266
Wei-Lin Chiang, Xuanqing Liu, Si Si, Yang Li, Samy Bengio, and Cho-Jui Hsieh. 2019 · 2019
Cited alongside, same era.
Traffic graph convolutional recurrent neural network: A deep learning framework for network-scale traffic learning and forecasting
Zhiyong Cui, Kristian Henrickson, Ruimin Ke, and Yinhai Wang. 2019 · 2019
Cited alongside, same era.
Hanrui Wang, Kuan Wang, Jiacheng Yang, Linxiao Shen, Nan Sun, Hae-Seung Lee, and Song Han. 2020c · 2020
Later among the works it cites.
GNNAdvisor: An Efficient Runtime System for GNN Acceleration on GPUs
Yuke Wang, Boyuan Feng, Gushu Li, Shuangchen Li, Lei Deng, Yuan Xie, and Yufei Ding. 2020a · 2020
Later among the works it cites.
Grid-gcn for fast and scalable point cloud learning. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 5661–5670
Qiangeng Xu, Xudong Sun, Cho-Ying Wu, Panqu Wang, and Ulrich Neumann. 2020 · 2020
Later among the works it cites.
Hygcn: A gcn accelerator with hybrid architecture. In 2020 IEEE International Symposium on High Performance Computer Architecture (HPCA) . IEEE, 15–29
Mingyu Yan, Lei Deng, Xing Hu, Ling Liang, Yujing Feng, Xiaochun Ye, Zhimin Zhang, Dongrui Fan, and Yuan Xie. 2020 · 2020
Later among the works it cites.
GraphACT: Accelerating GCN Training on CPU-FPGA Heterogeneous Platforms. In FPGA ’20: The 2020 ACM/SIGDA International Symposium on Field-Programmable Gate Arrays, Seaside, CA, USA, February 23-25, 2020 , Stephen Neuendorffer and Lesley Shannon (Eds.). ACM, 255–265
Hanqing Zeng and Viktor K. Prasanna. 2020 · 2020
Later among the works it cites.
GraphSAINT: Graph Sampling Based Inductive Learning Method. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 2020 . OpenReview.net
Hanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan, and Viktor K. Prasanna. 2020 · 2020
Later among the works it cites.
Multi-scale dynamic graph convolution network for point clouds classification
Zhengli Zhai, Xin Zhang, and Luyao Yao. 2020 · 2020
Later among the works it cites.
Hardware acceleration of large scale GCN inference. In 2020 IEEE 31st International Conference on Application-specific Systems, Architectures and Processors (ASAP) . IEEE, 61–68
Bingyi Zhang, Hanqing Zeng, and Viktor Prasanna. 2020 · 2020
Later among the works it cites.
FAFNIR: Accelerating Sparse Gathering by Using Efficient Near-Memory Intelligent Reduction. In IEEE International Symposium on High-Performance Computer Architecture, HPCA 2021, Seoul, South Korea, February 27 - March 3, 2021 . IEEE, 908–920
Bahar Asgari, Ramyad Hadidi, Jiashen Cao, Da Eun Shim, Sung Kyu Lim, and Hyesoon Kim. 2021 · 2021
Closest in time.
DGCL: an efficient communication library for distributed GNN training. In EuroSys ’21: Sixteenth European Conference on Computer Systems, Online Event, United Kingdom, April 26-28, 2021 , Antonio Barbalace, Pramod Bhatotia, Lorenzo Alvisi, and Cristian Cadar (Eds.). ACM, 130–144
Zhenkun Cai, Xiao Yan, Yidi Wu, Kaihao Ma, James Cheng, and Fan Yu. 2021 · 2021
Closest in time.
Towards efficient allocation of graph convolutional networks on hybrid computation-in-memory architecture
Jiaxian Chen, Guanquan Lin, Jiexin Chen, and Yi Wang. 2021 · 2021
Closest in time.
P3: Distributed Deep Graph Learning at Scale. In 15th { \{ USENIX } \} Symposium on Operating Systems Design and Implementation ( { \{ OSDI } \} 21) . 551–568
Swapnil Gandhi and Anand Padmanabha Iyer. 2021 · 2021
Closest in time.
I-GCN: A Graph Convolutional Network Accelerator with Runtime Locality Enhancement through Islandization. In MICRO-54: 54th Annual IEEE/ACM International Symposium on Microarchitecture . 1051–1063
Tong Geng, Chunshu Wu, Yongan Zhang, Cheng Tan, Chenhao Xie, Haoran You, Martin Herbordt, Yingyan Lin, and Ang Li. 2021 · 2021
Closest in time.
Near-Memory Processing in Action: Accelerating Personalized Recommendation with AxDIMM
Liu Ke, Xuan Zhang, Jinin So, Jong-Geon Lee, Shin-Haeng Kang, Sukhan Lee, Songyi Han, Yeongon Cho, Jin Hyun Kim, Yongsuk Kwon, et al · 2021
Closest in time.
Aquabolt-XL: Samsung HBM2-PIM with in-memory processing for ML accelerators and beyond. In 2021 IEEE Hot Chips 33 Symposium (HCS) . IEEE, 1–26
Jin Hyun Kim, Shin-haeng Kang, Sukhan Lee, Hyeonsu Kim, Woongjae Song, Yuhwan Ro, Seungwon Lee, David Wang, Hyunsung Shin, Bengseng Phuah, et al · 2021
Closest in time.
25.4 A 20nm 6GB Function-In-Memory DRAM, Based on HBM2 with a 1.2TFLOPS Programmable Computing Unit Using Bank-Level Parallelism, for Machine Learning Applications. In IEEE International Solid-State Circuits Conference, ISSCC 2021, San Francisco, CA, USA, February 13-22, 2021 . IEEE, 350–352
Young-Cheon Kwon, Suk Han Lee, Jaehoon Lee, Sang-Hyuk Kwon, Je-Min Ryu, Jong-Pil Son, Seongil O, Hak-soo Yu, Haesuk Lee, Soo Young Kim, Youngmin Cho, Jin Guk Kim, Jongyoon Choi, Hyunsung Shin, Jin Kim, BengSeng Phuah, HyoungMin Kim, Myeong Jun Song, Ahn Choi, Daeho Kim, Sooyoung Kim, Eun-Bong Kim, David Wang, Shinhaeng Kang, Yuhwan Ro, Seungwoo Seo, Joon-Ho Song, Jaeyoun Youn, Kyomin Sohn, and Nam Sung Kim. 2021 · 2021
Closest in time.
Task Parallelism-Aware Deep Neural Network Scheduling on Multiple Hybrid Memory Cube-Based Processing-in-Memory
Young Sik Lee and Tae Hee Han. 2021 · 2021
Closest in time.
GCNAX: A Flexible and Energy-efficient Accelerator for Graph Convolutional Neural Networks. In 2021 IEEE International Symposium on High-Performance Computer Architecture (HPCA) . 775–788
Jiajun Li, Ahmed Louri, Avinash Karanth, and Razvan Bunescu. 2021 · 2021
Closest in time.
ENMC: Extreme Near-Memory Classification via Approximate Screening. In MICRO-54: 54th Annual IEEE/ACM International Symposium on Microarchitecture . 1309–1322
Liu Liu, Jilan Lin, Zheng Qu, Yufei Ding, and Yuan Xie. 2021 · 2021
Closest in time.
Sampling Methods for Efficient Training of Graph Convolutional Networks: A Survey
Xin Liu, Mingyu Yan, Lei Deng, Guoqi Li, Xiaochun Ye, and Dongrui Fan. 2022 · 2021
Closest in time.
DistGNN: Scalable Distributed Training for Large-Scale Graph Neural Networks
Vasimuddin Md, Sanchit Misra, Guixiang Ma, Ramanarayan Mohanty, Evangelos Georganas, Alexander Heinecke, Dhiraj Kalamkar, Nesreen K Ahmed, and Sasikanth Avancha. 2021 · 2021
Closest in time.
Image segmentation using deep learning: A survey
Shervin Minaee, Yuri Y Boykov, Fatih Porikli, Antonio J Plaza, Nasser Kehtarnavaz, and Demetri Terzopoulos. 2021 · 2021
Closest in time.
Marius: Learning Massive Graph Embeddings on a Single Machine. In 15th { \{ USENIX } \} Symposium on Operating Systems Design and Implementation ( { \{ OSDI } \} 21) . 533–549
Jason Mohoney, Roger Waleffe, Henry Xu, Theodoros Rekatsinas, and Shivaram Venkataraman. 2021 · 2021
Closest in time.
I-GCN: Incremental Graph Convolution Network for Conversation Emotion Detection
Weizhi Nie, Rihao Chang, Minjie Ren, Yuting Su, and Anan Liu. 2021 · 2021
Closest in time.
TRiM: Enhancing Processor-Memory Interfaces with Scalable Tensor Reduction in Memory. In MICRO-54: 54th Annual IEEE/ACM International Symposium on Microarchitecture . 268–281
Jaehyun Park, Byeongho Kim, Sungmin Yun, Eojin Lee, Minsoo Rhu, and Jung Ho Ahn. 2021 · 2021
Closest in time.
Ten Lessons From Three Generations Shaped Google’s TPUv4i. In Annual International Symposium on Computer Architecture (ISCA)
Norman P.Jouppi, Doe Hyun Yoon, Matthew Ashcraft, and Mark Gottscho et al. 2021 · 2021
Closest in time.
Pu-gcn: Point cloud upsampling using graph convolutional networks. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11683–11692
Guocheng Qian, Abdulellah Abualshour, Guohao Li, Ali Thabet, and Bernard Ghanem. 2021 · 2021
Closest in time.
Cambricon-G: A Polyvalent Energy-efficient Accelerator for Dynamic Graph Neural Networks
Xinkai Song, Tian Zhi, Zhe Fan, Zhenxing Zhang, Xi Zeng, Wei Li, Xing Hu, Zidong Du, Qi Guo, and Yunji Chen. 2021 · 2021
Closest in time.
GNNerator: A Hardware/Software Framework for Accelerating Graph Neural Networks
Jacob R Stevens, Dipankar Das, Sasikanth Avancha, Bharat Kaul, and Anand Raghunathan. 2021 · 2021
Closest in time.
ABC-DIMM: Alleviating the Bottleneck of Communication in DIMM-based Near-Memory Processing with Inter-DIMM Broadcast. In 48th ACM/IEEE Annual International Symposium on Computer Architecture, ISCA 2021, Valencia, Spain, June 14-18, 2021 . IEEE, 237–250
Weiyi Sun, Zhaoshi Li, Shouyi Yin, Shaojun Wei, and Leibo Liu. 2021 · 2021
Closest in time.
Dorylus: Affordable, Scalable, and Accurate { \{ GNN } \} Training with Distributed { \{ CPU } \} Servers and Serverless Threads. In 15th { \{ USENIX } \} Symposium on Operating Systems Design and Implementation ( { \{ OSDI } \} 21) . 495–514
John Thorpe, Yifan Qiao, Jonathan Eyolfson, Shen Teng, Guanzhou Hu, Zhihao Jia, Jinliang Wei, Keval Vora, Ravi Netravali, Miryung Kim, et al · 2021
Closest in time.
SpaceA: Sparse Matrix Vector Multiplication on Processing-in-Memory Accelerator. In IEEE International Symposium on High-Performance Computer Architecture, HPCA 2021, Seoul, South Korea, February 27 - March 3, 2021 . IEEE, 570–583
Xinfeng Xie, Zheng Liang, Peng Gu, Abanti Basak, Lei Deng, Ling Liang, Xing Hu, and Yuan Xie. 2021 · 2021
Closest in time.
BlockGNN: Towards Efficient GNN Acceleration Using Block-Circulant Weight Matrices
Zhe Zhou, Bizhao Shi, Zhe Zhang, Yijin Guan, Guangyu Sun, and Guojie Luo. 2021 · 2021
Closest in time.
AliGraph: A Comprehensive Graph Neural Network Platform
Rong Zhu, Kun Zhao, Hongxia Yang, Wei Lin, Chang Zhou, Baole Ai, Yong Li, and Jingren Zhou. 2019 · 2094
Closest in time.