Fetching the paper…
Reading the bibliography…
Tensor algebra is widely used in many applications, such as scientific computing, machine learning, and data analytics.
Die grundlage der allgemeinen relativitätstheorie
Albert Einstein. 1923 · 1923
Earlier work this paper cites.
Mlog: Towards declarative in-database machine learning
Xupeng Li, Bin Cui, Yiru Chen, Wentao Wu, and Ce Zhang. 2017 · 1936
Earlier work this paper cites.
Doubly lexical orderings of matrices
Anna Lubiw. 1987 · 1987
Earlier work this paper cites.
Three partition refinement algorithms
Robert Paige and Robert E Tarjan. 1987 · 1987
Earlier work this paper cites.
ITPACKV 2D user’s guide
David R Kincaid, Thomas C Oppe, and David M Young. 1989 · 1989
Earlier work this paper cites.
Compilation Techniques for Sparse Matrix Computations (ICS ’93) . Association for Computing Machinery, New York, NY, USA, 416–424
Aart J. C. Bik and Harry A. G. Wijshoff. 1993 · 1993
Earlier work this paper cites.
Cilk: An efficient multithreaded runtime system
Robert D Blumofe, Christopher F Joerg, Bradley C Kuszmaul, Charles E Leiserson, Keith H Randall, and Yuli Zhou. 1995 · 1995
Earlier work this paper cites.
On improving the performance of sparse matrix-vector multiplication. In Proceedings Fourth International Conference on High-Performance Computing . IEEE, 66–71
James B White and Ponnuswamy Sadayappan. 1997 · 1997
Earlier work this paper cites.
FLAME: Formal linear algebra methods environment
John A Gunnels, Fred G Gustavson, Greg M Henry, and Robert A Van De Geijn. 2001 · 2001
Earlier work this paper cites.
Tensor contraction engine: Abstraction and automated parallel implementation of configuration-interaction, coupled-cluster, and many-body perturbation theories
So Hirata. 2003 · 2003
Earlier work this paper cites.
Performance models for evaluation and automatic tuning of symmetric sparse matrix-vector multiply. In International Conference on Parallel Processing, 2004. ICPP 2004. IEEE, 169–176
Benjamin C Lee, Richard W Vuduc, James W Demmel, and Katherine A Yelick. 2004 · 2004
Earlier work this paper cites.
Fast sparse matrix-vector multiplication by exploiting variable block structure. In International Conference on High Performance Computing and Communications . Springer, 807–816
Richard W Vuduc and Hyun-Jin Moon. 2005 · 2005
Earlier work this paper cites.
Automatic code generation for many-body electronic structure methods: the tensor contraction engine
Alexander A Auer, Gerald Baumgartner, David E Bernholdt, Alina Bibireata, Venkatesh Choppella, Daniel Cociorva, Xiaoyang Gao, Robert Harrison, Sriram Krishnamoorthy, Sandhya Krishnan, et al · 2006
Earlier work this paper cites.
Multiway analysis of epilepsy tensors
Evrim Acar, Canan Aykut-Bingol, Haluk Bingol, Rasmus Bro, and Bülent Yener. 2007 · 2007
Earlier work this paper cites.
Efficient MATLAB computations with sparse and factored tensors
Brett W. Bader and Tamara G. Kolda. 2007 · 2007
Earlier work this paper cites.
Optimization of sparse matrix-vector multiplication on emerging multicore platforms. In SC’07: Proceedings of the 2007 ACM/IEEE Conference on Supercomputing . IEEE, 1–12
Samuel Williams, Leonid Oliker, Richard Vuduc, John Shalf, Katherine Yelick, and James Demmel. 2007 · 2007
Earlier work this paper cites.
Efficient sparse matrix-vector multiplication on CUDA
Nathan Bell and Michael Garland. 2008 · 2008
Earlier work this paper cites.
On the representation and multiplication of hypersparse matrices. In 2008 IEEE International Symposium on Parallel and Distributed Processing . IEEE, 1–11
Aydin Buluc and John R Gilbert. 2008 · 2008
Earlier work this paper cites.
Scalable tensor decompositions for multi-aspect data mining. In 2008 Eighth IEEE international conference on data mining . IEEE, 363–372
Tamara G Kolda and Jimeng Sun. 2008 · 2008
Earlier work this paper cites.
Implementing sparse matrix-vector multiplication on throughput-oriented processors. In Proceedings of the conference on high performance computing networking, storage and analysis
Nathan Bell and Michael Garland. 2009 · 2009
Earlier work this paper cites.
Tensor decompositions and applications
Tamara G Kolda and Brett W Bader. 2009 · 2009
Earlier work this paper cites.
Learning optimal ranking with tensor factorization for tag recommendation. In Proceedings of the 15th ACM SIGKDD international conference on Knowledge discovery and data mining . 727–736
Steffen Rendle, Leandro Balby Marinho, Alexandros Nanopoulos, and Lars Schmidt-Thieme. 2009 · 2009
Earlier work this paper cites.
Automatic Performance Tuning of SpMV on GPGPU
Xianyi Zhang, Yunquan Zhang, Xiangzheng Sun, Fangfang Liu, Shengfei Liu, Yuxin Tang, and Yucheng Li. 2009 · 2009
Earlier work this paper cites.
Model-driven autotuning of sparse matrix-vector multiply on GPUs
Jee W Choi, Amik Singh, and Richard W Vuduc. 2010 · 2010
Earlier work this paper cites.
Automatically tuning sparse matrix-vector multiplication for GPU architectures. In International Conference on High-Performance Embedded Architectures and Compilers . Springer, 111–125
Alexander Monakov, Anton Lokhmotov, and Arutyun Avetisyan. 2010 · 2010
Earlier work this paper cites.
NWChem: A comprehensive and scalable open-source solution for large scale molecular simulations
Marat Valiev, Eric J Bylaska, Niranjan Govind, Karol Kowalski, Tjerk P Straatsma, Hubertus JJ Van Dam, Dunyou Wang, Jarek Nieplocha, Edoardo Apra, Theresa L Windus, et al · 2010
Earlier work this paper cites.
The University of Florida sparse matrix collection
Timothy A Davis and Yifan Hu. 2011 · 2011
Cited alongside, same era.
Optimization of sparse matrix-vector multiplication with variant CSR on GPUs. In 2011 IEEE 17th International Conference on Parallel and Distributed Systems . IEEE, 165–172
Xiaowen Feng, Hai Jin, Ran Zheng, Kan Hu, Jingxiang Zeng, and Zhiyuan Shao. 2011 · 2011
Cited alongside, same era.
CSX: an extended compression format for spmv on shared memory systems
Kornilios Kourtis, Vasileios Karakasis, Georgios Goumas, and Nectarios Koziris. 2011 · 2011
Cited alongside, same era.
Efficient and scalable computations with sparse tensors. In 2012 IEEE Conference on High Performance Extreme Computing . IEEE, 1–6
Muthu Baskaran, Benoît Meister, Nicolas Vasilache, and Richard Lethin. 2012 · 2012
Cited alongside, same era.
National Institute of Standards and Technology
2013 · 2013
Cited alongside, same era.
Analyzing tensor power method dynamics in overcomplete regime
Animashree Anandkumar, Rong Ge, and Majid Janzamin. 2017 · 2017
Later among the works it cites.
taco: A Tool to Generate Tensor Algebra Kernels. In 2017 32nd IEEE/ACM International Conference on Automated Software Engineering (ASE) . 943–948
Fredrik Kjolstad, Stephen Chou, David Lugato, Shoaib Kamil, and Saman Amarasinghe. 2017a · 2017
Later among the works it cites.
The Tensor Algebra Compiler
Fredrik Kjolstad, Shoaib Kamil, Stephen Chou, David Lugato, and Saman Amarasinghe. 2017b · 2017
Later among the works it cites.
The tensor algebra compiler
Fredrik Kjolstad, Shoaib Kamil, Stephen Chou, David Lugato, and Saman Amarasinghe. 2017c · 2017
Later among the works it cites.
A unified optimization approach for sparse tensor operations on gpus. In 2017 IEEE international conference on cluster computing (CLUSTER) . IEEE, 47–57
Bangtian Liu, Chengyao Wen, Anand D Sarwate, and Maryam Mehri Dehnavi. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
New implementation of high-level correlated methods using a general block tensor library for high-performance electronic structure calculations
Evgeny Epifanovsky, Michael Wormit, Tomasz Kuś, Arie Landau, Dmitry Zuev, Kirill Khistyaev, Prashant Manohar, Ilya Kaliman, Andreas Dreuw, and Anna I Krylov. 2013 · 2013
Cited alongside, same era.
SMAT: an input adaptive auto-tuner for sparse matrix-vector multiplication. In Proceedings of the 34th ACM SIGPLAN conference on Programming language design and implementation . 117–126
Jiajia Li, Guangming Tan, Mingyu Chen, and Ninghui Sun. 2013 · 2013
Cited alongside, same era.
AdELL: An adaptive warp-balancing ELL format for efficient sparse matrix-vector multiplication on GPUs. In 2013 42nd international conference on parallel processing . IEEE, 11–20
Marco Maggioni and Tanya Berger-Wolf. 2013 · 2013
Cited alongside, same era.
Elemental: A new framework for distributed memory dense matrix computations
Jack Poulson, Bryan Marker, Robert A Van de Geijn, Jeff R Hammond, and Nichols A Romero. 2013 · 2013
Cited alongside, same era.
Low-overhead load-balanced scheduling for sparse tensor computations. In 2014 IEEE High Performance Extreme Computing Conference (HPEC) . IEEE
Muthu Baskaran, Benoit Meister, and Richard Lethin. 2014 · 2014
Cited alongside, same era.
Exploiting symmetry in tensors for high performance: Multiplication with symmetric tensors
Martin D Schatz, Tze Meng Low, Robert A van de Geijn, and Tamara G Kolda. 2014 · 2014
Cited alongside, same era.
A massively parallel tensor contraction framework for coupled-cluster computations
Edgar Solomonik, Devin Matthews, Jeff R Hammond, John F Stanton, and James Demmel. 2014 · 2014
Cited alongside, same era.
Tensor factorization toward precision medicine
Yuan Luo, Fei Wang, and Peter Szolovits. 2017 · 2017
Later among the works it cites.
Tensor decomposition for signal processing and machine learning
Nicholas D Sidiropoulos, Lieven De Lathauwer, Xiao Fu, Kejun Huang, Evangelos E Papalexakis, and Christos Faloutsos. 2017 · 2017
Later among the works it cites.
FROSTT: The Formidable Repository of Open Sparse Tensors and Tools
Shaden Smith, Jee W. Choi, Jiajia Li, Richard Vuduc, Jongsoo Park, Xing Liu, and George Karypis. 2017 · 2017
Later among the works it cites.
Accelerating the tucker decomposition with compressed sparse tensors. In European Conference on Parallel Processing . Springer, 653–668
Shaden Smith and George Karypis. 2017 · 2017
Later among the works it cites.
HiCOO: hierarchical storage of sparse tensors. In SC18: International Conference for High Performance Computing, Networking, Storage and Analysis . IEEE, 238–252
Jiajia Li, Jimeng Sun, and Richard Vuduc. 2018 · 2018
Later among the works it cites.
Cvr: Efficient vectorization of spmv on x86 processors. In Proceedings of the 2018 International Symposium on Code Generation and Optimization . 149–162
Biwei Xie, Jianfeng Zhan, Xu Liu, Wanling Gao, Zhen Jia, Xiwen He, and Lixin Zhang. 2018 · 2018
Later among the works it cites.
Design principles for sparse matrix multiplication on the GPU. In European Conference on Parallel Processing . Springer
Carl Yang, Aydın Buluç, and John D Owens. 2018 · 2018
Later among the works it cites.
Tiramisu: A polyhedral compiler for expressing fast and portable code. In 2019 IEEE/ACM International Symposium on Code Generation and Optimization (CGO) . IEEE, 193–205
Riyadh Baghdadi, Jessica Ray, Malek Ben Romdhane, Emanuele Del Sozzo, Abdurrahman Akkas, Yunming Zhang, Patricia Suriana, Shoaib Kamil, and Saman Amarasinghe. 2019 · 2019
Later among the works it cites.
A code generator for high-performance tensor contractions on gpus. In 2019 IEEE/ACM International Symposium on Code Generation and Optimization (CGO) . IEEE
Jinsung Kim, Aravind Sukumaran-Rajam, Vineeth Thumma, Sriram Krishnamoorthy, Ajay Panyala, Louis-Noël Pouchet, Atanas Rountev, and Ponnuswamy Sadayappan. 2019 · 2019
Later among the works it cites.
Tensor algebra compilation with workspaces. In 2019 IEEE/ACM International Symposium on Code Generation and Optimization (CGO) . IEEE, 180–192
Fredrik Kjolstad, Peter Ahrens, Shoaib Kamil, and Saman Amarasinghe. 2019 · 2019
Later among the works it cites.
Efficient and effective sparse tensor reordering. In Proceedings of the ACM International Conference on Supercomputing . 227–237
Jiajia Li, Bora Uçar, Ümit V Çatalyürek, Jimeng Sun, Kevin Barker, and Richard Vuduc. 2019 · 2019
Later among the works it cites.
An efficient mixed-mode representation of sparse tensors. In Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis . 1–25
Israt Nisa, Jiajia Li, Aravind Sukumaran-Rajam, Prasant Singh Rawat, Sriram Krishnamoorthy, and Ponnuswamy Sadayappan. 2019 · 2019
Later among the works it cites.
Tensornetwork: A library for physics and machine learning
Chase Roberts, Ashley Milsted, Martin Ganahl, Adam Zalcman, Bruce Fontaine, Yijian Zou, Jack Hidary, Guifre Vidal, and Stefan Leichenauer. 2019 · 2019
Later among the works it cites.
Tensor completion algorithms in big data analytics
Qingquan Song, Hancheng Ge, James Caverlee, and Xia Hu. 2019 · 2019
Later among the works it cites.
Load-balancing sparse matrix vector product kernels on GPUs
Hartwig Anzt, Terry Cojean, Chen Yen-Chen, Jack Dongarra, Goran Flegar, Pratik Nayak, Stanimire Tomov, Yuhsiang M Tsai, and Weichung Wang. 2020 · 2020
Later among the works it cites.
aeSpTV: An adaptive and efficient framework for sparse tensor-vector product kernel on a high-performance computing platform
Yuedan Chen, Guoqing Xiao, M Tamer Özsu, Chubo Liu, Albert Y Zomaya, and Tao Li. 2020 · 2020
Later among the works it cites.
The ITensor Software Library for tensor network calculations
Matthew Fishman, Steven R White, and E Miles Stoudenmire. 2020 · 2020
Later among the works it cites.
A novel data transformation and execution strategy for accelerating sparse matrix multiplication on GPUs. In Proceedings of the 25th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming . 376–388
Peng Jiang, Changwan Hong, and Gagan Agrawal. 2020 · 2020
Later among the works it cites.
MLIR: A Compiler Infrastructure for the End of Moore’s Law
Chris Lattner, Jacques Pienaar, Mehdi Amini, Uday Bondhugula, River Riddle, Albert Cohen, Tatiana Shpeisman, Andy Davis, Nicolas Vasilache, and Oleksandr Zinenko. 2020 · 2020
Later among the works it cites.
A survey of the usages of deep learning for natural language processing
Daniel W Otter, Julian R Medina, and Jugal K Kalita. 2020 · 2020
Later among the works it cites.