Fetching the paper…
Reading the bibliography…
Symmetric tensor operations arise in a wide variety of computations.
Integral transformations. A bottleneck in molecular quantum mechanical calculations
C.F Bender · 1972
Earlier work this paper cites.
Basic linear algebra subprograms for Fortran usage
C. L. Lawson, R. J. Hanson, D. R. Kincaid, and F. T. Krogh · 1979
Earlier work this paper cites.
An extended set of FORTRAN basic linear algebra subprograms
Jack J. Dongarra, Jeremy Du Croz, Sven Hammarling, and Richard J. Hanson · 1988
Earlier work this paper cites.
A set of level 3 basic linear algebra subprograms
Jack J. Dongarra, Jeremy Du Croz, Sven Hammarling, and Iain Duff · 1990
Earlier work this paper cites.
ScaLAPACK: A scalable linear algebra library for distributed memory concurrent computers
J. Choi, J. J. Dongarra, R. Pozo, and D. W. Walker · 1992
Earlier work this paper cites.
Optimizing matrix multiply using PHiPAC: a portable, high-performance, ANSI C coding methodology
J. Bilmes, K. Asanovic, C.W. Chin, and J. Demmel · 1997
Earlier work this paper cites.
Using PLAPACK: Parallel Linear Algebra Package
Robert A. van de Geijn · 1997
Earlier work this paper cites.
Automatically tuned linear algebra software
R. Clint Whaley and Jack J. Dongarra · 1998
Earlier work this paper cites.
LAPACK Users’ guide (third ed.)
E. Anderson, Z. Bai, C. Bischof, L. S. Blackford, J. Demmel, Jack J. Dongarra, J. Du Croz, S. Hammarling, A. Greenbaum, A. McKenney, and D. Sorensen · 1999
Earlier work this paper cites.
FLAME: Formal Linear Algebra Methods Environment
John A. Gunnels, Fred G. Gustavson, Greg M. Henry, and Robert A. van de Geijn · 2001
Earlier work this paper cites.
Notes on equilibria in symmetric games
Shih fen Cheng, Daniel M. Reeves, Yevgeniy Vorobeychik, and Michael P. Wellman · 2004
Earlier work this paper cites.
An API for manipulating matrices stored by blocks
Tze Meng Low and Robert van de Geijn · 2004
Earlier work this paper cites.
Synthesis of high-performance parallel programs for a class of ab initio quantum chemistry models
G. Baumgartner, A. Auer, D.E. Bernholdt, A. Bibireata, V. Choppella, D. Cociorva, X. Gao, R.J. Harrison, S. Hirata, S. Krishnamoorthy, S. Krishnan, C. Lam, Q. Lu, M. Nooijen, R.M. Pitzer, J. Ramanujam, P. Sadayappan, and A. Sibiryakov · 2005
Earlier work this paper cites.
SPIRAL: Code generation for DSP transforms
Markus Püschel, José M. F. Moura, Jeremy Johnson, David Padua, Manuela Veloso, Bryan W. Singer, Jianxin Xiong, Franz Franchetti, Aca Gačić, Yevgen Voronenko, Kang Chen, Robert W. Johnson, and Nick Rizzolo · 2005
Earlier work this paper cites.
Four-index integral transformation exploiting symmetry
Shigeyoshi Yamamoto and Umpei Nagashima · 2005
Cited alongside, same era.
Algorithm 862: MATLAB tensor classes for fast algorithm prototyping
Brett W. Bader and Tamara G. Kolda · 2006
Cited alongside, same era.
High-performance implementation of the level-3 BLAS
Kazushige Goto and Robert van de Geijn · 2008
Cited alongside, same era.
Anatomy of high-performance matrix multiplication
Kazushige Goto and Robert A. van de Geijn · 2008
Cited alongside, same era.
Numerical linear algebra on emerging architectures: The PLASMA and MAGMA projects
E. Agullo, J. Demmel, J. Dongarra, B. Hadri, J. Kurzak, J. Langou, H. Ltaief, P. Luszczek, and S. Tomov · 2009
Cited alongside, same era.
Level-3 BLAS on a GPU: Picking the low hanging fruit. FLAME Working Note #37
Francisco D. Igual, Gregorio Quintana-Ortí, and Robert van de Geijn · 2009
Matlab tensor toolbox version 2.5 [Online]
Brett W. Bader, Tamara G. Kolda, et al · 2012
Later among the works it cites.
PLS toolbox [Online]
Eigenvector Research, Inc · 2012
Later among the works it cites.
The FLAME approach: From dense linear algebra algorithms to high-performance multi-accelerator implementations
Francisco D. Igual, Ernie Chan, Enrique S. Quintana-Ortí, Gregorio Quintana-Ortí, Robert A. van de Geijn, and Field G. Van Zee · 2012
Later among the works it cites.
Block tensor unfoldings
Stefan Ragnarsson and Charles F. Van Loan · 2012
Later among the works it cites.
A preliminary analysis of Cyclops Tensor Framework
Edgar Solomonik, Jeff Hammond, and James Demmel · 2012
Later among the works it cites.
BLIS: A framework for generating BLAS-like libraries
Field G. Van Zee and Robert A. van de Geijn. FLAME Working Note #66 · 2012
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Tensor decompositions and applications
T. G. Kolda and B. W. Bader · 2009
Cited alongside, same era.
Solving “large” dense matrix problems on multi-core processors and GPUs
Mercedes Marqués, Gregorio Quintana-Ortí, Enrique S. Quintana-Ortí, and Robert van de Geijn · 2009
Cited alongside, same era.
Programming matrix algorithms-by-blocks for thread-level parallelism
Gregorio Quintana-Ortí, Enrique S. Quintana-Ortí, Robert A. van de Geijn, Field G. Van Zee, and Ernie Chan · 2009
Cited alongside, same era.
Multilinear models: Iterative methods
G. Tomasi and R. Bro · 2009
Cited alongside, same era.
The libflame library for dense matrix computations
Field G. Van Zee, Ernie Chan, Robert van de Geijn, Enrique S. Quintana-Ortí, and Gregorio Quintana-Ortí · 2009
Cited alongside, same era.
libflame
Field G. Van Zee · 2009
Cited alongside, same era.
Model-driven level 3 BLAS performance optimization on Loongson 3A processor
Zhang Xianyi, Wang Qian, and Zhang Yunquan · 2012
Later among the works it cites.
New implementation of high-level correlated methods using a general block tensor library for high-performance electronic structure calculations
Evgeny Epifanovsky, Michael Wormit, Tomasz Kus, Arie Landau, Dmitry Zuev, Kirill Khistyaev, Prashant Manohar, Ilya Kaliman, Andreas Dreuw, and Anna I. Krylov · 2013
Closest in time.
Jacobi algorithm for the best low multilinear rank approximation of symmetric tensors
M. Ishteva, P. Absil, and P. Van Dooren · 2013
Closest in time.
Fast alternating ls algorithms for high order candecomp/parafac tensor factorizations
Anh-Huy Phan, P. Tichavsky, and A. Cichocki · 2013
Closest in time.
Elemental: A new framework for distributed memory dense matrix computations
Jack Poulson, Bryan Marker, Robert A. van de Geijn, Jeff R. Hammond, and Nichols A. Romero · 2013
Closest in time.
Block tensors and symmetric embeddings
Stefan Ragnarsson and Charles F. Van Loan · 2013
Closest in time.
Monotonically convergent algorithms for symmetric tensor approximation
Phillip A. Regalia · 2013
Closest in time.