Fetching the paper…
Reading the bibliography…
The present panorama of HPC architectures is extremely heterogeneous, ranging from traditional multi-core CPU processors, supporting a wide class of applications but delivering moderate computing performance, to many-core GPUs, exploiting aggressive data-parallelism and delivering higher performances for streaming computing applications.
K. G. Wilson, Physical Review D
1974
Earlier work this paper cites.
D. Weingarten and D. Petcher, Physics Letters B
1981
Earlier work this paper cites.
P. Weisz, Nuclear Physics B
1983
Earlier work this paper cites.
G. Curci, P. Menotti and G. Paffuti, Physics Letters B
1983
Earlier work this paper cites.
P. De Forcrand, D. Lellouch and C. Roiesnel, Journal of Computational Physics
1985
Earlier work this paper cites.
M. Albanese et al
1987
Earlier work this paper cites.
S. Duane, A. D. Kennedy, B. J. Pendleton and D. Roweth, Physics letters B
1987
Earlier work this paper cites.
T. A. Degrand and P. Rossi, Computer Physics Communications
1990
Earlier work this paper cites.
J. Sexton and D. Weingarten, Nuclear Physics B
1992
Earlier work this paper cites.
B. Jegerlehner, arXiv preprint hep-lat/9612014
1996
Earlier work this paper cites.
C. Bernard et al
2002
Earlier work this paper cites.
I. Omelyan, I. Mryglod and R. Folk, Physical Review E
2002
Earlier work this paper cites.
I. Omelyan, I. Mryglod and R. Folk, Computer Physics Communications
2003
Earlier work this paper cites.
doi:10.1109/SC.2004.46
P. A. Boyle, D. Chen, N. H. Christ, M. Clark, S. Cohen, Z. Dong, A. Gara, B. Joo, C. Jung, L. Levkova, X. Liao, G. Liu, R. D. Mawhinney, S. Ohta, K. Petrov, T. Wettig, A. Yamaguchi and C. Cristian, Qcdoc: A 10 teraflops computer for tightly-coupled calculations, in Supercomputing, 2004. Proceedings of the ACM/IEEE SC2004 Conference · 2004
Earlier work this paper cites.
C. Morningstar and M. Peardon, Physical Review D
2004
Earlier work this paper cites.
M. Clark, A. Kennedy and Z. Sroczynski, arXiv preprint hep-lat/0409133
2004
Earlier work this paper cites.
G. Bilardi, A. Pietracaprina, G. Pucci, F. Schifano and R. Tripiccione, Lecture Notes in Computer Science
2005
Earlier work this paper cites.
F. Belletti, S. F. Schifano, R. Tripiccione, F. Bodin, P. Boucaud, J. Micheli, O. Pene, N. Cabibbo, S. De Luca, A. Lonardo, D. Rossetti, P. Vicini, M. Lukyanov, L. Morin, N. Paschedag, H. Simma, V. Morenas, D. Pleiter and F. Rapuano, Computing in Science and Engineering
2006
Earlier work this paper cites.
A. Kennedy, arXiv preprint hep-lat/0607038
2006
Cited alongside, same era.
T. DeGrand and C. DeTar, Lattice methods for quantum chromodynamics
2006
Cited alongside, same era.
C. Urbach, K. Jansen, A. Shindler and U. Wenger, Computer Physics Communications
2006
Cited alongside, same era.
T. Takaishi and P. De Forcrand, Physical Review E
2006
Cited alongside, same era.
G. I. Egri, Z. Fodor, C. Hoelbling, S. D. Katz, D. Nogradi and K. K. Szabo, Comput. Phys. Commun
2007
Cited alongside, same era.
M. Clark and A. Kennedy, Physical review letters
2007
Cited alongside, same era.
M. Clark and A. Kennedy, Physical Review D
doi:10.1109/SC.2014.73
O. Villa, D. R. Johnson, M. O’Connor, E. Bolotin, D. Nellans, J. Luitjens, N. Sakharnykh, P. Wang, P. Micikevicius, A. Scudiero, S. W. Keckler and W. J. Dally, Scaling the power wall: A path to exascale, in Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis · 2014
Later among the works it cites.
S. Wienke, C. Terboven, J. Beyer and M. Müller, Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics)
2014
Later among the works it cites.
doi:10.3204/DESY-PROC-2014-05/30
O. Philipsen, C. Pinke, A. Sciarra and M. Bach, CL & 2 \&2 QCD - Lattice QCD based on OpenCL, in Proceedings, GPU Computing in High-Energy Physics (GPUHEP2014): Pisa, Italy, September 10-12, 2014 · 2014
Later among the works it cites.
doi:10.1109/WACCPD.2014.11
J. Kraus, M. Schlottke, A. Adinetz and D. Pleiter, Accelerating a c++ cfd code with openacc, in Accelerator Programming using Directives (WACCPD), 2014 First Workshop on · 2014
Later among the works it cites.
NVIDIA’s Next Generation CUDA™ Compute Architecture: Kepler™ GK110/210 (2014)
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2007
Cited alongside, same era.
V. Simoncini and D. B. Szyld, Numerical Linear Algebra with Applications
2007
Cited alongside, same era.
K. Barros, R. Babich, R. Brower, M. A. Clark and C. Rebbi, PoS LATTICE 2008
2008
Cited alongside, same era.
H. Baier, H. Boettiger, M. Drochner, N. Eicker, U. Fischer, Z. Fodor, A. Frommer, C. Gomez, G. Goldrian, S. Heybrock, D. Hierl, M. Hüsken, T. Huth, B. Krill, J. Lauritsen, T. Lippert, T. Maurer, B. Mendl, N. Meyer, A. Nobile, I. Ouda, M. Pivanti, D. Pleiter, M. Ries, A. Schäfer, H. Schick, F. Schifano, H. Simma, S. Solbrig, T. Streuer, K.-H. Sulanke, R. Tripiccione, J.-S. Vogt, T. Wettig and F. Winter, Computer Science - Research and Development
2010
Cited alongside, same era.
R. Babich, M. A. Clark and B. Joo, Parallelizing the QUDA Library for Multi-GPU Calculations in Lattice Quantum Chromodynamics, in SC 10 (Supercomputing 2010) New Orleans, Louisiana, November 13-19, 2010
2010
Cited alongside, same era.
M. A. Clark, R. Babich, K. Barros, R. C. Brower and C. Rebbi, Computer Physics Communications
2010
Cited alongside, same era.
doi:10.1109/SE4HPCS.2015.9
C. Bonati, E. Calore, S. Coscetti, M. D’Elia, M. Mesiti, F. Negro, S. F. Schifano and R. Tripiccione, Development of scientific software for HPC architectures using OpenACC: the case of LQCD, in The 2015 International Workshop on Software Engineering for High Performance Computing in Science (SE4HPCS) · 2015
Later among the works it cites.
doi:10.1145/2832105.2832111
S. Blair, C. Albing, A. Grund and A. Jocksch, Accelerating an mpi lattice boltzmann code using openacc, in Proceedings of the Second Workshop on Accelerator Programming Using Directives · 2015
Later among the works it cites.
doi:10.1016/B978-0-12-803819-2.00023-9
B. Joó, M. Smelyanskiy, D. D. Kalamkar and K. Vaidyanathan, Wilson dslash kernel from lattice QCD optimization, in High Performance Parallelism Pearls Volume 2: Multicore and Many-core Programming Approaches · 2015
Later among the works it cites.
doi:10.1016/B978-0-12-803819-2.00023-9
B. Joó, M. Smelyanskiy, D. D. Kalamkar and K. Vaidyanathan, Chapter 9 - wilson dslash kernel from lattice QCD optimization, in High Performance Parallelism Pearls · 2015
Later among the works it cites.
doi:10.1109/SE-HPCCSE.2016.8
D. Pflüger, M. Mehl, J. Valentin, F. Lindner, D. Pfander, S. Wagner, D. Graziotin and Y. Wang, The scalability-efficiency/maintainability-portability trade-off in simulation software engineering: Examples and a preliminary systematic literature review, in Proceedings of the Fourth International Workshop on Software Engineering for HPC in Computational Science and Engineering · 2016
Later among the works it cites.
doi:10.1109/WACCPD.2016.9
M. G. Lopez, V. V. Larrea, W. Joubert, O. Hernandez, A. Haidar, S. Tomov and J. Dongarra, Towards achieving performance portability using directives for accelerators, in Proceedings of the Third International Workshop on Accelerator Programming Using Directives · 2016
Later among the works it cites.
E. Calore, A. Gabbana, J. Kraus, S. F. Schifano and R. Tripiccione, Concurrency and Computation: Practice and Experience
2016
Later among the works it cites.
P. Majumdar, Journal of Physics: Conference Series
2016
Later among the works it cites.
C. Patrignani, P. D. Group et al
2016
Later among the works it cites.
NVIDIA Tesla P100 v1.1 (2016)
2016
Later among the works it cites.
doi:10.1007/978-3-319-32149-3_6
E. Calore, N. Demo, S. F. Schifano and R. Tripiccione, Experience on vectorizing lattice boltzmann kernels for multi- and many-core architectures, in Parallel Processing and Applied Mathematics: 11th International Conference, PPAM 2015, Krakow, Poland, September 6-9, 2015. Revised Selected Papers, Part I · 2016
Later among the works it cites.
Revised Selected Papers
B. Joó, D. D. Kalamkar, T. Kurth, K. Vaidyanathan and A. Walden, Optimizing Wilson-Dirac Operator and Linear Solvers for Intel KNL · 2016
Later among the works it cites.