Fetching the paper…
Reading the bibliography…
The gap between the cost of moving data and the cost of computing continues to grow, making it ever harder to design iterative solvers on extreme-scale architectures.
M. R. Hestenes and E. Stiefel, “Methods of conjugate gradients for solving linear systems,” Journal of research of the National Bureau of Standards , vol. 49, pp. 409–436, 1952
1952
Earlier work this paper cites.
K. Wilson, “Quarks and strings on a lattice,” in New Phenomena in Subnuclear Physics , ser. The Subnuclear Series, A. Zichichi, Ed. Springer US, 1977, vol. 13, pp. 69–142. [Online]. Available: http://dx.doi.org/10.1007/978-1-4613-4208-3_6
1977
Earlier work this paper cites.
B. Sheikholeslami and R. Wohlert, “Improved Continuum Limit Lattice Action for QCD with Wilson Fermions,” Nucl. Phys. , vol. B259, p. 572, 1985
1985
Earlier work this paper cites.
Y. Saad and M. Schultz, “GMRES: A Generalized Minimal Residual Algorithm for Solving Nonsymmetric Linear Systems,” SIAM Journal on Scientific and Statistical Computing , vol. 7, no. 3, pp. 856–869, 1986. [Online]. Available: http://epubs.siam.org/doi/abs/10.1137/0907058
1986
Earlier work this paper cites.
S. Duane, A. D. Kennedy, B. Pendleton, and D. Roweth, “Hybrid Monte Carlo,” Phys. Lett. , vol. B195, pp. 216–222, 1987
1987
Earlier work this paper cites.
H. A. van der Vorst, “BI-CGSTAB: A Fast and Smoothly Converging Variant of BI-CG for the Solution of Nonsymmetric Linear Systems,” SIAM J. Sci. Stat. Comput. , vol. 13, no. 2, pp. 631–644, Mar. 1992. [Online]. Available: http://dx.doi.org/10.1137/0913035
1992
Earlier work this paper cites.
Y. Saad, Iterative Methods for Sparse Linear Systems, Second Edition , 2nd ed. Society for Industrial and Applied Mathematics, 2003
2003
Earlier work this paper cites.
M. Lüscher, “Solution of the Dirac equation in lattice QCD using a domain decomposition method,” Comput. Phys. Commun. , vol. 156, pp. 209–220, 2004
2004
Earlier work this paper cites.
——, “Schwarz-preconditioned HMC algorithm for two-flavour lattice QCD,” Comput. Phys. Commun. , vol. 165, pp. 199–220, 2005
2005
Earlier work this paper cites.
R. G. Edwards and B. Joo, “The Chroma software system for lattice QCD,” Nucl. Phys. Proc. Suppl. , vol. 140, p. 832, 2005
2005
Cited alongside, same era.
M. Lüscher, “Local coherence and deflation of the low quark modes in lattice QCD,” JHEP , vol. 0707, p. 081, 2007
2007
Cited alongside, same era.
K. Bergman et al. , “ExaScale Computing Study: Technology Challenges in Achieving Exascale Systems Peter Kogge, Editor & Study Lead,” 2008
2008
Cited alongside, same era.
M. Lüscher, “Computational strategies in lattice QCD,” in Modern Perspectives in Lattice QCD: Quantum Field Theory and High Performance Computing , L. Lellouch, R. Sommer, B. Svetitsky, A. Vladikas, and L. F. Cugliandolo, Eds., vol. Les Houches 2009, Session XCIII, 2009, pp. 331–399
2009
Cited alongside, same era.
S. Dürr, Z. Fodor, C. Hölbling, R. Hoffmann, S. Katz et al. , “Scaling study of dynamical smeared-link clover fermions,” Phys. Rev. , vol. D79, p. 014501, 2009
2010
Later among the works it cites.
R. Babich, M. A. Clark, B. Joó, G. Shi, R. C. Brower, and S. Gottlieb, “Scaling Lattice QCD Beyond 100 GPUs,” in Proceedings of 2011 International Conference for High Performance Computing, Networking, Storage and Analysis , ser. SC ’11. New York, NY, USA: ACM, 2011, pp. 70:1–70:11. [Online]. Available: http://doi.acm.org/10.1145/2063384.2063478
2011
Later among the works it cites.
2012
Later among the works it cites.
B. Joó, D. D. Kalamkar, K. Vaidyanathan, M. Smelyanskiy, K. Pamnany, V. W. Lee, P. Dubey, and I. Watson, William, “Lattice QCD on Intel Xeon Phi Coprocessors,” in Supercomputing , ser. Lecture Notes in Computer Science, J. M. Kunkel, T. Ludwig, and H. W. Meuer, Eds. Springer Berlin Heidelberg, 2013, vol. 7905, pp. 40–54. [Online]. Available: http://dx.doi.org/10.1007/978-3-642-38750-0_4
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2009
Cited alongside, same era.
A. Nguyen, N. Satish, J. Chhugani, C. Kim, and P. Dubey, “3.5-D Blocking Optimization for Stencil Computations on Modern CPUs and GPUs,” in Proceedings of the 2010 ACM/IEEE International Conference for High Performance Computing, Networking, Storage and Analysis , ser. SC ’10. Washington, DC, USA: IEEE Computer Society, 2010, pp. 1–13. [Online]. Available: http://dx.doi.org/10.1109/SC.2010.2
2010
Cited alongside, same era.
R. Babich, J. Brannick, R. Brower, M. Clark, T. Manteuffel et al. , “Adaptive multigrid algorithm for the lattice Wilson-Dirac operator,” Phys. Rev. Lett. , vol. 105, p. 201602, 2010
2010
Cited alongside, same era.
J. Osborn, R. Babich, J. Brannick, R. Brower, M. Clark et al. , “Multigrid solver for clover fermions,” PoS , vol. LATTICE2010, p. 037, 2010
2010
Cited alongside, same era.
H. A. Schwarz, “Über einen Grenzübergang durch alternierendes Verfahren,” in Vierteljahrsschrift der Naturforschenden Gesellschaft in Zürich , 1870, vol. 15, pp. 272–286
Cited in the paper.
2013
Later among the works it cites.
G. Bali, P. Bruns, S. Collins, M. Deka, B. Gläßle et al. , “Nucleon mass and sigma term from lattice QCD with two light fermion flavors,” Nucl. Phys. , vol. B866, pp. 1–25, 2013
2013
Later among the works it cites.
2013
Later among the works it cites.
2013
Later among the works it cites.
K. Vaidyanathan, K. Pamnany, D. D. Kalamkar, A. Heinecke, M. Smelyanskiy, J. Park, D. Kim, A. Shet, B. Kaul, B. Joo, and P. Dubey, “Improving Communication Performance and Scalability of Native Applications on Intel Xeon Phi Coprocessor Clusters,” in IPDPS 2014 (28th IEEE International Parallel & Distributed Processing Symposium) . to be published
2014
Closest in time.