Fetching the paper…
Reading the bibliography…
This paper addresses the problem of providing a novel approach to sourcing significant training data for LLMs focused on science and engineering.
OpenMP ARB (Architecture Review Boards), “OpenMP 4.0 Reference Guide – C/C++ (October 2013 PDF).” Available from: https://www.openmp.org/resources/refguides/
2013
Earlier work this paper cites.
P. Lewis, et al., “Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks.” Advances in Neural Information Processing Systems (NeurIPS 2020), vol 33, pp 9459–9474, 2020
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
P. E. Black, “Ratcliff/Obershelp pattern recognition,” in Dictionary of Algorithms and Data Structures [online], Paul E. Black, ed. 8 January 2021. Available from: https://www.nist.gov/dads/HTML/ratcliffObershelp.html
2021
Earlier work this paper cites.
Z. Jin and J. S. Vetter, “A Benchmark Suite for Improving Performance Portability of the SYCL Programming Model.” 2023 IEEE International Symposium on Performance Analysis of Systems and Software (ISPASS), 2023, pp. 325-327
2023
Earlier work this paper cites.
Z. Luo et al., “WizardCoder: Empowering Code Large Language Models with Evol-Instruct.” arXiv, Jun. 14, 2023. doi: 10.48550/arXiv.2306.08568
2023
Earlier work this paper cites.
G. Serapio-García et al., “Personality Traits in Large Language Models.” arXiv, Sep. 21, 2023. doi: 10.48550/arXiv.2307.00184
2023
Earlier work this paper cites.
O. Khattab et al., “DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines.” arXiv, Oct. 05, 2023. doi: 10.48550/arXiv.2310.03714
2023
Earlier work this paper cites.
https://www.hpcwire.com/2023/11/13/training-of-1-trillion-parameter-scientific-ai-begins/
2023
Cited alongside, same era.
B. Lei et al., “Creating a Dataset for High-Performance Computing Code Translation using LLMs: A Bridge Between OpenMP Fortran and C++,” 2023 IEEE High Performance Extreme Computing Conference (HPEC), 2023, pp. 1–7
2023
Cited alongside, same era.
OpenAI et al., “GPT-4 Technical Report.” arXiv, Mar. 04, 2024. doi: 10.48550/arXiv.2303.08774
2024
Cited alongside, same era.
Mistral AI Gdzm, 2024 https://mistral.ai/news/codestral/
2024
Cited alongside, same era.
A. Lozhkov et al., “StarCoder 2 and The Stack v2: The Next Generation.” arXiv, Feb. 29, 2024. doi: 10.48550/arXiv.2402.19173
2024
Cited alongside, same era.
C. Ling et al.,“Domain Specialization as the Key to Make Large Language Models Disruptive: A Comprehensive Survey.” arXiv, Mar. 29, 2024. doi: 10.48550/arXiv.2305.18703
2024
Closest in time.
D. Nichols, J. H. Davis, Z. Xie, A. Rajaram, and A. Bhatele, “Can Large Language Models Write Parallel Code?” In The 33rd International Symposium on High-Performance Parallel and Distributed Computing (HPDC ’24), June 3–7, 2024, Pisa, Italy. ACM, New York, NY, USA, 14 pages. https://doi.org/10.1145/3625549.3658689
2024
Closest in time.
NVIDIA et al., “Nemotron-4 340B Technical Report.” arXiv, Jun. 17, 2024. doi: 10.48550/arXiv.2406.11704
2024
Closest in time.
MetaAI. Introducing meta llama 3: The most capable openly available llm to date. https://ai.meta.com/blog/meta-llama-3/, 2024
2024
Closest in time.
Qwen-Team. Hello qwen2. https://qwenlm.github.io/blog/qwen2, 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
B. Rozière et al., “Code Llama: Open Foundation Models for Code.” arXiv, Jan. 31, 2024. doi: 10.48550/arXiv.2308.12950
2024
Cited alongside, same era.
DeepSeek-AI et al., “DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence.” arXiv, Jun. 17, 2024. doi: 10.48550/arXiv.2406.11931
2024
Cited alongside, same era.
Trillion Parameter Consortium (TPC) https://tpc.dev
Cited in the paper.
Mistral-AI-Team. Mistral 8x22b. https://mistral.ai/news/mixtral-8x22b, 2024b
Cited in the paper.
NVIDIA. “CUDA C++ Programming Guide, Release 12.5.” Available from: https://docs.nvidia.com/cuda/pdf/CUDA_C_Programming_Guide.pdf
Cited in the paper.
https://en.wikipedia.org/wiki/GPT-4
Cited in the paper.
16-bit floating-point precision, https://en.wikipedia.org/wiki/Half-precision_floating-point_format
Cited in the paper.
2024
Closest in time.
R. Battle and T. Gollapudi, “The Unreasonable Effectiveness of Eccentric Automatic Prompts.” arXiv, Feb. 20, 2024. doi: 10.48550/arXiv.2402.10949
2024
Closest in time.
Y. Tian et al., “Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing.” arXiv, Apr. 18, 2024. doi: 10.48550/arXiv.2404.12253
2024
Closest in time.