Fetching the paper…
Reading the bibliography…
In this paper, we introduce SecQA, a novel dataset tailored for evaluating the performance of Large Language Models (LLMs) in the domain of computer security.
Measuring Massive Multitask Language Understanding
Hendrycks, D.; Burns, C.; Basart, S.; Zou, A.; Mazeika, M.; Song, D.; and Steinhardt, J. 2020 · 2009
Earlier work this paper cites.
V2W-BERT: A Framework for Effective Hierarchical Multiclass Classification of Software Vulnerabilities
Das, S. S.; Serra, E.; Halappanavar, M.; Pothen, A.; and Al-Shaer, E. 2021 · 2021
Earlier work this paper cites.
A Framework for Few-Shot Language Model Evaluation
Gao, L.; Tow, J.; Biderman, S.; Black, S.; DiPofi, A.; Foster, C.; Golding, L.; Hsu, J.; McDonell, K.; Muennighoff, N.; et al. 2021 · 2021
Earlier work this paper cites.
MalBERT: Using Transformers for Cybersecurity and Malicious Software Detection
Rahali, A.; and Akhloufi, M. A. 2021 · 2021
Earlier work this paper cites.
CVSS-BERT: Explainable Natural Language Processing to Determine the Severity of a Computer Security Vulnerability from its Description
Shahid, M. R.; and Debar, H. 2021 · 2021
Earlier work this paper cites.
CyNER: A Python Library for Cybersecurity Named Entity Recognition
Alam, M. T.; Bhusal, D.; Park, Y.; and Rastogi, N. 2022 · 2022
Earlier work this paper cites.
Mapping Linux Shell Commands to MITRE ATT&CK using NLP-Based Approach
Andrew, Y.; Lim, C.; and Budiarto, E. 2022 · 2022
Earlier work this paper cites.
Holistic Evaluation of Language Models
Liang, P.; Bommasani, R.; Lee, T.; Tsipras, D.; Soylu, D.; Yasunaga, M.; Zhang, Y.; Narayanan, D.; Wu, Y.; Kumar, A.; et al. 2022 · 2022
Cited alongside, same era.
Training Language Models to Follow Instructions with Human Feedback
Ouyang, L.; Wu, J.; Jiang, X.; Almeida, D.; Wainwright, C.; Mishkin, P.; Zhang, C.; Agarwal, S.; Slama, K.; Ray, A.; et al. 2022 · 2022
Cited alongside, same era.
APTNER: A Specific Dataset for NER Missions in Cyber Threat Intelligence Field
Wang, X.; He, S.; Xiong, Z.; Wei, X.; Jiang, Z.; Chen, S.; and Jiang, J. 2022 · 2022
Cited alongside, same era.
Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
bench authors, B. 2023 · 2023
Cited alongside, same era.
Revolutionizing Cyber Threat Detection with Large Language Models
Ferrag, M. A.; Ndhlovu, M.; Tihanyi, N.; Cordeiro, L. C.; Debbah, M.; and Lestable, T. 2023 · 2023
Liu, Z.; and Buford, J. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
MalBERTv2: Code Aware BERT-Based Model for Malware Identification
Rahali, A.; and Akhloufi, M. A. 2023 · 2023
Closest in time.
Computer Systems Security
Tolboom, R. 2023 · 2023
Closest in time.
Llama 2: Open Foundation and Fine-Tuned Chat Models
Touvron, H.; Martin, L.; Stone, K.; Albert, P.; Almahairi, A.; Babaei, Y.; Bashlykov, N.; Batra, S.; Bhargava, P.; Bhosale, S.; et al. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Jiang, A. Q.; Sablayrolles, A.; Mensch, A.; Bamford, C.; Chaplot, D. S.; Casas, D. d. l.; Bressand, F.; Lengyel, G.; Lample, G.; Saulnier, L.; et al. 2023 · 2023
Cited alongside, same era.
Challenges and Applications of Large Language Models
Kaddour, J.; Harris, J.; Mozes, M.; Bradley, H.; Raileanu, R.; and McHardy, R. 2023 · 2023
Cited alongside, same era.
Tunstall, L.; Beeching, E.; Lambert, N.; Rajani, N.; Rasul, K.; Belkada, Y.; Huang, S.; von Werra, L.; Fourrier, C.; Habib, N.; et al. 2023 · 2023
Closest in time.
Judging LLM-as-a-judge with MT-Bench and Chatbot Arena
Zheng, L.; Chiang, W.-L.; Sheng, Y.; Zhuang, S.; Wu, Z.; Zhuang, Y.; Lin, Z.; Li, Z.; Li, D.; Xing, E. P.; Zhang, H.; Gonzalez, J. E.; and Stoica, I. 2023 · 2023
Closest in time.