Evaluation of LLM Chatbots for OSINT-based Cyber Threat Awareness

Shafee, Samaneh; Bessani, Alysson; Ferreira, Pedro M.

Computer Science > Cryptography and Security

arXiv:2401.15127 (cs)

[Submitted on 26 Jan 2024 (v1), last revised 19 Apr 2024 (this version, v3)]

Title:Evaluation of LLM Chatbots for OSINT-based Cyber Threat Awareness

Authors:Samaneh Shafee, Alysson Bessani, Pedro M. Ferreira

View PDF HTML (experimental)

Abstract:Knowledge sharing about emerging threats is crucial in the rapidly advancing field of cybersecurity and forms the foundation of Cyber Threat Intelligence (CTI). In this context, Large Language Models are becoming increasingly significant in the field of cybersecurity, presenting a wide range of opportunities. This study surveys the performance of ChatGPT, GPT4all, Dolly, Stanford Alpaca, Alpaca-LoRA, Falcon, and Vicuna chatbots in binary classification and Named Entity Recognition (NER) tasks performed using Open Source INTelligence (OSINT). We utilize well-established data collected in previous research from Twitter to assess the competitiveness of these chatbots when compared to specialized models trained for those tasks. In binary classification experiments, Chatbot GPT-4 as a commercial model achieved an acceptable F1 score of 0.94, and the open-source GPT4all model achieved an F1 score of 0.90. However, concerning cybersecurity entity recognition, all evaluated chatbots have limitations and are less effective. This study demonstrates the capability of chatbots for OSINT binary classification and shows that they require further improvement in NER to effectively replace specially trained models. Our results shed light on the limitations of the LLM chatbots when compared to specialized models, and can help researchers improve chatbots technology with the objective to reduce the required effort to integrate machine learning in OSINT-based CTI tools.

Subjects:	Cryptography and Security (cs.CR); Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2401.15127 [cs.CR]
	(or arXiv:2401.15127v3 [cs.CR] for this version)
	https://doi.org/10.48550/arXiv.2401.15127

Submission history

From: Samaneh Shafee [view email]
[v1] Fri, 26 Jan 2024 13:15:24 UTC (683 KB)
[v2] Wed, 13 Mar 2024 23:51:13 UTC (1,191 KB)
[v3] Fri, 19 Apr 2024 09:40:04 UTC (1,192 KB)

Computer Science > Cryptography and Security

Title:Evaluation of LLM Chatbots for OSINT-based Cyber Threat Awareness

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Cryptography and Security

Title:Evaluation of LLM Chatbots for OSINT-based Cyber Threat Awareness

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators