TensorX
返回文献探索

Paper · arXiv 2409.11276

Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments

Maria Rigaki, Carlos Catania, Sebastian Garcia

11 upvotesSeptember 17, 2024arXiv 预印本
AI 摘要

Hackphyr, a locally fine-tuned LLM with 7 billion parameters, performs comparably to larger commercial models like GPT-4 and outperforms GPT-3.5-turbo and Q-learning agents in cybersecurity tasks using a custom dataset and runs on a single GPU.

Large Language ModelsLLMsred-team agentnetwork securityfine-tunedGPT-4GPT-3.5-turboQ-learning agentstask-specific cybersecurity datasetparameter-efficient fine-tuning

Abstract

Large Language Models (LLMs) have shown remarkable potential across various domains, including cybersecurity. Using commercial cloud-based LLMs may be undesirable due to privacy concerns, costs, and network connectivity constraints. In this paper, we present Hackphyr, a locally fine-tuned LLM to be used as a red-team agent within network security environments. Our fine-tuned 7 billion parameter model can run on a single GPU card and achieves performance comparable with much larger and more powerful commercial models such as GPT-4. Hackphyr clearly outperforms other models, including GPT-3.5-turbo, and baselines, such as Q-learning agents in complex, previously unseen scenarios. To achieve this performance, we generated a new task-specific cybersecurity dataset to enhance the base model's capabilities. Finally, we conducted a comprehensive analysis of the agents' behaviors that provides insights into the planning abilities and potential shortcomings of such agents, contributing to the broader understanding of LLM-based agents in cybersecurity contexts

北京市昌平区探索星信息技术及软件开发工作室

京ICP备2026059466号
Hackphyr: A Local Fine-Tuned LLM Agent for Network Security Environments | TensorX