arrow
Return

Efficient Network Traffic Analysis Using Large-Parameter LLMs on Consumer-Grade GPUs

delete2025-11-23
delete0
delete
OA
AI
W
Wang, Qingxuan
Z
Zhihua Wang
D
Duo Chen
L
Lizhao You *
DOI:10.3390/math13233754delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
With the growth of network scale and the sophistication of cyberattacks, traditional learning-based traffic analysis methods struggle to maintain generalization. While Large Language Model (LLM)-based approaches offer improved generalization, they suffer from low training and inference efficiency on consumer-grade GPU platforms-typical in resource-constrained deployment scenarios. As a result, existing LLM-based methods often rely on small-parameter models, which limit their effectiveness. To overcome these limitations, we propose to use a large-parameter LLM-based algorithm for network traffic analysis that enhances both generalization and performance. We further introduce two key techniques to enable practical deployment and improve efficiency on consumer-grade GPUs: (a) a traffic-to-text mapping strategy that allows LLMs to process raw network traffic, coupled with a LoRA-based fine-tuning mechanism to improve adaptability across downstream tasks while reducing training overhead; and (b) a sparsity-aware inference acceleration mechanism that employs a hot-cold neuron allocation strategy to alleviate hardware bottlenecks and predicts inactive neurons to skip redundant computations. Experimental results on a consumer-grade NVIDIA RTX A6000 GPU show that our method outperforms existing LLM-based approaches by 6-8% in accuracy across various network traffic analysis tasks, benefiting from the adoption of large-parameter models. Furthermore, our approach achieves up to a 4.07x improvement in inference efficiency compared with llama.cpp, demonstrating both the effectiveness and practicality of the proposed design for real-world network traffic analysis applications.
Keywords:
network security
network traffic analysis
large language model
LoRA fine-tuning
sparsity-aware acceleration
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Mathematics cover
Mathematics
IF:
2.2
Papers:
2.9K
Citations:
3.6W

Organization

C
China Southern Power Grid
Scholars:
3.4K
Papers: 2.4K
Citations: 8
X
Xiamen University
Scholars:
5.3K
Papers: 1.7K
Citations: 6.2W