NetBench: A Large-Scale and Comprehensive Network Traffic Benchmark Dataset for Foundation Models

Qian, Chen; Li, Xiaochang; Wang, Qineng; Zhou, Gang; Shao, Huajie

Computer Science > Networking and Internet Architecture

arXiv:2403.10319 (cs)

[Submitted on 15 Mar 2024 (v1), last revised 19 Mar 2024 (this version, v2)]

Title:NetBench: A Large-Scale and Comprehensive Network Traffic Benchmark Dataset for Foundation Models

Authors:Chen Qian, Xiaochang Li, Qineng Wang, Gang Zhou, Huajie Shao

View PDF HTML (experimental)

Abstract:In computer networking, network traffic refers to the amount of data transmitted in the form of packets between internetworked computers or Cyber-Physical Systems. Monitoring and analyzing network traffic is crucial for ensuring the performance, security, and reliability of a network. However, a significant challenge in network traffic analysis is to process diverse data packets including both ciphertext and plaintext. While many methods have been adopted to analyze network traffic, they often rely on different datasets for performance evaluation. This inconsistency results in substantial manual data processing efforts and unfair comparisons. Moreover, some data processing methods may cause data leakage due to improper separation of training and testing data. To address these issues, we introduce the NetBench, a large-scale and comprehensive benchmark dataset for assessing machine learning models, especially foundation models, in both network traffic classification and generation tasks. NetBench is built upon seven publicly available datasets and encompasses a broad spectrum of 20 tasks, including 15 classification tasks and 5 generation tasks. Furthermore, we evaluate eight State-Of-The-Art (SOTA) classification models (including two foundation models) and two generative models using our benchmark. The results show that foundation models significantly outperform the traditional deep learning methods in traffic classification. We believe NetBench will facilitate fair comparisons among various approaches and advance the development of foundation models for network traffic. Our benchmark is available at this https URL.

Subjects:	Networking and Internet Architecture (cs.NI); Cryptography and Security (cs.CR)
Cite as:	arXiv:2403.10319 [cs.NI]
	(or arXiv:2403.10319v2 [cs.NI] for this version)
	https://doi.org/10.48550/arXiv.2403.10319

Submission history

From: Xiaochang Li [view email]
[v1] Fri, 15 Mar 2024 14:09:54 UTC (893 KB)
[v2] Tue, 19 Mar 2024 03:36:53 UTC (898 KB)

Computer Science > Networking and Internet Architecture

Title:NetBench: A Large-Scale and Comprehensive Network Traffic Benchmark Dataset for Foundation Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Networking and Internet Architecture

Title:NetBench: A Large-Scale and Comprehensive Network Traffic Benchmark Dataset for Foundation Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators