Active_Learning_Framework_techrxiv_2022.pdf (480.37 kB)
Download file

Active Learning Framework to Automate Network Traffic Classification

Download (480.37 kB)
posted on 2022-12-06, 04:48 authored by Jaroslav Pesek, Dominik SoukupDominik Soukup, Tomáš Čejka

Recent network traffic classification methods benefit from machine learning (ML) technology. However, there are many challenges due to use of ML, such as: lack of high-quality annotated datasets, data-drifts and other effects causing aging of datasets and ML models, high volumes of network traffic etc. This paper argues that it is necessary to augment traditional workflows of ML training&deployment and adapt Active Learning concept on network traffic analysis. The paper presents a novel Active Learning Framework (ALF) to address this topic. ALF provides prepared software components that can be used to deploy an active learning loop and maintain an ALF instance that continuously evolves a dataset and ML model automatically. The resulting solution is deployable for IP flow-based analysis of high-speed (100 Gb/s) networks, and also supports research experiments on different strategies and methods for annotation, evaluation, dataset optimization, etc. Finally, the paper lists some research challenges that emerge from the first experiments with ALF in practice.


Sharing and Automation for Privacy Preserving Attack Neutralization

European Commission

Find out more...



Email Address of Submitting Author

ORCID of Submitting Author


Submitting Author's Institution

Czech Technical University in Prague

Submitting Author's Country

  • Czech Republic

Usage metrics