English
Related papers

Related papers: Bias in the Shadows: Explore Shortcuts in Encrypte…

200 papers

Dataset bias, where data points are skewed to certain concepts, is ubiquitous in machine learning datasets. Yet, systematically identifying these biases is challenging without costly, fine-grained attribute annotations. We present…

Computer Vision and Pattern Recognition · Computer Science 2025-10-31 Jinho Choi , Hyesu Lim , Steffen Schneider , Jaegul Choo

Visual data from the Web power image classifiers, which often underpin many web services, such as recommendation and content moderation. However, the raw Web data often contain spurious correlations and social biases, and neural networks…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Jungwook Seo , Yoonsik Park , Changmin Lee , Sungyong Baik

Internet traffic classification has become more important with rapid growth of current Internet network and online applications. There have been numerous studies on this topic which have led to many different approaches. Most of these…

Network traffic classification (NTC) is vital for efficient network management, security, and performance optimization, particularly with 5G/6G technologies. Traditional methods, such as deep packet inspection (DPI) and port-based…

Networking and Internet Architecture · Computer Science 2025-09-30 Ehsan Eslami , Walaa Hamouda

Language models are prone to dataset biases, known as shortcuts and spurious correlations in data, which often result in performance drop on new data. We present a new debiasing framework called ``FairFlow'' that mitigates dataset biases by…

Machine Learning · Computer Science 2025-03-25 Jiali Cheng , Hadi Amiri

Computer vision datasets frequently contain spurious correlations between task-relevant labels and (easy to learn) latent task-irrelevant attributes (e.g. context). Models trained on such datasets learn "shortcuts" and underperform on…

Computer Vision and Pattern Recognition · Computer Science 2023-10-02 Sriram Yenamandra , Pratik Ramesh , Viraj Prabhu , Judy Hoffman

We study fairness in supervised few-shot meta-learning models that are sensitive to discrimination (or bias) in historical data. A machine learning model trained based on biased data tends to make unfair predictions for users from minority…

Machine Learning · Computer Science 2020-09-25 Chen Zhao , Feng Chen

Monitoring network traffic to identify content, services, and applications is an active research topic in network traffic control systems. While modern firewalls provide the capability to decrypt packets, this is not appealing for privacy…

Networking and Internet Architecture · Computer Science 2021-06-25 Niloofar Bayat , Weston Jackson , Derrick Liu

Network traffic classification, a task to classify network traffic and identify its type, is the most fundamental step to improve network services and manage modern networks. Classical machine learning and deep learning method have…

Networking and Internet Architecture · Computer Science 2021-07-09 Yao Peng , Meirong He , Yu Wang

The accuracy and fairness of perception systems in autonomous driving are essential, especially for vulnerable road users such as cyclists, pedestrians, and motorcyclists who face significant risks in urban driving environments. While…

Computer Vision and Pattern Recognition · Computer Science 2025-05-22 Dewant Katare , David Solans Noguero , Souneil Park , Nicolas Kourtellis , Marijn Janssen , Aaron Yi Ding

Supervised machine learning techniques rely on labeled data to achieve high task performance, but this requires the labels to capture some meaningful differences in the underlying data structure. For training network intrusion detection…

Cryptography and Security · Computer Science 2025-09-12 Meghan Wilkinson , Robert H Thomson

In contrast to previous surveys, the present work is not focused on reviewing the datasets used in the network security field. The fact is that many of the available public labeled datasets represent the network behavior just for a…

Cryptography and Security · Computer Science 2022-01-03 Jorge Guerra , Carlos Catania , Eduardo Veas

Self-supervised masked modeling shows promise for encrypted traffic classification by masking and reconstructing raw bytes. Yet recent work reveals these methods fail to reduce reliance on labeled data despite costly pretraining: under…

Networking and Internet Architecture · Computer Science 2026-05-12 Sizhe Huang , Zitong Li , Shujie Yang

Network traffic includes data transmitted across a network, such as web browsing and file transfers, and is organized into packets (small units of data) and flows (sequences of packets exchanged between two endpoints). Classifying encrypted…

Cryptography and Security · Computer Science 2024-12-23 Xu-Yang Chen , Lu Han , De-Chuan Zhan , Han-Jia Ye

Recent advancements in deep learning have significantly enhanced the performance and efficiency of traffic classification in networking systems. However, the lack of transparency in their predictions and decision-making has made network…

Networking and Internet Architecture · Computer Science 2025-09-23 Riya Ponraj , Ram Durairajan , Yu Wang

Recent works find that AI algorithms learn biases from data. Therefore, it is urgent and vital to identify biases in AI algorithms. However, the previous bias identification pipeline overly relies on human experts to conjecture potential…

Computer Vision and Pattern Recognition · Computer Science 2021-10-05 Zhiheng Li , Chenliang Xu

Machine learning (ML) is promising in accurately detecting malicious flows in encrypted network traffic; however, it is challenging to collect a training dataset that contains a sufficient amount of encrypted malicious data with correct…

Cryptography and Security · Computer Science 2023-09-12 Yuqi Qing , Qilei Yin , Xinhao Deng , Yihao Chen , Zhuotao Liu , Kun Sun , Ke Xu , Jia Zhang , Qi Li

Labeled datasets reflect the biases of their annotation pipelines, which sometimes introduce label bias: group-conditional label errors that cause systematic performance disparities across demographic subgroups. Label bias in image…

Computer Vision and Pattern Recognition · Computer Science 2026-05-11 Aditya Parikh , Stella Frank , Sneha Das , Aasa Feragen

The widespread adoption of deep-learning models in data-driven applications has drawn attention to the potential risks associated with biased datasets and models. Neglected or hidden biases within datasets and models can lead to unexpected…

Machine Learning · Computer Science 2026-01-27 Md Sahidullah , Hye-jin Shim , Rosa Gonzalez Hautamäki , Tomi H. Kinnunen

In this paper, we present three datasets that have been built from network traffic traces using ASNM features, designed in our previous work. The first dataset was built using a state-of-the-art dataset called CDX 2009, while the remaining…

Cryptography and Security · Computer Science 2020-07-23 Ivan Homoliak , Petr Hanacek