中文
相关论文

相关论文: Bias in the Shadows: Explore Shortcuts in Encrypte…

200 篇论文

Traffic classification associates packet streams with known application labels, which is vital for network security and network management. With the rise of NAT, port dynamics, and encrypted traffic, it is increasingly challenging to obtain…

机器学习 · 计算机科学 2021-10-20 Bo Pang , Yongquan Fu , Siyuan Ren , Ye Wang , Qing Liao , Yan Jia

The massive growth of network traffic data leads to a large volume of datasets. Labeling these datasets for identifying intrusion attacks is very laborious and error-prone. Furthermore, network traffic data have complex time-varying…

密码学与安全 · 计算机科学 2022-04-11 Amardeep Singh , Julian Jang-Jaccard

Mobile Internet has profoundly reshaped modern lifestyles in various aspects. Encrypted Traffic Classification (ETC) naturally plays a crucial role in managing mobile Internet, especially with the explosive growth of mobile apps using…

密码学与安全 · 计算机科学 2023-09-07 Xiang Li , Juncheng Guo , Qige Song , Jiang Xie , Yafei Sang , Shuyuan Zhao , Yongzheng Zhang

With the growing significance of network security, the classification of encrypted traffic has emerged as an urgent challenge. Traditional byte-based traffic analysis methods are constrained by the rigid granularity of information and fail…

密码学与安全 · 计算机科学 2025-01-08 Haozhen Zhang , Haodong Yue , Xi Xiao , Le Yu , Qing Li , Zhen Ling , Ye Zhang

Machine learning model bias can arise from dataset composition: correlated sensitive features can distort the downstream classification model's decision boundary and lead to performance differences along these features. Existing de-biasing…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Miao Zhang , Zee fryer , Ben Colman , Ali Shahriyari , Gaurav Bharaj

Recent research has revealed that deep neural networks often take dataset biases as a shortcut to make decisions rather than understand tasks, leading to failures in real-world applications. In this study, we focus on the spurious…

计算与语言 · 计算机科学 2023-06-23 Yanrui Du , Jing Yan , Yan Chen , Jing Liu , Sendong Zhao , Qiaoqiao She , Hua Wu , Haifeng Wang , Bing Qin

Benchmark datasets play an important role in evaluating Natural Language Understanding (NLU) models. However, shortcuts -- unwanted biases in the benchmark datasets -- can damage the effectiveness of benchmark datasets in revealing models'…

人机交互 · 计算机科学 2023-01-16 Zhihua Jin , Xingbo Wang , Furui Cheng , Chunhui Sun , Qun Liu , Huamin Qu

In a single-slot recommendation system, users are only exposed to one item at a time, and the system cannot collect user feedback on multiple items simultaneously. Therefore, only pointwise modeling solutions can be adopted, focusing solely…

信息检索 · 计算机科学 2025-06-03 Chao Wang , Yue Zheng , Yujing Zhang , Yan Feng , Zhe Wang , Xiaowei Shi , An You , Yu Chen

Subspace clustering algorithms are used for understanding the cluster structure that explains the dataset well. These methods are extensively used for data-exploration tasks in various areas of Natural Sciences. However, most of these…

机器学习 · 计算机科学 2022-11-15 Ashutosh Singh , Ashish Singh , Aria Masoomi , Tales Imbiriba , Erik Learned-Miller , Deniz Erdogmus

Bottleneck identification is a challenging task in network analysis, especially when the network is not fully specified. To address this task, we develop a unified online learning framework based on combinatorial semi-bandits that performs…

机器学习 · 计算机科学 2023-03-07 Fazeleh Hoseini , Niklas Åkerblom , Morteza Haghir Chehreghani

Traffic classification, a technique for assigning network flows to predefined categories, has been widely deployed in enterprise and carrier networks. With the massive adoption of mobile devices, encryption is increasingly used in mobile…

网络与互联网体系结构 · 计算机科学 2025-09-03 Kun Qiu , Ying Wang , Baoqian Li , Wenjun Zhu

Traffic classification is vital for cybersecurity, yet encrypted traffic poses significant challenges. We present PacketCLIP, a multi-modal framework combining packet data with natural language semantics through contrastive pretraining and…

密码学与安全 · 计算机科学 2025-03-06 Ryozo Masukawa , Sanggeon Yun , Sungheon Jeong , Wenjun Huang , Yang Ni , Ian Bryant , Nathaniel D. Bastian , Mohsen Imani

The advent of transformer-based language models has reshaped how AI systems process and generate text. In software engineering (SE), these models now support diverse activities, accelerating automation and decision-making. Yet, evidence…

软件工程 · 计算机科学 2026-01-12 Gianmario Voria , Moses Openja , Foutse Khomh , Gemma Catolino , Fabio Palomba

Encrypted network traffic Classification tackles the problem from different approaches and with different goals. One of the common approaches is using Machine learning or Deep Learning-based solutions on a fixed number of classes, leading…

机器学习 · 计算机科学 2024-03-20 Amir Lukach , Ran Dubin , Amit Dvir , Chen Hajaj

The primary objective of an anonymity tool is to protect the anonymity of its users through the implementation of strong encryption and obfuscation techniques. As a result, it becomes very difficult to monitor and identify users activities…

密码学与安全 · 计算机科学 2023-11-29 Javeriah Saleem , Rafiqul Islam , Zahidul Islam

A recent study has shown that large-scale visual datasets are very biased: they can be easily classified by modern neural networks. However, the concrete forms of bias among these datasets remain unclear. In this study, we propose a…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Boya Zeng , Yida Yin , Zhuang Liu

In recent years, the clandestine nature of darknet activities has presented an escalating challenge to cybersecurity efforts, necessitating sophisticated methods for the detection and classification of network traffic associated with these…

机器学习 · 计算机科学 2024-08-31 Anjali Sureshkumar Nair , Prashant Nitnaware

Feedforward neural networks (FNNs) can be viewed as non-linear regression models, where covariates enter the model through a combination of weighted summations and non-linear functions. Although these models have some similarities to the…

统计方法学 · 统计学 2024-05-02 Andrew McInerney , Kevin Burke

In recent years there has been a dramatic increase in the number of malware attacks that use encrypted HTTP traffic for self-propagation or communication. Antivirus software and firewalls typically will not have access to encryption keys,…

密码学与安全 · 计算机科学 2023-12-11 Anish Singh Shekhawat , Fabio Di Troia , Mark Stamp

Multi-class unsupervised anomaly detection (MUAD) has garnered growing research interest, as it seeks to develop a unified model for anomaly detection across multiple classes, i.e., eliminating the need to train separate models for distinct…

人工智能 · 计算机科学 2026-03-31 Peng Tang , Xiaobin Hu , Tingcheng Li , Yang Nan , Tobias Lasser , Hongwei Bran Li