中文
相关论文

相关论文: Stacking classifiers for anti-spam filtering of e-…

200 篇论文

Simulation-based inference has been popular for amortized Bayesian computation. It is typical to have more than one posterior approximation, from different inference algorithms, different architectures, or simply the randomness of…

统计方法学 · 统计学 2024-03-04 Yuling Yao , Bruno Régaldo-Saint Blancard , Justin Domke

Matrix factorization is one of the best approaches for collaborative filtering, because of its high accuracy in presenting users and items latent factors. The main disadvantages of matrix factorization are its complexity, and being very…

机器学习 · 计算机科学 2017-08-10 Mostafa A. Shehata , Mohammad Nassef , Amr A. Badr

Smishing, also known as SMS phishing, is a type of fraudulent communication in which an attacker disguises SMS communications to deceive a target into providing their sensitive data. Smishing attacks use a variety of tactics; however, they…

密码学与安全 · 计算机科学 2024-04-30 Daniel Timko , Muhammad Lutfor Rahman

Online reviews play a crucial role in helping consumers evaluate and compare products and services. However, review hosting sites are often targeted by opinion spamming. In recent years, many such sites have put a great deal of effort in…

社会与信息网络 · 计算机科学 2016-11-22 Huayi Li , Geli Fei , Shuai Wang , Bing Liu , Weixiang Shao , Arjun Mukherjee , Jidong Shao

Misclassifications in spam and phishing detection are very harmful, as false negatives expose users to attacks while false positives degrade trust. Existing uncertainty-based detectors can flag potential errors, but possibly be deceived and…

人工智能 · 计算机科学 2026-02-18 Qi Zhang , Dian Chen , Lance M. Kaplan , Audun Jøsang , Dong Hyun Jeong , Feng Chen , Jin-Hee Cho

To date, most studies on spam have focused only on the spamming phase of the spam cycle and have ignored the harvesting phase, which consists of the mass acquisition of email addresses. It has been observed that spammers conceal their…

社会与信息网络 · 计算机科学 2013-05-02 Kevin S. Xu , Mark Kliger , Yilun Chen , Peter J. Woolf , Alfred O. Hero

Misspelled words of the malicious kind work by changing specific keywords and are intended to thwart existing automated applications for cyber-environment control such as harassing content detection on the Internet and email spam detection.…

计算与语言 · 计算机科学 2019-01-24 Hongyu Gong , Yuchen Li , Suma Bhat , Pramod Viswanath

Stack filters are a special case of non-linear filters. They have a good performance for filtering images with different types of noise while preserving edges and details. A stack filter decomposes an input image into stacks of binary…

计算机视觉与模式识别 · 计算机科学 2013-06-11 María Elena Buemi , Alejandro C. Frery , Heitor S. Ramos

We provide a simple method for improving the performance of the recently introduced learned Bloom filters, by showing that they perform better when the learned function is sandwiched between two Bloom filters.

数据结构与算法 · 计算机科学 2018-03-06 Michael Mitzenmacher

Ensembling methods are well known for improving prediction accuracy. However, they are limited in the sense that they cannot discriminate among component models effectively. In this paper, we propose stacking with auxiliary features that…

计算与语言 · 计算机科学 2016-05-30 Nazneen Fatema Rajani , Raymond J. Mooney

Collaborative filtering is amongst the most preferred techniques when implementing recommender systems. Recently, great interest has turned towards parallel and distributed implementations of collaborative filtering algorithms. This work is…

信息检索 · 计算机科学 2014-09-10 Efthalia Karydi , Konstantinos G. Margaritis

The spread of unwanted or malicious content through social media has become a major challenge. Traditional examples of this include social network spam, but an important new concern is the propagation of fake news through social media. A…

多智能体系统 · 计算机科学 2018-01-26 Sixie Yu , Yevgeniy Vorobeychik , Scott Alfeld

Mobile network operators implement firewalls to stop illicit messages, but scammers find ways to evade detection. Previous work has looked into SMS texts that are blocked by these firewalls. However, there is little insight into SMS texts…

密码学与安全 · 计算机科学 2025-11-11 Sharad Agarwal , Guillermo Suarez-Tangil , Marie Vasek

Text classification is a task of automatic classification of text into one of the predefined categories. The problem of text classification has been widely studied in different communities like natural language processing, data mining and…

计算与语言 · 计算机科学 2014-06-24 Reshma Prasad , Mary Priya Sebastian

Web spam is a big challenge for quality of search engine results. It is very important for search engines to detect web spam accurately. In this paper we present 32 low cost quality factors to classify spam and ham pages on real time basis.…

信息检索 · 计算机科学 2014-10-09 Ashish Chandra , Mohammad Suaib , Dr. Rizwan Beg

Gradient sparsification is a widely adopted solution for reducing the excessive communication traffic in distributed deep learning. However, most existing gradient sparsifiers have relatively poor scalability because of considerable…

机器学习 · 计算机科学 2023-07-17 Daegun Yoon , Sangyoon Oh

Stacking (or stacked generalization) is an ensemble learning method with one main distinctiveness from the rest: even though several base models are trained on the original data set, their predictions are further used as input data for one…

机器学习 · 计算机科学 2024-04-19 Ilya Ploshchik , Angelos Chatzimparmpas , Andreas Kerren

Recommendation and collaborative filtering systems are important in modern information and e-commerce applications. As these systems are becoming increasingly popular in the industry, their outputs could affect business decision making,…

机器学习 · 计算机科学 2016-10-07 Bo Li , Yining Wang , Aarti Singh , Yevgeniy Vorobeychik

Negative sampling is essential for implicit-feedback-based collaborative filtering, which is used to constitute negative signals from massive unlabeled data to guide supervised learning. The state-of-the-art idea is to utilize hard negative…

信息检索 · 计算机科学 2023-08-14 Yuhan Zhao , Rui Chen , Riwei Lai , Qilong Han , Hongtao Song , Li Chen

Generalized Additive Models (GAMs) have quickly become the leading choice for inherently-interpretable machine learning. However, unlike uninterpretable methods such as DNNs, they lack expressive power and easy scalability, and are hence…

机器学习 · 计算机科学 2022-10-20 Abhimanyu Dubey , Filip Radenovic , Dhruv Mahajan