中文
相关论文

相关论文: Off-Policy Evaluation of Probabilistic Identity Da…

200 篇论文

The new information and communication technology providers collect increasing amounts of personal data, a lot of which is user generated. Unless use policies are privacy-friendly, this leaves users vulnerable to privacy risks such as…

计算机与社会 · 计算机科学 2020-05-20 Jana Korunovska , Bernadette Kamleitner , Sarah Spiekermann

This work studies the problem of batch off-policy evaluation for Reinforcement Learning in partially observable environments. Off-policy evaluation under partial observability is inherently prone to bias, with risk of arbitrarily large…

机器学习 · 计算机科学 2019-11-26 Guy Tennenholtz , Shie Mannor , Uri Shalit

Bias in data can have unintended consequences that propagate to the design, development, and deployment of machine learning models. In the financial services sector, this can result in discrimination from certain financial instruments and…

密码学与安全 · 计算机科学 2019-11-12 Reginald Bryant , Celia Cintas , Isaac Wambugu , Andrew Kinai , Komminist Weldemariam

Misinformation is a growing societal threat, and susceptibility to misinformative claims varies across demographic groups due to differences in underlying beliefs. As Large Language Models (LLMs) are increasingly used to simulate human…

计算与语言 · 计算机科学 2026-05-27 Angana Borah , Zohaib Khan , Rada Mihalcea , Verónica Pérez-Rosas

We propose the first boosting algorithm for off-policy learning from logged bandit feedback. Unlike existing boosting methods for supervised learning, our algorithm directly optimizes an estimate of the policy's expected reward. We analyze…

机器学习 · 计算机科学 2023-05-03 Ben London , Levi Lu , Ted Sandler , Thorsten Joachims

Importance sampling (IS) is often used to perform off-policy policy evaluation but is prone to several issues, especially when the behavior policy is unknown and must be estimated from data. Significant differences between the target and…

机器学习 · 计算机科学 2021-11-23 Anton Matsson , Fredrik D. Johansson

For an autonomous agent, executing a poor policy may be costly or even dangerous. For such agents, it is desirable to determine confidence interval lower bounds on the performance of any given policy without executing said policy. Current…

人工智能 · 计算机科学 2018-09-25 Josiah P. Hanna , Peter Stone , Scott Niekum

Personal informatics (PI) systems, powered by smartphones and wearables, enable people to lead healthier lifestyles by providing meaningful and actionable insights that break down barriers between users and their health information. Today,…

计算机与社会 · 计算机科学 2023-07-27 Sofia Yfantidou , Pavlos Sermpezis , Athena Vakali , Ricardo Baeza-Yates

AI-based face recognition, i.e., the re-identification of individuals within images, is an already well established technology for video surveillance, for user authentication, for tagging photos of friends, etc. This paper demonstrates that…

密码学与安全 · 计算机科学 2022-01-26 Stefan Vamosi , Michael Platzer , Thomas Reutterer

We develop a generic data-driven method for estimator selection in off-policy policy evaluation settings. We establish a strong performance guarantee for the method, showing that it is competitive with the oracle estimator, up to a constant…

机器学习 · 计算机科学 2020-08-25 Yi Su , Pavithra Srinath , Akshay Krishnamurthy

DeepFake detection has so far been dominated by ``artifact-driven'' methods and the detection performance significantly degrades when either the type of image artifacts is unknown or the artifacts are simply too hard to find. In this work,…

计算机视觉与模式识别 · 计算机科学 2020-12-08 Xiaoyi Dong , Jianmin Bao , Dongdong Chen , Weiming Zhang , Nenghai Yu , Dong Chen , Fang Wen , Baining Guo

Learning policies on data synthesized by models can in principle quench the thirst of reinforcement learning algorithms for large amounts of real experience, which is often costly to acquire. However, simulating plausible experience de novo…

Offline evaluation plays a central role in benchmarking recommender systems when online testing is impractical or risky. However, it is susceptible to two key sources of bias: exposure bias, where users only interact with items they are…

信息检索 · 计算机科学 2025-08-12 Bruno L. Pereira , Alan Said , Rodrygo L. T. Santos

We linked names and contact information to publicly available profiles in the Personal Genome Project. These profiles contain medical and genomic information, including details about medications, procedures and diseases, and demographic…

计算机与社会 · 计算机科学 2013-04-30 Latanya Sweeney , Akua Abu , Julia Winn

Digital images are ubiquitous in our modern lives, with uses ranging from social media to news, and even scientific papers. For this reason, it is crucial evaluate how accurate people are when performing the task of identify doctored…

图形学 · 计算机科学 2016-01-14 Victor Schetinger , Manuel M. Oliveira , Roberto da Silva , Tiago J. Carvalho

Large Language Model (LLM) outputs often vary across user sociodemographic attributes, leading to disparities in factual accuracy, utility, and safety, even for objective questions where demographic information is irrelevant. Unlike prior…

计算与语言 · 计算机科学 2026-01-15 Miao Zhang , Kelly Chen , Md Mehrab Tanjim , Rumi Chunara

We address the problem of training conversion prediction models in advertising domains under privacy constraints, where direct links between ad clicks and conversions are unavailable. Motivated by privacy-preserving browser APIs and the…

机器学习 · 计算机科学 2026-02-09 Lorne Applebaum , Robert Busa-Fekete , August Y. Chen , Claudio Gentile , Tomer Koren , Aryan Mokhtari

We provide a sharp identification region for discrete choice models where consumers' preferences are not necessarily complete even if only aggregate choice data is available. Behavior is modeled using an upper and a lower utility for each…

计量经济学 · 经济学 2025-02-19 Luca Rigotti , Arie Beresteanu

Our ability to control the flow of sensitive personal information to online systems is key to trust in personal privacy on the internet. We ask how to detect, assess and defend user privacy in the face of search engine personalisation? We…

密码学与安全 · 计算机科学 2016-09-27 Pól Mac Aonghusa , Douglas J. Leith

Training data are critical in face recognition systems. However, labeling a large scale face data for a particular domain is very tedious. In this paper, we propose a method to automatically and incrementally construct datasets from massive…

计算机视觉与模式识别 · 计算机科学 2016-11-28 Shengyong Ding , Junyu Wu , Wei Xu , Hongyang Chao