中文
相关论文

相关论文: Optimal and Private Learning from Human Response D…

200 篇论文

We study statistical risk minimization problems under a privacy model in which the data is kept confidential even from the learner. In this local privacy framework, we establish sharp upper and lower bounds on the convergence rates of…

机器学习 · 统计学 2013-10-11 John C. Duchi , Michael I. Jordan , Martin J. Wainwright

Personalized medicine has gained much popularity recently as a way of providing better healthcare by tailoring treatments to suit individuals. Our research, motivated by the UK INTERVAL blood donation trial, focuses on estimating the…

统计方法学 · 统计学 2023-02-24 Yuejia Xu , Angela M. Wood , David J. Roberts , Brian D. M. Tom

Aligning human preference and value is an important requirement for contemporary foundation models. State-of-the-art techniques such as Reinforcement Learning from Human Feedback (RLHF) often consist of two stages: 1) supervised fine-tuning…

人工智能 · 计算机科学 2024-10-29 Jiaxiang Li , Siliang Zeng , Hoi-To Wai , Chenliang Li , Alfredo Garcia , Mingyi Hong

Protecting individual privacy is crucial when releasing sensitive data for public use. While data de-identification helps, it is not enough. This paper addresses parameter estimation in scenarios where data are perturbed using the…

统计方法学 · 统计学 2024-03-13 Qinglong Tian , Jiwei Zhao

The Intelligent Fault Diagnosis of rotating machinery currently proposes some captivating challenges. Although results achieved by artificial intelligence and deep learning constantly improve, this field is characterized by several open…

信号处理 · 电气工程与系统科学 2022-07-26 Eugenio Brusa , Cristiana Delprete , Luigi Gianpio Di Maggio

The Ising model, originally developed as a spin-glass model for ferromagnetic elements, has gained popularity as a network-based model for capturing dependencies in agents' outputs. Its increasing adoption in healthcare and the social…

统计方法学 · 统计学 2024-01-31 Abhinav Chakraborty , Anirban Chatterjee , Abhinandan Dalal

In this paper, the inverse reinforcement learning (IRL) problem is addressed to reconstruct the unknown cost function underlying an observed optimal policy in a model-free manner, whose online adaptation with completely off-policy system…

最优化与控制 · 数学 2025-11-20 Yibei Li , Yuexin Cao , Zhixin Liu , Lihua Xie

Private information retrieval (PIR) is a privacy setting that allows a user to download a required message from a set of messages stored in a system of databases without revealing the index of the required message to the databases. PIR was…

信息论 · 计算机科学 2023-04-28 Sajani Vithana , Zhusheng Wang , Sennur Ulukus

Given a large number of covariates $Z$, we consider the estimation of a high-dimensional parameter $\theta$ in an individualized linear threshold $\theta^T Z$ for a continuous variable $X$, which minimizes the disagreement between…

统计理论 · 数学 2019-05-28 Huijie Feng , Yang Ning , Jiwei Zhao

Ultra-low-resolution Infrared (IR) array sensors offer a low-cost, energy-efficient, and privacy-preserving solution for people counting, with applications such as occupancy monitoring. Previous work has shown that Deep Learning (DL) can…

计算机视觉与模式识别 · 计算机科学 2023-12-06 Chen Xie , Francesco Daghero , Yukai Chen , Marco Castellano , Luca Gandolfi , Andrea Calimera , Enrico Macii , Massimo Poncino , Daniele Jahier Pagliari

Learning a reward model (RM) from human preferences has been an important component in aligning large language models (LLMs). The canonical setup of learning RMs from pairwise preference data is rooted in the classic Bradley-Terry (BT)…

机器学习 · 计算机科学 2024-11-21 Shang Liu , Yu Pan , Guanting Chen , Xiaocheng Li

Privacy-preserving data analysis is a rising challenge in contemporary statistics, as the privacy guarantees of statistical methods are often achieved at the expense of accuracy. In this paper, we investigate the tradeoff between…

机器学习 · 统计学 2020-11-11 T. Tony Cai , Yichen Wang , Linjun Zhang

Transparency and explainability are two important aspects to be considered when employing black-box machine learning models in high-stake applications. Providing counterfactual explanations is one way of catering this requirement. However,…

信息论 · 计算机科学 2025-08-06 Shreya Meel , Mohamed Nomeir , Pasan Dissanayake , Sanghamitra Dutta , Sennur Ulukus

Randomized response, as a basic building-block for differentially private mechanism, has given rise to great interest and found various potential applications in science communities. In this work, we are concerned with three-elements…

密码学与安全 · 计算机科学 2021-12-15 Fei Ma , Ping Wang

Inverse reinforcement learning (IRL) aims to infer an agent's preferences (represented as a reward function $R$) from their behaviour (represented as a policy $\pi$). To do this, we need a behavioural model of how $\pi$ relates to $R$. In…

机器学习 · 计算机科学 2024-03-12 Joar Skalse , Alessandro Abate

The current literature on memorization in Natural Language Models, especially Large Language Models (LLMs), poses severe security and privacy risks, as models tend to memorize personally identifying information (PIIs) from training data. We…

计算与语言 · 计算机科学 2026-02-19 Kunj Joshi , David A. Smith

Spectral algorithms are an important building block in machine learning and graph algorithms. We are interested in studying when such algorithms can be applied directly to provide optimal solutions to inference tasks. Previous works by…

数据结构与算法 · 计算机科学 2022-10-13 Souvik Dhara , Julia Gaudio , Elchanan Mossel , Colin Sandon

To promote precision medicine, individualized treatment regimes (ITRs) are crucial for optimizing the expected clinical outcome based on patient-specific characteristics. However, existing ITR research has primarily focused on scenarios…

统计方法学 · 统计学 2024-02-20 Chang Wang , Lu Wang

The Internet of things (IoT) is a rapidly advancing area of technology that has quickly become more widespread in recent years. With greater numbers of everyday objects being connected to the Internet, many different innovations have been…

网络与互联网体系结构 · 计算机科学 2022-02-08 Zachary Menter , Wei Tee , Rushit Dave

We study a problem of privacy-preserving mechanism design. A data collector wants to obtain data from individuals to perform some computations. To relieve the privacy threat to the contributors, the data collector adopts a…

计算机科学与博弈论 · 计算机科学 2019-11-12 Guocheng Liao , Xu Chen , Jianwei Huang
‹ 上一页 1 8 9 10 下一页 ›