中文
相关论文

相关论文: Testing Preferential Domains Using Sampling

200 篇论文

Some examples are easier for humans to classify than others. The same should be true for deep neural networks (DNNs). We use the term example perplexity to refer to the level of difficulty of classifying an example. In this paper, we…

机器学习 · 计算机科学 2022-03-18 Nevin L. Zhang , Weiyan Xie , Zhi Lin , Guanfang Dong , Xiao-Hui Li , Caleb Chen Cao , Yunpeng Wang

We study the general problem of testing whether an unknown distribution belongs to a specified family of distributions. More specifically, given a distribution family $\mathcal{P}$ and sample access to an unknown discrete distribution…

数据结构与算法 · 计算机科学 2017-08-09 Clément L. Canonne , Ilias Diakonikolas , Alistair Stewart

Quality by design in pharmaceutical manufacturing hinges on computational methods and tools that are capable of accurate quantitative prediction of the design space. This paper investigates Bayesian approaches to design space…

Level-1 Consensus is a property of a preference-profile. Intuitively, it means that there exists a preference relation which induces an ordering of all other preferences such that frequent preferences are those that are more similar to it.…

计算机科学与博弈论 · 计算机科学 2017-12-20 Mor Nitzan , Shmuel Nitzan , Erel Segal-Halevi

We study a public decision problem in which a finite society selects a public-good level from a closed interval. Agents either have single-peaked preferences or are completely indifferent over the interval; the latter capture abstention or…

理论经济学 · 经济学 2026-03-19 Parikshit De , Abinash Panda , Anup Pramanik

Starting with a set of weighted items, we want to create a generic sample of a certain size that we can later use to estimate the total weight of arbitrary subsets. For this purpose, we propose priority sampling which tested on Internet…

数据结构与算法 · 计算机科学 2007-05-23 Nick Duffield , Carsten Lund , Mikkel Thorup

Learning the preferences of a human improves the quality of the interaction with the human. The number of queries available to learn preferences maybe limited especially when interacting with a human, and so active learning is a must. One…

机器学习 · 计算机科学 2020-02-18 Sriram Gopalakrishnan , Utkarsh Soni

The paper delineates a proper statistical setting for defining the sampling design for a small area estimation problem. This problem is often treated only via indirect estimation using the values of the variable of interest also from…

统计方法学 · 统计学 2023-03-16 Piero Demetrio Falorsi , Stefano Falorsi , Vincenzo Nardelli , Paolo Righi

What advantage do \emph{sequential} procedures provide over batch algorithms for testing properties of unknown distributions? Focusing on the problem of testing whether two distributions $\mathcal{D}_1$ and $\mathcal{D}_2$ on $\{1,\dots,…

数据结构与算法 · 计算机科学 2022-05-13 Omar Fawzi , Nicolas Flammarion , Aurélien Garivier , Aadil Oufkir

We revisit the distributed hypothesis testing (or hypothesis testing with communication constraints) problem from the viewpoint of privacy. Instead of observing the raw data directly, the transmitter observes a sanitized or randomized…

信息论 · 计算机科学 2019-06-26 Atefeh Gilani , Selma Belhadj Amor , Sadaf Salehkalaibar , Vincent Y. F. Tan

Estimating properties of discrete distributions is a fundamental problem in statistical learning. We design the first unified, linear-time, competitive, property estimator that for a wide class of properties and for all underlying…

机器学习 · 统计学 2019-04-02 Yi Hao , Alon Orlitsky , Ananda T. Suresh , Yihong Wu

Preferential sampling provides a formal modeling specification to capture the effect of bias in a set of sampling locations on inference when a geostatistical model is used to explain observed responses at the sampled locations. In…

统计方法学 · 统计学 2022-02-21 Shinichiro Shirota , Alan E. Gelfand

This paper establishes problem-specific sample complexity lower bounds for linear system identification problems. The sample complexity is defined in the PAC framework: it corresponds to the time it takes to identify the system parameters…

系统与控制 · 计算机科学 2019-03-26 Yassir Jedra , Alexandre Proutiere

We give new characterizations of the sample complexity of answering linear queries (statistical queries) in the local and central models of differential privacy: *In the non-interactive local model, we give the first approximate…

数据结构与算法 · 计算机科学 2019-11-20 Alexander Edmonds , Aleksandar Nikolov , Jonathan Ullman

Domain generalization methods aim to learn models robust to domain shift with data from a limited number of source domains and without access to target domain samples during training. Popular domain alignment methods for domain…

机器学习 · 计算机科学 2022-06-17 Wenyu Zhang , Mohamed Ragab , Chuan-Sheng Foo

Proximal nested sampling was introduced recently to open up Bayesian model selection for high-dimensional problems such as computational imaging. The framework is suitable for models with a log-convex likelihood, which are ubiquitous in the…

统计方法学 · 统计学 2023-07-31 Jason D. McEwen , Tobías I. Liaudat , Matthew A. Price , Xiaohao Cai , Marcelo Pereyra

This paper explores a new class of incomplete preferences -- termed ``connected preferences'' -- in which maximal domains of comparability are topologically connected. We provide necessary and sufficient conditions for continuous…

理论经济学 · 经济学 2026-05-01 Leandro Gorno , Alessandro Rivello

Knowledge of the domain of applicability of a machine learning model is essential to ensuring accurate and reliable model predictions. In this work, we develop a new and general approach of assessing model domain and demonstrate that our…

材料科学 · 物理学 2025-03-25 Lane E. Schultz , Yiqi Wang , Ryan Jacobs , Dane Morgan

Top monotonicity is a relaxation of various well-known domain restrictions such as single-peaked and single-crossing for which negative impossibility results are circumvented and for which the median-voter theorem still holds. We examine…

计算机科学与博弈论 · 计算机科学 2014-06-03 Haris Aziz

Measuring the similarity between two different sentential arguments is an important task in argument mining. However, one of the challenges in this field is that the dataset must be annotated using expertise in a variety of topics, making…

计算与语言 · 计算机科学 2021-02-22 ChaeHun Park , Sangwoo Seo