中文
相关论文

相关论文: On uncertainty and information properties of ranke…

200 篇论文

Random matrix ensembles (RME) of quantum statistical Hamiltonian operators, {\em e.g.} Gaussian random matrix ensembles (GRME) and Ginibre random matrix ensembles (Ginibre RME), found applications in literature in study of following quantum…

统计力学 · 物理学 2007-05-23 Maciej M. Duras

In learning with noisy labels, the sample selection approach is very popular, which regards small-loss data as correctly labeled during training. However, losses are generated on-the-fly based on the model being trained with noisy labels,…

机器学习 · 计算机科学 2021-06-02 Xiaobo Xia , Tongliang Liu , Bo Han , Mingming Gong , Jun Yu , Gang Niu , Masashi Sugiyama

Ranked lists are frequently used by information retrieval (IR) systems to present results believed to be relevant to the users information need. Fairness is a relatively new but important aspect of these rankings to measure, joining a rich…

信息检索 · 计算机科学 2022-01-11 Amifa Raj , Michael D. Ekstrand

Estimating the uncertainty in deep neural network predictions is crucial for many real-world applications. A common approach to model uncertainty is to choose a parametric distribution and fit the data to it using maximum likelihood…

机器学习 · 计算机科学 2022-11-28 Ali Harakeh , Jordan Hu , Naiqing Guan , Steven L. Waslander , Liam Paull

Unsupervised performance estimation, or evaluating how well models perform on unlabeled data is a difficult task. Recently, a method was proposed by Garg et al. [2022] which performs much better than previous methods. Their method relies on…

机器学习 · 计算机科学 2023-06-21 Muhammad Maaz , Rui Qiao , Yiheng Zhou , Renxian Zhang

Large language models (LLMs) are increasingly deployed in settings where the available context is incomplete or degraded. We argue that an LLM generating answers under incomplete context can be viewed as an implicit imputer, and evaluated…

机器学习 · 统计学 2026-05-14 Stef van Buuren

Modern statistical estimation is often performed in a distributed setting where each sample belongs to a single user who shares their data with a central server. Users are typically concerned with preserving the privacy of their samples,…

We consider the problem of defining the significance of an itemset. We say that the itemset is significant if we are surprised by its frequency when compared to the frequencies of its sub-itemsets. In other words, we estimate the frequency…

机器学习 · 计算机科学 2019-04-30 Nikolaj Tatti

Shannon entropy is often a quantity of interest to linguists studying the communicative capacity of human language. However, entropy must typically be estimated from observed data because researchers do not have access to the underlying…

计算与语言 · 计算机科学 2022-04-06 Aryaman Arora , Clara Meister , Ryan Cotterell

Almost every software system provides configuration options to tailor the system to the target platform and application scenario. Often, this configurability renders the analysis of every individual system configuration infeasible. To…

软件工程 · 计算机科学 2016-02-17 Flávio Medeiros , Christian Kästner , Márcio Ribeiro , Rohit Gheyi , Sven Apel

It is pointed out that the case for Shannon entropy and von Neumann entropy, as measures of uncertainty in quantum mechanics, is not as bleak as suggested in quant-ph/0006087. The main argument of the latter is based on one particular…

量子物理 · 物理学 2007-05-23 Michael J. W. Hall

The performance of a machine learning system is usually evaluated by using i.i.d.\ observations with true labels. However, acquiring ground truth labels is expensive, while obtaining unlabeled samples may be cheaper. Stratified sampling can…

机器学习 · 计算机科学 2019-07-29 Tiancheng Yu , Xiyu Zhai , Suvrit Sra

For better or for worse, rankings of institutions, such as universities, schools and hospitals, play an important role today in conveying information about relative performance. They inform policy decisions and budgets, and are often…

统计理论 · 数学 2010-11-11 Peter Hall , Hugh Miller

This paper studies the complexity of estimating Renyi divergences of discrete distributions: $p$ observed from samples and the baseline distribution $q$ known \emph{a priori}. Extending the results of Acharya et al. (SODA'15) on estimating…

信息论 · 计算机科学 2017-02-09 Maciej Skorski

Kernel random matrices have attracted a lot of interest in recent years, from both practical and theoretical standpoints. Most of the theoretical work so far has focused on the case were the data is sampled from a low-dimensional structure.…

统计理论 · 数学 2010-11-12 Noureddine El Karoui

Boolean formulae compactly encode huge, constrained search spaces. Thus, variability-intensive systems are often encoded with Boolean formulae. The search space of a variability-intensive system is usually too large to explore without…

计算机科学中的逻辑 · 计算机科学 2025-03-19 Olivier Zeyen , Maxime Cordy , Martin Gubri , Gilles Perrouin , Mathieu Acher

Probability theory is fundamental for modeling uncertainty, with traditional probabilities being real and non-negative. Complex probability extends this concept by allowing complex-valued probabilities, opening new avenues for analysis in…

信息论 · 计算机科学 2025-03-07 Chan Li , Hejun Xu , Zhu Cao

Shannon entropy was defined for probability distributions and then its using was expanded to measure the uncertainty of knowledge for systems with complete information. In this article, it is proposed to extend the using of Shannon entropy…

信息论 · 计算机科学 2017-09-15 Vasile Patrascu

Explicit information seeking is essential to human problem-solving in practical environments characterized by incomplete information and noisy dynamics. When the true environmental state is not directly observable, humans seek information…

人工智能 · 计算机科学 2025-10-03 Djengo Cyun-Jyun Fang , Tsung-Wei Ke

The Shannon entropy, and related quantities such as mutual information, can be used to quantify uncertainty and relevance. However, in practice, it can be difficult to compute these quantities for arbitrary probability distributions,…

统计计算 · 统计学 2017-10-11 Brendon J. Brewer
‹ 上一页 1 8 9 10 下一页 ›