中文
相关论文

相关论文: Information Measures: the Curious Case of the Bina…

200 篇论文

We discuss inequalities holding between the vocabulary size, i.e., the number of distinct nonterminal symbols in a grammar-based compression for a string, and the excess length of the respective universal code, i.e., the code-based analog…

信息论 · 计算机科学 2020-03-11 Lukasz Debowski

Non-binary codes correcting multiple deletions have recently attracted a lot of attention. In this work, we focus on multiplicity-free codes, a family of non-binary codes where all symbols are distinct. Our main contribution is a new…

信息论 · 计算机科学 2025-08-06 Michael Schaller , Beatrice Toesca , Van Khu Vu

We study hypothesis testing under communication constraints, where each sample is quantized before being revealed to a statistician. Without communication constraints, it is well known that the sample complexity of simple binary hypothesis…

统计理论 · 数学 2023-12-19 Ankit Pensia , Varun Jog , Po-Ling Loh

Many real-world questions appear deceptively simple yet implicitly demand two capabilities: (i) systematic coverage of a bounded knowledge universe and (ii) compositional set-based reasoning over that universe, a phenomenon we term "the tip…

人工智能 · 计算机科学 2026-04-21 Xiao Zhang , Qianru Meng , Yongjian Chen , Yumeng Wang , Johan Bos

Unambiguous measurements play an important role in quantum information, with applications ranging from quantum key distribution to quantum state reconstruction. Recently, such measurements have also been used in quantum algorithms based on…

量子物理 · 物理学 2025-11-05 Quentin Buzet , André Chailloux

The Data Aggregation Problem occurs when a large collection of data takes on a higher security level than any of its individual component records. Traditional approaches of breaking up the data and restricting access on a "need to know"…

密码学与安全 · 计算机科学 2011-05-18 William R. Lorimer

We establish bounds on the KL divergence between two multivariate Gaussian distributions in terms of the Hamming distance between the edge sets of the corresponding graphical models. We show that the KL divergence is bounded below by a…

信息论 · 计算机科学 2015-04-06 Varun Jog , Po-Ling Loh

In this paper we have considered a single inequality having 11 known divergence measures. This inequality include measures like: Jeffryes-Kullback-Leiber J-divergence, Jensen-Shannon divergence (Burbea-Rao, 1982), arithmetic-geometric mean…

信息论 · 计算机科学 2012-05-08 Inder Jeet Taneja

Large language models are increasingly relied upon as sources of information, but their propensity for generating false or misleading statements with high confidence poses risks for users and society. In this paper, we confront the critical…

Recently, Chen and Sbert proposed a general divergence measure. This report presents some interim findings about the question whether the divergence measure is a metric or not. It has been postulated that (i) the measure might be a metric…

信息论 · 计算机科学 2021-01-18 Min Chen , Mateu Sbert

A key task in managing distributed, sensitive data is to measure the extent to which a distribution changes. Understanding this drift can effectively support a variety of federated learning and analytics tasks. However, in many practical…

机器学习 · 计算机科学 2024-12-02 Mary Scott , Sayan Biswas , Graham Cormode , Carsten Maple

Bias-variance decompositions are widely used to understand the generalization performance of machine learning models. While the squared error loss permits a straightforward decomposition, other loss functions - such as zero-one loss or…

机器学习 · 计算机科学 2026-01-27 Tom Heskes

With the growing use of ML in highly consequential domains, quantifying disparity with respect to protected attributes, e.g., gender, race, etc., is important. While quantifying disparity is essential, sometimes the needs of an occupation…

信息论 · 计算机科学 2021-08-10 Sanghamitra Dutta , Praveen Venkatesh , Piotr Mardziel , Anupam Datta , Pulkit Grover

To ensure stability of learning, state-of-the-art generalized policy iteration algorithms augment the policy improvement step with a trust region constraint bounding the information loss. The size of the trust region is commonly determined…

机器学习 · 计算机科学 2018-04-05 Boris Belousov , Jan Peters

We study codes that are list-decodable under insertions and deletions. Specifically, we consider the setting where a codeword over some finite alphabet of size $q$ may suffer from $\delta$ fraction of adversarial deletions and $\gamma$…

信息论 · 计算机科学 2018-02-26 Bernhard Haeupler , Amirbehshad Shahrasbi , Madhu Sudan

We deal with countable alphabet locally compact random subshifts of finite type (the latter merely meaning that the symbol space is generated by an incidence matrix) under the absence of Big Images Property and under the absence of uniform…

动力系统 · 数学 2015-09-02 Volker Mayer , Mariusz Urbanski

Bregman divergences generalize measures such as the squared Euclidean distance and the KL divergence, and arise throughout many areas of machine learning. In this paper, we focus on the problem of approximating an arbitrary Bregman…

机器学习 · 统计学 2020-11-04 Ali Siahkamari , Xide Xia , Venkatesh Saligrama , David Castanon , Brian Kulis

Despite the fact that large language models (LLMs) show exceptional skill in instruction following tasks, this strength can turn into a vulnerability when the models are required to disregard certain instructions. Instruction-following…

计算与语言 · 计算机科学 2025-08-12 Yerin Hwang , Yongil Kim , Jahyun Koo , Taegwan Kang , Hyunkyung Bae , Kyomin Jung

Large language models (LLMs) are increasingly used to meet user information needs, but their effectiveness in dealing with user queries that contain various types of ambiguity remains unknown, ultimately risking user trust and satisfaction.…

计算与语言 · 计算机科学 2024-06-04 Tong Zhang , Peixin Qin , Yang Deng , Chen Huang , Wenqiang Lei , Junhong Liu , Dingnan Jin , Hongru Liang , Tat-Seng Chua

Large language models (LLMs) achieve impressive results on advanced mathematics benchmarks but sometimes fail on basic arithmetic tasks, raising the question of whether they have truly grasped fundamental arithmetic rules or are merely…

计算与语言 · 计算机科学 2025-09-18 Yang Yan , Yu Lu , Renjun Xu , Zhenzhong Lan