English
Related papers

Related papers: Rethinking GSPO: The Perplexity-Entropy Equivalenc…

200 papers

Given an event log as a collection of recorded real-world process traces, process mining aims to automatically construct a process model that is both simple and provides a useful explanation of the traces. Conformance checking techniques…

Artificial Intelligence · Computer Science 2020-08-27 Artem Polyvyanyy , Alistair Moffat , Luciano García-Bañuelos

Our capacity to process information depends on the computational power at our disposal. Information theory captures our ability to distinguish states or communicate messages when it is unconstrained with unrivaled beauty and elegance. For…

Quantum Physics · Physics 2026-04-08 Johannes Jakob Meyer , Asad Raza , Jacopo Rizzo , Lorenzo Leone , Sofiene Jerbi , Jens Eisert

We propose a general approach to construct weighted likelihood estimating equations with the aim of obtain robust estimates. The weight, attached to each score contribution, is evaluated by comparing the statistical data depth at the model…

Methodology · Statistics 2018-02-16 Claudio Agostinelli

We carry out a numerical study of the bi-partite entanglement entropy in the gapped regime of two paradigmatic quantum spin chain models: the Ising chain in an external magnetic field and the anti-ferromagnetic XXZ model. The universal…

High Energy Physics - Theory · Physics 2013-10-30 Emanuele Levi , Olalla A. Castro-Alvaredo , Benjamin Doyon

Quantum entropy and skew information play important roles in quantum information science. They are defined by the trace of the positive operators so that the trace inequalities often have important roles to develop the mathematical theory…

Functional Analysis · Mathematics 2010-08-23 Shigeru Furuichi

Importance weighting is a general way to adjust Monte Carlo integration to account for draws from the wrong distribution, but the resulting estimate can be highly variable when the importance ratios have a heavy right tail. This routinely…

Computation · Statistics 2024-04-12 Aki Vehtari , Daniel Simpson , Andrew Gelman , Yuling Yao , Jonah Gabry

In this article, we propose two classes of relative information measures based on extropy, viz., the generalized extropy similarity ratio (GESR) and generalized extropy divergence ratio (GEDR), that measure the similarity and discrepancy…

Methodology · Statistics 2025-08-20 Saranya P. , Sunoj S. M

We consider a probability distribution depending on a real parameter $x$. As functions of $x$, the R\'enyi entropy and the Tsallis entropy can be expressed in terms of the associated index of coincidence $S(x)$. We establish recurrence…

Classical Analysis and ODEs · Mathematics 2019-10-31 Alexandra Maduta , Diana Otrocol , Ioan Rasa

We introduce an ambidextrous view of stochastic dynamical systems, comparing their forward-time and reverse-time representations and then integrating them into a single time-symmetric representation. The perspective is useful theoretically,…

Statistical Mechanics · Physics 2015-05-13 Christopher J. Ellison , John R. Mahoney , James P. Crutchfield

We introduce a family of scale-invariant entropy statistics derived from logarithmically aggregated distance distributions of point processes, with prime numbers serving as a motivating example. The construction associates to each finite…

Methodology · Statistics 2026-04-06 Mohamed Gewily

Wasserstein Policy Optimization (WPO) is a recently proposed reinforcement learning algorithm that leverages Wasserstein gradient flows to optimize stochastic policies in continuous action spaces. Despite its empirical success, the…

Machine Learning · Computer Science 2026-05-22 David Šiška , Yufei Zhang

Reinforcement learning with verifiable rewards has shown notable effectiveness in enhancing large language models (LLMs) reasoning performance, especially in mathematics tasks. However, such improvements often come with reduced outcome…

Artificial Intelligence · Computer Science 2026-02-03 Chenyi Li , Yuan Zhang , Bo Wang , Guoqing Ma , Wei Tang , Haoyang Huang , Nan Duan

The entropy of a graph is an information-theoretic quantity which expresses the complexity of a graph \cite{DM1,M}. After Shannon introduced the definition of entropy to information and communication, many generalizations of the entropy…

Combinatorics · Mathematics 2014-11-26 Xueliang Li , Zhongmei Qin , Meiqin Wei , Ivan Gutman , Matthias Dehmer

A unified combinatorial definition of the information content and entropy of different types of patterns, compatible with the traditional concepts of information and entropy, going beyond the limitations of Shannon information interpretable…

Information Theory · Computer Science 2025-01-22 Zsolt Pocze

Reinforcement learning (RL) plays an increasingly important role in enhancing the reasoning capabilities of large language models (LLMs), yet stable and performant policy optimization remains challenging. Token-level importance ratios often…

Machine Learning · Computer Science 2025-12-02 Chang Gao , Chujie Zheng , Xiong-Hui Chen , Kai Dang , Shixuan Liu , Bowen Yu , An Yang , Shuai Bai , Jingren Zhou , Junyang Lin

Large language models frequently exhibit suboptimal performance on low resource languages, primarily due to inefficient subword segmentation and systemic training data imbalances. In this paper, we propose Variable Entropy Policy…

Computation and Language · Computer Science 2026-03-20 Chonghan Liu , Yimin Du , Qi An , Xin He , Cunqi Zhai , Fei Tan , Weijia Lin , Xiaochun Gong , Yongchao Deng , Shousheng Jia , Xiangzheng Zhang

Some essential conceptual aspects that will fill some logical gaps of the frame to interpret the gravity as an entropic force was investigated, we focus on some crucial issues that didn't emphasized in Verlinde's original…

General Relativity and Quantum Cosmology · Physics 2014-02-21 Qiao-Jun Cao

Information theoretic quantities play a central role in machine learning. The recent surge in the complexity of data and models has increased the demand for accurate estimation of these quantities. However, as the dimension grows the…

Machine Learning · Statistics 2024-05-21 Viktor Nilsson , Anirban Samaddar , Sandeep Madireddy , Pierre Nyquist

Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as an indispensable paradigm for enhancing reasoning in Large Language Models (LLMs). However, standard policy optimization methods, such as Group Relative Policy…

Machine Learning · Computer Science 2026-02-09 Pengyi Li , Elizaveta Goncharova , Andrey Kuznetsov , Ivan Oseledets

Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better solutions. We test this claim in a resource-constrained setting by applying GRPO with…

Machine Learning · Computer Science 2026-04-09 Suraj Yadav , Siddharth Yadav , Parth Goyal
‹ Prev 1 8 9 10 Next ›