English
Related papers

Related papers: Partial Recovery for Top-$k$ Ranking: Optimality o…

200 papers

This paper studies human preference learning based on partially revealed choice behavior and formulates the problem as a generalized Bradley-Terry-Luce (BTL) ranking model that accounts for heterogeneous preferences. Specifically, we assume…

Methodology · Statistics 2025-09-03 Jianqing Fan , Hyukjun Kwon , Xiaonan Zhu

A fundamental problem in statistics and machine learning is to estimate a function $f$ from possibly noisy observations of its point samples. The goal is to design a numerical algorithm to construct an approximation $\hat f$ to $f$ in a…

Statistics Theory · Mathematics 2025-05-30 Ronald DeVore , Robert D. Nowak , Rahul Parhi , Guergana Petrova , Jonathan W. Siegel

Orthogonal group synchronization aims to recover orthogonal group elements from their noisy pairwise measurements. It has found numerous applications including computer vision, imaging science, and community detection. Due to the orthogonal…

Statistics Theory · Mathematics 2025-02-21 Ziliang Samuel Zhong , Shuyang Ling

In online ranking, a learning algorithm sequentially ranks a set of items and receives feedback on its ranking in the form of relevance scores. Since obtaining relevance scores typically involves human annotation, it is of great interest to…

Machine Learning · Computer Science 2024-04-15 Mingyuan Zhang , Ambuj Tewari

We study matrix estimation problems arising in reinforcement learning (RL) with low-rank structure. In low-rank bandits, the matrix to be recovered specifies the expected arm rewards, and for low-rank Markov Decision Processes (MDPs), it…

Machine Learning · Computer Science 2023-10-31 Stefan Stojanovic , Yassir Jedra , Alexandre Proutiere

The $\Delta_3(L)$ statistic of Random Matrix Theory is defined as the average of a set of random numbers $\{\delta\}$, derived from a spectrum. The distribution $p(\delta)$ of these random numbers is used as the basis of a maximum…

Nuclear Theory · Physics 2015-03-19 Declan Mulhall

This paper addresses the challenges of aligning large language models (LLMs) with human values via preference learning (PL), focusing on incomplete and corrupted data in preference datasets. We propose a novel method for robustly and…

Artificial Intelligence · Computer Science 2025-10-30 Son The Nguyen , Niranjan Uma Naresh , Theja Tulabandhula

Recently, Rendle has warned that the use of sampling-based top-$k$ metrics might not suffice. This throws a number of recent studies on deep learning-based recommendation algorithms, and classic non-deep-learning algorithms using such a…

Information Retrieval · Computer Science 2021-06-22 Dong Li , Ruoming Jin , Jing Gao , Zhi Liu

Pairwise ranking systems based on Maximum Likelihood Estimation (MLE), such as the Bradley-Terry model, are widely used to aggregate preferences from pairwise comparisons. However, their robustness under strategic data manipulation remains…

Machine Learning · Computer Science 2026-04-21 Junyi Yao , Zihao Zheng , Jiayu Long

We consider the problem of ranking a set of items from pairwise comparisons in the presence of features associated with the items. Recent works have established that $O(n\log(n))$ samples are needed to rank well when there is no feature…

Machine Learning · Computer Science 2021-02-10 Aadirupa Saha , Arun Rajkumar

We study maximum likelihood estimation (MLE) in the generalized group orbit recovery model, where each observation is generated by applying a random group action and a known, fixed linear operator to an unknown signal, followed by additive…

Statistics Theory · Mathematics 2025-09-30 Sheng Xu , Anderson Ye Zhang , Amit Singer

We study the problem of efficiently recovering the matching between an unlabelled collection of $n$ points in $\mathbb{R}^d$ and a small random perturbation of those points. We consider a model where the initial points are i.i.d. standard…

Data Structures and Algorithms · Computer Science 2021-07-13 Dmitriy Kunisky , Jonathan Niles-Weed

Ranking LLMs via pairwise human feedback underpins current leaderboards for open-ended tasks, such as creative writing and problem-solving. We analyze ~89K comparisons in 116 languages from 52 LLMs from Arena, and show that the best-fit…

Machine Learning · Computer Science 2026-05-08 Jai Moondra , Ayela Chughtai , Bhargavi Lanka , Swati Gupta

The recent paper \cite{GSZ2023} on estimation and inference for top-ranking problem in Bradley-Terry-Lice (BTL) model presented a surprising result: component-wise estimation and inference can be done under much weaker conditions on the…

Statistics Theory · Mathematics 2025-06-08 Vladimir Spokoiny

In this paper, we examine the problem of partial inference in the context of structured prediction. Using a generative model approach, we consider the task of maximizing a score function with unary and pairwise potentials in the space of…

Machine Learning · Computer Science 2023-06-08 Chuyang Ke , Jean Honorio

[Abridged] - Spectral Retrieval is a plug-in re-ranking stage that interpolates between per-token MaxSim and mean-pool retrieval through a multi-scale sinc convolution over token embeddings. In standard dense retrieval each document is one…

Information Retrieval · Computer Science 2026-05-26 Andrea Morandi

Spectral learning recently generated lots of excitement in machine learning, largely because it is the first known method to produce consistent estimates (under suitable conditions) for several latent variable models. In contrast, maximum…

Machine Learning · Computer Science 2014-06-19 Han Zhao , Pascal Poupart

We describe $k$-MLE, a fast and efficient local search algorithm for learning finite statistical mixtures of exponential families such as Gaussian mixture models. Mixture models are traditionally learned using the expectation-maximization…

Machine Learning · Computer Science 2016-11-15 Frank Nielsen

Targeted maximum likelihood estimators (TMLEs) are asymptotically optimal among regular, asymptotically linear estimators. In small samples, however, we may be far from "asymptopia" and not reap the benefits of optimality. Here we propose a…

Methodology · Statistics 2025-02-04 Noel Pimentel , Alejandro Schuler , Mark van der Laan

This work considers Maximum Likelihood Estimation (MLE) of a Toeplitz structured covariance matrix. In this regard, an equivalent reformulation of the MLE problem is introduced and two iterative algorithms are proposed for the optimization…

Signal Processing · Electrical Eng. & Systems 2021-10-26 Augusto Aubry , Prabhu Babu , Antonio De Maio , Rikhabchand Jyothi