English
Related papers

Related papers: Roy's largest root under rank-one alternatives:The…

200 papers

Several Radio Resource Management (RRM) use cases can be framed as sequential decision planning problems, where an agent (the base station, typically) makes decisions that influence the network utility and state. While Reinforcement…

Networking and Internet Architecture · Computer Science 2024-05-31 Lorenzo Maggi , Matthew Andrews , Ryo Koblitz

The problem of symmetric rank-one approximation of symmetric tensors is important in Independent Components Analysis, also known as Blind Source Separation, as well as polynomial optimization. We analyze the symmetric rank-one approximation…

Computation · Statistics 2011-12-14 Michael James O'Hara

In this paper, we study a novel episodic risk-sensitive Reinforcement Learning (RL) problem, named Iterated CVaR RL, which aims to maximize the tail of the reward-to-go at each step, and focuses on tightly controlling the risk of getting…

Machine Learning · Computer Science 2023-05-12 Yihan Du , Siwei Wang , Longbo Huang

The Matrix-based Renyi's entropy enables us to directly measure information quantities from given data without the costly probability density estimation of underlying distributions, thus has been widely adopted in numerous statistical…

Machine Learning · Statistics 2022-05-17 Yuxin Dong , Tieliang Gong , Shujian Yu , Chen Li

Locating a target is key in many applications, namely in high-stakes real-world scenarios, like detecting humans or obstacles in vehicular networks. In scenarios where precise statistics of the measurement noise are unavailable,…

Optimization and Control · Mathematics 2022-08-17 João Domingos , Cláudia Soares , João Xavier

The smallest eigenvectors of the graph Laplacian are well-known to provide a succinct representation of the geometry of a weighted graph. In reinforcement learning (RL), where the weighted graph may be interpreted as the state transition…

Machine Learning · Computer Science 2018-10-11 Yifan Wu , George Tucker , Ofir Nachum

The Robbins-Monro stochastic approximation algorithm is a foundation of many algorithmic frameworks for reinforcement learning (RL), and often an efficient approach to solving (or approximating the solution to) complex optimal control…

Optimization and Control · Mathematics 2019-03-19 Andrey Bernstein , Yue Chen , Marcello Colombino , Emiliano Dall'Anese , Prashant Mehta , Sean Meyn

We consider the asymptotic fluctuation behavior of the largest eigenvalue of certain sample covariance matrices in the asymptotic regime where both dimensions of the corresponding data matrix go to infinity. More precisely, let $X$ be an…

Probability · Mathematics 2009-09-29 Noureddine El Karoui

We consider the distributed detection problem of a temporally correlated random radio source signal using a wireless sensor network capable of measuring the energy of the received signals. It is well-known that optimal tests in the…

Signal Processing · Electrical Eng. & Systems 2023-12-20 Juan Augusto Maya , Leonardo Rey Vega , Andrea M. Tonello

We study the problem of maximizing a spectral risk measure of a given output function which depends on several underlying variables, whose individual distributions are known but whose joint distribution is not. We establish and exploit an…

Optimization and Control · Mathematics 2022-11-16 Hamza Ennaji , Quentin Mérigot , Luca Nenna , Brendan Pass

We study the problem of assigning transmission ranges to radio stations placed arbitrarily in a $d$-dimensional ($d$-D) Euclidean space in order to achieve a strongly connected communication network with minimum total power consumption. The…

Computational Geometry · Computer Science 2015-02-17 Paz Carmi , Lilach Chaitman-Yerushalmi

In this article, we study a Radio Resource Allocation (RRA) that was formulated as a non-convex optimization problem whose main aim is to maximize the spectral efficiency subject to satisfaction guarantees in multiservice wireless systems.…

The study of complex networks has been one of the most active fields in science in recent decades. Spectral properties of networks (or graphs that represent them) are of fundamental importance. Researchers have been investigating these…

Combinatorics · Mathematics 2018-09-25 Daniel Montealegre , Van Vu

Maximum eigenvalue detection (MED) is an important application of random matrix theory in spectrum sensing and signal detection. However, in small signal-to-noise ratio environment, the maximum eigenvalue of the representative signal is at…

Signal Processing · Electrical Eng. & Systems 2018-03-28 Lin Zheng , Robert C. Qiu , Qing Feng , Xuebin Li

The prevailing paradigm for training large reasoning models--combining Supervised Fine-Tuning (SFT) with Reinforcement Learning with Verifiable Rewards (RLVR)--is fundamentally constrained by its reliance on high-quality, human-annotated…

Machine Learning · Computer Science 2026-03-24 Yuanfu Wang , Zhixuan Liu , Xiangtian Li , Chaochao Lu , Chao Yang

Learning to rank (LTR) plays a crucial role in various Information Retrieval (IR) tasks. Although supervised LTR methods based on fine-grained relevance labels (e.g., document-level annotations) have achieved significant success, their…

Information Retrieval · Computer Science 2025-08-21 Yiteng Tu , Zhichao Xu , Tao Yang , Weihang Su , Yujia Zhou , Yiqun Liu , Fen Lin , Qin Liu , Qingyao Ai

We study the problem of estimating a rank one signal matrix from an observed matrix generated by corrupting the signal with additive rotationally invariant noise. We develop a new class of approximate message-passing algorithms for this…

Statistics Theory · Mathematics 2025-09-09 Rishabh Dudeja , Songbin Liu , Junjie Ma

Reinforcement learning (RL) is an important field of research in machine learning that is increasingly being applied to complex optimization problems in physics. In parallel, concepts from physics have contributed to important advances in…

Machine Learning · Computer Science 2023-05-11 Argenis Arriojas , Jacob Adamczyk , Stas Tiomkin , Rahul V. Kulkarni

In this paper, we develop a penalized realized variance (PRV) estimator of the quadratic variation (QV) of a high-dimensional continuous It\^{o} semimartingale. We adapt the principle idea of regularization from linear regression to…

Econometrics · Economics 2026-01-28 Kim Christensen , Mikkel Slot Nielsen , Mark Podolskij

We establish large deviation principles for the largest eigenvalue of large random matrices with variance profiles. For $N \in \mathbb N$, we consider random $N \times N$ symmetric matrices $H^N$ which are such that…

Probability · Mathematics 2024-03-25 Raphaël Ducatez , Alice Guionnet , Jonathan Husson