English
Related papers

Related papers: Quadratic Upper Bound for Recursive Teaching Dimen…

200 papers

Recurrent neural networks (RNN) are powerful tools to explain how attractors may emerge from noisy, high-dimensional dynamics. We study here how to learn the ~N^(2) pairwise interactions in a RNN with N neurons to embed L manifolds of…

Disordered Systems and Neural Networks · Physics 2020-02-05 Aldo Battista , Rémi Monasson

We continue the study of model-independent constraints on the unitary Conformal Field Theories in 4-Dimensions, initiated in arXiv:0807.0004. Our main result is an improved upper bound on the dimension \Delta of the leading scalar operator…

High Energy Physics - Theory · Physics 2015-03-13 Vyacheslav S. Rychkov , Alessandro Vichi

Convex codes were recently introduced as models for neural codes in the brain. Any convex code $\C$ has an associated minimal embedding dimension $d(\C)$, which is the minimal Euclidean space dimension such that the code can be realized by…

Combinatorics · Mathematics 2016-12-23 Carina Curto , Ramón Vera

Klee's measure problem (computing the volume of the union of $n$ axis-parallel boxes in $\mathbb{R}^d$) is well known to have $n^{\frac{d}{2}\pm o(1)}$-time algorithms (Overmars, Yap, SICOMP'91; Chan FOCS'13). Only recently, a conditional…

Computational Geometry · Computer Science 2023-03-16 Egor Gorbachev , Marvin Künnemann

Recently, rubrics have been used to guide LLM judges in capturing subjective, nuanced, multi-dimensional human preferences, and have been extended from evaluation to reward signals for reinforcement fine-tuning (RFT). However, rubric…

We study time-inhomogeneous episodic reinforcement learning (RL) under general function approximation and sparse rewards. We design a new algorithm, Variance-weighted Optimistic $Q$-Learning (VO$Q$L), based on $Q$-learning and bound its…

Machine Learning · Computer Science 2022-12-13 Alekh Agarwal , Yujia Jin , Tong Zhang

We give an upper bound for the degree of rational curves in a family that covers a given birational ruled surface in projective space. The upper bound is stated in terms of the degree, sectional genus and arithmetic genus of the surface. We…

Algebraic Geometry · Mathematics 2021-03-09 Niels Lubbes

In safety-critical applications of reinforcement learning such as healthcare and robotics, it is often desirable to optimize risk-sensitive objectives that account for tail outcomes rather than expected reward. We prove the first regret…

Machine Learning · Computer Science 2022-10-12 O. Bastani , Y. J. Ma , E. Shen , W. Xu

The Upper Bound Theorem for convex polytopes implies that the $p$-th Betti number of the \v{C}ech complex of any set of $N$ points in $\mathbb R^d$ and any radius satisfies $\beta_{p} = O(N^{m})$, with $m = \min \{ p+1, \lceil d/2 \rceil…

Combinatorics · Mathematics 2023-10-24 Herbert Edelsbrunner , János Pach

The recent outburst of context-dependent knowledge on the Semantic Web (SW) has led to the realization of the importance of the quads in the SW community. Quads, which extend a standard RDF triple, by adding a new parameter of the `context'…

Logic in Computer Science · Computer Science 2014-06-05 Mathew Joseph , Gabriel Kuper , Luciano Serafini

Reading comprehension (RC) is a challenging task that requires synthesis of information across sentences and multiple turns of reasoning. Using a state-of-the-art RC model, we empirically investigate the performance of single-turn and…

Computation and Language · Computer Science 2017-11-10 Yelong Shen , Xiaodong Liu , Kevin Duh , Jianfeng Gao

Temporal difference (TD) learning is a popular algorithm for policy evaluation in reinforcement learning, but the vanilla TD can substantially suffer from the inherent optimization variance. A variance reduced TD (VRTD) algorithm was…

Machine Learning · Computer Science 2020-01-13 Tengyu Xu , Zhe Wang , Yi Zhou , Yingbin Liang

Constrained Markov Decision Processes are a class of stochastic decision problems in which the decision maker must select a policy that satisfies auxiliary cost constraints. This paper extends upper confidence reinforcement learning for…

Machine Learning · Computer Science 2020-01-28 Liyuan Zheng , Lillian J. Ratliff

Deep representation learning methods struggle with continual learning, suffering from both catastrophic forgetting of useful units and loss of plasticity, often due to rigid and unuseful units. While many methods address these two issues…

Machine Learning · Computer Science 2024-05-02 Mohamed Elsayed , A. Rupam Mahmood

We study the recovery of multiple high-dimensional signals from two noisy, correlated modalities: a spiked matrix and a spiked tensor sharing a common low-rank structure. This setting generalizes classical spiked matrix and tensor models,…

Machine Learning · Statistics 2025-06-04 Hugo Tabanelli , Pierre Mergny , Lenka Zdeborova , Florent Krzakala

In the target tracking and its engineering applications, recursive state estimation of the target is of fundamental importance. This paper presents a recursive performance bound for dynamic estimation and filtering problem, in the framework…

Applications · Statistics 2015-06-04 Huisi Tong , Hao Zhang , Huadong Meng , Xiqin Wang

To date, the tightest upper and lower-bounds for the active learning of general concept classes have been in terms of a parameter of the learning problem called the splitting index. We provide, for the first time, an efficient algorithm…

Machine Learning · Computer Science 2017-06-12 Christopher Tosh , Sanjoy Dasgupta

We consider coordinate descent methods on convex quadratic problems, in which exact line searches are performed at each iteration. (This algorithm is identical to Gauss-Seidel on the equivalent symmetric positive definite linear system.) We…

Optimization and Control · Mathematics 2020-01-14 Stephen J. Wright , Ching-Pei Lee

Let $\Lambda$ be an artin algebra. We give an upper bound for the dimension of the bounded derived category of the category $\mod \Lambda$ of finitely generated right $\Lambda$-modules in terms of the projective and injective dimensions of…

Rings and Algebras · Mathematics 2020-04-30 Junling Zheng , Zhaoyong Huang

We study the problem of learning-augmented predictive linear quadratic control. Our goal is to design a controller that balances \textit{"consistency"}, which measures the competitive ratio when predictions are accurate, and…

Systems and Control · Electrical Eng. & Systems 2025-04-08 Tongxin Li , Ruixiao Yang , Guannan Qu , Guanya Shi , Chenkai Yu , Adam Wierman , Steven H. Low
‹ Prev 1 4 5 6 7 8 10 Next ›