中文
相关论文

相关论文: Quadratic Upper Bound for Recursive Teaching Dimen…

200 篇论文

Recurrent neural networks (RNN) are powerful tools to explain how attractors may emerge from noisy, high-dimensional dynamics. We study here how to learn the ~N^(2) pairwise interactions in a RNN with N neurons to embed L manifolds of…

无序系统与神经网络 · 物理学 2020-02-05 Aldo Battista , Rémi Monasson

We continue the study of model-independent constraints on the unitary Conformal Field Theories in 4-Dimensions, initiated in arXiv:0807.0004. Our main result is an improved upper bound on the dimension \Delta of the leading scalar operator…

高能物理 - 理论 · 物理学 2015-03-13 Vyacheslav S. Rychkov , Alessandro Vichi

Convex codes were recently introduced as models for neural codes in the brain. Any convex code $\C$ has an associated minimal embedding dimension $d(\C)$, which is the minimal Euclidean space dimension such that the code can be realized by…

组合数学 · 数学 2016-12-23 Carina Curto , Ramón Vera

Klee's measure problem (computing the volume of the union of $n$ axis-parallel boxes in $\mathbb{R}^d$) is well known to have $n^{\frac{d}{2}\pm o(1)}$-time algorithms (Overmars, Yap, SICOMP'91; Chan FOCS'13). Only recently, a conditional…

计算几何 · 计算机科学 2023-03-16 Egor Gorbachev , Marvin Künnemann

Recently, rubrics have been used to guide LLM judges in capturing subjective, nuanced, multi-dimensional human preferences, and have been extended from evaluation to reward signals for reinforcement fine-tuning (RFT). However, rubric…

We study time-inhomogeneous episodic reinforcement learning (RL) under general function approximation and sparse rewards. We design a new algorithm, Variance-weighted Optimistic $Q$-Learning (VO$Q$L), based on $Q$-learning and bound its…

机器学习 · 计算机科学 2022-12-13 Alekh Agarwal , Yujia Jin , Tong Zhang

We give an upper bound for the degree of rational curves in a family that covers a given birational ruled surface in projective space. The upper bound is stated in terms of the degree, sectional genus and arithmetic genus of the surface. We…

代数几何 · 数学 2021-03-09 Niels Lubbes

In safety-critical applications of reinforcement learning such as healthcare and robotics, it is often desirable to optimize risk-sensitive objectives that account for tail outcomes rather than expected reward. We prove the first regret…

机器学习 · 计算机科学 2022-10-12 O. Bastani , Y. J. Ma , E. Shen , W. Xu

The Upper Bound Theorem for convex polytopes implies that the $p$-th Betti number of the \v{C}ech complex of any set of $N$ points in $\mathbb R^d$ and any radius satisfies $\beta_{p} = O(N^{m})$, with $m = \min \{ p+1, \lceil d/2 \rceil…

组合数学 · 数学 2023-10-24 Herbert Edelsbrunner , János Pach

The recent outburst of context-dependent knowledge on the Semantic Web (SW) has led to the realization of the importance of the quads in the SW community. Quads, which extend a standard RDF triple, by adding a new parameter of the `context'…

计算机科学中的逻辑 · 计算机科学 2014-06-05 Mathew Joseph , Gabriel Kuper , Luciano Serafini

Reading comprehension (RC) is a challenging task that requires synthesis of information across sentences and multiple turns of reasoning. Using a state-of-the-art RC model, we empirically investigate the performance of single-turn and…

计算与语言 · 计算机科学 2017-11-10 Yelong Shen , Xiaodong Liu , Kevin Duh , Jianfeng Gao

Temporal difference (TD) learning is a popular algorithm for policy evaluation in reinforcement learning, but the vanilla TD can substantially suffer from the inherent optimization variance. A variance reduced TD (VRTD) algorithm was…

机器学习 · 计算机科学 2020-01-13 Tengyu Xu , Zhe Wang , Yi Zhou , Yingbin Liang

Constrained Markov Decision Processes are a class of stochastic decision problems in which the decision maker must select a policy that satisfies auxiliary cost constraints. This paper extends upper confidence reinforcement learning for…

机器学习 · 计算机科学 2020-01-28 Liyuan Zheng , Lillian J. Ratliff

Deep representation learning methods struggle with continual learning, suffering from both catastrophic forgetting of useful units and loss of plasticity, often due to rigid and unuseful units. While many methods address these two issues…

机器学习 · 计算机科学 2024-05-02 Mohamed Elsayed , A. Rupam Mahmood

We study the recovery of multiple high-dimensional signals from two noisy, correlated modalities: a spiked matrix and a spiked tensor sharing a common low-rank structure. This setting generalizes classical spiked matrix and tensor models,…

机器学习 · 统计学 2025-06-04 Hugo Tabanelli , Pierre Mergny , Lenka Zdeborova , Florent Krzakala

In the target tracking and its engineering applications, recursive state estimation of the target is of fundamental importance. This paper presents a recursive performance bound for dynamic estimation and filtering problem, in the framework…

应用统计 · 统计学 2015-06-04 Huisi Tong , Hao Zhang , Huadong Meng , Xiqin Wang

To date, the tightest upper and lower-bounds for the active learning of general concept classes have been in terms of a parameter of the learning problem called the splitting index. We provide, for the first time, an efficient algorithm…

机器学习 · 计算机科学 2017-06-12 Christopher Tosh , Sanjoy Dasgupta

We consider coordinate descent methods on convex quadratic problems, in which exact line searches are performed at each iteration. (This algorithm is identical to Gauss-Seidel on the equivalent symmetric positive definite linear system.) We…

最优化与控制 · 数学 2020-01-14 Stephen J. Wright , Ching-Pei Lee

Let $\Lambda$ be an artin algebra. We give an upper bound for the dimension of the bounded derived category of the category $\mod \Lambda$ of finitely generated right $\Lambda$-modules in terms of the projective and injective dimensions of…

环与代数 · 数学 2020-04-30 Junling Zheng , Zhaoyong Huang

We study the problem of learning-augmented predictive linear quadratic control. Our goal is to design a controller that balances \textit{"consistency"}, which measures the competitive ratio when predictions are accurate, and…

系统与控制 · 电气工程与系统科学 2025-04-08 Tongxin Li , Ruixiao Yang , Guannan Qu , Guanya Shi , Chenkai Yu , Adam Wierman , Steven H. Low