中文
相关论文

相关论文: Minimum Width for Deep, Narrow MLP: A Diffeomorphi…

200 篇论文

Deep neural networks (DNNs) achieve remarkable predictive performance but remain difficult to interpret, largely due to overparameterization that obscures the minimal structure required for interpretation. Here we introduce DeepIn, a…

统计方法学 · 统计学 2026-03-26 Zhiyao Tan , Liu Li , Huazhen Lin

For continuous functions, midpoint convexity characterizes convex functions. By considering discrete versions of midpoint convexity, several types of discrete convexities of functions, including integral convexity, L$^\natural$-convexity…

最优化与控制 · 数学 2020-02-03 Akihisa Tamura , Kazuya Tsurumi

In millimeter wave cellular communication, fast and reliable beam alignment via beam training is crucial to harvest sufficient beamforming gain for the subsequent data transmission. In this paper, we establish fundamental limits in…

信息论 · 计算机科学 2017-05-22 Chunshan Liu , Min Li , Stephen V. Hanly , Iain B. Collings , Philip Whiting

In this paper, we study the differential inclusion associated to the minimal surface system for two-dimensional graphs in $\mathbb{R}^{2 + n}$. We prove regularity of $W^{1,2}$ solutions and a compactness result for approximate solutions of…

偏微分方程分析 · 数学 2020-03-18 Riccardo Tione

Face recognition has achieved great progress owing to the fast development of the deep neural network in the past a few years. As an important part of deep neural networks, a number of the loss functions have been proposed which…

计算机视觉与模式识别 · 计算机科学 2020-02-05 Xin Wei , Hui Wang , Bryan Scotney , Huan Wan

While machine learning is widely used to optimize wireless networks, training a separate model for each task in communication and localization is becoming increasingly unsustainable due to the significant costs associated with training and…

信号处理 · 电气工程与系统科学 2025-11-20 Mohammad Cheraghinia , Eli De Poorter , Jaron Fontaine , Kwang Soon Kim , Merouane Debbah , Adnan Shahid

Low-dimensional embeddings are essential for machine learning tasks involving graphs, such as node classification, link prediction, community detection, network visualization, and network compression. Although recent studies have identified…

机器学习 · 计算机科学 2025-03-04 Nikolaos Nakis , Niels Raunkjær Holm , Andreas Lyhne Fiehn , Morten Mørup

This paper introduces a measure, called Lipschitz widths, of the optimal performance possible of certain nonlinear methods of approximation. It discusses their relation to entropy numbers and other well known widths such as the Kolmogorov…

数值分析 · 数学 2021-11-03 Guergana Petrova , Przemyslaw Wojtaszczyk

The weight decay regularization term is widely used during training to constrain expressivity, avoid overfitting, and improve generalization. Historically, this concept was borrowed from the SVM maximum margin principle and extended to…

机器学习 · 计算机科学 2021-10-12 Berry Weinstein , Shai Fine , Yacov Hel-Or

A realisation of a metric $d$ on a finite set $X$ is a weighted graph $(G,w)$ whose vertex set contains $X$ such that the shortest-path distance between elements of $X$ considered as vertices in $G$ is equal to $d$. Such a realisation…

组合数学 · 数学 2015-02-10 Sven Herrmann , Jack Koolen , Alice Lesser , Vincent Moulton , Taoyang Wu

Learning in Deep Neural Networks (DNN) takes place by minimizing a non-convex high-dimensional loss function, typically by a stochastic gradient descent (SGD) strategy. The learning process is observed to be able to find good minimizers…

机器学习 · 计算机科学 2020-03-12 Carlo Baldassi , Fabrizio Pittorino , Riccardo Zecchina

In the signal processing and statistics literature, the minimum description length (MDL) principle is a popular tool for choosing model complexity. Successful examples include signal denoising and variable selection in linear regression,…

信号处理 · 电气工程与系统科学 2022-01-28 Zhenyu Wei , Raymond K. W. Wong , Thomas C. M. Lee

We study the approximation of shift-invariant or equivariant functions by deep fully convolutional networks from the dynamical systems perspective. We prove that deep residual fully convolutional networks and their continuous-layer…

机器学习 · 计算机科学 2023-05-19 Ting Lin , Zuowei Shen , Qianxiao Li

Recent theoretical work suggested upper bounds on the operating bandwidths of flat lenses. Here, we show how these bounds can be circumvented via a multi-level diffractive lens (MDL) of diameter = 100 mm, focal length = 200 mm, device…

光学 · 物理学 2022-01-03 Apratim Majumder , Monjurul Meem , Nicole Brimhall , Rajesh Menon

We show that deep narrow Boltzmann machines are universal approximators of probability distributions on the activities of their visible units, provided they have sufficiently many hidden layers, each containing the same number of units as…

机器学习 · 统计学 2015-04-13 Guido Montufar

Recently, we introduced an approach for more easily interpreting searches for resonances at the LHC - and to aid in distinguishing between realistic and unrealistic alternatives for potential signals. This `simplfied limits' approach was…

高能物理 - 唯象学 · 物理学 2017-10-04 R. Sekhar Chivukula , Pawin Ittisamai , Kirtimaan Mohan , Elizabeth H. Simmons

Measuring the similarity of images is a fundamental problem to computer vision for which no universal solution exists. While simple metrics such as the pixel-wise L2-norm have been shown to have significant flaws, they remain popular. One…

计算机视觉与模式识别 · 计算机科学 2022-07-07 Oskar Sjögren , Gustav Grund Pihlgren , Fredrik Sandin , Marcus Liwicki

Depth is widely viewed as a central contributor to the success of deep neural networks, whereas standard neural network approximation theory typically provides guarantees only for the final output and leaves the role of intermediate layers…

机器学习 · 计算机科学 2026-04-23 Shijun Zhang , Zuowei Shen , Yuesheng Xu

This paper investigates approximation-theoretic aspects of the in-context learning capability of the transformers in representing a family of noisy linear dynamical systems. Our first theoretical result establishes an upper bound on the…

机器学习 · 计算机科学 2025-10-22 Frank Cole , Yuxuan Zhao , Yulong Lu , Tianhao Zhang

Consider an $s$-dimensional function being evaluated at $n$ points of a low discrepancy sequence (LDS), where the objective is to approximate the one-dimensional functions that result from integrating out $(s-1)$ variables. Here, the…

数值分析 · 数学 2019-11-11 Chaitanya Joshi , Paul T. Brown , Stephen Joe