English
Related papers

Related papers: Stable non-minimal fixed points of threshold-linea…

200 papers

Large language models (LLMs) display strong comprehensive abilities, yet the internal mechanisms that support these behaviors remain insufficiently understood. In this work, we show that across a wide range of open-weight Transformers, a…

Machine Learning · Computer Science 2026-05-29 Xiangtian Ji , Yuxin Chen , Zhengzhou Cai , Xiang Wang , An Zhang , Tat-Seng Chua

We present a new algorithm to generate minimal, stable, and symbolic corrections to an input that will cause a neural network with ReLU activations to change its output. We argue that such a correction is a useful way to provide feedback to…

Machine Learning · Computer Science 2018-09-03 Xin Zhang , Armando Solar-Lezama , Rishabh Singh

We are interested in fixed points in Boolean networks, {\em i.e.} functions $f$ from $\{0,1\}^n$ to itself. We define the subnetworks of $f$ as the restrictions of $f$ to the subcubes of $\{0,1\}^n$, and we characterizes a class…

Discrete Mathematics · Computer Science 2014-12-05 Adrien Richard

In deep learning (DL) the instability phenomenon is widespread and well documented, most commonly using the classical measure of stability, the Lipschitz constant. While a small Lipchitz constant is traditionally viewed as guarantying…

Machine Learning · Computer Science 2024-01-17 Z. N. D. Liu , A. C. Hansen

In this paper, we study the trainability of rectified linear unit (ReLU) networks. A ReLU neuron is said to be dead if it only outputs a constant for any input. Two death states of neurons are introduced; tentative and permanent death. A…

Machine Learning · Computer Science 2020-10-23 Yeonjong Shin , George Em Karniadakis

We examine the stability of loss-minimizing training processes that are used for deep neural networks (DNN) and other classifiers. While a classifier is optimized during training through a so-called loss function, the performance of…

Analysis of PDEs · Mathematics 2020-10-05 Leonid Berlyand , Pierre-Emmanuel Jabin , C. Alex Safsten

There are a number of hypotheses underlying the existence of adversarial examples for classification problems. These include the high-dimensionality of the data, high codimension in the ambient space of the data manifolds of interest, and…

Machine Learning · Computer Science 2024-04-15 Brian Bell , Michael Geyer , David Glickenstein , Keaton Hamm , Carlos Scheidegger , Amanda Fernandez , Juston Moore

We investigate the loss surface of neural networks. We prove that even for one-hidden-layer networks with "slightest" nonlinearity, the empirical risks have spurious local minima in most cases. Our results thus indicate that in general "no…

Machine Learning · Computer Science 2019-05-29 Chulhee Yun , Suvrit Sra , Ali Jadbabaie

It is commonly recognized that the expressiveness of deep neural networks is contingent upon a range of factors, encompassing their depth, width, and other relevant considerations. Currently, the practical performance of the majority of…

Machine Learning · Computer Science 2023-11-08 Xuan Qi , Yi Wei

A residual network (or ResNet) is a standard deep neural net architecture, with state-of-the-art performance across numerous applications. The main premise of ResNets is that they allow the training of each layer to focus on fitting just…

Machine Learning · Computer Science 2018-09-28 Ohad Shamir

We prove that if a continuous piecewise-smooth map on $\mathbb{R}^n$ is comprised of two linear functions, has a bounded orbit, and satisfies a certain non-degeneracy condition, then it has a fixed point. The result has important…

Dynamical Systems · Mathematics 2024-12-17 David J. W. Simpson

We prove that half spaces are the only stable nonlocal $s$-minimal cones in $\mathbb{R}^3$, for $s\in(0,1)$ sufficiently close to $1$. This is the first classification result of stable $s$-minimal cones in dimension higher than two. Its…

Analysis of PDEs · Mathematics 2017-10-25 Xavier Cabre , Eleonora Cinti , Joaquim Serra

In this work, we address the need for efficient and formally stable Recurrent Neural Networks (RNNs) in environments with limited computational resources by analyzing the stability of the Minimal Gated Unit (MGU) network, a lightweight…

Optimization and Control · Mathematics 2026-03-04 Stefano De Carli , Davide Previtali , Mirko Mazzoleni , Fabio Previdi

We study the generalization of two-layer ReLU neural networks in a univariate nonparametric regression problem with noisy labels. This is a problem where kernels (\emph{e.g.} NTK) are provably sub-optimal and benign overfitting does not…

Machine Learning · Computer Science 2024-06-12 Dan Qiao , Kaiqi Zhang , Esha Singh , Daniel Soudry , Yu-Xiang Wang

This paper investigates the tilt stability of local minimizers for nonlinear programs under the relaxed constant rank constraint qualification in finite dimensions. By employing a neighborhood primal-dual approach and extending calculus…

Optimization and Control · Mathematics 2025-08-12 Nguyen Huy Chieu , Nguyen Thi Quynh Trang , Nguyen Thi Hai Yen

Let P be a set of n points in $\mathbb{R}^d$. A point x is said to be a centerpoint of P if x is contained in every convex object that contains more than $dn\over d+1$ points of P. We call a point x a strong centerpoint for a family of…

Computational Geometry · Computer Science 2012-08-15 Pradeesha Ashok , Umair Azmi , Sathish Govindarajan

Algebraic neural networks (AlgNNs) are composed of a cascade of layers each one associated to and algebraic signal model, and information is mapped between layers by means of a nonlinearity function. AlgNNs provide a generalization of…

Machine Learning · Computer Science 2020-10-23 Alejandro Parada-Mayorga , Alejandro Ribeiro

Small-world networks are highly clustered networks with small distances among the nodes. There are many biological neural networks that present this kind of connections. There are no special weightings in the connections of most existing…

Disordered Systems and Neural Networks · Physics 2009-11-10 Chunguang Li , Guanrong Chen

Mini-batch sub-sampling in neural network training is unavoidable, due to growing data demands, memory-limited computational resources such as graphical processing units (GPUs), and the dynamics of on-line learning. In this study we…

Machine Learning · Statistics 2020-04-07 Dominic Kafka , Daniel Wilke

We present a fixed point theorem for a class of (potentially) non-monotonic functions over specially structured complete lattices. The theorem has as a special case the Knaster-Tarski fixed point theorem when restricted to the case of…

Logic in Computer Science · Computer Science 2015-02-10 Zoltán Ésik , Panos Rondogiannis
‹ Prev 1 3 4 5 6 7 10 Next ›