English
Related papers

Related papers: Variations on the Chebyshev-Lagrange Activation Fu…

200 papers

This study investigates imposing hard inequality constraints on the outputs of convolutional neural networks (CNN) during training. Several recent works showed that the theoretical and practical advantages of Lagrangian optimization over…

Computer Vision and Pattern Recognition · Computer Science 2023-08-31 Hoel Kervadec , Jose Dolz , Jing Yuan , Christian Desrosiers , Eric Granger , Ismail Ben Ayed

We study deep neural networks with polynomial activations, particularly their expressive power. For a fixed architecture and activation degree, a polynomial neural network defines an algebraic map from weights to polynomials. The image of…

Machine Learning · Computer Science 2019-05-30 Joe Kileel , Matthew Trager , Joan Bruna

For a non-generic, yet dense subset of $C^1$ expanding Markov maps of the interval we prove the existence of uncountably many Lyapunov optimizing measures which are ergodic, fully supported and have positive entropy. These measures are…

Dynamical Systems · Mathematics 2017-08-29 Mao Shinoda , Hiroki Takahasi

Deep Neural Networks (DNNs) have obtained impressive performance across tasks, however they still remain as black boxes, e.g., hard to theoretically analyze. At the same time, Polynomial Networks (PNs) have emerged as an alternative method…

Computer Vision and Pattern Recognition · Computer Science 2023-03-27 Grigorios G Chrysos , Bohan Wang , Jiankang Deng , Volkan Cevher

Many activation functions have been proposed in the past, but selecting an adequate one requires trial and error. We propose a new methodology of designing activation functions within a neural network at each layer. We call this technique…

Machine Learning · Statistics 2017-02-28 Mark Harmon , Diego Klabjan

In current textbooks the use of Chebyshev nodes with Newton interpolation is advocated as the most efficient numerical interpolation method in terms of approximation accuracy and computational effort. However, we show numerically that the…

Numerical Analysis · Mathematics 2016-09-29 Michael Breuß , Friedemann Kemm , Oliver Vogel

Activation functions play a vital role in the training of Convolutional Neural Networks. For this reason, to develop efficient and performing functions is a crucial problem in the deep learning community. Key to these approaches is to…

Computer Vision and Pattern Recognition · Computer Science 2020-09-23 Gianluca Maguolo , Loris Nanni , Stefano Ghidoni

Low-rank decomposition has emerged as a vital tool for enhancing parameter efficiency in neural network architectures, gaining traction across diverse applications in machine learning. These techniques significantly lower the number of…

Machine Learning · Computer Science 2025-03-18 Yiping Ji , Hemanth Saratchandran , Cameron Gordon , Zeyu Zhang , Simon Lucey

We study Chebyshev filter diagonalization as a tool for the computation of many interior eigenvalues of very large sparse symmetric matrices. In this technique the subspace projection onto the target space of wanted eigenvectors is…

The development of effective initialization methods requires an understanding of random neural networks. In this work, a rigorous probabilistic analysis of deep unbiased Leaky ReLU networks is provided. We prove a Law of Large Numbers and a…

Machine Learning · Statistics 2026-02-12 Constantin Kogler , Tassilo Schwarz , Samuel Kittle

Linearized shallow neural networks that are constructed by fixing the hidden-layer parameters have recently shown strong performance in solving partial differential equations (PDEs). Such models, widely used in the random feature method…

Numerical Analysis · Mathematics 2026-01-21 Tong Mao , Jinchao Xu , Xiaofeng Xu

We study the realization map of deep ReLU networks, focusing on when a function determines its parameters up to scaling and permutation. To analyze hidden redundancies beyond these standard symmetries, we introduce a framework based on…

Machine Learning · Computer Science 2026-05-21 Moritz Grillo , Guido Montúfar

This paper presents a new approach for assembling graph neural networks based on framelet transforms. The latter provides a multi-scale representation for graph-structured data. We decompose an input graph into low-pass and high-pass…

Machine Learning · Computer Science 2021-06-22 Xuebin Zheng , Bingxin Zhou , Junbin Gao , Yu Guang Wang , Pietro Lio , Ming Li , Guido Montufar

We study the complexity of functions computable by deep feedforward neural networks with piecewise linear activations in terms of the symmetries and the number of linear regions that they have. Deep networks are able to sequentially map…

Machine Learning · Statistics 2014-06-10 Guido Montúfar , Razvan Pascanu , Kyunghyun Cho , Yoshua Bengio

We propose a novel activation function that implements piece-wise orthogonal non-linear mappings based on permutations. It is straightforward to implement, and very computationally efficient, also it has little memory requirements. We…

Neural and Evolutionary Computing · Computer Science 2017-02-02 Artem Chernodub , Dimitri Nowicki

We train an artificial neural network which distinguishes chaotic and regular dynamics of the two-dimensional Chirikov standard map. We use finite length trajectories and compare the performance with traditional numerical methods which need…

Machine Learning · Computer Science 2020-04-24 Woo Seok Lee , Sergej Flach

The nonlocal-based blocks are designed for capturing long-range spatial-temporal dependencies in computer vision tasks. Although having shown excellent performance, they still lack the mechanism to encode the rich, structured information…

Computer Vision and Pattern Recognition · Computer Science 2021-08-18 Lei Zhu , Qi She , Duo Li , Yanye Lu , Xuejing Kang , Jie Hu , Changhu Wang

Activation functions influence behavior and performance of DNNs. Nonlinear activation functions, like Rectified Linear Units (ReLU), Exponential Linear Units (ELU) and Scaled Exponential Linear Units (SELU), outperform the linear…

Neural and Evolutionary Computing · Computer Science 2019-02-05 Alberto Marchisio , Muhammad Abdullah Hanif , Semeen Rehman , Maurizio Martina , Muhammad Shafique

Activation functions play a key role in providing remarkable performance in deep neural networks, and the rectified linear unit (ReLU) is one of the most widely used activation functions. Various new activation functions and improvements on…

Machine Learning · Computer Science 2019-08-27 Yang Liu , Jianpeng Zhang , Chao Gao , Jinghua Qu , Lixin Ji

In this paper, we introduce the class of $(\beta,\gamma)$-Chebyshev functions and corresponding points, which can be seen as a family of {\it generalized} Chebyshev polynomials and points. For the $(\beta,\gamma)$-Chebyshev functions, we…

Numerical Analysis · Mathematics 2021-11-23 Stefano De Marchi , Giacomo Elefante , Francesco Marchetti
‹ Prev 1 4 5 6 7 8 10 Next ›