中文
相关论文

相关论文: Elementary superexpressive activations

200 篇论文

Objective: Brain is a fantastic organ that helps creature adapting to the environment. Network is the most essential structure of brain, but the capability of a simple network is still not very clear. In this study, we try to expound some…

神经元与认知 · 定量生物学 2019-11-05 Xiang Zou , Lie Yao , Donghua Zhao , Liang Chen , Ying Mao

In this paper, we study approximation properties of single hidden layer neural networks with weights varying on finitely many directions and thresholds from an open interval. We obtain a necessary and at the same time sufficient measure…

机器学习 · 计算机科学 2023-04-05 Vugar Ismailov , Ekrem Savas

In this paper, we explain the universal approximation capabilities of deep residual neural networks through geometric nonlinear control. Inspired by recent work establishing links between residual networks and control systems, we provide a…

机器学习 · 计算机科学 2024-02-12 Paulo Tabuada , Bahman Gharesifard

We introduce a new framework for manipulating and interacting with deep generative models that we call network bending. We present a comprehensive set of deterministic transformations that can be inserted as distinct layers into the…

计算机视觉与模式识别 · 计算机科学 2021-03-15 Terence Broad , Frederic Fol Leymarie , Mick Grierson

Analysis and manipulation of trained neural networks is a challenging and important problem. We propose a symbolic representation for piecewise-linear neural networks and discuss its efficient computation. With this representation, one can…

机器学习 · 计算机科学 2019-08-21 Matthew Sotoudeh , Aditya V. Thakur

A key question in neuroscience is at which level functional meaning emerges from biophysical phenomena. In most vertebrate systems, precise functions are assigned at the level of neural populations, while single-neurons are deemed…

神经元与认知 · 定量生物学 2017-03-17 Wieland Brendel , Ralph Bourdoukan , Pietro Vertechi , Christian K. Machens , Sophie Denéve

Subshifts are sets of colorings of $\mathbb{Z}^d$ defined by families of forbidden patterns. In a given subshift, the extender set of a finite pattern is the set of all its admissible completions. Since soficity of $\mathbb{Z}$ subshifts is…

离散数学 · 计算机科学 2025-10-03 Antonin Callard , Léo Paviet Salomon , Pascal Vanier

We construct examples of finitely presented simple groups whose Dehn functions are at least exponential. To the best of our knowledge, these are the first such examples known. Our examples arise from R\"over-Nekrashevych groups, using…

群论 · 数学 2024-07-12 Matthew C. B. Zaremsky

Flexible models for probability distributions are an essential ingredient in many machine learning tasks. We develop and investigate a new class of probability distributions, which we call a Squared Neural Family (SNEFY), formed by squaring…

机器学习 · 计算机科学 2023-10-27 Russell Tsuchida , Cheng Soon Ong , Dino Sejdinovic

Data-driven constitutive modeling frameworks based on neural networks and classical representation theorems have recently gained considerable attention due to their ability to easily incorporate constitutive constraints and their excellent…

软凝聚态物质 · 物理学 2023-08-23 Jan N. Fuhg , Nikolaos Bouklas , Reese E. Jones

We introduce a class of trainable nonlinear operators based on semirings that are suitable for use in neural networks. These operators generalize the traditional alternation of linear operators with activation functions in neural networks.…

机器学习 · 计算机科学 2025-03-11 Bart M. N. Smets , Peter D. Donker , Jim W. Portegies

Maximum entropy models are the least structured probability distributions that exactly reproduce a chosen set of statistics measured in an interacting network. Here we use this principle to construct probabilistic models which describe the…

神经元与认知 · 定量生物学 2014-01-28 Gašper Tkačik , Olivier Marre , Dario Amodei , Elad Schneidman , William Bialek , Michael J Berry

We extend the hyperplane arrangement framework for neural network expressivity from the braid to discriminantal arrangements. Compatible piecewise linear functions are characterized by circuit relations and admit a matroidal description via…

组合数学 · 数学 2026-04-06 Pragnya Das

Artificial neurons with arbitrarily complex internal structure are introduced. The neurons can be described in terms of a set of internal variables, a set activation functions which describe the time evolution of these variables and a set…

神经与进化计算 · 计算机科学 2007-05-23 G. A. Kohring

The standard approaches to neural network implementation yield powerful function approximation capabilities but are limited in their abilities to learn meta representations and reason probabilistic uncertainties in their predictions.…

机器学习 · 计算机科学 2023-10-05 Saurav Jha , Dong Gong , Xuesong Wang , Richard E. Turner , Lina Yao

We study the use of binary activated neural networks as interpretable and explainable predictors in the context of regression tasks on tabular data; more specifically, we provide guarantees on their expressiveness, present an approach based…

机器学习 · 计算机科学 2024-06-11 Benjamin Leblanc , Pascal Germain

Modern neural networks are often quite wide, causing large memory and computation costs. It is thus of great interest to train a narrower network. However, training narrow neural nets remains a challenging task. We ask two theoretical…

机器学习 · 计算机科学 2022-10-24 Jiawei Zhang , Yushun Zhang , Mingyi Hong , Ruoyu Sun , Zhi-Quan Luo

We consider active, semi-supervised learning in an offline transductive setting. We show that a previously proposed error bound for active learning on undirected weighted graphs can be generalized by replacing graph cut with an arbitrary…

机器学习 · 计算机科学 2012-02-20 Andrew Guillory , Jeff A. Bilmes

We give a large family of almost perfect nonlinear (APN) permutations of finite vector spaces of every odd dimension divisible by three. We also give APN functions that are not bijective on even dimensions and related highly nonlinear…

组合数学 · 数学 2026-05-19 Faruk Göloğlu , Lukas Kölsch

In this paper, we prove that in the overparametrized regime, deep neural network provide universal approximations and can interpolate any data set, as long as the activation function is locally in $L^1(\RR)$ and not an affine function.…

机器学习 · 计算机科学 2024-04-26 Vlad-Raul Constantinescu , Ionel Popescu