English
Related papers

Related papers: Minimal model of permutation symmetry in unsupervi…

200 papers

In a previous report we have evaluated analytically the mutual information between the firing rates of N independent units and a set of continuous+discrete stimuli, for finite N and in the limit of large noise. Here, we extend the analysis…

Disordered Systems and Neural Networks · Physics 2009-11-07 Valeria Del Prete , Alessandro Treves

Collaborative learning through latent shared feature representations enables heterogeneous clients to train personalized models with improved performance and reduced sample complexity. Despite empirical success and extensive study, the…

Machine Learning · Computer Science 2025-11-25 Xiaochun Niu , Lili Su , Jiaming Xu , Pengkun Yang

We propose a method for learning embeddings for few-shot learning that is suitable for use with any number of ways and any number of shots (shot-free). Rather than fixing the class prototypes to be the Euclidean average of sample…

Machine Learning · Computer Science 2020-04-23 Avinash Ravichandran , Rahul Bhotika , Stefano Soatto

Training neural networks means solving a high-dimensional optimization problem. Normally the goal is to minimize a loss function that depends on what is called the network function, or in other words the function that gives the network…

Machine Learning · Computer Science 2022-11-15 Umberto Michelucci

This paper introduces a new learning paradigm termed Neural Metamorphosis (NeuMeta), which aims to build self-morphable neural networks. Contrary to crafting separate models for different architectures or sizes, NeuMeta directly learns the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-17 Xingyi Yang , Xinchao Wang

We study orbifolds by permutations of two identical N=2 minimal models within the Gepner construction of four dimensional heterotic strings. This is done using the new N=2 supersymmetric permutation orbifold building blocks we have recently…

High Energy Physics - Theory · Physics 2011-08-03 M. Maio , A. N. Schellekens

Understanding the mechanisms behind neural network optimization is crucial for improving network design and performance. While various optimization techniques have been developed, a comprehensive understanding of the underlying principles…

Machine Learning · Computer Science 2024-09-13 Jun-Jie Zhang , Nan Cheng , Fu-Peng Li , Xiu-Cheng Wang , Jian-Nan Chen , Long-Gang Pang , Deyu Meng

Non-Hermitian systems offer new platforms for unusual physical properties that can be flexibly manipulated by redistribution of the real and imaginary parts of refractive indices, whose presence breaks conventional wave propagation…

Optics · Physics 2022-04-29 W. W. Ahmed , M. Farhat , K. Staliunas , X. Zhang , Y. Wu

Statistical inference from high-dimensional data with low-dimensional structures has recently attracted lots of attention. In machine learning, deep generative modeling approaches implicitly estimate distributions of complex objects by…

Statistics Theory · Mathematics 2022-02-21 Rong Tang , Yun Yang

Recent advances in self-supervised learning have attracted significant attention from both machine learning and neuroscience. This is primarily because self-supervised methods do not require annotated supervisory information, making them…

Neurons and Cognition · Quantitative Biology 2025-12-05 Asaki Kataoka , Yoshihiro Nagano , Masafumi Oizumi

The bias-variance trade-off is a central concept in supervised learning. In classical statistics, increasing the complexity of a model (e.g., number of parameters) reduces bias but also increases variance. Until recently, it was commonly…

Machine Learning · Statistics 2022-03-25 Jason W. Rocks , Pankaj Mehta

Double descent is a surprising phenomenon in machine learning, in which as the number of model parameters grows relative to the number of data, test error drops as models grow ever larger into the highly overparameterized (data…

Recent analysis on the training dynamics of Transformers has unveiled an interesting characteristic: the training loss plateaus for a significant number of training steps, and then suddenly (and sharply) drops to near--optimal values. To…

Machine Learning · Computer Science 2024-10-30 Pulkit Gopalani , Ekdeep Singh Lubana , Wei Hu

Recent empirical studies have demonstrated that diffusion models can effectively learn the image distribution and generate new samples. Remarkably, these models can achieve this even with a small number of training samples despite a large…

Machine Learning · Computer Science 2025-07-08 Peng Wang , Huijie Zhang , Zekai Zhang , Siyi Chen , Yi Ma , Qing Qu

This study investigates the dynamics of alternating minimization applied to a bilinear regression task with normally distributed covariates, under the asymptotic system size limit where the number of parameters and observations diverge at…

Optimization and Control · Mathematics 2025-02-03 Koki Okajima , Takashi Takahashi

Biological visual systems learn from limited experience, unlike deep learning models that rely on millions of training images. What learning principles make this possible? We tested whether efficient coding, the idea that neural…

Computer Vision and Pattern Recognition · Computer Science 2026-05-20 Ananya Passi , Brian S. Robinson , Michael F. Bonner

Suppose that two large, multi-dimensional data sets are each noisy measurements of the same underlying random process, and principle components analysis is performed separately on the data sets to reduce their dimensionality. In some…

Machine Learning · Statistics 2024-06-27 Donniell E. Fishkind , Cencheng Shen , Youngser Park , Carey E. Priebe

We study the problem of learning overcomplete HMMs---those that have many hidden states but a small output alphabet. Despite having significant practical importance, such HMMs are poorly understood with no known positive or negative results…

Machine Learning · Computer Science 2018-06-29 Vatsal Sharan , Sham Kakade , Percy Liang , Gregory Valiant

Graphical models are a rich language for describing high-dimensional distributions in terms of their dependence structure. While there are algorithms with provable guarantees for learning undirected graphical models in a variety of…

Machine Learning · Computer Science 2018-11-07 Guy Bresler , Frederic Koehler , Ankur Moitra , Elchanan Mossel

Large Language Models (LLMs) are typically trained on data mixtures: most data come from web scrapes, while a small portion is curated from high-quality sources with dense domain-specific knowledge. In this paper, we show that when training…

Machine Learning · Computer Science 2026-05-12 Xinran Gu , Kaifeng Lyu , Jiazheng Li , Jingzhao Zhang
‹ Prev 1 8 9 10 Next ›