English
Related papers

Related papers: Simplified Information Geometry Approach for Massi…

200 papers

We present a rigorous convergence analysis of a new method for density-based topology optimization that provides point-wise bound preserving design updates and faster convergence than other popular first-order topology optimization methods.…

Optimization and Control · Mathematics 2025-02-25 Brendan Keith , Dohyun Kim , Boyan S. Lazarov , Thomas M. Surowiec

The stochastic gradient descent (SGD) algorithm has been widely used in statistical estimation for large-scale data due to its computational and memory efficiency. While most existing works focus on the convergence of the objective function…

Machine Learning · Statistics 2023-11-02 Xi Chen , Jason D. Lee , Xin T. Tong , Yichen Zhang

The information exponent ([BAGJ21]) and its extensions -- which are equivalent to the lowest degree in the Hermite expansion of the link function (after a potential label transform) for Gaussian single-index models -- have played an…

Machine Learning · Computer Science 2025-10-07 Yunwei Ren , Jason D. Lee

Using the generalized entropies which depend on two parameters we propose a set of quantitative characteristics derived from the Information Geometry based on these entropies. Our aim, at this stage, is modest, as we are first constructing…

Mathematical Physics · Physics 2018-02-14 Demetris P. K. Ghikas , Fotios Oikonomou

This article is concerned with the numerical solution of subspace optimization problems, consisting of minimizing a smooth functional over the set of orthogonal projectors of fixed rank. Such problems are encountered in particular in…

Numerical Analysis · Mathematics 2022-10-17 Eric Cancès , Gaspard Kemlin , Antoine Levitt

We consider a high-dimensional monotone single index model (hdSIM), which is a semiparametric extension of a high-dimensional generalize linear model (hdGLM), where the link function is unknown, but constrained with monotone and…

Statistics Theory · Mathematics 2021-05-18 Ran Dai , Hyebin Song , Rina Foygel Barber , Garvesh Raskutti

In this paper, we provide novel algorithms with identifiability guarantees for simplex-structured matrix factorization (SSMF), a generalization of nonnegative matrix factorization. Current state-of-the-art algorithms that provide…

Machine Learning · Computer Science 2021-05-12 Maryam Abdolali , Nicolas Gillis

The analysis of second-order optimization methods based either on sub-sampling, randomization or sketching has two serious shortcomings compared to the conventional Newton method. The first shortcoming is that the analysis of the iterates…

Optimization and Control · Mathematics 2024-04-05 Nick Tsipinakis , Panos Parpas

Graph matching finds the correspondence of nodes across two correlated graphs and lies at the core of many applications. When graph side information is not available, the node correspondence is estimated on the sole basis of network…

Machine Learning · Computer Science 2022-02-08 Weijie Liu , Chao Zhang , Nenggan Zheng , Hui Qian

The information geometry of the 2-manifold of gamma probability density functions provides a framework in which pseudorandom number generators may be evaluated using a neighbourhood of the curve of exponential density functions. The process…

Computation · Statistics 2009-07-13 C. T. J. Dodson

Federated learning aggregates model updates from distributed clients, but standard first order methods such as FedAvg apply the same scalar weight to all parameters from each client. Under non-IID data, these uniformly weighted updates can…

Machine Learning · Computer Science 2026-01-21 Zhipeng Chang , Ting He , Wenrui Hao

We propose a stochastic conditional gradient method (CGM) for minimizing convex finite-sum objectives formed as a sum of smooth and non-smooth terms. Existing CGM variants for this template either suffer from slow convergence rates, or…

Machine Learning · Computer Science 2022-04-19 Gideon Dresdner , Maria-Luiza Vladarean , Gunnar Rätsch , Francesco Locatello , Volkan Cevher , Alp Yurtsever

Convex-composite optimization, which minimizes an objective function represented by the sum of a differentiable function and a convex one, is widely used in machine learning and signal/image processing. Fast Iterative Shrinkage Thresholding…

Optimization and Control · Mathematics 2022-05-12 Hiroki Tanabe , Ellen H. Fukuda , Nobuo Yamashita

With a unified belief propagation (BP) and mean field (MF) framework, we propose an iterative message passing receiver, which performs joint channel state and noise precision (the reciprocal of noise variance) estimation and decoding for…

Information Theory · Computer Science 2017-11-13 Zhengdao Yuan , Chuanzong Zhang , Zhongyong Wang , Qinghua Guo , Sheng Wu , and Xingye Wang

Stochastic variance-reduced algorithms such as Stochastic Average Gradient (SAG) and SAGA, and their deterministic counterparts like the Incremental Aggregated Gradient (IAG) method, have been extensively studied in large-scale machine…

Machine Learning · Computer Science 2026-05-22 Feng Zhu , Robert W. Heath , Aritra Mitra

We study the minimization of smooth, possibly nonconvex functions over the positive orthant, a key setting in Poisson inverse problems, using the exponentiated gradient (EG) method. Interpreting EG as Riemannian gradient descent (RGD) with…

Optimization and Control · Mathematics 2025-04-08 Yara Elshiaty , Ferdinand Vanmaele , Stefania Petra

We consider the problem of minimizing a functional over a parametric family of probability measures, where the parameterization is characterized via a push-forward structure. An important application of this problem is in training…

Machine Learning · Statistics 2020-11-10 Zebang Shen , Zhenfu Wang , Alejandro Ribeiro , Hamed Hassani

Motivated by the success of Sinkhorn's algorithm for entropic optimal transport, we study convergence properties of iterative proportional fitting procedures (IPFP) used to solve more general information projection problems. We establish…

Optimization and Control · Mathematics 2025-04-14 Stephan Eckstein , Aziz Lakhal

Diffusion models have achieved remarkable success in synthesizing complex static and temporal visuals, a breakthrough largely driven by Classifier-Free Guidance (CFG). However, despite its pivotal role in aligning generated content with…

Computer Vision and Pattern Recognition · Computer Science 2026-04-30 Haosen Li , Wenshuo Chen , Lei Wang , Shaofeng Liang , Bowen Tian , Soning Lai , Yutao Yue

We introduce a novel method for solving density-based topology optimization problems: Sigmoidal Mirror descent with a Projected Latent variable (SiMPL). The SiMPL method (pronounced as ``the simple method'') optimizes a design using only…

Optimization and Control · Mathematics 2025-02-25 Dohyun Kim , Boyan Stefanov Lazarov , Thomas M. Surowiec , Brendan Keith