English
Related papers

Related papers: Relative gradient optimization of the Jacobian ter…

200 papers

Inverse problems arise in a number of domains such as medical imaging, remote sensing, and many more, relying on the use of advanced signal and image processing approaches -- such as sparsity-driven techniques -- to determine their…

Machine Learning · Computer Science 2019-02-01 Jaweria Amjad , Zhaoyan Lyu , Miguel R. D. Rodrigues

We study the problem of learning similarity functions over very large corpora using neural network embedding models. These models are typically trained using SGD with sampling of random observed and unobserved pairs, with a number of…

Machine Learning · Statistics 2018-07-20 Walid Krichene , Nicolas Mayoraz , Steffen Rendle , Li Zhang , Xinyang Yi , Lichan Hong , Ed Chi , John Anderson

Jacobian-Enhanced Neural Networks (JENN) are densely connected multi-layer perceptrons, whose training process is modified to predict partial derivatives accurately. Their main benefit is better accuracy with fewer training points compared…

Machine Learning · Computer Science 2026-01-01 Steven H. Berguin

We propose a gradient-based Jacobi algorithm for a class of maximization problems on the unitary group, with a focus on approximate diagonalization of complex matrices and tensors by unitary transformations. We provide weak convergence…

Optimization and Control · Mathematics 2020-07-13 Konstantin Usevich , Jianze Li , Pierre Comon

We introduce a new training algorithm for deep neural networks that utilize random complex exponential activation functions. Our approach employs a Markov Chain Monte Carlo sampling procedure to iteratively train network layers, avoiding…

Machine Learning · Computer Science 2025-03-07 Owen Davis , Gianluca Geraci , Mohammad Motamed

For large nonlinear least squares loss functions in machine learning we exploit the property that the number of model parameters typically exceeds the data in one batch. This implies a low-rank structure in the Hessian of the loss, which…

Machine Learning · Computer Science 2021-07-13 Johannes J. Brust

We describe, implement and test a novel method for training neural networks to estimate the Jacobian matrix $J$ of an unknown multivariate function $F$. The training set is constructed from finitely many pairs $(x,F(x))$ and it contains no…

Machine Learning · Computer Science 2022-04-04 Frédéric Latrémolière , Sadananda Narayanappa , Petr Vojtěchovský

There is growing interest in learning Fourier domain sampling strategies (particularly for magnetic resonance imaging, MRI) using optimization approaches. For non-Cartesian sampling patterns, the system models typically involve non-uniform…

Image and Video Processing · Electrical Eng. & Systems 2023-02-07 Guanhua Wang , Jeffrey A. Fessler

We propose a new way of constructing invertible neural networks by combining simple building blocks with a novel set of composition rules. This leads to a rich set of invertible architectures, including those similar to ResNets. Inversion…

Machine Learning · Computer Science 2019-10-30 Yang Song , Chenlin Meng , Stefano Ermon

During neural network training, the sharpness of the Hessian matrix of the training loss rises until training is on the edge of stability. As a result, even nonstochastic gradient descent does not accurately model the underlying dynamical…

Machine Learning · Statistics 2024-06-04 Mark Lowell , Catharine Kastner

We address the problem of distributed convex unconstrained optimization over networks characterized by asynchronous and possibly lossy communications. We analyze the case where the global cost function is the sum of locally coupled local…

Optimization and Control · Mathematics 2020-10-06 Marco Todescato , Nicoletta Bof , Guido Cavraro , Ruggero Carli , Luca Schenato

Direct image-to-image alignment that relies on the optimization of photometric error metrics suffers from limited convergence range and sensitivity to lighting conditions. Deep learning approaches has been applied to address this problem by…

Computer Vision and Pattern Recognition · Computer Science 2018-12-27 Lei Han , Mengqi Ji , Lu Fang , Matthias Nießner

We introduce a novel data-driven approach aimed at designing high-quality shape deformations based on a coarse localized input signal. Unlike previous data-driven methods that require a global shape encoding, we observe that…

Graphics · Computer Science 2024-10-14 Ramana Sundararaman , Nicolas Donati , Simone Melzi , Etienne Corman , Maks Ovsjanikov

Most machine learning methods require tuning of hyper-parameters. For kernel ridge regression with the Gaussian kernel, the hyper-parameter is the bandwidth. The bandwidth specifies the length scale of the kernel and has to be carefully…

Machine Learning · Statistics 2023-12-04 Oskar Allerbo , Rebecka Jörnsten

Gradients of neural networks can be computed efficiently for any architecture, but some applications require differential operators with higher time complexity. We describe a family of restricted neural network architectures that allow…

Machine Learning · Computer Science 2019-12-10 Ricky T. Q. Chen , David Duvenaud

Topological learning is a wide research area aiming at uncovering the mutual spatial relationships between the elements of a set. Some of the most common and oldest approaches involve the use of unsupervised competitive neural networks.…

Machine Learning · Statistics 2021-11-03 Pietro Barbiero , Gabriele Ciravegna , Vincenzo Randazzo , Giansalvo Cirrincione

When optimizing over-parameterized models, such as deep neural networks, a large set of parameters can achieve zero training error. In such cases, the choice of the optimization algorithm and its respective hyper-parameters introduces…

Machine Learning · Computer Science 2019-12-06 Gauthier Gidel , Francis Bach , Simon Lacoste-Julien

Projection-based model reduction has become a popular approach to reduce the cost associated with integrating large-scale dynamical systems so they can be used in many-query settings such as optimization and uncertainty quantification. For…

Numerical Analysis · Mathematics 2020-08-26 Han Gao , Jian-Xun Wang , Matthew J. Zahr

Generative networks implicitly approximate complex densities from their sampling with impressive accuracy. However, because of the enormous scale of modern datasets, this training process is often computationally expensive. We cast…

Machine Learning · Computer Science 2020-03-03 Vincent Schellekens , Laurent Jacques

Understanding why gradient-based training in deep networks exhibits strong implicit bias remains challenging, in part because tractable singular-value dynamics are typically available only for balanced deep linear models. We propose an…

Machine Learning · Computer Science 2026-02-17 Nathanaël Haas , François Gatine , Augustin M Cosse , Zied Bouraoui