English
Related papers

Related papers: Deep Kuratowski Embedding Neural Networks for Wass…

200 papers

Solving large scale Optimal Transport (OT) in machine learning typically relies on sampling measures to obtain a tractable discrete problem. While the discrete solver's accuracy is controllable, the rate of convergence of the discretization…

Machine Learning · Statistics 2026-02-05 Ferdinand Genans , Olivier Wintenberger

Deep convolutional neural networks (CNNs) are broadly considered to be state-of-the-art generic end-to-end image classification systems. However, they are known to underperform when training data are limited and thus require data…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Mohammad Shifat E Rabbi , Yan Zhuang , Shiying Li , Abu Hasnat Mohammad Rubaiyat , Xuwang Yin , Gustavo K. Rohde

Deep learning (deep structured learning, hierarchi- cal learning or deep machine learning) is a branch of machine learning based on a set of algorithms that attempt to model high- level abstractions in data by using multiple processing…

Machine Learning · Computer Science 2018-01-30 Yuriy Kochura , Sergii Stirenko , Yuri Gordienko

The ``differentiability gap'' presents a primary bottleneck in Earth system deep learning: since models cannot be trained directly on non-differentiable scientific metrics and must rely on smooth proxies (e.g., MSE), they often fail to…

Machine Learning · Computer Science 2026-04-14 Filippo Quarenghi , Ryan Cotsakis , Tom Beucler

Given two labeled data-sets $\mathcal{S}$ and $\mathcal{T}$, we design a simple and efficient greedy algorithm to reweigh the loss function such that the limiting distribution of the neural network weights that result from training on…

Machine Learning · Statistics 2025-02-26 Pratik Worah

We propose using the Wasserstein loss for training in inverse problems. In particular, we consider a learned primal-dual reconstruction scheme for ill-posed inverse problems using the Wasserstein distance as loss function in the learning.…

Computer Vision and Pattern Recognition · Computer Science 2017-10-31 Jonas Adler , Axel Ringh , Ozan Öktem , Johan Karlsson

This paper studies the problem of computing a linear approximation of quadratic Wasserstein distance $W_2$. In particular, we compute an approximation of the negative homogeneous weighted Sobolev norm whose connection to Wasserstein…

Numerical Analysis · Mathematics 2022-03-02 Philip Greengard , Jeremy G. Hoskins , Nicholas F. Marshall , Amit Singer

DeepONets have recently been proposed as a framework for learning nonlinear operators mapping between infinite dimensional Banach spaces. We analyze DeepONets and prove estimates on the resulting approximation and generalization errors. In…

Numerical Analysis · Mathematics 2022-01-14 Samuel Lanthaler , Siddhartha Mishra , George Em Karniadakis

It is challenging to design an equalizer for the complex time-frequency doubly-selective channel. In this paper, we employ the deep unfolding approach to establish an equalizer for the underwater acoustic (UWA) orthogonal frequency division…

Signal Processing · Electrical Eng. & Systems 2023-06-02 Hao Zhao , Cui Yang , Yalu Xu , Fei Ji , Miaowen Wen , Yankun Chen

The demand of artificial intelligent adoption for condition-based maintenance strategy is astonishingly increased over the past few years. Intelligent fault diagnosis is one critical topic of maintenance solution for mechanical systems.…

Machine Learning · Computer Science 2022-06-17 Cheng Cheng , Beitong Zhou , Guijun Ma , Dongrui Wu , Ye Yuan

In many domains of computer vision, generative adversarial networks (GANs) have achieved great success, among which the family of Wasserstein GANs (WGANs) is considered to be state-of-the-art due to the theoretical contributions and…

Computer Vision and Pattern Recognition · Computer Science 2018-09-06 Jiqing Wu , Zhiwu Huang , Janine Thoma , Dinesh Acharya , Luc Van Gool

Regression loss design is an essential topic for oriented object detection. Due to the periodicity of the angle and the ambiguity of width and height definition, traditional L1-distance loss and its variants have been suffered from the…

Computer Vision and Pattern Recognition · Computer Science 2023-12-13 Yuke Zhu , Yumeng Ruan , Zihua Xiong , Sheng Guo

Disentangling polysemantic neurons is at the core of many current approaches to interpretability of large language models. Here we attempt to study how disentanglement can be used to understand performance, particularly under weight…

Machine Learning · Computer Science 2025-02-27 Shashata Sawmya , Linghao Kong , Ilia Markov , Dan Alistarh , Nir Shavit

Persistence diagrams (PD)s play a central role in topological data analysis, and are used in an ever increasing variety of applications. The comparison of PD data requires computing comparison metrics among large sets of PDs, with metrics…

Computational Geometry · Computer Science 2024-02-23 Rolando Kindelan Nuñez , Mircea Petrache , Mauricio Cerda , Nancy Hitschfeld

Deep neural networks (DNNs) have provided brilliant performance across various tasks. However, this success often comes at the cost of unnecessarily large model sizes, high computational demands, and substantial memory footprints.…

Machine Learning · Computer Science 2025-11-26 Shaharyar Ahmed Khan Tareen , Filza Khan Tareen

Structure-preserved denoising of 3D magnetic resonance imaging (MRI) images is a critical step in medical image analysis. Over the past few years, many algorithms with impressive performances have been proposed. In this paper, inspired by…

Medical Physics · Physics 2019-05-07 Maosong Ran , Jinrong Hu , Yang Chen , Hu Chen , Huaiqiang Sun , Jiliu Zhou , Yi Zhang

An increasing number of machine learning tasks deal with learning representations from set-structured data. Solutions to these problems involve the composition of permutation-equivariant modules (e.g., self-attention, or individual…

Machine Learning · Computer Science 2021-03-09 Navid Naderializadeh , Soheil Kolouri , Joseph F. Comer , Reed W. Andrews , Heiko Hoffmann

Stacking Gaussian Processes severely diminishes the model's ability to detect outliers, which when combined with non-zero mean functions, further extrapolates low non-parametric variance to low training data density regions. We propose a…

Machine Learning · Statistics 2022-02-02 Sebastian Popescu , David Sharp , James Cole , Ben Glocker

Multiple marginal matching problem aims at learning mappings to match a source domain to multiple target domains and it has attracted great attention in many applications, such as multi-domain image translation. However, addressing this…

Machine Learning · Computer Science 2019-11-05 Jiezhang Cao , Langyuan Mo , Yifan Zhang , Kui Jia , Chunhua Shen , Mingkui Tan

The Wasserstein distance from optimal mass transport (OMT) is a powerful mathematical tool with numerous applications that provides a natural measure of the distance between two probability distributions. Several methods to incorporate OMT…

Machine Learning · Computer Science 2023-10-31 Jung Hun Oh , Rena Elkin , Anish Kumar Simhal , Jiening Zhu , Joseph O Deasy , Allen Tannenbaum