中文
相关论文

相关论文: Pruned Wasserstein Index Generation Model and wigp…

200 篇论文

Transfer learning is a popular strategy to leverage external knowledge and improve statistical efficiency, particularly with a limited target sample. We propose a novel knowledge-guided Wasserstein Distributionally Robust Optimization…

机器学习 · 计算机科学 2025-02-13 Zitao Wang , Ziyuan Wang , Molei Liu , Nian Si

Wasserstein Gradient Flows (WGF) with respect to specific functionals have been widely used in the machine learning literature. Recently, neural networks have been adopted to approximate certain intractable parts of the underlying…

机器学习 · 计算机科学 2024-01-26 Huminhao Zhu , Fangyikang Wang , Chao Zhang , Hanbin Zhao , Hui Qian

We consider the distributional connection between the lossy compressed representation of a high-dimensional signal $X$ using a random spherical code and the observation of $X$ under an additive white Gaussian noise (AWGN). We show that the…

信息论 · 计算机科学 2021-12-14 Alon Kipnis , Galen Reeves

This paper presents a population synthesis model that utilizes the Wasserstein Generative-Adversarial Network (WGAN) for training on incomplete microsamples. By using a mask matrix to represent missing values, the study proposes a WGAN…

机器学习 · 计算机科学 2025-10-02 Tanay Rastogi , Daniel Jonsson , Anders Karlström

Gaussian process (GP) regression is widely used for uncertainty quantification, yet the standard formulation assumes noise-free covariates. When inputs are measured with error, this errors-in-variables (EIV) setting can lead to…

统计方法学 · 统计学 2026-03-19 Hengrui Luo , Xiaoye S. Li , Yang Liu , Marcus Noack , Ji Qiang , Mark D. Risser

Intent recognition (IR) for speech commands is essential for artificial intelligence (AI) assistant systems; however, most existing approaches are limited to short commands and are predominantly developed for English. This paper addresses…

计算与语言 · 计算机科学 2025-08-11 Theresa Pekarek Rosin , Burak Can Kaplan , Stefan Wermter

Recent advancements in large language models have intensified the need for efficient and deployable models within limited inference budgets. Structured pruning pipelines have shown promise in token efficiency compared to training…

计算与语言 · 计算机科学 2025-03-11 Yixiao Li , Xianzhi Du , Ajay Jaiswal , Tao Lei , Tuo Zhao , Chong Wang , Jianyu Wang

We present a weak formulation and discretization of the system discovery problem from noisy measurement data. This method of learning differential equations from data fits into a new class of algorithms that replace pointwise derivative…

数值分析 · 数学 2022-05-03 Daniel A. Messenger , David M. Bortz

Discrete diffusion models have emerged as a powerful paradigm for generative modeling on sequence data; however, the information-theoretic principles governing their reverse processes remain significantly less understood than those of their…

机器学习 · 计算机科学 2026-02-10 Alberto Foresti , Mustapha Bounoua , Giulio Franzese , Luca Ambrogioni , Pietro Michiardi

Point processes are becoming very popular in modeling asynchronous sequential data due to their sound mathematical foundation and strength in modeling a variety of real-world phenomena. Currently, they are often characterized via intensity…

机器学习 · 计算机科学 2017-05-24 Shuai Xiao , Mehrdad Farajtabar , Xiaojing Ye , Junchi Yan , Le Song , Hongyuan Zha

This paper presents a pruning technique which can be used to reduce the number of paths searched in rule-based bag generators of the type proposed by \cite{poznanskietal95} and \cite{popowich95}. Pruning the search space in these generators…

cmp-lg · 计算机科学 2008-02-03 Arturo Trujillo , Simon Berry

Gaussian graphical models are widely used to infer dependence structures. Bayesian methods are appealing to quantify uncertainty associated with structural learning, i.e., the plausibility of conditional independence statements given the…

统计方法学 · 统计学 2025-11-05 Deborah Sulem , Jack Jewson , David Rossell

In this paper, we propose new sampling approaches for the Shrinkage Inverse-Wishart (SIW) distribution, a generalized family of the Inverse-Wishart distribution originally proposed by Berger et al. (2020, Annals of Statistics). It offers a…

统计方法学 · 统计学 2025-11-14 Yiye Jiang

We propose a new statistical model, the spiked transport model, which formalizes the assumption that two probability distributions differ only on a low-dimensional subspace. We study the minimax rate of estimation for the Wasserstein…

统计理论 · 数学 2019-09-18 Jonathan Niles-Weed , Philippe Rigollet

This study addresses the challenge of inaccurate gradients in computing the empirical Fisher Information Matrix during neural network pruning. We introduce SWAP, a formulation of Entropic Wasserstein regression (EWR) for pruning,…

人工智能 · 计算机科学 2024-02-21 Lei You , Hei Victor Cheng

Score-based generative models are shown to achieve remarkable empirical performances in various applications such as image generation and audio synthesis. However, a theoretical understanding of score-based diffusion models is still…

机器学习 · 计算机科学 2022-12-14 Dohyun Kwon , Ying Fan , Kangwook Lee

This paper studies the optimization of the KL functional on the Wasserstein space of probability measures, and develops a sampling framework based on Wasserstein gradient descent (WGD). We identify two important subclasses of the…

统计计算 · 统计学 2026-02-04 Van Chien Ta , Thi Mai Hong Chu , Minh-Ngoc Tran

Large language models have demonstrated capabilities in text generation, while their increasing parameter scales present challenges in computational and memory efficiency. Post-training sparsity (PTS), which reduces model cost by removing…

计算与语言 · 计算机科学 2026-02-26 Minhao Jiang , Zhikai Li , Xuewen Liu , Jing Zhang , Mengjuan Chen , Qingyi Gu

Weight pruning is an effective technique to reduce the model size and inference time for deep neural networks in real-world deployments. However, since magnitudes and relative importance of weights are very different for different layers of…

机器学习 · 计算机科学 2021-05-05 Xiao Zhou , Weizhong Zhang , Hang Xu , Tong Zhang

In this paper, we provide a theoretical understanding of word embedding and its dimensionality. Motivated by the unitary-invariance of word embedding, we propose the Pairwise Inner Product (PIP) loss, a novel metric on the dissimilarity…

机器学习 · 计算机科学 2018-12-12 Zi Yin , Yuanyuan Shen