中文
相关论文

相关论文: Minimal Polynomial and Reduced Rank Extrapolation …

200 篇论文

Mixture-of-Experts (MoE) architectures have emerged as a promising approach to scale Large Language Models (LLMs). MoE boosts the efficiency by activating a subset of experts per token. Recent works show that fine-grained experts…

This paper presents a comprehensive overview of several multidimensional reduction methods focusing on Multidimensional Principal Component Analysis (MPCA), Multilinear Orthogonal Neighborhood Preserving Projection (MONPP), Multidimensional…

数值分析 · 数学 2026-01-05 Mohamed El Guide , Alaa El Ichi , Khalide Jbilou , Lothar Reichel , Hessah Alqahtani

Matrix completion and extrapolation (MCEX) are dealt with here over reproducing kernel Hilbert spaces (RKHSs) in order to account for prior information present in the available data. Aiming at a faster and low-complexity solver, the task is…

机器学习 · 统计学 2019-10-02 Pere Giménez-Febrer , Alba Pagès-Zamora , Georgios B. Giannakis

Wisely utilizing the internal and external learning methods is a new challenge in super-resolution problem. To address this issue, we analyze the attributes of two methodologies and find two observations of their recovered details: 1) they…

计算机视觉与模式识别 · 计算机科学 2023-02-20 Shuang Wang , Bo Yue , Xuefeng Liang , Peiyuan Ji , Licheng Jiao

Given an input matrix polynomial whose coefficients are floating point numbers, we consider the problem of finding the nearest matrix polynomial which has rank at most a specified value. This generalizes the problem of finding a nearest…

符号计算 · 计算机科学 2017-12-13 Mark Giesbrecht , Joseph Haraldson , George Labahn

The Minimum Eccentricity Shortest Path (MESP) Problem consists in determining a shortest path (a path whose length is the distance between its extremities) of minimum eccentricity in a graph. It was introduced by Dragan and Leitert [9] who…

计算复杂性 · 计算机科学 2016-09-16 Etienne Birmelé , Fabien De Montgolfier , Léo Planche

This paper presents two new greedy sensor placement algorithms, named minimum nonzero eigenvalue pursuit (MNEP) and maximal projection on minimum eigenspace (MPME), for linear inverse problems, with greater emphasis on the MPME algorithm…

数据结构与算法 · 计算机科学 2016-11-08 Chaoyang Jiang , Yeng Chai Soh , Hua Li

Inverse problems are pervasive mathematical methods in inferring knowledge from observational and experimental data by leveraging simulations and models. Unlike direct inference methods, inverse problem approaches typically require many…

计算物理 · 物理学 2019-12-20 Sheroze Sheriffdeen , Jean C. Ragusa , Jim E. Morel , Marvin L. Adams , Tan Bui-Thanh

Reed-Muller codes encode an $m$-variate polynomial of degree $r$ by evaluating it on all points in $\{0,1\}^m$. We denote this code by $RM(m,r)$. The minimal distance of $RM(m,r)$ is $2^{m-r}$ and so it cannot correct more than half that…

信息论 · 计算机科学 2015-08-28 Ramprasad Saptharishi , Amir Shpilka , Ben Lee Volk

We present MoE-MLA-RoPE, a novel architecture combination that combines Mixture of Experts (MoE) with Multi-head Latent Attention (MLA) and Rotary Position Embeddings (RoPE) for efficient language modeling. Our approach addresses the…

人工智能 · 计算机科学 2025-08-05 Sushant Mehta , Raj Dandekar , Rajat Dandekar , Sreedath Panat

In this paper, we consider approximating the parameter-to-solution maps of parametric partial differential equations (PPDEs) using deep neural networks (DNNs). We propose an efficient approach combining reduced collocation methods (RCMs)…

数值分析 · 数学 2025-08-18 Guanhang Lei , Zhen Lei , Lei Shi , Chenyu Zeng

Vector extrapolation methods are widely used in large-scale simulation studies, and numerous extrapolation-based acceleration techniques have been developed to enhance the convergence of linear and nonlinear fixed-point iterative methods.…

数值分析 · 数学 2026-02-03 Abdellatif Mouhssine

We introduce the telescopic relative entropy (TRE), which is a new regularisation of the relative entropy related to smoothing, to overcome the problem that the relative entropy between pure states is either zero or infinity and therefore…

数学物理 · 物理学 2011-04-28 Koenraad M. R. Audenaert

Parameter-dependent models arise in many contexts such as uncertainty quantification, sensitivity analysis, inverse problems or optimization. Parametric or uncertainty analyses usually require the evaluation of an output of a model for many…

数值分析 · 数学 2018-10-22 Anthony Nouy

In [1] is proposed a simplified DeC method, that, when combined with the residual distribution (RD) framework, allows to construct a high order, explicit FE scheme with continuous approximation avoiding the inversion of the mass matrix for…

数值分析 · 数学 2022-11-17 Rémi Abgrall , Elise Le Mélédo , Philipp Öffner , Davide Torlo

In this paper, we consider a class of nonsmooth nonconvex optimization problems whose objective is the sum of a block relative smooth function and a proper and lower semicontinuous block separable function. Although the analysis of block…

最优化与控制 · 数学 2022-04-27 Le Thi Khanh Hien , Duy Nhat Phan , Nicolas Gillis , Masoud Ahookhosh , Panagiotis Patrinos

Low Rank Approximation is among most fundamental subjects of numerical linear algebra having important applications to various areas of modern computing and %they range from machine learning theory and %neural networks to data mining and…

数值分析 · 数学 2018-09-25 Victor Y. Pan , Qi Luan , John Svadlenka , Liang Zhao

Reinforcement Learning with Verifiable Rewards (RLVR) elicits long chain-of-thought reasoning in large language models (LLMs), but outcome-based rewards lead to coarse-grained advantage estimation. While existing approaches improve RLVR via…

计算与语言 · 计算机科学 2026-01-08 Fei Wu , Zhenrong Zhang , Qikai Chang , Jianshu Zhang , Quan Liu , Jun Du

Regularization of neural machine translation is still a significant problem, especially in low-resource settings. To mollify this problem, we propose regressing word embeddings (ReWE) as a new regularization technique in a system that is…

计算与语言 · 计算机科学 2019-04-05 Inigo Jauregi Unanue , Ehsan Zare Borzeshi , Nazanin Esmaili , Massimo Piccardi

Multi-relational learning has received lots of attention from researchers in various research communities. Most existing methods either suffer from superlinear per-iteration cost, or are sensitive to the given ranks. To address both issues,…

机器学习 · 计算机科学 2016-01-19 Fanhua Shang , James Cheng , Hong Cheng