中文
相关论文

相关论文: Exponentially Weighted l_2 Regularization Strategy…

200 篇论文

The paper deals with generalized functional regression. The aim is to estimate the influence of covariates on observations, drawn from an exponential distribution. The link considered has a semiparametric expression: if we are interested in…

统计理论 · 数学 2013-09-20 Irène Gannaz

Policy gradient methods are a powerful family of reinforcement learning algorithms for continuous control that optimize a policy directly. However, standard first-order methods often converge slowly. Second-order methods can accelerate…

系统与控制 · 电气工程与系统科学 2025-11-05 Amirreza Valaei , Arash Bahari Kordabad , Sadegh Soudjani

Regression analysis is employed to examine and quantify the relationships between input variables and a dependent and continuous output variable. It is widely used for predictive modelling in fields such as finance, healthcare, and…

机器学习 · 计算机科学 2025-10-16 Ashish Bhatia , Renato Cordeiro de Amorim , Vito De Feo

We propose a new algorithmic framework for constrained compressed sensing models that admit nonconvex sparsity-inducing regularizers including the log-penalty function as objectives, and nonconvex loss functions such as the Cauchy loss…

最优化与控制 · 数学 2022-06-17 Shuqin Sun , Ting Kei Pong

Motivated by value function estimation in reinforcement learning, we study statistical linear inverse problems, i.e., problems where the coefficients of a linear system to be solved are observed in noise. We consider penalized estimators,…

机器学习 · 计算机科学 2012-07-03 Bernardo Avila Pires , Csaba Szepesvari

This paper introduces a fuzzy reinforcement learning framework, Enhanced-FQL($\lambda$), that integrates novel Fuzzified Eligibility Traces (FET) and Segmented Experience Replay (SER) into fuzzy Q-learning with the Fuzzified Bellman…

机器学习 · 计算机科学 2026-04-14 Mohsen Jalaeian-Farimani , Xiong Xiong , Luca Bascetta

We propose a novel method to model nonlinear regression problems by adapting the principle of penalization to Partial Least Squares (PLS). Starting with a generalized additive model, we expand the additive component of each variable in…

统计理论 · 数学 2010-08-13 Nicole Kraemer , Anne-Laure Boulesteix , Gerhard Tutz

Robust Principal Component Analysis (RPCA) is a fundamental technique for decomposing data into low-rank and sparse components, which plays a critical role for applications such as image processing and anomaly detection. Traditional RPCA…

机器学习 · 计算机科学 2024-12-20 Kexin Li , You-wei Wen , Xu Xiao , Mingchao Zhao

In the second part of our study we introduce the concept of global extended exactness of penalty and augmented Lagrangian functions, and derive the localization principle in the extended form. The main idea behind the extended exactness…

最优化与控制 · 数学 2018-11-26 M. V. Dolgopolik

Lyapunov functions with exponential weights have been used successfully as a powerful tool for the stability analysis of hyperbolic systems of balance laws. In this paper we extend the class of weight functions to a family of hyperbolic…

最优化与控制 · 数学 2024-10-02 Martin Gugat

In high-dimensional statistics, the Lasso is a cornerstone method for simultaneous variable selection and parameter estimation. However, its reliance on the squared loss function renders it highly sensitive to outliers and heavy-tailed…

机器学习 · 统计学 2025-11-20 The Tien Mai

Emergent Large Language Models (LLMs) use their extraordinary performance and powerful deduction capacity to discern from traditional language models. However, the expenses of computational resources and storage for these LLMs are stunning,…

计算与语言 · 计算机科学 2024-06-25 Yifei Gao , Jie Ou , Lei Wang , Yuting Xiao , Zhiyuan Xiang , Ruiting Dai , Jun Cheng

In this paper, we propose a novel heuristic algorithm for constructing a Type-2 Fuzzy Set of the Linear Linguistic Regression (T2F-LLR) model, designed to address uncertainty and vagueness in real-world decision-making. We consider a…

This paper introduces a flexible regularization approach that reduces point estimation risk of group means stemming from e.g. categorical regressors, (quasi-)experimental data or panel data models. The loss function is penalized by adding…

计量经济学 · 经济学 2019-01-08 Phillip Heiler , Jana Mareckova

In this paper, we use the advantage of large-scale systems modeling based on the type-2 fuzzy Takagi-Sugeno model to cover the uncertainties caused by large-scale systems modeling. The advantage of using membership function information is…

系统与控制 · 电气工程与系统科学 2021-09-01 Mojtaba Asadi Jokar , Iman Zamani , Mohamad Manthouri , Mohammad Sarbaz

The quantization of large language models (LLMs) has been a prominent research area aimed at enabling their lightweight deployment in practice. Existing research about LLM's quantization has mainly explored the interplay between weights and…

计算与语言 · 计算机科学 2025-05-16 Yifei Gao , Jie Ou , Lei Wang , Jun Cheng , Mengchu Zhou

Reinforcement learning (RL) algorithms for real-world robotic applications need a data-efficient learning process and the ability to handle complex, unknown dynamical systems. These requirements are handled well by model-based and…

机器人学 · 计算机科学 2017-06-20 Yevgen Chebotar , Karol Hausman , Marvin Zhang , Gaurav Sukhatme , Stefan Schaal , Sergey Levine

Full-waveform inversion (FWI) is a powerful seismic imaging technique used to estimate high-resolution physical properties of subsurface structures by minimizing the misfit between observed and modeled seismic data. FWI is inherently a…

地球物理 · 物理学 2025-12-19 Kamal Aghazade , Toktam Zand , Ali Gholami

Sparse regularization techniques are well-established in machine learning, yet their application in neural networks remains challenging due to the non-differentiability of penalties like the $L_1$ norm, which is incompatible with stochastic…

机器学习 · 计算机科学 2025-02-10 Chris Kolb , Tobias Weber , Bernd Bischl , David Rügamer

Reinforcement learning (RL) for exponential-utility optimization in discounted Markov decision processes (MDPs) lacks principled value-based algorithms. We address this gap in the fixed risk-aversion setting. Building on the Bellman-type…

机器学习 · 计算机科学 2026-05-11 Gugan Thoppe , L. A. Prashanth , Ankur Naskar , Sanjay Bhat