中文
相关论文

相关论文: Design and Implementation of Hedge Algebra Control…

200 篇论文

Recent studies observe that reinforcement learning with verifiable rewards (RLVR) reliably improves pass@1 on reasoning tasks, yet often fails to yield comparable gains in pass@k, raising the question of whether RLVR genuinely enables large…

机器学习 · 计算机科学 2026-05-20 Chanuk Lee , Minki Kang , Sung Ju Hwang

Semantic communication (SemCom) aims to achieve high fidelity information delivery under low communication consumption by only guaranteeing semantic accuracy. Nevertheless, semantic communication still suffers from unexpected channel…

系统与控制 · 电气工程与系统科学 2024-03-26 Fei Ni , Rongpeng Li , Zhifeng Zhao , Honggang Zhang

Research in quantitative finance has demonstrated that reinforcement learning (RL) methods have delivered promising outcomes in the context of hedging financial portfolios. For example, hedging a portfolio of European options using RL…

计算工程、金融与科学 · 计算机科学 2024-07-16 Anil Sharma , Freeman Chen , Jaesun Noh , Julio DeJesus , Mario Schlener

Cross-modal retrieval aims to search for instances, which are semantically related to the query through the interaction of different modal data. Traditional solutions utilize a single-tower or dual-tower framework to explicitly compute the…

This article introduces a novel framework for data-driven linear quadratic regulator (LQR) design. First, we introduce a reinforcement learning paradigm for on-policy data-driven LQR, where exploration and exploitation are simultaneously…

系统与控制 · 电气工程与系统科学 2024-02-23 Marco Borghesi , Alessandro Bosso , Giuseppe Notarstefano

The hardware-friendly implementation of transcendental functions remains a longstanding challenge in design automation. These functions, which cannot be expressed as finite combinations of algebraic operations, pose significant complexity…

新兴技术 · 计算机科学 2026-01-13 Mehran Moghadam , Sercan Aygun , M. Hassan Najafi

We propose and analyze a randomization scheme for a general class of impulse control problems. The solution to this randomized problem is characterized as the fixed point of a compound operator which consists of a regularized nonlocal…

最优化与控制 · 数学 2026-05-26 Haoyang Cao , Yuchao Dong , Zhouhao Yang

Recent advances in reinforcement learning from human feedback (RLHF) and preference optimization have substantially improved the usability, coherence, and safety of large language models. However, recurring behaviors such as performative…

人工智能 · 计算机科学 2026-05-13 William Parris

We introduce a novel extension to robust control theory that explicitly addresses uncertainty in the value function's gradient, a form of uncertainty endemic to applications like reinforcement learning where value functions are…

机器学习 · 计算机科学 2025-07-22 Qian Qi

In this paper, the Model Predictive Control (MPC) and Moving Horizon Estimator (MHE) strategies using a data-driven approach to learn a Takagi-Sugeno (TS) representation of the vehicle dynamics are proposed to solve autonomous driving…

系统与控制 · 电气工程与系统科学 2020-04-30 Eugenio Alcalá , Olivier Sename , Vicenç Puig , Joseba Quevedo

The central notion of this work is that of a functor between categories of finitely presented modules over so-called computable rings, i.e. rings R where one can algorithmically solve inhomogeneous linear equations with coefficients in R.…

交换代数 · 数学 2016-12-06 Mohamed Barakat , Daniel Robertz

We present a heuristic policy and performance bound for risk-sensitive convex stochastic control that generalizes linear-exponential-quadratic regulator (LEQR) theory. Our heuristic policy extends standard, risk-neutral model predictive…

最优化与控制 · 数学 2022-05-30 Nicholas Moehle

Coherent control, aka quantum control, is a central concept in quantum computing that is attracting increasing attention from both the quantum foundations and quantum software communities. Defining coherent control in the presence of…

计算机科学中的逻辑 · 计算机科学 2026-03-02 Kathleen Barsse , Romain Péchoux , Simon Perdrix

A computationally efficient reformulation of the rigid tube model predictive control is developed. A unique feature of the derived formulation is the utilization of the implicit set representations. This novel formulation does not require…

最优化与控制 · 数学 2023-08-24 Saša V. Raković

The concept of controlling non-linear systems is one the significant fields in scientific researches for the purpose of which intelligent approaches can provide desirable applicability. Fuzzy systems are systems with ambiguous definition…

系统与控制 · 计算机科学 2013-08-23 Qasem Abdollah Nezhad , Javad Palizvan Zand , Samira Shah Hoseini

Inductive cold-start recommendation remains the "Achilles' Heel" of industrial academic platforms, where thousands of new scholars join daily without historical interaction records. While recent Generative Graph Models (e.g., HiGPT, OFA)…

信息检索 · 计算机科学 2026-01-07 Zhexiang Li

We present a model-based globally convergent policy gradient method (PGM) for linear quadratic Gaussian (LQG) control. Firstly, we establish equivalence between optimizing dynamic output feedback controllers and designing a static feedback…

最优化与控制 · 数学 2024-02-27 Tomonori Sadamoto , Fumiya Nakamata

Controllable generative models have been widely used to improve the realism of synthetic visual content. However, such models must handle control conditions and content generation computational requirements, resulting in generally low…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Lin Liu , Huixia Ben , Shuo Wang , Jinda Lu , Junxiang Qiu , Shengeng Tang , Yanbin Hao

Quantum control optimization algorithms are routinely used to generate optimal quantum gates or efficient quantum state transfers. However, there are two main challenges in designing efficient optimization algorithms, namely overcoming the…

量子物理 · 物理学 2022-02-02 Priya Batra , M. Harshanth Ram , T. S. Mahesh

Human cognition excels at symbolic reasoning, deducing abstract rules from limited samples. This has been explained using symbolic and connectionist approaches, inspiring the development of a neuro-symbolic architecture that combines both…

人工智能 · 计算机科学 2024-05-24 Mohamed Mejri , Chandramouli Amarnath , Abhijit Chatterjee