中文
相关论文

相关论文: Initialization-driven neural generation and traini…

200 篇论文

Mixed optimal stopping and stochastic control problems define variational inequalities with non-linear Hamilton-Jacobi-Bellman (HJB) operators, whose numerical solution is notoriously difficult and lack of reliable benchmarks. We first use…

最优化与控制 · 数学 2025-05-27 Yun Zhao , Harry Zheng

Compressing large Neural Networks (NN) by quantizing the parameters, while maintaining the performance is highly desirable due to reduced memory and time complexity. In this work, we cast NN quantization as a discrete labelling problem, and…

计算机视觉与模式识别 · 计算机科学 2019-08-21 Thalaiyasingam Ajanthan , Puneet K. Dokania , Richard Hartley , Philip H. S. Torr

This work considers the problem of approximating initial condition and time-dependent optimal control and trajectory surfaces using multivariable Fourier series. A modified Augmented Lagrangian algorithm for translating the optimal control…

最优化与控制 · 数学 2023-12-14 Gabriel Nicolosi , Terry Friesz , Christopher Griffin

A promising approach to optimal control of nonlinear systems involves iteratively linearizing the system and solving an optimization problem at each time instant to determine the optimal control input. Since this approach relies on online…

最优化与控制 · 数学 2025-01-30 Anran Li , John P. Swensen , Mehdi Hosseinzadeh

We propose a game-based formulation for learning dimensionality-reducing representations of feature vectors, when only a prior knowledge on future prediction tasks is available. In this game, the first player chooses a representation, and…

机器学习 · 计算机科学 2024-03-12 Neria Uzan , Nir Weinberger

We consider a deterministic mean field games problem in which a typical agent solves an optimal control problem where the dynamics is affine with respect to the control and the cost functional has a growth which is polynomial with respect…

最优化与控制 · 数学 2023-05-03 Justina Gianatti , Francisco J. Silva , Ahmad Zorkot

We address finding the semi-global solutions to optimal feedback control and the Hamilton--Jacobi--Bellman (HJB) equation. Using the solution of an HJB equation, a feedback optimal control law can be implemented in real-time with minimum…

最优化与控制 · 数学 2016-06-17 Wei Kang , Lucas C. Wilcox

A Hamiltonian algorithm, both theoretical and numerical, to obtain the reduced equations implementing Pontryagine's Maximum Principle for singular linear-quadratic optimal control problems is presented. This algorithm is inspired on the…

最优化与控制 · 数学 2012-04-13 M. Delgado-Tellez , A. Ibort

The topics treated in this thesis are inherently two-fold. The first part considers the problem of a market maker optimally setting bid/ask quotes over a finite time horizon, to maximize her expected utility. The intensities of the orders…

最优化与控制 · 数学 2020-09-15 Diego Zabaljauregui

Optimal planning with respect to learned neural network (NN) models in continuous action and state spaces using mixed-integer linear programming (MILP) is a challenging task for branch-and-bound solvers due to the poor linear relaxation of…

人工智能 · 计算机科学 2019-07-29 Buser Say , Scott Sanner , Sylvie Thiébaux

We study the optimal investment-consumption problem for a member of defined contribution plan during the decumulation phase. For a fixed annuitization time, to achieve higher final annuity, we consider a variable consumption rate. Moreover,…

投资组合管理 · 定量金融 2020-08-18 Hassan Dadashi

This paper investigates the so-called reward-balancing methods, a novel class of algorithms for solving discounted-return reinforcement learning (RL) problems. These methods consist of iteratively adjusting the reward function to transform…

最优化与控制 · 数学 2026-04-23 Simone Baroncini , Bahman Gharesifard , Giuseppe Notarstefano

We study a particle approximation for one-dimensional first-order Mean-Field-Games (MFGs) with local interactions with planning conditions. Our problem comprises a system of a Hamilton-Jacobi equation coupled with a transport equation. As…

最优化与控制 · 数学 2021-09-07 Marco Di Francesco , Serikbolsyn Duisembay , Diogo Aguiar Gomes , Ricardo Ribeiro

In this paper, we mainly focus on solving high-dimensional stochastic Hamiltonian systems with boundary condition, which is essentially a Forward Backward Stochastic Differential Equation (FBSDE in short), and propose a novel method from…

最优化与控制 · 数学 2021-12-13 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

We provide a data-driven framework for optimal control of a continuous-time stochastic dynamical system. The proposed framework relies on the linear operator theory involving linear Perron-Frobenius (P-F) and Koopman operators. Our first…

最优化与控制 · 数学 2022-02-04 Umesh Vaidya , Duvan Tellez-Castro

We study the problem of learning optimal policies in finite-horizon Markov Decision Processes (MDPs) using low-rank reinforcement learning (RL) methods. In finite-horizon MDPs, the policies, and therefore the value functions (VFs) are not…

机器学习 · 计算机科学 2026-05-14 Sergio Rozada , Jose Luis Orejuela , Antonio G. Marques

Mean-field games (MFGs) have shown strong modeling capabilities for large systems in various fields, driving growth in computational methods for mean-field game problems. However, high order methods have not been thoroughly investigated. In…

数值分析 · 数学 2023-08-16 Guosheng Fu , Siting Liu , Stanley Osher , Wuchen Li

This paper studies an optimal stochastic impulse control problem in a finite horizon with a decision lag, by which we mean that after an impulse is made, a fixed number units of time has to be elapsed before the next impulse is allowed to…

最优化与控制 · 数学 2021-02-09 Chang Li , Jiongmin Yong

This study investigates a stochastic production planning problem with a running cost composed of quadratic production costs and inventory-dependent costs. The objective is to minimize the expected cost until production stops when inventory…

最优化与控制 · 数学 2025-05-20 Dragos-Patru Covei

In this paper, we consider a class of continuous-time, continuous-space stochastic optimal control problems. Building upon recent advances in Markov chain approximation methods and sampling-based algorithms for deterministic path planning,…

机器人学 · 计算机科学 2012-02-27 Vu Anh Huynh , Sertac Karaman , Emilio Frazzoli
‹ 上一页 1 8 9 10 下一页 ›