中文
相关论文

相关论文: Lie Bracket Approximation of Extremum Seeking Syst…

200 篇论文

Estimation of the degree of stability and the bounds of solutions to non-autonomous nonlinear systems present major concerns in numerous applied problems. Yet, current techniques are frequently yield overconservative conditions which are…

动力系统 · 数学 2020-12-29 Mark A. Pinsky

We study the problem of pure exploration in matching markets under uncertain preferences, where the goal is to identify a stable matching with confidence parameter $\delta$ and minimal sample complexity. Agents learn preferences via…

计算机科学与博弈论 · 计算机科学 2025-09-19 Tejas Pagare , Agniv Bandyopadhyay , Sandeep Juneja

Mappings to structured output spaces (strings, trees, partitions, etc.) are typically learned using extensions of classification algorithms to simple graphical structures (eg., linear chains) in which search and parameter estimation can be…

机器学习 · 计算机科学 2009-07-07 Hal Daumé , Daniel Marcu

In this article we study control problems with systems that are governed by ordinary differential equations whose vector fields depend linearly in the time derivatives of some components of the control. The remaining components are…

最优化与控制 · 数学 2012-10-17 Maria Soledad Aronna , Franco Rampazzo

Determining the reachable set for a given nonlinear control system is crucial for system control and planning. However, computing such a set is impossible if the system's dynamics are not fully known. This paper is motivated by a scenario…

最优化与控制 · 数学 2021-08-26 Taha Shafa , Melkior Ornik

The conventional definition of extremality of a finite collection of sets is extended by replacing a fixed point (extremal point) in the intersection of the sets by a collection of sequences of points in the individual sets with the…

最优化与控制 · 数学 2025-07-22 Nguyen Duy Cuong , Alexander Y. Kruger

We present evidence of substantial benefit from efficient exploration in gathering human feedback to improve large language models. In our experiments, an agent sequentially generates queries while fitting a reward model to the feedback…

机器学习 · 计算机科学 2024-06-06 Vikranth Dwaracherla , Seyed Mohammad Asghari , Botao Hao , Benjamin Van Roy

Maximum likelihood constraint inference is a powerful technique for identifying unmodeled constraints that affect the behavior of a demonstrator acting under a known objective function. However, it was originally formulated only for…

机器人学 · 计算机科学 2021-09-13 Kaylene C. Stocking , David L. McPherson , Robert P. Matthew , Claire J. Tomlin

Problem of damping of an arbitrary number of linear oscillators under common bounded control is considered. We are looking for a feedback control steering the system to the equilibrium. The obtained control is asymptotically optimal: the…

最优化与控制 · 数学 2016-12-02 Alexander Ovseevich , Aleksey Fedorov

This paper is concerned with the robust tracking control of linear uncertain systems, whose unknown system parameters and disturbances are bounded within ellipsoidal sets. We propose an adaptive robust control that can actively learn the…

系统与控制 · 电气工程与系统科学 2023-08-08 Xuehui Ma , Shiliang Zhang , Yushuai Li , Fucai Qian , Tingwen Huang

Iterative trajectory optimization techniques for non-linear dynamical systems are among the most powerful and sample-efficient methods of model-based reinforcement learning and approximate optimal control. By leveraging time-variant local…

系统与控制 · 电气工程与系统科学 2019-08-01 Onur Celik , Hany Abdulsamad , Jan Peters

This paper introduces a new approach for continual planning and model learning in relational, non-stationary stochastic environments. Such capabilities are essential for the deployment of sequential decision-making systems in the uncertain…

人工智能 · 计算机科学 2024-07-24 Rushang Karia , Pulkit Verma , Alberto Speranzon , Siddharth Srivastava

We consider the problem of pure exploration with subset-wise preference feedback, which contains $N$ arms with features. The learner is allowed to query subsets of size $K$ and receives feedback in the form of a noisy winner. The goal of…

机器学习 · 计算机科学 2021-04-13 Shubham Gupta , Aadirupa Saha , Sumeet Katariya

We introduce the extremal range, a local statistic for studying the spatial extent of extreme events in random fields on $\mathbb{R}^d$. Conditioned on exceedance of a high threshold at a location $s$, the extremal range at $s$ is the…

统计理论 · 数学 2024-11-06 Ryan Cotsakis , Elena Di Bernardino , Thomas Opitz

Large Language Models (LLMs) have become integral components in various autonomous agent systems. In this study, we present an exploration-based trajectory optimization approach, referred to as ETO. This learning method is designed to…

计算与语言 · 计算机科学 2024-07-11 Yifan Song , Da Yin , Xiang Yue , Jie Huang , Sujian Li , Bill Yuchen Lin

Presentation bias is one of the key challenges when learning from implicit feedback in search engines, as it confounds the relevance signal. While it was recently shown how counterfactual learning-to-rank (LTR) approaches…

信息检索 · 计算机科学 2018-12-14 Aman Agarwal , Ivan Zaitsev , Xuanhui Wang , Cheng Li , Marc Najork , Thorsten Joachims

Extremum seeking control (ESC) often employs perturbation-based estimates of derivatives for some sensor field or cost function. These estimates are generally obtained by simply multiplying the output of a single-unit sensor by some…

系统与控制 · 电气工程与系统科学 2025-09-23 Dylan James-Kavanaugh , Patrick McNamee , Qixu Wang , Zahra Nili Ahmadabadi

Exploration in unknown environments is a fundamental problem in reinforcement learning and control. In this work, we study task-guided exploration and determine what precisely an agent must learn about their environment in order to complete…

机器学习 · 计算机科学 2021-07-13 Andrew Wagenmaker , Max Simchowitz , Kevin Jamieson

A learning approach for optimal feedback gains for nonlinear continuous time control systems is proposed and analysed. The goal is to establish a rigorous framework for computing approximating optimal feedback gains using neural networks.…

最优化与控制 · 数学 2020-08-27 Karl Kunisch , Daniel Walter

Accurate reconstruction of the environment is a central goal of Simultaneous Localization and Mapping (SLAM) systems. However, the agent's trajectory can significantly affect estimation accuracy. This paper presents a new method to model…

机器人学 · 计算机科学 2025-06-24 Sebastian Sansoni , Javier Gimenez , Gastón Castro , Santiago Tosetti , Flavio Craparo
‹ 上一页 1 8 9 10 下一页 ›