中文
相关论文

相关论文: Structured Policy Representation: Imposing Stabili…

200 篇论文

Forecasting physical signals in long time range is among the most challenging tasks in Partial Differential Equations (PDEs) research. To circumvent limitations of traditional solvers, many different Deep Learning methods have been…

机器学习 · 计算机科学 2023-06-09 Leon Migus , Julien Salomon , Patrick Gallinari

Robotic manipulation behavior should be robust to disturbances that violate high-level task-structure. Such robustness can be achieved by constantly monitoring the environment to observe the discrete high-level state of the task. This is…

机器人学 · 计算机科学 2022-05-10 Manuel Baum , Oliver Brock

We study systems of interacting reinforced stochastic processes, where agents' decisions evolve under reinforcement, network-mediated interactions, and environmental influences. In competitive environments with irreducible networks, we…

概率论 · 数学 2025-09-18 Michele Aleandri , Paolo Dai Pra , Ida Germana Minelli

Understanding and interacting with everyday physical scenes requires rich knowledge about the structure of the world, represented either implicitly in a value or policy function, or explicitly in a transition model. Here we introduce a new…

When neural networks are used to model dynamics, properties such as stability of the dynamics are generally not guaranteed. In contrast, there is a recent method for learning the dynamics of autonomous systems that guarantees global…

机器学习 · 计算机科学 2022-03-21 Kenji Kashima , Ryota Yoshiuchi , Yu Kawano

Causal representation learning promises to extend causal models to hidden causal variables from raw entangled measurements. However, most progress has focused on proving identifiability results in different settings, and we are not aware of…

机器学习 · 计算机科学 2025-02-04 Dingling Yao , Caroline Muller , Francesco Locatello

Practitioners often rely on compute-intensive domain randomization to ensure reinforcement learning policies trained in simulation can robustly transfer to the real world. Due to unmodeled nonlinearities in the real system, however, even…

机器学习 · 计算机科学 2020-02-27 Gabriel I. Fernandez , Colin Togashi , Dennis W. Hong , Lin F. Yang

We consider kinetic systems and prove their stability working in weighted spaces in which the systems are symmetric. We prove stability for various explicit and implicit semi-discrete and fully discrete schemes. The applications include…

数值分析 · 数学 2017-08-07 F. Patricia Medina , Malgorzata Peszynska

We investigate the important problem of certifying stability of reinforcement learning policies when interconnected with nonlinear dynamical systems. We show that by regulating the input-output gradients of policies, strong guarantees of…

系统与控制 · 计算机科学 2018-10-30 Ming Jin , Javad Lavaei

Deep networks are commonly used to model dynamical systems, predicting how the state of a system will evolve over time (either autonomously or in response to control inputs). Despite the predictive power of these systems, it has been…

机器学习 · 计算机科学 2020-01-20 Gaurav Manek , J. Zico Kolter

In the theory of dynamic programming, an optimal policy is a policy whose lifetime value dominates that of all other policies from every possible initial condition in the state space. This raises a natural question: when does optimality…

最优化与控制 · 数学 2025-05-13 John Stachurski , Jingni Yang , Ziyue Yang

Learning robust and generalizable world models is crucial for enabling efficient and scalable robotic control in real-world environments. In this work, we introduce a novel framework for learning world models that accurately capture…

机器人学 · 计算机科学 2025-12-16 Chenhao Li , Andreas Krause , Marco Hutter

Notwithstanding the usefulness of system dynamics in analyzing complex policy problems, policy design is far from straightforward and in many instances trial-and-error driven. To address this challenge, we propose to combine system dynamics…

物理与社会 · 物理学 2017-11-15 Lukas Schoenenberger , Radu Tanase

Deep structured models are widely used for tasks like semantic segmentation, where explicit correlations between variables provide important prior information which generally helps to reduce the data needs of deep nets. However, current…

机器学习 · 计算机科学 2018-11-02 Colin Graber , Ofer Meshi , Alexander Schwing

Sociability is essential for modern robots to increase their acceptability in human environments. Traditional techniques use manually engineered utility functions inspired by observing pedestrian behaviors to achieve social navigation.…

机器人学 · 计算机科学 2023-04-26 Yigit Yildirim , Emre Ugur

The advancement of robots, particularly those functioning in complex human-centric environments, relies on control solutions that are driven by machine learning. Understanding how learning-based controllers make decisions is crucial since…

机器学习 · 计算机科学 2023-11-14 Tsun-Hsuan Wang , Wei Xiao , Tim Seyde , Ramin Hasani , Daniela Rus

We present CREST, an approach for causal reasoning in simulation to learn the relevant state space for a robot manipulation policy. Our approach conducts interventions using internal models, which are simulations with approximate dynamics…

机器人学 · 计算机科学 2022-03-15 Tabitha Edith Lee , Jialiang Zhao , Amrita S. Sawhney , Siddharth Girdhar , Oliver Kroemer

We present a structured neural network architecture that is inspired by linear time-varying dynamical systems. The network is designed to mimic the properties of linear dynamical systems which makes analysis and control simple. The…

机器人学 · 计算机科学 2018-08-06 Alexander Broad , Ian Abraham , Todd Murphey , Brenna Argall

There has been a long-standing and at times fractious debate whether complex and large systems can be stable. In ecology, the so-called `diversity-stability debate' arose because mathematical analyses of ecosystem stability were either…

动力系统 · 数学 2015-09-02 Paul Kirk , Delphine M. Y. Rolando , Adam L. MacLean , Michael P. H. Stumpf

In bipartite matching problems, agents on two sides of a graph want to be paired according to their preferences. The stability of a matching depends on these preferences, which in uncertain environments also reflect agents' beliefs about…

计算机科学与博弈论 · 计算机科学 2025-11-10 Jonathan Shaki , Jiarui Gan , Sarit Kraus