English
Related papers

Related papers: Multistep Inverse Is Not All You Need

200 papers

We address the problem of agile 1v1 quadrotor pursuit-evasion, where a pursuer and an evader learn to outmaneuver each other through reinforcement learning (RL). Such settings face two major challenges: non-stationarity, since each agent's…

Robotics · Computer Science 2025-09-16 Alejandro Sanchez Roncero , Yixi Cai , Olov Andersson , Petter Ogren

In reinforcement learning (RL), world models serve as internal simulators, enabling agents to predict environment dynamics and future outcomes in order to make informed decisions. While previous approaches leveraging discrete latent spaces,…

Machine Learning · Computer Science 2025-03-04 Aidan Scannell , Mohammadreza Nakhaei , Kalle Kujanpää , Yi Zhao , Kevin Sebastian Luck , Arno Solin , Joni Pajarinen

This paper proposes a hard-constrained unsupervised learning framework for rapidly solving the non-linear and non-convex AC optimal power flow (AC-OPF) problem in real-time operation. Without requiring ground-truth AC-OPF solutions,…

Systems and Control · Electrical Eng. & Systems 2026-02-09 Kejun Chen , Bernard Knueven , Wesley Jones

Recent advances in learning aligned multimodal representations have been primarily driven by training large neural networks on massive, noisy paired-modality datasets. In this work, we ask whether it is possible to achieve similar results…

Machine Learning · Computer Science 2022-10-11 Elan Rosenfeld , Preetum Nakkiran , Hadi Pouransari , Oncel Tuzel , Fartash Faghri

This paper targets the problem of image set-based face verification and identification. Unlike traditional single media (an image or video) setting, we encounter a set of heterogeneous contents containing orderless images and videos. The…

Computer Vision and Pattern Recognition · Computer Science 2019-08-06 Xiaofeng Liu , B. V. K Vijaya Kumar , Chao Yang , Qingming Tang , Jane You

In this paper we present a hierarchical multi-rate control architecture for nonlinear autonomous systems operating in partially observable environments. Control objectives are expressed using syntactically co-safe Linear Temporal Logic…

Systems and Control · Electrical Eng. & Systems 2022-07-04 Ugo Rosolia , Andrew Singletary , Aaron D. Ames

We consider least squares semidefinite programming (LSSDP) where the primal matrix variable must satisfy given linear equality and inequality constraints, and must also lie in the intersection of the cone of symmetric positive semidefinite…

Optimization and Control · Mathematics 2015-05-26 Defeng Sun , Kim-Chuan Toh , Liuqin Yang

Non-stationary domains, that change in unpredicted ways, are a challenge for agents searching for optimal policies in sequential decision-making problems. This paper presents a combination of Markov Decision Processes (MDP) with Answer Set…

Artificial Intelligence · Computer Science 2017-06-06 Leonardo A. Ferreira , Reinaldo A. C. Bianchi , Paulo E. Santos , Ramon Lopez de Mantaras

This paper designs traffic signal control policies for a network of signalized intersections without knowing the demand and parameters. Within a model predictive control (MPC) framework, control policies consist of an algorithm that…

Systems and Control · Electrical Eng. & Systems 2025-03-17 Zhexian Li , Ketan Savla

The goal of imitation learning is for an apprentice to learn how to behave in a stochastic environment by observing a mentor demonstrating the correct behavior. Accurate prior knowledge about the correct behavior can reduce the need for…

Machine Learning · Computer Science 2012-06-26 Umar Syed , Robert E. Schapire

Reactive (memoryless) policies are sufficient in completely observable Markov decision processes (MDPs), but some kind of memory is usually necessary for optimal control of a partially observable MDP. Policies with finite memory can be…

Artificial Intelligence · Computer Science 2013-01-30 Nicolas Meuleau , Leonid Peshkin , Kee-Eung Kim , Leslie Pack Kaelbling

This work presents a novel posterior inference method for models with intractable evidence and likelihood functions. Error-guided likelihood-free MCMC, or EG-LF-MCMC in short, has been developed for scientific applications, where a…

Machine Learning · Statistics 2021-04-27 Volodimir Begy , Erich Schikuta

Quadrotors are increasingly used in the evolving field of aerial robotics for their agility and mechanical simplicity. However, inherent uncertainties, such as aerodynamic effects coupled with quadrotors' operation in dynamically changing…

Robotics · Computer Science 2024-05-24 Yuliang Gu , Sheng Cheng , Naira Hovakimyan

Synthesizing controllable motion for a character using deep learning has been a promising approach due to its potential to learn a compact model without laborious feature engineering. To produce dynamic motion from weak control signals such…

Computer Vision and Pattern Recognition · Computer Science 2023-03-06 Lintao Wang , Kun Hu , Lei Bai , Yu Ding , Wanli Ouyang , Zhiyong Wang

Reduced-order models are indispensable for multi-query or real-time problems. However, there are still many challenges to constructing efficient ROMs for time-dependent parametrized problems. Using a linear reduced space is inefficient for…

Numerical Analysis · Mathematics 2023-11-17 Junming Duan , Jan S. Hesthaven

We investigate the possibility of forcing a self-supervised model trained using a contrastive predictive loss to extract slowly varying latent representations. Rather than producing individual predictions for each of the future…

The iterative selection of examples for labeling in active machine learning is conceptually similar to feedback channel coding in information theory: in both tasks, the objective is to seek a minimal sequence of actions to encode…

Machine Learning · Statistics 2021-03-02 Gregory Canal , Matthieu Bloch , Christopher Rozell

Multistep traffic forecasting on road networks is a crucial task in successful intelligent transportation system applications. To capture the complex non-stationary temporal dynamics and spatial dependency in multistep traffic-condition…

Machine Learning · Computer Science 2018-10-30 Zhengchao Zhang , Meng Li , Xi Lin , Yinhai Wang , Fang He

Detecting regime shifts in chaotic time series is hard because observation-space signals are entangled with intrinsic variability. We propose Parameter--Space Changepoint Detection (Param--CPD), a two--stage framework that first amortizes…

Machine Learning · Computer Science 2025-12-09 Xiangbo Deng , Cheng Chen , Peng Yang

Change detection is of fundamental importance when analyzing data streams. Detecting changes both quickly and accurately enables monitoring and prediction systems to react, e.g., by issuing an alarm or by updating a learning algorithm.…

Machine Learning · Computer Science 2024-01-17 Marco Heyden , Edouard Fouché , Vadim Arzamasov , Tanja Fenn , Florian Kalinke , Klemens Böhm