English
Related papers

Related papers: Oblivious and Semi-Oblivious Boundedness for Exist…

200 papers

We study fundamental reachability problems on pseudo-orbits of linear dynamical systems. Pseudo-orbits can be viewed as a model of computation with limited precision and pseudo-reachability can be thought of as a robust version of classical…

Logic in Computer Science · Computer Science 2022-07-07 Julian D'Costa , Toghrul Karimov , Rupak Majumdar , Joël Ouaknine , Mahmoud Salamati , James Worrell

This paper studies semiparametric contextual bandits, a generalization of the linear stochastic bandit problem where the reward for an action is modeled as a linear function of known action features confounded by an non-linear…

Machine Learning · Statistics 2018-07-17 Akshay Krishnamurthy , Zhiwei Steven Wu , Vasilis Syrgkanis

We study an online classification problem with partial feedback in which individuals arrive one at a time from a fixed but unknown distribution, and must be classified as positive or negative. Our algorithm only observes the true label of…

Machine Learning · Computer Science 2020-04-17 Yahav Bechavod , Katrina Ligett , Aaron Roth , Bo Waggoner , Zhiwei Steven Wu

This note aims at providing a rather informal and hopefully accessible overview of the fairly long and technical work [4]. In that paper, the authors established new global-in-time existence results for admissible solutions of nonlinear…

Analysis of PDEs · Mathematics 2024-05-06 Laura V. Spinolo , Fabio Ancona , Andrea Marson

We study a pursuit-evasion problem which can be viewed as an extension of the keep-away game. In the game, pursuer(s) will attempt to intersect or catch the evader, while the evader can visit a fixed set of locations, which we denote as the…

Robotics · Computer Science 2022-06-17 Weifu Wang , Ping Li

This paper investigates the challenges of optimal online policy learning under missing data. State-of-the-art algorithms implicitly assume that rewards are always observable. I show that when rewards are missing at random, the Upper…

Econometrics · Economics 2025-07-29 Filippo Palomba

This paper mainly studies nonnegativity decision of forms based on variable substitutions. Unlike existing research, the paper regards simplex subdivisions as new perspectives to study variable substitutions, gives some subdivisions of the…

Symbolic Computation · Computer Science 2009-12-23 Xiaorong Hou , Song Xu

We consider online learning problems under a partial observability model capturing situations where the information conveyed to the learner is between full information and bandit feedback. In the simplest variant, we assume that in addition…

Machine Learning · Computer Science 2026-04-28 Tomas Kocak , Gergely Neu , Michal Valko , Remi Munos

Logically constrained term rewriting is a rewriting framework that supports built-in data structures such as integers and bit vectors. Recently, constrained terms play a key role in various analyses and applications of logically constrained…

Logic in Computer Science · Computer Science 2025-12-16 Kanta Takahata , Jonas Schöpf , Naoki Nishida , Takahito Aoto

We investigate reflected random walks in the quarter plane, with particular emphasis on the time spent along the reflection boundary axes. Assuming the drift of the random walk lies within the cone, the local time converges -- without the…

Probability · Mathematics 2025-07-08 Viet Hung Hoang , Kilian Raschel

Robustness of machine learning models to various adversarial and non-adversarial corruptions continues to be of interest. In this paper, we introduce the notion of the boundary thickness of a classifier, and we describe its connection with…

This paper addresses the nonovershooting control problem for strict-feedback nonlinear systems with unknown control direction. We propose a method that integrates extremum seeking with Lie bracket-based design to achieve approximately…

Systems and Control · Electrical Eng. & Systems 2026-01-16 Kaixin Lu , Ziliang Lyu , Yanfang Mo , Yiguang Hong , Haoyong Yu

Tracking of reference signals is addressed in the context of a class of nonlinear controlled systems modelled by $r$-th order functional differential equations, encompassing inter alia systems with unknown "control direction" and dead-zone…

Optimization and Control · Mathematics 2021-01-18 Thomas Berger , Achim Ilchmann , Eugene P Ryan

The maximum surveillance of a target which is holding course is considered, wherein an observer vehicle aims to maximize the time that a faster target remains within a fixed-range of the observer. This entails two coupled phases: an…

Optimization and Control · Mathematics 2022-09-26 Isaac E. Weintraub , Alexander Von Moll , Eloy Garcia , David W. Casbeer , Meir Pachter

This paper introduces a class of objects called decision rules that map infinite sequences of alternatives to a decision space. These objects can be used to model situations where a decision maker encounters alternatives in a sequence such…

Theoretical Economics · Economics 2022-09-12 Bhavook Bhardwaj , Siddharth Chatterjee

In Euclidean space there is a trivial upper bound on the maximum length of a compound "walk" built up of variable-length jumps, and a considerably less trivial lower bound on its minimum length. The existence of this non-trivial lower bound…

Mathematical Physics · Physics 2013-09-19 Petarpa Boonserm , Matt Visser

We study contextual bandits in the presence of a stage-wise constraint when the constraint must be satisfied both with high probability and in expectation. We start with the linear case where both the reward function and the stage-wise…

Machine Learning · Computer Science 2025-08-22 Aldo Pacchiano , Mohammad Ghavamzadeh , Peter Bartlett

We study a pursuit-evasion game between two players with car-like dynamics and sensing limitations by formalizing it as a partially observable stochastic zero-sum game. The partial observability caused by the sensing constraints is…

Robotics · Computer Science 2025-06-17 Burak M. Gonultas , Volkan Isler

An important step in the Markov reward approach to error bounds on stationary performance measures of Markov chains is to bound the bias terms. Affine functions have been successfully used for these bounds for various models, but there are…

Probability · Mathematics 2019-01-04 Xinwei Bai , Jasper Goseling

We consider a class of pursuit-evasion differential games in which the evader has continuous access to the pursuer's location, but not vice-versa. There is a remote sensor (e.g., a radar station) that can sense the evader's location upon a…

Systems and Control · Electrical Eng. & Systems 2023-07-04 Dipankar Maity
‹ Prev 1 8 9 10 Next ›