English
Related papers

Related papers: On the equivalence of internal and external habit …

200 papers

Finding the optimal policy for multi-period perishable inventory systems requires solving computationally-expensive stochastic dynamic programs (DP). To avoid the difficulty of solving DP models, we propose a framework that uses an…

Optimization and Control · Mathematics 2021-07-28 Katsunobu Sasanuma , Mohammad Delasay , Christine Pitocco , Alan Scheller-Wolf , Thomas Sexton

We address the problem of approximate model minimization for MDPs in which the state is partitioned into endogenous and (much larger) exogenous components. An exogenous state variable is one whose dynamics are independent of the agent's…

Machine Learning · Computer Science 2019-10-01 Rohan Chitnis , Tomás Lozano-Pérez

The regression discontinuity (RD) design is widely used for program evaluation with observational data. The primary focus of the existing literature has been the estimation of the local average treatment effect at the existing treatment…

Methodology · Statistics 2024-09-05 Yi Zhang , Eli Ben-Michael , Kosuke Imai

Consider first a memoryless population model described by the usual branching process with a given mean reproduction matrix on a finite space of types. Motivated by the consequences of atavism in Evolutionary Biology, we are interested in a…

Probability · Mathematics 2025-01-03 Jean Bertoin

Recent works in end-to-end control for autonomous driving have investigated the use of vision-based exteroceptive perception. Inspired by such results, we propose a new end-to-end memory-based neural architecture for robot steering and…

Robotics · Computer Science 2022-05-25 Sergio Paniego Blanco , Sakshay Mahna , Utkarsh A. Mishra , JoseMaria Canas

The most commonly developed inventory models are the classical economic order quantity model, is governed by the integer order differential equations. We want to come out from the traditional thought i.e. classical order inventory model…

Optimization and Control · Mathematics 2019-04-18 Rituparna Pakhira , Uttam Ghosh , Susmita Sarkar , Vishnu Narayan Mishra

We show that bootstrap methods based on the positivity of probability measures provide a systematic framework for studying both synchronous and asynchronous nonequilibrium stochastic processes on infinite lattices. First, we formulate…

Statistical Mechanics · Physics 2025-11-12 Minjae Cho

In tasks aiming for long-term returns, planning becomes essential. We study generative modeling for planning with datasets repurposed from offline reinforcement learning. Specifically, we identify temporal consistency in the absence of…

Machine Learning · Computer Science 2025-08-19 Deqian Kong , Dehong Xu , Minglu Zhao , Bo Pang , Jianwen Xie , Andrew Lizarraga , Yuhao Huang , Sirui Xie , Ying Nian Wu

Evolution equations are derived for the amplitudes of associative memories: heterogeneous states stored in the connectivity of distributed systems with non-local interactions. The resulting coupled amplitude equations describe the…

Neurons and Cognition · Quantitative Biology 2025-10-20 Akke Mats Houben

This paper considers a class of convex optimization problems where both, the objective function and the constraints, have a continuously varying dependence on time. Our goal is to develop an algorithm to track the optimal solution as it…

Optimization and Control · Mathematics 2015-10-07 Mahyar Fazlyab , Santiago Paternain , Victor M. Preciado , Alejandro Ribeiro

We introduce a model of infinite horizon linear dynamic optimization and obtain results concerning existence of solution and satisfaction of the competitive condition and transversality condition being unconditionally sufficient for…

Optimization and Control · Mathematics 2025-06-23 Somdeb Lahiri

Species experience both internal feedbacks with endogenous factors such as trait evolution and external feedbacks with exogenous factors such as weather. These feedbacks can play an important role in determining whether populations persist…

Populations and Evolution · Quantitative Biology 2016-12-21 Swati Patel , Sebastian J Schreiber

This study proposes an end-to-end algorithm for policy learning in causal inference. We observe data consisting of covariates, treatment assignments, and outcomes, where only the outcome corresponding to the assigned treatment is observed.…

Econometrics · Economics 2025-12-30 Masahiro Kato

As the relative power, performance, and area (PPA) impact of embedded memories continues to grow, proper parameterization of each of the thousands of memories on a chip is essential. When the parameters of all memories of a product are…

Neural and Evolutionary Computing · Computer Science 2022-05-17 Felix Last , Ceren Yeni , Ulf Schlichtmann

Policy learning for partially observed control tasks requires policies that can remember salient information from past observations. In this paper, we present a method for learning policies with internal memory for high-dimensional,…

Machine Learning · Computer Science 2015-09-24 Marvin Zhang , Zoe McCarthy , Chelsea Finn , Sergey Levine , Pieter Abbeel

We investigate how to exploit structural similarities of an individual's potential outcomes (POs) under different treatments to obtain better estimates of conditional average treatment effects in finite samples. Especially when it is…

Machine Learning · Statistics 2021-10-26 Alicia Curth , Mihaela van der Schaar

The ability to flexibly compose previously acquired skills to execute intelligent behaviors is a hallmark of natural intelligence. Such compositional flexibility is often attributed to context-dependent gating mechanisms that determine how…

Optimization and Control · Mathematics 2026-05-18 Francesca Rossi , Veronica Centorrino , Francesco Bullo , Giovanni Russo

The notion of \emph{policy regret} in online learning is a well defined? performance measure for the common scenario of adaptive adversaries, which more traditional quantities such as external regret do not take into account. We revisit the…

Machine Learning · Computer Science 2020-03-24 Raman Arora , Michael Dinitz , Teodor V. Marinov , Mehryar Mohri

Causal inference methods are widely applied in the fields of medicine, policy, and economics. Central to these applications is the estimation of treatment effects to make decisions. Current methods make binary yes-or-no decisions based on…

Machine Learning · Computer Science 2020-04-24 Will Y. Zou , Smitha Shyam , Michael Mui , Mingshi Wang , Jan Pedersen , Zoubin Ghahramani

Under non-exponential discounting, we develop a dynamic theory for stopping problems in continuous time. Our framework covers discount functions that induce decreasing impatience. Due to the inherent time inconsistency, we look for…

Optimization and Control · Mathematics 2017-03-13 Yu-Jui Huang , Adrien Nguyen-Huu