English
Related papers

Related papers: Discounting and Impatience

200 papers

Relative temporal-difference (TD) learning was introduced to mitigate the slow convergence of TD methods when the discount factor approaches one by subtracting a baseline from the temporal-difference update. While this idea has been studied…

Machine Learning · Computer Science 2026-04-08 Masoud S. Sakha , Rushikesh Kamalapurkar , Sean Meyn

The changes in user preferences can originate from substantial reasons, like personality shift, or transient and circumstantial ones, like seasonal changes in item popularities. Disregarding these temporal drifts in modelling user…

Information Retrieval · Computer Science 2018-03-01 F. Zafari , I. Moser , T. Baarslag

AI systems are often used to make or contribute to important decisions in a growing range of applications, including criminal justice, hiring, and medicine. Since these decisions impact human lives, it is important that the AI systems act…

Artificial Intelligence · Computer Science 2021-03-16 Duncan C McElfresh , Lok Chan , Kenzie Doyle , Walter Sinnott-Armstrong , Vincent Conitzer , Jana Schaich Borg , John P Dickerson

Temporal difference (TD) methods constitute a class of methods for learning predictions in multi-step prediction problems, parameterized by a recency factor lambda. Currently the most important application of these methods is to temporal…

Artificial Intelligence · Computer Science 2008-02-03 P. Cichosz

We now set up Constraint Closure in a manner consistent with Temporal and Configurational Relationalism. This requires modifying the Dirac Algorithm - which addresses the Constraint Closure Problem facet of the Problem of Time piecemeal -…

General Relativity and Quantum Cosmology · Physics 2019-07-10 Edward Anderson

Recent studies have demonstrated the great power of Transformer models for time series forecasting. One of the key elements that lead to the transformer's success is the channel-independent (CI) strategy to improve the training robustness.…

Machine Learning · Computer Science 2024-02-19 Wang Xue , Tian Zhou , Qingsong Wen , Jinyang Gao , Bolin Ding , Rong Jin

In this paper, we consider the classic stochastic (dynamic) knapsack problem, a fundamental mathematical model in revenue management, with general time-varying random demand. Our main goal is to study the optimal policies, which can be…

Optimization and Control · Mathematics 2018-07-19 Yingdong Lu

Motivated by pricing in ad exchange markets, we consider the problem of robust learning of reserve prices against strategic buyers in repeated contextual second-price auctions. Buyers' valuations for an item depend on the context that…

Machine Learning · Computer Science 2020-02-27 Negin Golrezaei , Adel Javanmard , Vahab Mirrokni

We use a controlled laboratory experiment to study the causal impact of income decreases within a time period on redistribution decisions at the end of that period, in an environment where we keep fixed the sum of incomes over the period.…

General Economics · Economics 2021-07-08 Nickolas Gagnon , Riccardo D. Saulle , Henrik W. Zaunbrecher

In contextual dynamic pricing, a seller sequentially prices goods based on contextual information. Buyers will purchase products only if the prices are below their valuations. The goal of the seller is to design a pricing strategy that…

Machine Learning · Statistics 2025-02-14 Matilde Tullii , Solenne Gaucher , Nadav Merlis , Vianney Perchet

Desharnais, Gupta, Jagadeesan and Panangaden introduced a family of behavioural pseudometrics for probabilistic transition systems. These pseudometrics are a quantitative analogue of probabilistic bisimilarity. Distance zero captures…

Logic in Computer Science · Computer Science 2015-07-01 Franck van Breugel , Babita Sharma , James Worrell

In this paper, we investigate the effects of applying generalised (non-exponential) discounting on a long-run impulse control problem for a Feller-Markov process. We show that the optimal value of the discounted problem is the same as the…

Optimization and Control · Mathematics 2024-04-22 Damian Jelito , Łukasz Stettner

With the widespread application of machine learning in financial risk management, conventional wisdom suggests that longer training periods and more feature variables contribute to improved model performance. This paper, focusing on…

Statistical Finance · Quantitative Finance 2025-01-03 Chengyue Huang , Yahe Yang

Service platforms must determine rules for matching heterogeneous demand (customers) and supply (workers) that arrive randomly over time and may be lost if forced to wait too long for a match. Our objective is to maximize the cumulative…

Optimization and Control · Mathematics 2023-12-19 Angelos Aveklouris , Levi DeValve , Maximiliano Stock , Amy R. Ward

Models of human behavior for prediction and collaboration tend to fall into two categories: ones that learn from large amounts of data via imitation learning, and ones that assume human behavior to be noisily-optimal for some reward…

Artificial Intelligence · Computer Science 2022-04-25 Cassidy Laidlaw , Anca Dragan

Human decision-making differs due to variation in both incentives and available information. This constitutes a substantial challenge for the evaluation of whether and how machine learning predictions can improve decision outcomes. We…

General Economics · Economics 2020-11-24 Michael Allan Ribers , Hannes Ullrich

We consider the dividend maximization problem including a ruin penalty in a diffusion environment. The additional penalty term is motivated by a constraint on dividend strategies. Intentionally, we use different discount rates for the…

Optimization and Control · Mathematics 2022-04-20 Josef Anton Strini , Stefan Thonhauser

Time inconsistency is prevalent in dynamic choice problems: a plan of actions to be taken in the future that is optimal for an agent today may not be optimal for the same agent in the future. If the agent is aware of this intra-personal…

Optimization and Control · Mathematics 2021-05-06 Xue Dong He , Xun Yu Zhou

In recent years, there is growing need and interest in formalizing and reasoning about the quality of software and hardware systems. As opposed to traditional verification, where one handles the question of whether a system satisfies, or…

Logic in Computer Science · Computer Science 2014-11-20 Shaull Almagor , Udi Boker , Orna Kupferman

In this paper, we study the assortment optimization problem faced by many online retailers such as Amazon. We develop a \emph{cascade multinomial logit model}, based on the classic multinomial logit model, to capture the consumers'…

Machine Learning · Computer Science 2020-07-15 Shaojie Tang , Jing Yuan