English
Related papers

Related papers: Neural Demand Estimation with Habit Formation and …

200 papers

We study a sequential resource allocation problem where a decision maker selects subsets of agents at each period to maximize overall outcomes without prior knowledge of individual-level effects. Our framework applies to settings such as…

Machine Learning · Computer Science 2025-08-29 Katherine B. Adams , Justin J. Boutilier , Qinyang He , Yonatan Mintz

Reinforcement learning (RL) is a dominant paradigm for improving the reasoning abilities of large language models, yet its effectiveness varies across tasks and compute budgets. We propose a \emph{relative-budget} theory explaining this…

Machine Learning · Computer Science 2026-02-03 Akifumi Wachi , Hirota Kinoshita , Shokichi Takakura , Rei Higuchi , Taiji Suzuki

We study online bilateral trade, where a learner facilitates repeated exchanges between a buyer and a seller to maximize the Gain From Trade (GFT), i.e., the social welfare. In doing so, the learner must guarantee not to subsidize the…

Computer Science and Game Theory · Computer Science 2026-02-06 Anna Lunghi , Mattia Piccinato , Matteo Castiglioni , Alberto Marchesi

We develop a unified model in which AI adoption in financial markets generates systemic risk through three mutually reinforcing channels: performative prediction, algorithmic herding, and cognitive dependency. Within an extended rational…

Computational Finance · Quantitative Finance 2026-04-07 Shuchen Meng , Xupeng Chen

This paper studies a composite problem involving the decision making of the optimal entry time and dynamic consumption afterwards. In stage-1, the investor has access to full market information subjecting to some information costs and needs…

Optimization and Control · Mathematics 2021-07-05 Yue Yang , Xiang Yu

We consider model-free reinforcement learning (RL) in non-stationary Markov decision processes. Both the reward functions and the state transition functions are allowed to vary arbitrarily over time as long as their cumulative variations do…

Machine Learning · Computer Science 2022-08-23 Weichao Mao , Kaiqing Zhang , Ruihao Zhu , David Simchi-Levi , Tamer Başar

In the context of structured nonconvex optimization, we estimate the increase in minimum value for a decision that is robust to parameter perturbations as compared to the value of a nominal problem. The estimates rely on detailed…

Optimization and Control · Mathematics 2022-11-22 Johannes O. Royset

We observe nominal price rigidity in tobacco markets across China. The monopolistic seller responds by adjusting product assortments, which remain unobserved by the analyst. We develop and estimate a logit demand model that incorporates…

Econometrics · Economics 2025-01-30 Hui Liu , Yao Luo

State-of-the-art model-based reinforcement learning methods train policies on imagined rollouts. These rollouts are trajectories generated by a learned dynamics model and are scored by a learned reward model, but without querying the true…

Machine Learning · Computer Science 2026-05-13 Nadav Timor , Ravid Shwartz-Ziv , Micah Goldblum , Yann LeCun , David Harel

We develop a framework for difference-in-differences designs with staggered treatment adoption and heterogeneous causal effects. We show that conventional regression-based estimators fail to provide unbiased estimates of relevant estimands…

Econometrics · Economics 2024-01-18 Kirill Borusyak , Xavier Jaravel , Jann Spiess

Statistical spoken dialogue systems have the attractive property of being able to be optimised from data via interactions with real users. However in the reinforcement learning paradigm the dialogue manager (agent) often requires…

Machine Learning · Computer Science 2015-08-19 Pei-Hao Su , David Vandyke , Milica Gasic , Nikola Mrksic , Tsung-Hsien Wen , Steve Young

Working in the framework of morphoelasticity, we develop a model of neurite growth in response to elastic deformation. We decompose the applied stretch into an elastic component and a growth component, and adopt an observationally-motivated…

Biological Physics · Physics 2019-12-13 Madeleine Anthonisen , Peter Grutter

Learning user preferences for products based on their past purchases or reviews is at the cornerstone of modern recommendation engines. One complication in this learning task is that some users are more likely to purchase products or review…

Information Retrieval · Computer Science 2023-03-08 Wanning Chen , Mohsen Bayati

We establish explicit socially optimal rules for an irreversible investment deci- sion with time-to-build and uncertainty. Assuming a price sensitive demand function with a random intercept, we provide comparative statics and economic…

Mathematical Finance · Quantitative Finance 2014-06-03 René Aid , Salvatore Federico , Huyên Pham , Bertrand Villeneuve

Bike-sharing systems are a rapidly developing mode of transportation and provide an efficient alternative to passive, motorized personal mobility. The asymmetric nature of bike demand causes the need for rebalancing bike stations, which is…

Optimization and Control · Mathematics 2021-08-03 Daniele Gammelli , Yihua Wang , Dennis Prak , Filipe Rodrigues , Stefan Minner , Francisco Camara Pereira

We develop a behavioral asset pricing model in which agents trade in a market with information friction. Profit-maximizing agents switch between trading strategies in response to dynamic market conditions. Due to noisy private information…

Trading and Market Microstructure · Quantitative Finance 2019-05-02 Zhentao Shi , Huanhuan Zheng

Since batch algorithms suffer from lack of proficiency in confronting model mismatches and disturbances, this contribution proposes an adaptive scheme based on continuous Lyapunov function for online robot dynamic identification. This paper…

Robotics · Computer Science 2022-10-28 Pedram Agand , Mahdi Aliyari Shoorehdeli

We consider a power utility maximization problem with additive habits in a framework of discrete-time markets and random endowments. For certain classes of incomplete markets, we establish estimates for the optimal consumption stream in…

Portfolio Management · Quantitative Finance 2011-08-16 Roman Muraviev

Whether stochastic or parametric, the Pareto/NBD model can only be utilized for an in-sample prediction rather than an out-of-sample prediction. This research thus provides a neural network based extension of the Pareto/NBD model to…

Applications · Statistics 2019-11-06 Shao-Ming Xie

This paper studies the continuous time utility maximization problem on consumption with addictive habit formation in incomplete semimartingale markets. Introducing the set of auxiliary state processes and the modified dual space, we embed…

Portfolio Management · Quantitative Finance 2015-05-29 Xiang Yu