English
Related papers

Related papers: Selecting the Most Effective Nudge: Evidence from …

200 papers

Epidemics of infectious diseases posing a serious risk to human health have occurred throughout history. During recent epidemics there has been much debate about policy, including how and when to impose restrictions on behaviour.…

Theoretical Economics · Economics 2024-04-08 Simon K. Schnyder , John J. Molina , Ryoichi Yamamoto , Matthew S. Turner

Competitive multi-agent reinforcement learning in imperfect-information games requires agents to act under partial observability and against adversarial opponents, necessitating stochastic policies. While self-play reinforcement learning…

Machine Learning · Computer Science 2026-05-20 Zhiyuan Fan , Gabriele Farina

Chemotherapy dose optimization can be formulated as a dynamic treatment regime, requiring sequential decisions under uncertainty that must balance tumor suppression against toxicity. However, most reinforcement learning approaches assume…

Machine Learning · Computer Science 2026-05-05 Firas Mohamed Elamine Kiram , Imane Youkana , Rachida Saouli , Gian Antonio Susto , Laid Kahloul

Background: Policy evaluation studies that assess how state-level policies affect health-related outcomes are foundational to health and social policy research. The relative ability of newer analytic methods to address confounding, a key…

Reinforcement learning (RL) with group relative policy optimization (GRPO) has become a widely adopted approach for enhancing the reasoning capabilities of multimodal large language models (MLLMs). While GRPO enables long-chain reasoning…

Artificial Intelligence · Computer Science 2026-03-03 Haowen Gao , Zhenyu Zhang , Liang Pang , Fangda Guo , Hongjian Dou , Guannan Lv , Shaoguo Liu , Tingting Gao , Huawei Shen , Xueqi Cheng

In high-energy physics, with the search for ever smaller signals in ever larger data sets, it has become essential to extract a maximum of the available information from the data. Multivariate classification methods based on machine…

Rewards and punishments in different forms are pervasive and present in a wide variety of decision-making scenarios. By observing the outcome of a sufficient number of repeated trials, one would gradually learn the value and usefulness of a…

Machine Learning · Computer Science 2019-06-25 Nikki Lijing Kuang , Clement H. C. Leung

Motivated by applications such as cloud platforms allocating GPUs to users or governments deploying mobile health units across competing regions, we study the dynamic allocation of a reusable resource to strategic agents with private…

Computer Science and Game Theory · Computer Science 2025-07-15 Yan Dai , Negin Golrezaei , Patrick Jaillet

The COVID-19 pandemic left its unique mark on the 21st century as one of the most significant disasters in history, triggering governments all over the world to respond with a wide range of interventions. However, these restrictions come…

Computational Engineering, Finance, and Science · Computer Science 2021-06-02 Assem Zhunis , Tung-Duong Mai , Sundong Kim

The adoption of pre-trained visual representations (PVRs), leveraging features from large-scale vision models, has become a popular paradigm for training visuomotor policies. However, these powerful representations can encode a broad range…

Since vision-based manipulation policies are typically trained from data gathered from a single viewpoint, their performance drops when the view changes during deployment. Naively aggregating demonstrations from numerous random views is not…

Robotics · Computer Science 2025-10-07 Sreevishakh Vasudevan , Som Sagar , Ransalu Senanayake

Conducting randomized experiments in education settings raises the question of how we can use machine learning techniques to improve educational interventions. Using Multi-Armed Bandits (MAB) algorithms like Thompson Sampling (TS) in…

Machine Learning · Computer Science 2022-08-11 Fernando J. Yanez , Angela Zavaleta-Bernuy , Ziwen Han , Michael Liut , Anna Rafferty , Joseph Jay Williams

Targeting and personalization policies can be used to improve outcomes beyond the uniform policy that assigns the best performing treatment in an A/B test to everyone. Personalization relies on the presence of heterogeneity of treatment…

Applications · Statistics 2025-12-12 Anya Shchetkina

The primary goal of randomized trials is to compare the effects of different interventions on some outcome of interest. In addition to the treatment assignment and outcome, data on baseline covariates, such as demographic characteristics or…

Applications · Statistics 2014-01-09 Alisa J. Stephens , Eric J. Tchetgen Tchetgen , Victor De Gruttola

Recent work has shown that reinforcement learning agents can develop policies that exploit spurious correlations between rewards and observations. This phenomenon, known as policy confounding, arises because the agent's policy influences…

Machine Learning · Computer Science 2025-06-16 Miguel Suau

Generative Recommendation (GR) reformulates recommendation as a next-token generation problem and has shown promise in industrial applications. However, extending GR to industrial advertising is non-trivial because the system must optimize…

There is increasing interest in allocating treatments based on observed individual characteristics: examples include targeted marketing, individualized credit offers, and heterogeneous pricing. Treatment personalization introduces…

Econometrics · Economics 2023-04-06 Evan Munro

Background: Recently, a high number of daily positive COVID-19 cases have been reported in regions with relatively high vaccination rates; hence, booster vaccination has become necessary. In addition, infections caused by the different…

Machine Learning · Computer Science 2022-08-26 Essam A. Rashed , Sachiko Kodera , Akimasa Hirata

We study A/B experiments that are designed to compare the performance of two recommendation algorithms. Prior work has observed that the stable unit treatment value assumption (SUTVA) often does not hold in large-scale recommendation…

Machine Learning · Statistics 2026-02-24 Shuangning Li , Chonghuan Wang , Jingyan Wang

The effectiveness of vaccination highly depends on the choice of individuals to vaccinate, even if the same number of individuals are vaccinated. Vaccinating individuals with high centrality measures such as betweenness centrality (BC) and…

Physics and Society · Physics 2021-12-30 Bukyoung Jhun
‹ Prev 1 4 5 6 7 8 10 Next ›