English
Related papers

Related papers: Selecting the Most Effective Nudge: Evidence from …

200 papers

Celebrities can significantly influence the public towards any desired outcome. In a bid to tackle an infectious disease, a leader (government) exploits such influence towards motivating a fraction of public to get vaccinated, sufficient…

Optimization and Control · Mathematics 2023-04-07 Vartika Singh , Veeraruna Kavitha

This paper introduces a method for constructing an upper bound for exploration policy using either the weighted variance of return sequences or the weighted temporal difference (TD) error. We demonstrate that the variance of the return…

Machine Learning · Computer Science 2020-11-18 Zerong Xi , Gita Sukthankar

Policy gradient methods have been successfully applied to many complex reinforcement learning problems. However, policy gradient methods suffer from high variance, slow convergence, and inefficient exploration. In this work, we introduce a…

Machine Learning · Computer Science 2017-04-11 Yang Liu , Prajit Ramachandran , Qiang Liu , Jian Peng

In order to meet regulatory approval, pharmaceutical companies often must demonstrate that new vaccines reduce the total risk of a post-infection outcome like transmission, symptomatic disease, severe illness, or death in randomized,…

Methodology · Statistics 2024-09-23 Rob Trangucci , Yang Chen , Jon Zelner

Modern deep policy gradient methods achieve effective performance on simulated robotic tasks, but they all require large replay buffers or expensive batch updates, or both, making them incompatible for real systems with resource-limited…

Machine Learning · Computer Science 2025-05-22 Gautham Vasan , Mohamed Elsayed , Alireza Azimi , Jiamin He , Fahim Shariar , Colin Bellinger , Martha White , A. Rupam Mahmood

Vote-boosting is a sequential ensemble learning method in which the individual classifiers are built on different weighted versions of the training data. To build a new classifier, the weight of each training instance is determined in terms…

Machine Learning · Computer Science 2018-02-22 Maryam Sabzevari , Gonzalo Martínez-Muñoz , Alberto Suárez

We present Reasons For and Against Vaccination (RFAV), a dataset for predicting reasons for and against vaccination, and scientific authorities used to justify them, annotated through nichesourcing and augmented using GPT4 and GPT3.5-Turbo.…

Computation and Language · Computer Science 2024-07-01 Damián Ariel Furman , Juan Junqueras , Z. Burçe Gümüslü , Edgar Altszyler , Joaquin Navajas , Ophelia Deroy , Justin Sulik

Motivated by applications in precision medicine and treatment effect heterogeneity, recent research has focused on estimating conditional average treatment effects (CATEs) using machine learning (ML). CATE estimates may represent…

Methodology · Statistics 2025-12-30 Oliver J. Hines , Karla Diaz-Ordaz , Stijn Vansteelandt

Vaccines have proven effective in mitigating the threat of severe infections and deaths during outbreaks of infectious diseases. However, vaccine hesitancy (VH) complicates disease spread prediction and healthcare resource assessment across…

Optimization and Control · Mathematics 2024-05-10 Hieu Bui , Sandra Eksioglu , Ruben Proano , Haoming Shen

Evaluating agent performance when outcomes are stochastic and agents use randomized strategies can be challenging when there is limited data available. The variance of sampled outcomes may make the simple approach of Monte Carlo sampling…

Artificial Intelligence · Computer Science 2017-01-23 Neil Burch , Martin Schmid , Matej Moravčík , Michael Bowling

Voluntary vaccination is effective to prevent infectious diseases from spreading. Both vaccination behavior and cognition of the vaccination risk play important roles in individual vaccination decision making. However, it is not clear how…

Dynamical Systems · Mathematics 2022-09-21 Yuan Liu , Bin Wu

Robust imitation learning using disturbance injections overcomes issues of limited variation in demonstrations. However, these methods assume demonstrations are optimal, and that policy stabilization can be learned via simple augmentations.…

Robotics · Computer Science 2022-05-10 Hirotaka Tahara , Hikaru Sasaki , Hanbit Oh , Brendan Michael , Takamitsu Matsubara

Limited Voting (LV) is an approval-based method for multi-winner elections where all ballots are required to have a same fixed size. While it appears to be used as voting method in corporate governance and has some political applications,…

Computer Science and Game Theory · Computer Science 2024-07-26 Maaike Venema-Los , Zoé Christoff , Davide Grossi

Treatment effects of stochastic policy shifts quantify differences in outcomes across counterfactual scenarios with varying treatment distributions. Stochastic policy shifts may be of interest in settings where it is unrealistic or…

Methodology · Statistics 2026-03-31 Michael Jetsupphasuk , Chenwei Fang , Didong Li , Michael G. Hudgens

Interactive reinforcement learning agents use human feedback or instruction to help them learn in complex environments. Often, this feedback comes in the form of a discrete signal that is either positive or negative. While informative, this…

Artificial Intelligence · Computer Science 2021-04-13 Tasmia Tasrin , Md Sultan Al Nahian , Habarakadage Perera , Brent Harrison

The single transferable vote (STV) is a system of preferential proportional voting employed in multi-seat elections. Each ballot cast by a voter is a (potentially partial) ranking over a set of candidates. The margin of victory, or simply…

Computer Science and Game Theory · Computer Science 2025-03-26 Michelle Blom , Alexander Ek , Peter J. Stuckey , Vanessa Teague , Damjan Vukcevic

Image augmentations applied during training are crucial for the generalization performance of image classifiers. Therefore, a large body of research has focused on finding the optimal augmentation policy for a given task. Yet, RandAugment…

Machine Learning · Computer Science 2021-12-06 Arno Blaas , Xavier Suau , Jason Ramapuram , Nicholas Apostoloff , Luca Zappella

In order to slow the spread of the CoViD-19 pandemic, governments around the world have enacted a wide set of policies limiting the transmission of the disease. Initially, these focused on non-pharmaceutical interventions; more recently,…

Populations and Evolution · Quantitative Biology 2021-06-24 Janoś Gabler , Tobias Raabe , Klara Röhrl , Hans-Martin von Gaudecker

Value iteration (VI) is a foundational dynamic programming method, important for learning and planning in optimal control and reinforcement learning. VI proceeds in batches, where the update to the value of each state must be completed…

Machine Learning · Computer Science 2022-11-29 Tian Tian , Kenny Young , Richard S. Sutton

Containing an epidemic at its origin is the most desirable mitigation. Epidemics have often originated in rural areas, with rural communities among the first affected. Disease dynamics in rural regions have received limited attention, and…

‹ Prev 1 8 9 10 Next ›