English
Related papers

Related papers: Information-Seeking Decision Strategies Mitigate R…

200 papers

The aim of a number of psychophysics tasks is to uncover how mammals make decisions in a world that is in flux. Here we examine the characteristics of ideal and near-ideal observers in a task of this type. We ask when and how performance…

Neurons and Cognition · Quantitative Biology 2019-10-10 Adrian E. Radillo , Alan Veliz-Cuba , Krešimir Josić , Zachary P. Kilpatrick

The objective of a reinforcement learning agent is to discover better actions through exploration. However, typical exploration techniques aim to maximize rewards, often incurring high costs in both exploration and learning processes. We…

Machine Learning · Computer Science 2024-12-24 Akane Tsuboya , Yu Kono , Tatsuji Takahashi

We study a game-theoretic information retrieval model in which strategic publishers aim to maximize their chances of being ranked first by the search engine while maintaining the integrity of their original documents. We show that the…

Computer Science and Game Theory · Computer Science 2025-07-31 Omer Madmon , Idan Pipano , Itamar Reinman , Moshe Tennenholtz

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

Machine Learning · Computer Science 2025-12-24 Peter N. Loxley

Foraging and acquiring of food is a delicate balance between managing the costs, both energy and social, and individual preferences. Previous research on the solitary foraging of free ranging dogs showed that they prioritized the…

Continuous adaptation to variable environments is crucial for the survival of living organisms. Here, we analyze how adaptation, forecasting, and resource mobilization towards a target state, termed actionability, interact to determine…

Cell Behavior · Quantitative Biology 2024-04-15 Jose M. G. Vilar , Leonor Saiz

A key challenge in satisficing planning is to use multiple heuristics within one heuristic search. An aggregation of multiple heuristic estimates, for example by taking the maximum, has the disadvantage that bad estimates of a single…

Artificial Intelligence · Computer Science 2021-04-13 David Speck , André Biedenkapp , Frank Hutter , Robert Mattmüller , Marius Lindauer

There is an increasingly urgent need to develop knowledge and practices to manage climate risks. For example, flood-risk information can inform household decisions such as purchasing a home or flood insurance. However, flood-risk estimates…

Applications · Statistics 2022-01-05 Courtney M. Cooper , Sanjib Sharma , Robert E. Nicholas , Klaus Keller

This paper analyzes a dynamic interaction between a fully rational, privately informed sender and a boundedly rational, uninformed receiver with memory constraints. The sender controls the flow of information, while the receiver designs a…

Theoretical Economics · Economics 2025-11-12 Qingmin Liu , Yuyang Miao

Decisions and the underlying rules are indispensable for driving process execution during runtime, i.e., for routing process instances at alternative branches based on the values of process data. Decision rules can comprise unary data…

Machine Learning · Computer Science 2021-09-16 Beate Scheibel , Stefanie Rinderle-Ma

Humans use social context to specify preferences over behaviors, i.e. their reward functions. Yet, algorithms for inferring reward models from preference data do not take this social learning view into account. Inspired by pragmatic human…

Machine Learning · Computer Science 2024-05-24 Andi Peng , Yuying Sun , Tianmin Shu , David Abel

Curiosity-based reward schemes can present powerful exploration mechanisms which facilitate the discovery of solutions for complex, sparse or long-horizon tasks. However, as the agent learns to reach previously unexplored spaces and the…

We demonstrate that the set of cost distributions under which the optimal strategy for maximizing compliance (or more generally, effort) in a binary choice environment is identical to the optimal strategy for maximizing the accuracy of the…

Theoretical Economics · Economics 2025-05-26 John W. Patty , Elizabeth Maggie Penn

One explanation for how people can plan efficiently despite limited cognitive resources is that we possess a set of adaptive planning strategies and know when and how to use them. But how are these strategies acquired? While previous…

Artificial Intelligence · Computer Science 2024-12-05 Ruiqi He , Falk Lieder

A decision maker typically (i) incorporates training data to learn about the relative effectiveness of treatments, and (ii) chooses an implementation mechanism that implies an ``optimal'' predicted outcome distribution according to some…

Econometrics · Economics 2025-05-29 Anders Bredahl Kock , David Preinerstorfer

This paper reviews exploration techniques in deep reinforcement learning. Exploration techniques are of primary importance when solving sparse reward problems. In sparse reward problems, the reward is rare, which means that the agent will…

Machine Learning · Computer Science 2022-05-03 Pawel Ladosz , Lilian Weng , Minwoo Kim , Hyondong Oh

How should dispersal strategies be chosen to increase the likelihood of survival of a species? We obtain the answer for the spatially extended versions of three well-known models of two competing species with unequal diffusivities. Though…

Populations and Evolution · Quantitative Biology 2020-07-08 Tapas Singha , Prasad Perlekar , Mustansir Barma

We consider a model of nomadic agents exploring and competing for time-varying location-specific resources, arising in crowdsourced transportation services, online communities, and in traditional location based economic activity. This model…

Computer Science and Game Theory · Computer Science 2016-02-23 Pu Yang , Krishnamurthy Iyer , Peter Frazier

Due to the complexity of many decision making problems, tree search algorithms often have inadequate information to produce accurate transition models. This results in ambiguities (uncertainties for which there are multiple plausible…

Robotics · Computer Science 2024-08-27 Jared J. Beard , R. Michael Butts , Yu Gu

Given two sources of evidence about a latent variable, one can combine the information from both by multiplying the likelihoods of each piece of evidence. However, when one or both of the observation models are misspecified, the…

Machine Learning · Computer Science 2021-03-24 Dmitrii Krasheninnikov , Rohin Shah , Herke van Hoof