English
Related papers

Related papers: Rats optimally accumulate and discount evidence in…

200 papers

We consider a model where an agent has a repeated decision to make and wishes to maximize their total payoff. Payoffs are influenced by an action taken by the agent, but also an unknown state of the world that evolves over time. Before…

Computer Science and Game Theory · Computer Science 2021-01-20 Nicole Immorlica , Ian Kash , Brendan Lucier

The aim of this paper is to study the reward based policy exploration problem in a supervised learning approach and enable robots to form complex movement trajectories in challenging reward settings and search spaces. For this, the…

Robotics · Computer Science 2020-11-10 M. Tuluhan Akbulut , Utku Bozdogan , Ahmet Tekden , Emre Ugur

Individual neurons often produce highly variable responses over nominally identical trials, reflecting a mixture of intrinsic "noise" and systematic changes in the animal's cognitive and behavioral state. Disentangling these sources of…

Neurons and Cognition · Quantitative Biology 2021-11-08 Alex H. Williams , Scott W. Linderman

Environmental stochasticity is known to be a destabilizing factor, increasing abundance fluctuations and extinction rates of populations. However, the stability of a community may benefit from the differential response of species to…

Populations and Evolution · Quantitative Biology 2016-02-10 Matan Danino , Nadav M. Shnerb , Sandro Azaele , William E. Kunin , David A. Kessler

Humans and animals show remarkable learning efficiency, adapting to new environments with minimal experience. This capability is not well captured by standard reinforcement learning algorithms that rely on incremental value updates. Rapid…

Artificial Intelligence · Computer Science 2025-12-03 Ching Fang , Kanaka Rajan

Stochastic systems with memory naturally appear in life science, economy, and finance. We take the modelling point of view of stochastic functional delay equations and we study these structures when the driving noises admit jumps. Our…

Probability · Mathematics 2016-06-01 D. R. Baños , F. Cordoni , G. Di Nunno , L. Di Persio , E. E. Røse

Reinforcement learning (RL) has been successfully applied to solve the problem of finding obstacle-free paths for autonomous agents operating in stochastic and uncertain environments. However, when the underlying stochastic dynamics of the…

Machine Learning · Computer Science 2024-10-29 Sheryl Paul , Jyotirmoy V. Deshmukh

A fundamental (and largely open) challenge in sequential decision-making is dealing with non-stationary environments, where exogenous environmental conditions change over time. Such problems are traditionally modeled as non-stationary…

Artificial Intelligence · Computer Science 2024-01-23 Baiting Luo , Yunuo Zhang , Abhishek Dubey , Ayan Mukhopadhyay

Many potential applications of reinforcement learning (RL) require guarantees that the agent will perform well in the face of disturbances to the dynamics or reward function. In this paper, we prove theoretically that maximum entropy…

Machine Learning · Computer Science 2022-05-06 Benjamin Eysenbach , Sergey Levine

Brains adapt to the statistical structure of their input. In the visual system, local light intensities change rapidly, the variance of the intensity changes more slowly, and the dynamic range of contrast itself changes more slowly still.…

Neurons and Cognition · Quantitative Biology 2025-09-03 Charles J. Edelson , Sima Setayeshgar , William Bialek , Rob R. de Ruyter van Steveninck

Real-world applications of reinforcement learning for recommendation and experimentation faces a practical challenge: the relative reward of different bandit arms can evolve over the lifetime of the learning agent. To deal with these…

Machine Learning · Computer Science 2022-06-29 Srivas Chennu , Andrew Maher , Jamie Martin , Subash Prabanantham

Deep Reinforcement Learning (DRL) policies have been shown to be vulnerable to small adversarial noise in observations. Such adversarial noise can have disastrous consequences in safety-critical environments. For instance, a self-driving…

Machine Learning · Computer Science 2024-03-28 Roman Belaire , Pradeep Varakantham , Thanh Nguyen , David Lo

We present an approach for autonomous sensor control for information gathering under partially observable, dynamic and sparsely sampled environments that maximizes information about entities present in that space. We describe our approach…

Artificial Intelligence · Computer Science 2023-05-24 J. Brian Burns , Aravind Sundaresan , Pedro Sequeira , Vidyasagar Sadhu

Reinforcement learning (RL) algorithms struggle with learning optimal policies for tasks where reward feedback is sparse and depends on a complex sequence of events in the environment. Probabilistic reward machines (PRMs) are finite-state…

Machine Learning · Computer Science 2025-10-20 Jan Corazza , Hadi Partovi Aria , Daniel Neider , Zhe Xu

Ecologists are interested in modeling the population growth of species in various ecosystems. Studying population dynamics can assist environmental managers in making better decisions for the environment. Traditionally, the sampling of…

Methodology · Statistics 2021-02-04 Rebecca E. Atanga , Edward L. Boone , Ryad A. Ghanam , Ben Stewart-Koster

As a robot's operational environment and tasks to perform within it grow in complexity, the explicit specification and balancing of optimization objectives to achieve a preferred behavior profile moves increasingly farther out of reach.…

Robotics · Computer Science 2026-03-10 Yi-Shiuan Tung , Gyanig Kumar , Wei Jiang , Bradley Hayes , Alessandro Roncone

We study the dynamics of a simple adaptive system in the presence of noise and periodic damping. The system is composed by two paths connecting a source and a sink, the dynamics is governed by equations that usually describe food search of…

Statistical Mechanics · Physics 2021-12-23 Frederic Folz , Kurt Mehlhorn , Giovanna Morigi

In deep Reinforcement Learning (RL), the learning rate critically influences both stability and performance, yet its optimal value shifts during training as the environment and policy evolve. Standard decay schedulers assume monotonic…

Machine Learning · Computer Science 2025-10-09 Henrique Donâncio , Antoine Barrier , Leah F. South , Florence Forbes

Information processing in neural populations is inherently constrained by metabolic resource limits and noise properties, with dynamics that are not accurately described by existing mathematical models. Recent data, for example, shows that…

Neural and Evolutionary Computing · Computer Science 2026-02-16 Yi-Chun Hung , Gregory Schwartz , Emily A. Cooper , Emma Alexander

Decisions in a group often result in imitation and aggregation, which are enhanced in panic, dangerous, stressful or negative situations. Current explanations of this enhancement are restricted to particular contexts, such as anti-predatory…

Neurons and Cognition · Quantitative Biology 2014-03-31 Alfonso Pérez-Escudero , Gonzalo G. de Polavieja
‹ Prev 1 8 9 10 Next ›