English
Related papers

Related papers: Rats optimally accumulate and discount evidence in…

200 papers

Discovering causal relationships between different variables from time series data has been a long-standing challenge for many domains such as climate science, finance, and healthcare. Given the complexity of real-world relationships and…

Machine Learning · Computer Science 2022-10-27 Wenbo Gong , Joel Jennings , Cheng Zhang , Nick Pawlowski

Data is scaling exponentially in fields ranging from genomics to neuroscience to economics. A central question is: can modern machine learning methods be applied to construct predictive models of natural systems like cells and brains based…

Statistical Mechanics · Physics 2018-08-17 Audrey Huang , Benjamin Sheldan , David A. Sivak , Matt Thomson

We consider the rates of noise-induced switching between the stable states of dissipative dynamical systems with delay and also the rates of noise-induced extinction, where such systems model population dynamics. We study a class of systems…

Statistical Mechanics · Physics 2015-01-27 Ira B. Schwartz , Lora Billings , Thomas W. Carr , Mark Dykman

Reinforcement learning (RL) with sparse and deceptive rewards is challenging because non-zero rewards are rarely obtained. Hence, the gradient calculated by the agent can be stochastic and without valid information. Recent studies that…

Machine Learning · Computer Science 2024-02-08 Guojian Wang , Faguo Wu , Xiao Zhang , Jianxiang Liu

The sequential nature of decision-making in financial asset trading aligns naturally with the reinforcement learning (RL) framework, making RL a common approach in this domain. However, the low signal-to-noise ratio in financial markets…

Machine Learning · Computer Science 2024-11-14 Sven Goluža , Tomislav Kovačević , Stjepan Begušić , Zvonko Kostanjčar

Most contextual bandit algorithms minimize regret against the best fixed policy, a questionable benchmark for non-stationary environments that are ubiquitous in applications. In this work, we develop several efficient contextual bandit…

Machine Learning · Computer Science 2019-04-05 Haipeng Luo , Chen-Yu Wei , Alekh Agarwal , John Langford

This paper focuses on the contextual optimization problem where a decision is subject to some uncertain parameters and covariates that have some predictive power on those parameters are available before the decision is made. More…

Optimization and Control · Mathematics 2024-08-12 Zhaoen Li , Maoqi Liu , Zhi-Hai Zhang

Contextual dueling bandit is used to model the bandit problems, where a learner's goal is to find the best arm for a given context using observed noisy human preference feedback over the selected arms for the past contexts. However,…

Machine Learning · Computer Science 2025-04-17 Arun Verma , Zhongxiang Dai , Xiaoqiang Lin , Patrick Jaillet , Bryan Kian Hsiang Low

In general, comprehension of any type of complex system depends on the resolution used to examine the phenomena occurring within it. However, identifying a priori, for example, the best time frequencies/scales to study a certain system…

Data Analysis, Statistics and Probability · Physics 2025-12-01 Domiziano Doria , Simone Martino , Matteo Becchi , Giovanni M. Pavan

In simple perceptual decisions the brain has to identify a stimulus based on noisy sensory samples from the stimulus. Basic statistical considerations state that the reliability of the stimulus information, i.e., the amount of noise in the…

Neurons and Cognition · Quantitative Biology 2015-09-08 Sebastian Bitzer , Stefan J. Kiebel

Progress has led to a detailed understanding of the neural mechanisms that underlie decision making in primates. However, less is known about why such mechanisms are present in the first place. Theory suggests that primate decision making…

Neurons and Cognition · Quantitative Biology 2026-01-21 Nathan J. Wispinski , Scott A. Stone , Anthony Singhal , Patrick M. Pilarski , Craig S. Chapman

Environmental noises cause the relaxation of quantum systems and decrease the precision of operations. Apprehending the relaxation mechanism via environmental noises is essential for building quantum technologies. Relaxations can be…

Quantum Physics · Physics 2024-04-08 Shingo Kukita , Haruki Kiya , Yasushi Kondo

In the real world, agents often have to operate in situations with incomplete information, limited sensing capabilities, and inherently stochastic environments, making individual observations incomplete and unreliable. Moreover, in many…

Machine Learning · Computer Science 2018-09-26 Akshat Agarwal , Abhinau Kumar , Kyle Dunovan , Erik Peterson , Timothy Verstynen , Katia Sycara

In complex environments, there are costs to both ignorance and perception. An organism needs to track fitness-relevant information about its world, but the more information it tracks, the more resources it must devote to memory and…

Neurons and Cognition · Quantitative Biology 2018-10-17 Sarah E. Marzen , Simon DeDeo

Active inference helps us simulate adaptive behavior and decision-making in biological and artificial agents. Building on our previous work exploring the relationship between active inference, well-being, resilience, and sustainability, we…

Artificial Intelligence · Computer Science 2024-06-13 Mahault Albarracin , Ines Hipolito , Maria Raffa , Paul Kinghorn

We study the problem of distributed task allocation inspired by the behavior of social insects, which perform task allocation in a setting of limited capabilities and noisy environment feedback. We assume that each task has a demand that…

Multiagent Systems · Computer Science 2018-05-15 Anna Dornhaus , Nancy Lynch , Frederik Mallmann-Trenn , Dominik Pajak , Tsvetomira Radeva

This paper studies reinforcement learning (RL) in doubly inhomogeneous environments under temporal non-stationarity and subject heterogeneity. In a number of applications, it is commonplace to encounter datasets generated by system dynamics…

Machine Learning · Statistics 2025-03-18 Liyuan Hu , Mengbing Li , Chengchun Shi , Zhenke Wu , Piotr Fryzlewicz

A rich line of recent work has studied distributionally robust learning approaches that seek to learn a hypothesis that performs well, in the worst-case, on many different distributions over a population. We argue that although the most…

Machine Learning · Computer Science 2024-05-10 Jabari Hastings , Christopher Jung , Charlotte Peale , Vasilis Syrgkanis

We study the problem of dynamic batch learning in high-dimensional sparse linear contextual bandits, where a decision maker, under a given maximum-number-of-batch constraint and only able to observe rewards at the end of each batch, can…

Machine Learning · Statistics 2022-07-19 Zhimei Ren , Zhengyuan Zhou

Domain adaptation is a common problem in robotics, with applications such as transferring policies from simulation to real world and lifelong learning. Performing such adaptation, however, requires informative data about the environment to…

Machine Learning · Computer Science 2021-03-15 Karol Arndt , Oliver Struckmeier , Ville Kyrki
‹ Prev 1 4 5 6 7 8 10 Next ›