English
Related papers

Related papers: Learning the Arrow of Time

200 papers

Recent successes combine reinforcement learning algorithms and deep neural networks, despite reinforcement learning not being widely applied to robotics and real world scenarios. This can be attributed to the fact that current…

Machine Learning · Computer Science 2020-09-01 Vinicius G. Goecks

Economists model knowledge use and acquisition as a cause-and-effect calculus associating observations made by a decision-maker about their world with possible underlying causes. Knowledge models are well-established for static contexts,…

Neural and Evolutionary Computing · Computer Science 2024-12-03 Abigail Devereaux , Roger Koppl

In this paper we introduce a new approach to discrete-time semi-Markov decision processes based on the sojourn time process. Different characterizations of discrete-time semi-Markov processes are exploited and decision processes are…

Probability · Mathematics 2022-04-21 Giacomo Ascione , Salvatore Cuomo

Apprenticeship learning crucially depends on effectively learning rewards, and hence control policies from user demonstrations. Of particular difficulty is the setting where the desired task consists of a number of sub-goals with temporal…

Robotics · Computer Science 2023-11-10 Aniruddh G. Puranic , Jyotirmoy V. Deshmukh , Stefanos Nikolaidis

In the understanding of the fundamental interactions, the origin of an arrow of time is viewed as problematic. However, quantum field theory has an arrow of causality, which tells us which time direction is the past lightcone and which is…

Quantum Physics · Physics 2020-10-15 John F. Donoghue , Gabriel Menezes

We study episodic reinforcement learning in Markov decision processes when the agent receives additional feedback per step in the form of several transition observations. Such additional observations are available in a range of tasks…

Machine Learning · Computer Science 2020-05-11 Christoph Dann , Yishay Mansour , Mehryar Mohri , Ayush Sekhari , Karthik Sridharan

Stochastic time-varying optimization is an integral part of learning in which the shape of the function changes over time in a non-deterministic manner. This paper considers multiple models of stochastic time variation and analyzes the…

Optimization and Control · Mathematics 2023-02-23 Ali Yekkehkhany , Han Feng , Donghao Ying , Javad Lavaei

Time has been an elusive concept to grasp. Although we do not yet understand it properly, there has been advances made in regards to how we can explain it. One such advance is the Page-Wootters mechanism. In this mechanism time is seen as…

Quantum Physics · Physics 2019-12-02 Leandro R. S. Mendes , Diogo O. Soares-Pinto

The training of autonomous agents often requires expensive and unsafe trial-and-error interactions with the environment. Nowadays several data sets containing recorded experiences of intelligent agents performing various tasks, spanning…

Machine Learning · Computer Science 2020-10-06 Giorgio Angelotti , Nicolas Drougard , Caroline Ponzoni Carvalho Chanel

We try to establish a unified information theoretic approach to learning and to explore some of its applications. First, we define {\em predictive information} as the mutual information between the past and the future of a time series,…

Data Analysis, Statistics and Probability · Physics 2007-05-23 Ilya Nemenman

Despite strong evidence for peer effects, little is known about how individuals balance intrinsic preferences and social learning in different choice environments. Using a combination of experiments and discrete choice modeling, we show…

General Economics · Economics 2024-02-29 Fabian Dvorak , Urs Fischbacher

To make informed decisions in natural environments that change over time, humans must update their beliefs as new observations are gathered. Studies exploring human inference as a dynamical process that unfolds in time have focused on…

Neurons and Cognition · Quantitative Biology 2022-03-03 Arthur Prat-Carrabin , Robert C. Wilson , Jonathan D. Cohen , Rava Azeredo da Silveira

In the classical herding literature, agents receive a private signal regarding a binary state of nature, and sequentially choose an action, after observing the actions of their predecessors. When the informativeness of private signals is…

Probability · Mathematics 2018-07-27 Wade Hann-Caruthers , Vadim V. Martynov , Omer Tamuz

Self-organization is ubiquitous in nature and mind. However, machine learning and theories of cognition still barely touch the subject. The hurdle is that general patterns are difficult to define in terms of dynamical equations and…

Artificial Intelligence · Computer Science 2023-02-07 Danilo Vasconcellos Vargas , Tham Yik Foong , Heng Zhang

A continuous-time Markov process is proposed to analyze how a group of humans solves a complex task, consisting in the search of the optimal set of decisions on a fitness landscape. Individuals change their opinions driven by two different…

Multiagent Systems · Computer Science 2015-12-21 Giuseppe Carbone , Ilaria Giannoccaro

We study the problem of learning the Markov order in categorical sequences that represent paths in a network, i.e. sequences of variable lengths where transitions between states are constrained to a known graph. Such data pose challenges…

Machine Learning · Computer Science 2020-07-07 Luka V. Petrović , Ingo Scholtes

To be capable of lifelong learning in a real-life environment, robots have to tackle multiple challenges. Being able to relate physical properties they may observe in their environment to possible interactions they may have is one of them.…

Artificial Intelligence · Computer Science 2020-09-24 Alexandre Manoury , Sao Mai Nguyen , Cédric Buche

Emergence of one-time-direction macroscopic evolution of a classical system of two mixed gases having different temperatures is derived and explained. The analysis performed at the microscopic level, where the time-symmetric laws of…

Classical Physics · Physics 2016-08-11 Krzysztof Rębilas

One effective approach for equipping artificial agents with sensorimotor skills is to use self-exploration. To do this efficiently is critical, as time and data collection are costly. In this study, we propose an exploration mechanism that…

Robotics · Computer Science 2021-02-18 Melisa Sener , Yukie Nagai , Erhan Oztop , Emre Ugur

In this paper, we are interested in optimal decisions in a partially observable Markov universe. Our viewpoint departs from the dynamic programming viewpoint: we are directly approximating an optimal strategic tree depending on the…

General Mathematics · Mathematics 2007-05-23 Frederic Dambreville