English
Related papers

Related papers: Performance of normative and approximate evidence …

200 papers

Learning near-optimal behaviour from an expert's demonstrations typically relies on the assumption that the learner knows the features that the true reward function depends on. In this paper, we study the problem of learning from…

Machine Learning · Computer Science 2019-03-28 Luis Haug , Sebastian Tschiatschek , Adish Singla

In classic reinforcement learning (RL) and decision making problems, policies are evaluated with respect to a scalar reward function, and all optimal policies are the same with regards to their expected return. However, many real-world…

Machine Learning · Computer Science 2023-11-02 Han Shao , Lee Cohen , Avrim Blum , Yishay Mansour , Aadirupa Saha , Matthew R. Walter

Recommendation systems (RSs) are increasingly used to guide job seekers on online platforms, yet the algorithms currently deployed are typically optimized for predictive objectives such as clicks, applications, or hires, rather than job…

The brain constantly turns large flows of sensory information into selective representations of the environment. It, therefore, needs to learn to process those sensory inputs that are most relevant for behaviour. It is not well understood…

Neurons and Cognition · Quantitative Biology 2023-01-10 Pouya Baniasadi

We develop a dynamical systems approach to prioritizing and selecting multiple recurring tasks with the aim of conferring a degree of deliberative goal selection to a mobile robot confronted with competing objectives. We take navigation as…

Dynamical Systems · Mathematics 2018-03-12 Paul B. Reverdy , Daniel E. Koditschek

Visual search is a fundamental natural task for humans and other animals. We investigated the decision processes humans use in covert (single-fixation) search with briefly presented displays having well-separated potential target locations.…

Neurons and Cognition · Quantitative Biology 2025-04-16 Anqi Zhang , Wilson S. Geisler

Individual neurons often produce highly variable responses over nominally identical trials, reflecting a mixture of intrinsic "noise" and systematic changes in the animal's cognitive and behavioral state. Disentangling these sources of…

Neurons and Cognition · Quantitative Biology 2021-11-08 Alex H. Williams , Scott W. Linderman

Dynamic task allocation is an essential requirement for multi-robot systems operating in unknown dynamic environments. It allows robots to change their behavior in response to environmental changes or actions of other robots in order to…

Robotics · Computer Science 2007-05-23 Kristina Lerman , Chris Jones , Aram Galstyan , Maja J Mataric

We develop dependent hierarchical normalized random measures and apply them to dynamic topic modeling. The dependency arises via superposition, subsampling and point transition on the underlying Poisson processes of these measures. The…

Machine Learning · Computer Science 2012-06-22 Changyou Chen , Nan Ding , Wray Buntine

Transferring reinforcement learning policies trained in physics simulation to the real hardware remains a challenge, known as the "sim-to-real" gap. Domain randomization is a simple yet effective technique to address dynamics discrepancies…

Robotics · Computer Science 2021-04-05 Ioannis Exarchos , Yifeng Jiang , Wenhao Yu , C. Karen Liu

In this paper, we learn dynamics models for parametrized families of dynamical systems with varying properties. The dynamics models are formulated as stochastic processes conditioned on a latent context variable which is inferred from…

Machine Learning · Computer Science 2024-10-08 Jan Achterhold , Joerg Stueckler

Understanding which features humans rely on -- in visually recognizing action similarity is a crucial step towards a clearer picture of human action perception from a learning and developmental perspective. In the present work, we…

Opinion dynamics is of paramount importance as it provides insights into the complex dynamics of opinion propagation and social relationship adjustment. It is assumed in most of the previous works that social relationships evolve much…

Physics and Society · Physics 2024-04-05 Xunlong Wang , Bin Wu

Policy learning utilizing observational data is pivotal across various domains, with the objective of learning the optimal treatment assignment policy while adhering to specific constraints such as fairness, budget, and simplicity. This…

Methodology · Statistics 2023-10-12 Pan Zhao , Antoine Chambaz , Julie Josse , Shu Yang

A major goal of computational neuroscience has been to explain how the primate ventral visual stream (VVS) transforms visual input into temporally evolving neural representations that support robust visual perception. Historically, most…

Neurons and Cognition · Quantitative Biology 2026-01-21 Matteo Dunnhofer , Maren Wehrheim , Hamidreza Ramezanpour , Sabine Muzellec , Kohitij Kar

Stochastic processes offer a flexible mathematical formalism to model and reason about systems. Most analysis tools, however, start from the premises that models are fully specified, so that any parameters controlling the system's dynamics…

Systems and Control · Computer Science 2017-01-11 Luca Bortolussi , Guido Sanguinetti

We consider a random walker whose motion is tethered around a focal point. We use two models that exhibit the same spatial dependence in the steady state but widely different dynamics. In one case, the walker is subject to a deterministic…

Statistical Mechanics · Physics 2019-01-11 Luca Giuggioli , Shamik Gupta , Matt Chase

The mimicking of human-like arm movement characteristics involves the consideration of three factors during control policy synthesis: (a) chosen task requirements, (b) inclusion of noise during movement execution and (c) chosen optimality…

Economic model predictive control and tracking model predictive control are two popular advanced process control strategies used in various of fields. Nevertheless, which one should be chosen to achieve better performance in the presence of…

Systems and Control · Electrical Eng. & Systems 2022-01-07 Jiangbang Liu , Song Bo , Benjamin Decardi-Nelson , Jinfeng Liu , Jingtao Hu , Tao Zou

Model-based Reinforcement Learning estimates the true environment through a world model in order to approximate the optimal policy. This family of algorithms usually benefits from better sample efficiency than their model-free counterparts.…

Machine Learning · Computer Science 2021-10-27 Valentin Charvet , Bjørn Sand Jensen , Roderick Murray-Smith