English
Related papers

Related papers: Speed and accuracy in a visual motion discriminati…

200 papers

A strong preference for novelty emerges in infancy and is prevalent across the animal kingdom. When incorporated into reinforcement-based machine learning algorithms, visual novelty can act as an intrinsic reward signal that vastly…

Neurons and Cognition · Quantitative Biology 2019-01-10 Andrew Jaegle , Vahid Mehrpour , Nicole Rust

In the real world, RL agents should be rewarded for fulfilling human preferences. We show that RL agents implicitly learn the preferences of humans in their environment. Training a classifier to predict if a simulated human's preferences…

Artificial Intelligence · Computer Science 2020-02-17 Nevan Wichers

We investigated motor skill learning using a path tracking task, where human subjects had to track various curved paths as fast as possible, in the absence of any external perturbations. Subjects became better with practice, producing…

Neurons and Cognition · Quantitative Biology 2024-06-06 Luke Bashford , Dmitry Kobak , Carsten Mehring

Recent experimental results based on multi-electrode and imaging techniques have reinvigorated the idea that large neural networks operate near a critical point, between order and disorder. However, evidence for criticality has relied on…

Neurons and Cognition · Quantitative Biology 2015-03-05 Thierry Mora , Stéphane Deny , Olivier Marre

One of the hallmarks of biological organisms is their ability to integrate disparate information sources to optimize their behavior in complex environments. How this capability can be quantified and related to the functional complexity of…

Populations and Evolution · Quantitative Biology 2011-11-08 Jeffrey Edlund , Nicolas Chaumont , Arend Hintze , Christof Koch , Giulio Tononi , Christoph Adami

Deep Reinforcement Learning (DRL) is a promising approach for teaching robots new behaviour. However, one of its main limitations is the need for carefully hand-coded reward signals by an expert. We argue that it is crucial to automate the…

Robotics · Computer Science 2021-08-09 Abdalkarim Mohtasib , Gerhard Neumann , Heriberto Cuayahuitl

Large language models (LLMs) have shown remarkable reasoning capabilities, yet aligning such abilities to small language models (SLMs) remains a challenge due to distributional mismatches and limited model capacity. Existing reasoning…

Computation and Language · Computer Science 2025-05-28 Yong Wu , Weihang Pan , Ke Li , Chen Binhui , Ping Li , Binbin Lin

Extracellular recordings of single neurons in primary and secondary somatosensory cortices of monkeys in vivo have shown that their firing rate can increase, decrease, or remain constant in different cells, as the external stimulus…

Neurons and Cognition · Quantitative Biology 2012-03-06 Jing Kang , Hugh P. C. Robinson , Jianfeng Feng

This paper presents a Dynamic Vision Sensor (DVS) based system for reasoning about high speed motion. As a representative scenario, we consider the case of a robot at rest reacting to a small, fast approaching object at speeds higher than…

Computer Vision and Pattern Recognition · Computer Science 2022-10-12 Anthony Bisulco , Fernando Cladera Ojeda , Volkan Isler , Daniel D. Lee

Learning from human feedback has shown to be a useful approach in acquiring robot reward functions. However, expert feedback is often assumed to be drawn from an underlying unimodal reward function. This assumption does not always hold…

Machine Learning · Computer Science 2021-10-20 Vivek Myers , Erdem Bıyık , Nima Anari , Dorsa Sadigh

Sensory neuroscience seeks to understand how the brain encodes natural environments. However, neural coding has largely been studied using simplified stimuli. In order to assess whether the brain's coding strategy depend on the stimulus…

Neurons and Cognition · Quantitative Biology 2007-05-23 Tatyana O. Sharpee , Hiroki Sugihara , Andrei V. Kurgansky , Sergei P. Rebrik , Michael P. Stryker , Kenneth D. Miller

Feedback controlled ratchets are thermal rectifiers that use information on the state of the system to operate. We study the effects of time delays in the feedback for a protocol that performs an instantaneous maximization of the…

Statistical Mechanics · Physics 2008-04-13 F. J. Cao , M. Feito

Mammalian sleep is characterized by multiple alternations between episodes of rapid-eye-movement sleep (REMS) and non-REM sleep (NREMS). While the mechanisms governing the timing of these ultradian NREMS-REMS cycles remain poorly…

A multi-modal machine learning system uses multiple unique data sources and types to improve its performance. This article proposes a system that combines results from several types of models, all of which are trained on different data…

Machine Learning · Computer Science 2024-02-05 Aaron Mullen , Samuel E. Armstrong , Jasmine Perdeh , Bjorn Bauer , Jeffrey Talbert , V. K. Cody Bumgardner

Humans have needs motivating their behavior according to intensity and context. However, we also create preferences associated with each action's perceived pleasure, which is susceptible to changes over time. This makes decision-making more…

Robotics · Computer Science 2024-09-05 Letícia Berto , Paula Costa , Alexandre Simões , Ricardo Gudwin , Esther Colombini

Deep reinforcement learning (deep RL) has been successful in learning sophisticated behaviors automatically; however, the learning process requires a huge number of trials. In contrast, animals can learn new tasks in just a few trials,…

Artificial Intelligence · Computer Science 2016-11-11 Yan Duan , John Schulman , Xi Chen , Peter L. Bartlett , Ilya Sutskever , Pieter Abbeel

Our goal in this paper is to plan the motion of a robot in a partitioned environment with dynamically changing, locally sensed rewards. We assume that arbitrary assumptions on the reward dynamics can be given. The robot aims to accomplish a…

Robotics · Computer Science 2012-08-30 Maria Svorenova , Jana Tumova , Jiri Barnat , Ivana Cerna

Recent advances in the field of machine learning have led to new ways for mobile robots to acquire advanced navigational capabilities. However, these learning-based methods raise the possibility that learned navigation behaviors may not…

Robotics · Computer Science 2024-10-01 Haresh Karnan

While Artificial Intelligence has successfully outperformed humans in complex combinatorial games (such as chess and checkers), humans have retained their supremacy in social interactions that require intuition and adaptation, such as…

Computers and Society · Computer Science 2014-04-22 Fatimah Ishowo-Oloko , Jacob Crandall , Manuel Cebrian , Sherief Abdallah , Iyad Rahwan

Designing reward functions for continuous-control robotics often leads to subtle misalignments or reward hacking, especially in complex tasks. Preference-based RL mitigates some of these pitfalls by learning rewards from comparative…

Artificial Intelligence · Computer Science 2025-03-19 Anukriti Singh , Amisha Bhaskar , Peihong Yu , Souradip Chakraborty , Ruthwik Dasyam , Amrit Bedi , Pratap Tokekar