中文
相关论文

相关论文: Recursion and evolution: Part II

200 篇论文

It is often difficult to hand-specify what the correct reward function is for a task, so researchers have instead aimed to learn reward functions from human behavior or feedback. The types of behavior interpreted as evidence of the reward…

机器学习 · 计算机科学 2020-12-14 Hong Jun Jeon , Smitha Milli , Anca D. Dragan

Adversarial examples are a challenging open problem for deep neural networks. We propose in this paper to add a penalization term that forces the decision function to be at in some regions of the input space, such that it becomes, at least…

机器学习 · 计算机科学 2019-03-06 Ismaïla Seck , Gaëlle Loosli , Stephane Canu

In this article we address the question whether it is possible to learn the differential equations describing the physical properties of a dynamical system, subject to non-conservative forces, from observations of its realspace…

机器学习 · 计算机科学 2021-07-30 Roger Alexander Müller , Jonathan Laflamme-Janssen , Jaime Camacaro , Carolina Bessega

Preference-based learning of reward functions, where the reward function is learned using comparison data, has been well studied for complex robotic tasks such as autonomous driving. Existing algorithms have focused on learning reward…

机器人学 · 计算机科学 2021-03-05 Sydney M. Katz , Amir Maleki , Erdem Bıyık , Mykel J. Kochenderfer

We describe a mechanism for biological learning and adaptation based on two simple principles: (I) Neuronal activity propagates only through the network's strongest synaptic connections (extremal dynamics), and (II) The strengths of active…

无序系统与神经网络 · 物理学 2009-10-31 Per Bak , Dante R Chialvo

Imitation is a key component of human social behavior, and is widely used by both children and adults as a way to navigate uncertain or unfamiliar situations. But in an environment populated by multiple heterogeneous agents pursuing…

神经元与认知 · 定量生物学 2023-05-15 Max Taylor-Davies , Stephanie Droop , Christopher G. Lucas

This article examines the effect of different other-regarding preference types on the emergence of altruistic punishment behavior from an evolutionary perspective. Our findings corroborate, complement, and interlink the experimental and…

物理与社会 · 物理学 2014-08-26 Moritz Hetzer , Didier Sornette

Complex adaptive systems have been the subject of much recent attention. It is by now well-established that members (`agents') tend to self-segregate into opposing groups characterized by extreme behavior. However, while different social…

凝聚态物理 · 物理学 2009-11-07 Shahar Hod , Ehud Nakar

Humans are skillful navigators: We aptly maneuver through new places, realize when we are back at a location we have seen before, and can even conceive of shortcuts that go through parts of our environments we have never visited. Current…

机器学习 · 计算机科学 2022-09-30 Tankred Saanum , Eric Schulz

Economic experiments reveal that humans value cooperation and fairness. Punishing unfair behavior is therefore common, and according to the theory of strong reciprocity, it is also directly related to rewarding cooperative behavior.…

物理与社会 · 物理学 2013-11-28 Attila Szolnoki , Matjaz Perc

This paper studies the potential of the return distribution for exploration in deterministic reinforcement learning (RL) environments. We study network losses and propagation mechanisms for Gaussian, Categorical and Gaussian mixture…

机器学习 · 计算机科学 2018-07-04 Thomas M. Moerland , Joost Broekens , Catholijn M. Jonker

We aim to design strategies for sequential decision making that adjust to the difficulty of the learning problem. We study this question both in the setting of prediction with expert advice, and for more general combinatorial decision…

机器学习 · 计算机科学 2015-03-02 Wouter M. Koolen , Tim van Erven

In this paper, we wish to investigate the dynamics of information transfer in evolutionary dynamics. We use information theoretic tools to track how much information an evolving population has obtained and managed to retain about different…

种群与进化 · 定量生物学 2021-04-09 Nicholas Guttenberg

In designing an intelligent system that must be able to explain its reasoning to a human user, or to provide generalizations that the human user finds reasonable, it may be useful to take into consideration psychological data on what types…

人工智能 · 计算机科学 2013-04-15 James E. Corter , Mark A. Gluck

Nature is in constant flux, so animals must account for changes in their environment when making decisions. How animals learn the timescale of such changes and adapt their decision strategies accordingly is not well understood. Recent…

神经元与认知 · 定量生物学 2018-12-24 Zachary P. Kilpatrick , William R. Holmes , Tahra L. Eissa , Krešimir Josić

In this paper we propose a novel gradient algorithm to learn a policy from an expert's observed behavior assuming that the expert behaves optimally with respect to some unknown reward function of a Markovian Decision Problem. The…

机器学习 · 计算机科学 2012-06-26 Gergely Neu , Csaba Szepesvari

Modern software systems provide many configuration options which significantly influence their non-functional properties. To understand and predict the effect of configuration options, several sampling and learning strategies have been…

Starting from a heuristic learning scheme for N-person games, we derive a new class of continuous-time learning dynamics consisting of a replicator-like drift adjusted by a penalty term that renders the boundary of the game's strategy space…

最优化与控制 · 数学 2014-04-08 Pierre Coucheney , Bruno Gaujal , Panayotis Mertikopoulos

Is it possible to generally construct a dynamical system to simulate a black system without recovering the equations of motion of the latter? Here we show that this goal can be approached by a learning machine. Trained by a set of…

机器学习 · 统计学 2022-04-15 Hong Zhao

Probabilistic modeling enables combining domain knowledge with learning from data, thereby supporting learning from fewer training instances than purely data-driven methods. However, learning probabilistic models is difficult and has not…

机器学习 · 计算机科学 2017-05-17 Avi Pfeffer