English
Related papers

Related papers: An accumulator model for primes and targets with i…

200 papers

Reinforcement learning suffers from limitations in real practices primarily due to the number of required interactions with virtual environments. It results in a challenging problem because we are implausible to obtain a local optimal…

Machine Learning · Computer Science 2024-10-28 Qizhen Wu , Kexin Liu , Lei Chen

We consider issues of time in automated trading strategies in simulated financial markets containing a single exchange with public limit order book and continuous double auction matching. In particular, we explore two effects: (i) reaction…

Multiagent Systems · Computer Science 2021-03-02 Henry Hanifan , Ben Watson , John Cartlidge , Dave Cliff

Training a high-dimensional simulated agent with an under-specified reward function often leads the agent to learn physically infeasible strategies that are ineffective when deployed in the real world. To mitigate these unnatural behaviors,…

Artificial Intelligence · Computer Science 2022-03-30 Alejandro Escontrela , Xue Bin Peng , Wenhao Yu , Tingnan Zhang , Atil Iscen , Ken Goldberg , Pieter Abbeel

In this paper, we leverage ideas from model-based control to address the sample efficiency problem of reinforcement learning (RL) algorithms. Accelerating learning is an active field of RL highly relevant in the context of time-varying…

Systems and Control · Electrical Eng. & Systems 2023-05-23 Ibrahim Ahmed , Marcos Quinones-Grueiro , Gautam Biswas

The comprehension of how local interactions arise in global collective behavior is of utmost importance in both biological and physical research. Traditional agent-based models often rely on static rules that fail to capture the dynamic…

Populations and Evolution · Quantitative Biology 2023-08-25 Jianan Li , Liang Li , Shiyu Zhao

Autonomous agents operating in continuous environments must decide not only what to do, but when to act. We introduce a lightweight adaptive temporal control system that learns the optimal interval between cognitive ticks from experience,…

Machine Learning · Computer Science 2026-03-27 Davide Di Gioia

The primary paradigm for multi-task training in natural language processing is to represent the input with a shared pre-trained language model, and add a small, thin network (head) per task. Given an input, a target head is the head that is…

Computation and Language · Computer Science 2021-09-07 Mor Geva , Uri Katz , Aviv Ben-Arie , Jonathan Berant

This paper considers the problem of offering a scarce object with a common unobserved quality to strategic agents in a priority queue. Each agent has a private signal over the quality of the object and observes the decisions made by other…

Computer Science and Game Theory · Computer Science 2024-05-01 Itai Ashlagi , Jamie Kang , Moran Koren , Faidra Monachou

We discuss the collective dynamics of self-propelled particles with selective attraction and repulsion interactions. Each particle, or individual, may respond differently to its neighbors depending on the sign of their relative velocity.…

Other Condensed Matter · Physics 2012-05-16 Pawel Romanczuk , Lutz Schimansky-Geier

In the same way that generative models today conduct most of their training in a self-supervised fashion, how can agentic models conduct their training in a self-supervised fashion, interactively exploring, learning, and preparing to…

Machine Learning · Computer Science 2025-10-21 Kathryn Wantlin , Chongyi Zheng , Benjamin Eysenbach

Estimation of the Average Treatment Effect (ATE) is often carried out in 2 steps, wherein the first step, the treatment and outcome are modeled, and in the second step the predictions are inserted into the ATE estimator. In the first steps,…

Methodology · Statistics 2023-07-21 Mehdi Rostami , Olli Saarela

Response time has attracted increased interest in educational and psychological assessment for, e.g., measuring test takers' processing speed, improving the measurement accuracy of ability, and understanding aberrant response behavior. Most…

Methodology · Statistics 2023-06-22 Ick Hoon Jin , Jonghyun Yun , Hyunjoo Kim , Minjeong Jeon

Large language models excel at complex instructions yet struggle to deviate from their helpful assistant persona, as post-training instills strong priors that resist conflicting instructions. We introduce system prompt strength, a…

Computation and Language · Computer Science 2026-01-13 Yijiang River Dong , Tiancheng Hu , Zheng Hui , Nigel Collier

We compare how well agents aggregate information in two repeated social learning environments. In the first setting agents have access to a public data set. In the second they have access to the same data, and also to the past actions of…

Theoretical Economics · Economics 2026-05-20 Marina Agranov , Gabriel Lopez-Moctezuma , Philipp Strack , Omer Tamuz

Positive feedback regulation is ubiquitous in cell signaling networks, often leading to binary outcomes in response to graded stimuli. However, the role of such feedbacks in clustering, and in spatial spreading of activated molecules, has…

Soft Condensed Matter · Physics 2015-05-13 Jayajit Das , Mehran Kardar , Arup K. Chakraborty

Despite increasing attention paid to the need for fast, scalable methods to analyze next-generation neuroscience data, comparatively little attention has been paid to the development of similar methods for behavioral analysis. Just as the…

Neurons and Cognition · Quantitative Biology 2017-11-02 Shariq Iqbal , John Pearson

In this paper we develop an encounter-based model of reaction-subdiffusion in a domain $\Omega$ with a partially absorbing interior trap $\calU\subset \Omega$. We assume that the particle can freely enter and exit $\calU$, but is only…

Statistical Mechanics · Physics 2023-03-21 Paul C Bressloff

It is well known that the employed triggering scheme has great impact on the control performance when control loops operate under scarce communication resources. Various practical and simulative works have demonstrated the potential of…

Systems and Control · Electrical Eng. & Systems 2023-01-16 David Meister , Frank Aurzada , Mikhail A. Lifshits , Frank Allgöwer

In this chapter we look at one of the canonical driving examples for multi-agent systems: average consensus. In this scenario, a group of agents seek to agree on the average of their initial states. Depending on the particular application,…

Optimization and Control · Mathematics 2016-09-23 Cameron Nowzari , Jorge Cortes , George J. Pappas

We introduce an interacting particle system that models the spread of an epidemic in terms of heterogeneous diffusive dynamics, rather than exogenous contact and transmission rates at the population level as in classical compartmental…

Probability · Mathematics 2026-05-20 Eliana Fausti , Andreas Sojmark
‹ Prev 1 4 5 6 7 8 10 Next ›