中文
相关论文

相关论文: Learning Best Response Strategies for Agents in Ad…

200 篇论文

Reinforcement learning algorithms describe how an agent can learn an optimal action policy in a sequential decision process, through repeated experience. In a given environment, the agent policy provides him some running and terminal…

理论经济学 · 经济学 2020-03-24 Arthur Charpentier , Romuald Elie , Carl Remlinger

Online user-generated content platforms allocate billions of dollars of promotional traffic through algorithms in two-sided marketplaces. To evaluate updates to these algorithms, platforms frequently rely on creator-side randomized…

计量经济学 · 经济学 2026-03-10 Ruohan Zhan , Shichao Han , Yuchen Hu , Zhenling Jiang

We study an agency problem between a leader (the principal) seeking to design an optimal incentive scheme to a follower (the agent) to increase the value of a risky project subjected to accidents and volatility uncertainty. The agency…

最优化与控制 · 数学 2026-05-11 Thibaut Mastrolia , Haoze Yan

Information is often stored in a distributed and proprietary form, and agents who own information are often self-interested and require incentives to reveal their information. Suitable mechanisms are required to elicit and aggregate such…

多智能体系统 · 计算机科学 2022-12-02 Wenlong Wang , Thomas Pfeiffer

This study explores the design of an efficient rebate policy in auction markets, focusing on a continuous-time setting with competition among market participants. In this model, a stock exchange collects transaction fees from auction…

交易与市场微观结构 · 定量金融 2025-01-23 Thibaut Mastrolia , Tianrui Xu

A contemporary feed application usually provides blended results of organic items and sponsored items~(ads) to users. Conventionally, ads are exposed at fixed positions. Such a static exposure strategy is inefficient due to ignoring users'…

人工智能 · 计算机科学 2022-10-20 Dagui Chen , Qi Yan , Chunjie Chen , Zhenzhe Zheng , Yangsu Liu , Zhenjia Ma , Chuan Yu , Jian Xu , Bo Zheng

We propose a neural network approach for solving high-dimensional optimal control problems. In particular, we focus on multi-agent control problems with obstacle and collision avoidance. These problems immediately become high-dimensional,…

最优化与控制 · 数学 2022-05-05 Derek Onken , Levon Nurbekyan , Xingjian Li , Samy Wu Fung , Stanley Osher , Lars Ruthotto

We propose a data-driven framework to enable the modeling and optimization of human-machine interaction processes, e.g., systems aimed at assisting humans in decision-making or learning, work-load allocation, and interactive advertising.…

机器学习 · 计算机科学 2019-03-19 Jiaxiao Zheng , Gustavo de Veciana

We study the consequences of information asymmetries and misaligned incentives in settings with multiple independent agents. We model an interaction between a Sender, who holds vital private information but cannot act, and a Receiver, who…

多智能体系统 · 计算机科学 2026-05-13 Nanda Kishore Sreenivas , Kate Larson

This paper considers the problem of autonomous multi-agent cooperative target search in an unknown environment using a decentralized framework under a no-communication scenario. The targets are considered as static targets and the agents…

机器人学 · 计算机科学 2020-03-13 Titas Bera , Rajarshi Bardhan , Sundaram Suresh

While many multiagent algorithms are designed for homogeneous systems (i.e. all agents are identical), there are important applications which require an agent to coordinate its actions without knowing a priori how the other agents behave.…

人工智能 · 计算机科学 2019-07-17 Stefano V. Albrecht , Subramanian Ramamoorthy

This paper studies the decentralized optimization and learning problem where multiple interconnected agents aim to learn an optimal decision function defined over a reproducing kernel Hilbert space by jointly minimizing a global objective…

机器学习 · 计算机科学 2021-07-01 Ping Xu , Yue Wang , Xiang Chen , Zhi Tian

Maximizing utility with a budget constraint is the primary goal for advertisers in real-time bidding (RTB) systems. The policy maximizing the utility is referred to as the optimal bidding strategy. Earlier works on optimal bidding strategy…

机器学习 · 计算机科学 2020-04-02 Aritra Ghosh , Saayan Mitra , Somdeb Sarkhel , Viswanathan Swaminathan

Imitation is widely observed in populations of decision-making agents. Using our recent convergence results for asynchronous imitation dynamics on networks, we consider how such networks can be efficiently driven to a desired equilibrium…

计算机科学与博弈论 · 计算机科学 2017-04-17 James Riehl , Pouria Ramazi , Ming Cao

We propose a novel framework that computes the corrective control efforts to ensure joint safety in multi-agent dynamical systems. This framework efficiently distributes the required corrective effort without revealing individual agents'…

系统与控制 · 电气工程与系统科学 2026-01-21 Johnathan Corbin , Sarah H. Q. Li , Jonathan Rogers

Controllers with a diagonal-plus-low-rank structure constitute a scalable class of controllers for multi-agent systems. Previous research has shown that diagonal-plus-low-rank control laws appear as the optimal solution to a class of…

系统与控制 · 计算机科学 2016-01-20 Daria Madjidian , Leonid Mirkin , Anders Rantzer

Stochastic control with both inherent random system noise and lack of knowledge on system parameters constitutes the core and fundamental topic in reinforcement learning (RL), especially under non-episodic situations where online learning…

系统与控制 · 电气工程与系统科学 2019-06-24 Xin Huang , Duan Li , Daniel Zhuoyu Long

We study the policy evaluation problem in multi-agent reinforcement learning where a group of agents, with jointly observed states and private local actions and rewards, collaborate to learn the value function of a given policy via local…

最优化与控制 · 数学 2021-11-08 Dongsheng Ding , Xiaohan Wei , Zhuoran Yang , Zhaoran Wang , Mihailo R. Jovanović

Imitation learning is an effective alternative approach to learn a policy when the reward function is sparse. In this paper, we consider a challenging setting where an agent and an expert use different actions from each other. We assume…

机器学习 · 计算机科学 2019-08-27 Konrad Zolna , Negar Rostamzadeh , Yoshua Bengio , Sungjin Ahn , Pedro O. Pinheiro

This paper introduces a novel approach to solve the coverage optimization problem in multi-agent systems. The proposed technique offers an optimal solution with a lower cost with respect to conventional Voronoi-based techniques by…

系统与控制 · 电气工程与系统科学 2025-08-12 Mohammadhasan Faghihi , Meysam Yadegar , Mohammadhosein Bakhtiaridoust , Nader Meskin , Javad Sharifi , Peng Shi
‹ 上一页 1 8 9 10 下一页 ›