中文
相关论文

相关论文: Enforcing Regulation Under Illicit Adaptation

200 篇论文

The objectives of option hedging/trading extend beyond mere protection against downside risks, with a desire to seek gains also driving agent's strategies. In this study, we showcase the potential of robust risk-aware reinforcement learning…

计算金融 · 定量金融 2023-12-27 David Wu , Sebastian Jaimungal

Highly automated robot ecologies (HARE), or societies of independent autonomous robots or agents, are rapidly becoming an important part of much of the world's critical infrastructure. As with human societies, regulation, wherein a…

人工智能 · 计算机科学 2017-10-31 Wen Shen , Alanoud Al Khemeiri , Abdulla Almehrezi , Wael Al Enezi , Iyad Rahwan , Jacob W. Crandall

A growing body of evidence has shown that incorporating behavioral economics principles into the design of financial incentive programs helps improve their cost-effectiveness, promote individuals' short-term engagement, and increase…

社会与信息网络 · 计算机科学 2020-10-28 Palakorn Achananuparp , Ee-Peng Lim , Vibhanshu Abhishek , Tianjiao Yun

This paper focuses on adaptive control of the discrete-time linear quadratic regulator (adaptive LQR). Recent literature has made significant contributions in proving non-asymptotic convergence rates, but existing approaches have a few…

系统与控制 · 电气工程与系统科学 2026-04-27 Peter A. Fisher , Anuradha M. Annaswamy

The last few years have witnessed substantial progress in the field of embodied AI where artificial agents, mirroring biological counterparts, are now able to learn from interaction to accomplish complex tasks. Despite this success,…

计算机视觉与模式识别 · 计算机科学 2022-01-04 Sarah Pratt , Luca Weihs , Ali Farhadi

Learning high-performance control policies that remain consistent with expert behavior is a fundamental challenge in robotics. Reinforcement learning can discover high-performing strategies but often departs from desirable human behavior,…

机器人学 · 计算机科学 2026-04-06 Siwei Ju , Jan Tauberschmidt , Oleg Arenz , Peter van Vliet , Jan Peters

Adaptive synchronization protocols for heterogeneous multi-agent network are investigated. The interaction between each of the agents is carried out through a directed graph. We highlight the lack of communication between agents and the…

系统与控制 · 电气工程与系统科学 2020-10-07 Miguel F. Arevalo-Castiblanco , Duvan A. Tellez-Castro , Jorge Sofrony , Eduardo Mojica-Nava

We study overpricing in a repeated game between two representative agents: a market maker, who controls market liquidity, and a market taker, who chooses trade quantities. Market prices evolve through the endogenous price impact of trades…

交易与市场微观结构 · 定量金融 2026-05-12 Luigi Foscari , Emanuele Guidotti , Nicolò Cesa-Bianchi , Tatjana Chavdarova , Alfio Ferrara

In many practical problems, a learning agent may want to learn the best action in hindsight without ever taking a bad action, which is significantly worse than the default production action. In general, this is impossible because the agent…

机器学习 · 统计学 2018-06-05 Sumeet Katariya , Branislav Kveton , Zheng Wen , Vamsi K. Potluru

Can a regulated, legal market for wildlife products protect species threatened by poaching? It is one of the most controversial ideas in biodiversity conservation. Perhaps the most convincing reason for legalizing wildlife trade is that…

种群与进化 · 定量生物学 2021-03-24 Matthew H. Holden , Jakeb Lockyer

In speculative markets, risk-free profit opportunities are eliminated by traders exploiting them. Markets are therefore often described as "informationally efficient", rapidly removing predictable price changes, and leaving only residual…

交易与市场微观结构 · 定量金融 2013-10-08 Felix Patzelt , Klaus R. Pawelzik

We study a novel multi-armed bandit problem that models the challenge faced by a company wishing to explore new strategies to maximize revenue whilst simultaneously maintaining their revenue above a fixed baseline, uniformly over time.…

机器学习 · 统计学 2016-02-16 Yifan Wu , Roshan Shariff , Tor Lattimore , Csaba Szepesvári

Runtime enforcement can be effectively used to improve the reliability of software applications. However, it often requires the definition of ad hoc policies and enforcement strategies, which might be expensive to identify and implement.…

软件工程 · 计算机科学 2020-10-14 Oliviero Riganelli , Daniela Micucci , Leonardo Mariani

We propose a framework that aligns Conditional Average Treatment Effect (CATE) estimation with profit maximization. Our method recognizes that, for customers with extreme treatment effects, additional estimation accuracy is unlikely to…

计量经济学 · 经济学 2026-04-21 Artem Timoshenko , Caio Waisman

Compliance plays a crucial role in manipulation, as it balances between the concurrent control of position and force under uncertainties. Yet compliance is often overlooked by today's visuomotor policies that solely focus on position…

机器人学 · 计算机科学 2025-03-10 Yifan Hou , Zeyi Liu , Cheng Chi , Eric Cousineau , Naveen Kuppuswamy , Siyuan Feng , Benjamin Burchfiel , Shuran Song

Meta-Reinforcement learning approaches aim to develop learning procedures that can adapt quickly to a distribution of tasks with the help of a few examples. Developing efficient exploration strategies capable of finding the most useful…

机器学习 · 计算机科学 2019-11-12 Swaminathan Gurumurthy , Sumit Kumar , Katia Sycara

In statistical process control, procedures are applied that require relatively strict conditions for their use. If such assumptions are violated, these methods become inefficient, leading to increased incidence of false signals. Therefore,…

其他统计学 · 统计学 2019-01-15 Gejza Dohnal

In this paper, we present an approach based on reinforcement learning for eye tracking data manipulation. It is based on two opposing agents, where one tries to classify the data correctly and the second agent looks for patterns in the…

机器学习 · 计算机科学 2020-10-05 Wolfgang Fuhl , Efe Bozkir , Enkelejda Kasneci

This paper presents a novel control strategy to herd groups of non-cooperative evaders by means of a team of robotic herders. In herding problems, the motion of the evaders is typically determined by strongly nonlinear and heterogeneous…

系统与控制 · 电气工程与系统科学 2022-06-14 Eduardo Sebastián , Eduardo Montijano , Carlos Sagüés

Privacy-enhancing technologies (PETs) represent a critical operational challenge for the online advertising industry, requiring substantial infrastructure investment while promising improved consumer privacy protection. Even when PETs may…

综合经济学 · 经济学 2025-12-22 Kinshuk Jerath , Klaus M. Miller , D. Daniel Sokol