中文
相关论文

相关论文: Sybil-proof Answer Querying Mechanism

200 篇论文

Correctness is an emergent property of systems where exposing error is cheaper than committing it. In dynamic, low-trust environments, autonomous AI agents benefit from delegating work to sub-agents, yet correctness cannot be assured…

计算机科学与博弈论 · 计算机科学 2025-12-03 David Shi , Kevin Joo

The act of bluffing confounds game designers to this day. The very nature of bluffing is even open for debate, adding further complication to the process of creating intelligent virtual players that can bluff, and hence play, realistically.…

人工智能 · 计算机科学 2007-05-23 Evan Hurwitz , Tshilidzi Marwala

In this work, we describe a self-replication-based mechanism for designing agents of increasing complexity. We demonstrate the validity of this approach by solving simple, standard evolutionary computation problems in simulation. In the…

神经与进化计算 · 计算机科学 2018-07-20 Thommen George Karimpanal

We study the problem of designing mechanisms for \emph{information acquisition} scenarios. This setting models strategic interactions between an uniformed \emph{receiver} and a set of informed \emph{senders}. In our model the senders…

计算机科学与博弈论 · 计算机科学 2023-06-13 Federico Cacciamani , Matteo Castiglioni , Nicola Gatti

Effective human-agent collaboration is increasingly prevalent in real-world applications. Current trends in such collaborations are predominantly unidirectional, with users providing instructions or posing questions to agents, where agents…

人工智能 · 计算机科学 2025-12-16 Emre Can Acikgoz , Jinoh Oh , Jie Hao , Joo Hyuk Jeon , Heng Ji , Dilek Hakkani-Tür , Gokhan Tur , Xiang Li , Chengyuan Ma , Xing Fan

We consider a multi-round auction setting motivated by pay-per-click auctions for Internet advertising. In each round the auctioneer selects an advertiser and shows her ad, which is then either clicked or not. An advertiser derives value…

数据结构与算法 · 计算机科学 2013-06-05 Moshe Babaioff , Yogeshwer Sharma , Aleksandrs Slivkins

We propose a generic reward shaping approach for improving the rate of convergence in reinforcement learning (RL), called Self Improvement Based REwards, or SIBRE. The approach is designed for use in conjunction with any existing RL…

机器学习 · 计算机科学 2020-12-22 Somjit Nath , Richa Verma , Abhik Ray , Harshad Khadilkar

Peer selection, the evaluation and selection of agents by their peers, is an important problem in the field of computational social choice; with applications to grading in massively online courses (MOOCs) and academic peer review. Current…

计算机科学与博弈论 · 计算机科学 2026-05-26 Harper Lyon , Omer Lev , Nicholas Mattei

Agentic reinforcement learning (RL) trains large language models to autonomously call tools during reasoning, with search as the most common application. These models excel at multi-step reasoning tasks, but their safety properties are not…

计算与语言 · 计算机科学 2025-10-21 Yushi Yang , Shreyansh Padarha , Andrew Lee , Adam Mahdi

Fraud continues to proliferate online, from phishing and ransomware to impersonation scams. Yet automated prevention approaches adapt slowly and may not reliably protect users from falling prey to new scams. To better combat online scams,…

人机交互 · 计算机科学 2026-02-02 Owen Hoffman , Kangze Peng , Sajid Kamal , Zehua You , Sukrit Venkatagiri

In order to prevent the disclosure of privacy-sensitive data, such as names and relations between users, social network graphs have to be anonymised before publication. Naive anonymisation of social network graphs often consists in deleting…

社会与信息网络 · 计算机科学 2019-09-04 Sjouke Mauw , Yunior Ramírez-Cruz , Rolando Trujillo-Rasua

This work addresses the problem of sharing partial information within social learning strategies. In traditional social learning, agents solve a distributed multiple hypothesis testing problem by performing two operations at each instant:…

信号处理 · 电气工程与系统科学 2022-12-07 Virginia Bordignon , Vincenzo Matta , Ali H. Sayed

We study the incentivized information acquisition problem, where a principal hires an agent to gather information on her behalf. Such a problem is modeled as a Stackelberg game between the principal and the agent, where the principal…

机器学习 · 计算机科学 2023-08-08 Siyu Chen , Jibang Wu , Yifan Wu , Zhuoran Yang

An abundance of literature has shown that the injection of noise into complex socio-economic systems can improve their resilience. This study aims to understand whether the same applies in the context of information diffusion in social…

物理与社会 · 物理学 2024-03-21 Diana Riazi , Giacomo Livan

This paper studies a dynamic model of information acquisition, in which information might be secretly manipulated. A principal must choose between a safe action with known payoff and a risky action with uncertain payoff, favoring the safe…

理论经济学 · 经济学 2023-04-14 Raphael Boleslavsky

AI agents have been boosted by large language models. AI agents can function as intelligent assistants and complete tasks on behalf of their users with access to tools and the ability to execute commands in their environments. Through…

密码学与安全 · 计算机科学 2024-12-19 Yifeng He , Ethan Wang , Yuyang Rong , Zifei Cheng , Hao Chen

Peer prediction mechanisms incentivize agents to truthfully report their signals even in the absence of verification by comparing agents' reports with those of their peers. In the detail-free multi-task setting, agents respond to multiple…

计算机科学与博弈论 · 计算机科学 2021-08-27 Grant Schoenebeck , Fang-Yi Yu

Common sense suggests that when individuals explain why they believe something, we can arrive at more accurate conclusions than when they simply state what they believe. Yet, there is no known mechanism that provides incentives to elicit…

计算机科学与博弈论 · 计算机科学 2025-02-20 Siddarth Srinivasan , Ezra Karger , Michiel Bakker , Yiling Chen

Reinforcement Learning (RL) bears the promise of being a game-changer in many applications. However, since most of the literature in the field is currently focused on opaque models, the use of RL in high-stakes scenarios, where…

机器学习 · 计算机科学 2025-01-22 Leonardo Lucio Custode , Giovanni Iacca

The Inverse Reinforcement Learning (\textit{IRL}) problem has seen rapid evolution in the past few years, with important applications in domains like robotics, cognition, and health. In this work, we explore the inefficacy of current IRL…

机器学习 · 计算机科学 2022-09-28 Raeid Saqur
‹ 上一页 1 8 9 10 下一页 ›