中文
相关论文

相关论文: Collusion-proof And Sybil-proof Reward Mechanisms …

200 篇论文

Individual decision-makers consume information revealed by the previous decision makers, and produce information that may help in future decisions. This phenomenon is common in a wide range of scenarios in the Internet economy, as well as…

计算机科学与博弈论 · 计算机科学 2019-05-06 Yishay Mansour , Aleksandrs Slivkins , Vasilis Syrgkanis

In this paper, we study stochastic coupon probing problem in social networks. Assume there is a social network and a set of coupons. We can offer coupons to some users adaptively and those users who accept the offer will act as seeds and…

社会与信息网络 · 计算机科学 2018-07-11 Shaojie Tang

We investigate the problem of sybil (fake account) detection in social networks from a graph algorithms perspective, where graph structural information is used to classify users as sybil and benign. We introduce the novel notion of user…

社会与信息网络 · 计算机科学 2025-01-29 Ali Safarpoor Dehkordi , Ahad N. Zehmakan

Collusion among autonomous agents poses a critical security threat in embodied multi-agent systems (MAS), where coordinated behaviors can deviate from global objectives and lead to real-world consequences. Existing defenses, primarily based…

密码学与安全 · 计算机科学 2026-04-28 Qi Liu , Xiaohui Chen , Zhihui Zhao , Yaowen Zheng , Dan Yu , Zehua Zhang , Limin Sun , Yongle Chen

Common sense suggests that when individuals explain why they believe something, we can arrive at more accurate conclusions than when they simply state what they believe. Yet, there is no known mechanism that provides incentives to elicit…

计算机科学与博弈论 · 计算机科学 2025-02-20 Siddarth Srinivasan , Ezra Karger , Michiel Bakker , Yiling Chen

Correctness is an emergent property of systems where exposing error is cheaper than committing it. In dynamic, low-trust environments, autonomous AI agents benefit from delegating work to sub-agents, yet correctness cannot be assured…

计算机科学与博弈论 · 计算机科学 2025-12-03 David Shi , Kevin Joo

The primary objective of Multi-Agent Pathfinding (MAPF) is to plan efficient and conflict-free paths for all agents. Traditional multi-agent path planning algorithms struggle to achieve efficient distributed path planning for multiple…

人工智能 · 计算机科学 2024-07-18 Zhenyu Song , Ronghao Zheng , Senlin Zhang , Meiqin Liu

We study the consequences of information asymmetries and misaligned incentives in settings with multiple independent agents. We model an interaction between a Sender, who holds vital private information but cannot act, and a Receiver, who…

多智能体系统 · 计算机科学 2026-05-13 Nanda Kishore Sreenivas , Kate Larson

We describe an experiment in large-scale autoformalization of algebraic topology in an Interactive Theorem Proving (ITP) environment, where the workload is distributed among multiple LLM-based coding agents. Rather than relying on static…

计算机科学中的逻辑 · 计算机科学 2026-03-10 Chad E. Brown , Cezary Kaliszyk , Josef Urban

In cooperative multi-agent reinforcement learning (MARL), how to design a suitable reward signal to accelerate learning and stabilize convergence is a critical problem. The global reward signal assigns the same global reward to all agents…

人工智能 · 计算机科学 2020-03-10 Hangyu Mao , Zhibo Gong , Zhen Xiao

In this paper, we study the problem of deceptive reinforcement learning to preserve the privacy of a reward function. Reinforcement learning is the problem of finding a behaviour policy based on rewards received from exploratory behaviour.…

机器学习 · 计算机科学 2021-02-08 Zhengshang Liu , Yue Yang , Tim Miller , Peta Masters

Bugs in popular distributed protocol implementations have been the source of many downtimes in popular internet services. We describe a randomized testing approach for distributed protocol implementations based on reinforcement learning.…

软件工程 · 计算机科学 2024-09-05 Andrea Borgarelli , Constantin Enea , Rupak Majumdar , Srinidhi Nagendra

In this paper we study the problem of social learning under multiple true hypotheses and self-interested agents which exchange information over a graph. In this setup, each agent receives data that might be generated from a different…

多智能体系统 · 计算机科学 2021-10-27 Konstantinos Ntemos , Virginia Bordignon , Stefan Vlaski , Ali H. Sayed

The reward signal plays a central role in defining the desired behaviors of agents in reinforcement learning (RL). Rewards collected from realistic environments could be perturbed, corrupted, or noisy due to an adversary, sensor error, or…

机器学习 · 计算机科学 2025-03-12 Xi Chen , Zhihui Zhu , Andrew Perrault

Sequential allocation is a simple and widely studied mechanism to allocate indivisible items in turns to agents according to a pre-specified picking sequence of agents. At each turn, the current agent in the picking sequence picks its most…

数据结构与算法 · 计算机科学 2019-09-17 Mingyu Xiao , Jiaxing Ling

Large Language Models (LLMs) have demonstrated potential in automating scientific ideation, yet current approaches relying on iterative prompting or complex multi-agent architectures often suffer from hallucination or computational…

Collaborative machine learning (ML) is an appealing paradigm to build high-quality ML models by training on the aggregated data from many parties. However, these parties are only willing to share their data when given enough incentives,…

机器学习 · 计算机科学 2020-10-27 Rachael Hwee Ling Sim , Yehong Zhang , Mun Choon Chan , Bryan Kian Hsiang Low

Sequential Social Dilemmas (SSDs) provide a key framework for studying how cooperation emerges when individual incentives conflict with collective welfare. In Multi-Agent Reinforcement Learning, these problems are often addressed by…

机器学习 · 计算机科学 2026-02-18 Alper Demir , Hüseyin Aydın , Kale-ab Abebe Tessera , David Abel , Stefano V. Albrecht

We consider a refinement to the notions of collusion-resistance in transaction fee mechanisms. In particular, we require that the collusion is by itself incentive-compatible and individually rational to all of its participants. We then…

计算机科学与博弈论 · 计算机科学 2024-12-31 Matheus V. X. Ferreira , Yotam Gafni , Max Resnick

We study a cost sharing problem derived from bug bounty programs, where agents gain utility by the amount of time they get to enjoy the cost shared information. Once the information is provided to an agent, it cannot be retracted. The goal,…

计算机科学与博弈论 · 计算机科学 2020-06-26 Mingyu Guo , Yong Yang , Muhammad Ali Babar