中文
相关论文

相关论文: Personalization as a Game: Equilibrium-Guided Gene…

200 篇论文

Various social contexts ranging from public goods provision to information collection can be depicted as games of strategic interactions, where a player's well-being depends on her own action as well as on the actions taken by her…

物理与社会 · 物理学 2015-04-01 Giulio Cimini , Claudio Castellano , Angel Sánchez

We present a general framework for evolutionary learning to emergent unbiased state representation without any supervision. Evolutionary frameworks such as self-play converge to bad local optima in case of multi-agent reinforcement learning…

机器学习 · 统计学 2023-02-03 Shohei Ohsawa

Pervasive health games have a potential to impact health-related behaviors. And, similar to other types of interventions, engagement and adherence in health games is the keystone for examining their short- and long-term effects. Many…

人机交互 · 计算机科学 2021-06-28 Magy Seif El-Nasr , Shree Durga , Mariya Shiyko , Carmen Sceppa

The incorporation of generative artificial intelligence into personal health applications presents a transformative opportunity for personalized, data-driven health and fitness guidance, yet also poses challenges related to user safety,…

Behavioral health interventions, delivered through digital platforms, have the potential to significantly improve health outcomes, through education, motivation, reminders, and outreach. We study the problem of optimizing personalized…

机器学习 · 计算机科学 2024-07-19 Jackie Baek , Justin J. Boutilier , Vivek F. Farias , Jonas Oddur Jonasson , Erez Yoeli

Adaptive game systems aim to enrich player experiences by dynamically adjusting game content in response to user data. While extensive research has addressed content personalization and player experience modeling, the integration of these…

人机交互 · 计算机科学 2025-05-05 Phil Lopes , Nuno Fachada , Maria Fonseca

Regulatory compliance in the pharmaceutical industry entails navigating through complex and voluminous guidelines, often requiring significant human resources. To address these challenges, our study introduces a chatbot model that utilizes…

计算与语言 · 计算机科学 2024-02-07 Jaewoong Kim , Moohong Min

Data holders, such as mobile apps, hospitals and banks, are capable of training machine learning (ML) models and enjoy many intelligence services. To benefit more individuals lacking data and models, a convenient approach is needed which…

密码学与安全 · 计算机科学 2020-12-22 Jiasi Weng , Jian Weng , Hongwei Huang , Chengjun Cai , Cong Wang

When a prediction algorithm serves a collection of users, disparities in prediction quality are likely to emerge. If users respond to accurate predictions by increasing engagement, inviting friends, or adopting trends, repeated learning…

机器学习 · 计算机科学 2025-11-27 Eden Saig , Nir Rosenfeld

Theory of Mind benchmarks for large language models typically produce aggregate scores without theoretical grounding, making it unclear whether high performance reflects strategic reasoning or surface-level heuristics. We introduce a…

计算机科学与博弈论 · 计算机科学 2026-03-12 Mateo Pechon-Elkins , Jon Chun

The design of coherent and efficient policies to address infectious diseases and their consequences requires to model not only epidemics dynamics, but also individual behaviors, as the latter has a strong influence on the former. In our…

物理与社会 · 物理学 2024-04-16 Louis Bremaud , Olivier Giraud , Denis Ullmo

Effective emotional support hinges on understanding users' emotions and needs to provide meaningful comfort during multi-turn interactions. Large Language Models (LLMs) show great potential for expressing empathy; however, they often…

计算与语言 · 计算机科学 2025-05-23 Jing Ye , Lu Xiang , Yaping Zhang , Chengqing Zong

Empirical Centroid Fictitious Play (ECFP) is a generalization of the well-known Fictitious Play (FP) algorithm designed for implementation in large-scale games. In ECFP, the set of players is subdivided into equivalence classes with players…

最优化与控制 · 数学 2015-04-03 Brian Swenson , Soummya Kar , Joao Xavier

Addressing the issues of dullness, low compliance, and lack of appeal in current digital mental health education and serious games for students and adolescents, this study proposes a novel, experience-centered framework for serious game…

人机交互 · 计算机科学 2026-04-20 Ting-Chen Hsu , Zheyuan Zhang , Ziyi Chen , Yuwen Liu , Yanjia Liu

Designing reward functions for efficiently guiding reinforcement learning (RL) agents toward specific behaviors is a complex task. This is challenging since it requires the identification of reward structures that are not sparse and that…

机器学习 · 计算机科学 2023-11-01 Dhawal Gupta , Yash Chandak , Scott M. Jordan , Philip S. Thomas , Bruno Castro da Silva

Interactions between people are the basis on which the structure of our society arises as a complex system and, at the same time, are the starting point of any physical description of it. In the last few years, much theoretical research has…

计算机科学与博弈论 · 计算机科学 2017-12-06 Mattia Mazzoli , Angel Sanchez

Alignment of large language models (LLMs) with human preferences typically relies on supervised reward models or external judges that demand abundant annotations. However, in fields that rely on professional knowledge, such as medicine and…

人工智能 · 计算机科学 2025-11-18 Yiyang Zhao , Huiyu Bai , Xuejiao Zhao

We study the complexity of computing equilibria in two classes of network games based on flows - fractional BGP (Border Gateway Protocol) games and fractional BBC (Bounded Budget Connection) games. BGP is the glue that holds the Internet…

计算机科学与博弈论 · 计算机科学 2008-12-05 Laura J. Poplawski , Rajmohan Rajaraman , Ravi Sundaram , Shang-Hua Teng

Federated learning promises significant sample-efficiency gains by pooling data across multiple agents, yet incentive misalignment is an obstacle: each update is costly to the contributor but boosts every participant. We introduce a…

计算机科学与博弈论 · 计算机科学 2026-02-02 Ariel D. Procaccia , Han Shao , Itai Shapira

We analyze independent policy-gradient (PG) learning in $N$-player linear-quadratic (LQ) stochastic differential games. Each player employs a distributed policy that depends only on its own state and updates the policy independently using…

最优化与控制 · 数学 2026-02-19 Philipp Plank , Yufei Zhang