中文
相关论文

相关论文: Seller-Side Experiments under Interference Induced…

200 篇论文

Online labor markets have great potential as platforms for conducting experiments, as they provide immediate access to a large and diverse subject pool and allow researchers to conduct randomized controlled trials. We argue that online…

人机交互 · 计算机科学 2025-03-19 John J. Horton , David G. Rand , Richard J. Zeckhauser

Interference is a ubiquitous problem in experiments conducted on two-sided content marketplaces, such as Douyin (China's analog of TikTok). In many cases, creators are the natural unit of experimentation, but creators interfere with each…

统计方法学 · 统计学 2023-05-05 Vivek F. Farias , Hao Li , Tianyi Peng , Xinyuyang Ren , Huawei Zhang , Andrew Zheng

Social media platform design often incorporates explicit signals of positive feedback. Some moderators provide positive feedback with the goal of positive reinforcement, but are often unsure of their ability to actually influence user…

人机交互 · 计算机科学 2025-02-14 Charlotte Lambert , Koustuv Saha , Eshwar Chandrasekharan

A/B tests, also known as randomized controlled experiments (RCTs), are the gold standard for evaluating the impact of new policies, products, or decisions. However, these tests can be costly in terms of time and resources, potentially…

机器学习 · 统计学 2025-01-03 Shima Nassiri , Mohsen Bayati , Joe Cooprider

Recommendation systems underlie a variety of online platforms. These recommendation systems and their users form a feedback loop, wherein the former aims to maximize user engagement through personalization and the promotion of popular…

信息检索 · 计算机科学 2025-04-11 Atefeh Mollabagher , Parinaz Naghizadeh

Counterfactual explanations are a prominent example of post-hoc interpretability methods in the explainable Artificial Intelligence research domain. They provide individuals with alternative scenarios and a set of recommendations to achieve…

人工智能 · 计算机科学 2021-01-20 Andrea Ferrario , Michele Loi

We test whether lying aversion can steer equilibrium selection in mechanism design. In a principal-worker environment, the direct mechanism admits two dominant-strategy equilibria: the designer's target and a worker-optimal outcome. We show…

综合经济学 · 经济学 2026-02-20 Alex L. Brown , Ethan Park , Rodrigo A. Velez

We study sequential bilateral trade where sellers and buyers valuations are completely arbitrary (i.e., determined by an adversary). Sellers and buyers are strategic agents with private valuations for the good and the goal is to design a…

计算机科学与博弈论 · 计算机科学 2024-10-11 Yossi Azar , Amos Fiat , Federico Fusco

Recommender Systems (RSs) aim to provide personalized recommendations for users. A newly discovered bias, known as sentiment bias, uncovers a common phenomenon within Review-based RSs (RRSs): the recommendation accuracy of users or items…

信息检索 · 计算机科学 2025-05-07 Le Pan , Yuanjiang Cao , Chengkai Huang , Wenjie Zhang , Lina Yao

Machine learning (ML) models play an increasingly prevalent role in many software engineering tasks. However, because most models are now powered by opaque deep neural networks, it can be difficult for developers to understand why the model…

软件工程 · 计算机科学 2021-11-11 Jürgen Cito , Isil Dillig , Vijayaraghavan Murali , Satish Chandra

Adaptive experimental design (AED) methods are increasingly being used in industry as a tool to boost testing throughput or reduce experimentation cost relative to traditional A/B/N testing methods. However, the behavior and guarantees of…

机器学习 · 计算机科学 2024-09-19 Tanner Fiez , Houssam Nassif , Yu-Cheng Chen , Sergio Gamez , Lalit Jain

Online advertising platforms host hundreds of thousands of A/B tests, but the platform's delivery algorithm routes each creative to the audience it predicts will engage. Every two-arm test therefore conflates the creative's effect with the…

计量经济学 · 经济学 2026-05-25 Pallavi Pal , Anjana Susarla

NLP-assisted solutions to support qualitative data analysis have gained considerable traction. However, no unified evaluation framework exists which can account for the many different settings in which qualitative researchers may employ…

计算与语言 · 计算机科学 2026-04-22 Alvin Po-Chun Chen , Rohan Das , Dananjay Srinivas , Alexandra Barry , Maksim Seniw , Maria Leonor Pacheco

Real-time hybrid testing is a method in which a substructure of the system is realised experimentally and the rest numerically. The two parts interact in real time to emulate the dynamics of the full system. Such experiments however are…

动力系统 · 数学 2024-06-04 Sandor Beregi , David A. W. Barton , Djamel Rezgui , Simon A. Neild

Programming is a fundamentally interactive process, yet coding assistants are often evaluated using static benchmarks that fail to measure how well models collaborate with users. We introduce an interactive evaluation pipeline to examine…

人机交互 · 计算机科学 2025-02-26 Jane Pan , Ryan Shar , Jacob Pfau , Ameet Talwalkar , He He , Valerie Chen

In two-sided platforms (e.g., video streaming or e-commerce), viewers and providers engage in interactive dynamics: viewers benefit from increases in provider populations, while providers benefit from increases in viewer population. Despite…

计算机科学与博弈论 · 计算机科学 2025-05-28 Haruka Kiyohara , Fan Yao , Sarah Dean

Current approaches to A/B testing in networks focus on limiting interference, the concern that treatment effects can "spill over" from treatment nodes to control nodes and lead to biased causal effect estimation. Prominent methods for…

机器学习 · 计算机科学 2020-04-16 Zahra Fatemi , Elena Zheleva

Pervasive personal communication technologies offer the potential for important social benefits for individual users, but also the potential for significant social difficulties and costs. In research on face-to-face social interaction,…

人机交互 · 计算机科学 2008-04-18 Paul M. Aoki , Allison Woodruff

Online controlled experimentation is widely adopted for evaluating new features in the rapid development cycle for web products and mobile applications. Measurement of the overall experiment sample is a common practice to quantify the…

人机交互 · 计算机科学 2022-01-27 Zhenyu Zhao , Yan He , Miao Chen

Top-N recommendation, which aims to learn user ranking-based preference, has long been a fundamental problem in a wide range of applications. Traditional models usually motivate themselves by designing complex or tailored architectures…

信息检索 · 计算机科学 2021-09-14 Mengyue Yang , Quanyu Dai , Zhenhua Dong , Xu Chen , Xiuqiang He , Jun Wang