中文
相关论文

相关论文: A/B Testing for Recommender Systems in a Two-sided…

200 篇论文

Online controlled experiments (A/B tests) are fundamental to data-driven decision-making in the digital economy. However, their real-world application is frequently compromised by two critical shortcomings: the use of statistically flawed…

应用统计 · 统计学 2025-09-30 Srijesh Pillai , Rajesh Kumar Chandrawat

Driven by the new economic opportunities created by the creator economy, an increasing number of content creators rely on and compete for revenue generated from online content recommendation platforms. This burgeoning competition reshapes…

信息检索 · 计算机科学 2024-04-30 Fan Yao , Yiming Liao , Mingzhe Wu , Chuanhao Li , Yan Zhu , James Yang , Qifan Wang , Haifeng Xu , Hongning Wang

In recommender systems, users rate items, and are subsequently served other product recommendations based on these ratings. Even though users usually rate a tiny percentage of the available items, the system tries to estimate unobserved…

社会与信息网络 · 计算机科学 2024-06-21 Benjamin Leinwand , Vladas Pipiras

Two-sided marketplaces such as eBay, Etsy and Taobao have two distinct groups of customers: buyers who use the platform to seek the most relevant and interesting item to purchase and sellers who view the same platform as a tool to reach out…

信息检索 · 计算机科学 2019-05-17 Andrew Stanton , Akhila Ananthram , Congzhe Su , Liangjie Hong

The standard A/B testing approaches are mostly based on t-test in large scale industry applications. These standard approaches however suffers from low statistical power in business settings, due to nature of small sample-size or…

统计方法学 · 统计学 2025-12-30 Changshuai Wei , Phuc Nguyen , Benjamin Zelditch , Joyce Chen

Randomised Controlled Trials (RCTs) are the gold standard for estimating treatment effects across many fields of science. Technology companies have adopted A/B-testing methods as a modern RCT counterpart, where end-users are randomly…

社会与信息网络 · 计算机科学 2024-09-20 Olivier Jeunen

During the last few decades, online controlled experiments (also known as A/B tests) have been adopted as a golden standard for measuring business improvements in industry. In our company, there are more than a billion users participating…

应用统计 · 统计学 2021-08-06 Tao Xiong , Yihan Bao , Penglei Zhao , Yong Wang

In industry, online randomized controlled experiment (a.k.a. A/B experiment) is a standard approach to measure the impact of a causal change. These experiments have small treatment effect to reduce the potential blast radius. As a result,…

计量经济学 · 经济学 2025-05-29 Tanmoy Das , Dohyeon Lee , Arnab Sinha

In a two-sided marketplace, network effects are crucial for competitiveness, and platforms need to retain users through advanced customer relationship management as much as possible. Maintaining numerous providers' stable and active…

信息检索 · 计算机科学 2024-07-23 Koya Ohashi , Sho Sekine , Deddy Jobson , Jie Yang , Naoki Nishimura , Noriyoshi Sukegawa , Yuichi Takano

Controlled experiments (A/B tests or randomized field experiments) are the de facto standard to make data-driven decisions when implementing changes and observing customer responses. The methodology to analyze such experiments should be…

应用统计 · 统计学 2020-03-06 Shafi Kamalbasha , Manuel J. A. Eugster

Social influence is ubiquitous in cultural markets, from book recommendations in Amazon, to song popularities in iTunes and the ranking of newspaper articles in the online edition of the New York Times to mention only a few. Yet social…

社会与信息网络 · 计算机科学 2015-05-26 Pascal Van Hentenryck , Andres Abeliuk , Franco Berbeglia , Gerardo Berbeglia

Marketers often use A/B testing as a tool to compare marketing treatments in a test stage and then deploy the better-performing treatment to the remainder of the consumer population. While these tests have traditionally been analyzed using…

应用统计 · 统计学 2020-12-03 Elea McDonnell Feit , Ron Berman

Underpowered studies (below 50% power) suffer from the winner's curse: A statistically significant positive estimate must exaggerate the true treatment effect to meet the significance threshold. A study by Dipayan Biswas, Annika Abell, and…

A/B testing is widexly used in the industry to optimize customer facing websites. Many companies employ experimentation specialists to facilitate and improve the process of A/B testing. Here, we present the application of A/B testing to…

信息检索 · 计算机科学 2024-06-25 Melanie J. I. Müller

Online platforms collect rich information about participants and then share some of this information back with them to improve market outcomes. In this paper we study the following information disclosure problem in two-sided markets: If a…

理论经济学 · 经济学 2023-09-01 Bar Light , Ramesh Johari , Gabriel Weintraub

Users and creators are two crucial components of recommender systems. Typical recommender systems focus on the user side, providing the most suitable items based on each user's request. In such scenarios, a few items receive a majority of…

信息检索 · 计算机科学 2025-03-03 Xiaoshuang Chen , Yibo Wang , Yao Wang , Husheng Liu , Kaiqiao Zhan , Ben Wang , Kun Gai

Experimentation is widely utilized for causal inference and data-driven decision-making across disciplines. In an A/B experiment, for example, an online business randomizes two different treatments (e.g., website designs) to their customers…

统计方法学 · 统计学 2025-01-15 Wenxuan Guo , JungHo Lee , Panos Toulis

Information retrieval systems, such as online marketplaces, news feeds, and search engines, are ubiquitous in today's digital society. They facilitate information discovery by ranking retrieved items on predicted relevance, i.e. likelihood…

计量经济学 · 经济学 2022-05-16 Rina Friedberg , Karthik Rajkumar , Jialiang Mao , Qian Yao , YinYin Yu , Min Liu

A/B tests, also known as randomized controlled experiments (RCTs), are the gold standard for evaluating the impact of new policies, products, or decisions. However, these tests can be costly in terms of time and resources, potentially…

机器学习 · 统计学 2025-01-03 Shima Nassiri , Mohsen Bayati , Joe Cooprider

Online crowdsourcing provides a scalable and inexpensive means to collect knowledge (e.g. labels) about various types of data items (e.g. text, audio, video). However, it is also known to result in large variance in the quality of recorded…

人机交互 · 计算机科学 2018-12-10 Yuan Jin , Mark Carman , Ye Zhu , Yong Xiang