中文
相关论文

相关论文: HALO: Hindsight-Augmented Learning for Online Auto…

200 篇论文

The real-time bidding (RTB), aka programmatic buying, has recently become the fastest growing area in online advertising. Instead of bulking buying and inventory-centric buying, RTB mimics stock exchanges and utilises computer algorithms to…

计算机科学与博弈论 · 计算机科学 2013-06-28 Shuai Yuan , Jun Wang , Xiaoxue Zhao

We propose a novel model for learned query optimization which provides query hints leading to better execution plans. The model addresses the three key challenges in learned hint-based query optimization: reliable hint recommendation…

数据库 · 计算机科学 2024-12-06 Sergey Zinchenko , Sergey Iazov

Maximizing utility with a budget constraint is the primary goal for advertisers in real-time bidding (RTB) systems. The policy maximizing the utility is referred to as the optimal bidding strategy. Earlier works on optimal bidding strategy…

机器学习 · 计算机科学 2020-04-02 Aritra Ghosh , Saayan Mitra , Somdeb Sarkhel , Viswanathan Swaminathan

We study the problem of auction design for advertising platforms that face strategic advertisers who are bidding across platforms. Each advertiser's goal is to maximize their total value or conversions while satisfying some constraint(s)…

计算机科学与博弈论 · 计算机科学 2024-05-07 Gagan Aggarwal , Andres Perlroth , Ariel Schvartzman , Mingfei Zhao

Model-based Reinforcement Learning (MBRL) is a promising framework for learning control in a data-efficient manner. MBRL algorithms can be fairly complex due to the separate dynamics modeling and the subsequent planning algorithm, and as a…

Online advertisements are a primary revenue source for e-commerce platforms. Traditional advertising models are store-centric, selecting winning stores through auction mechanisms. Recently, a new approach known as joint advertising has…

计算机科学与博弈论 · 计算机科学 2025-07-11 Zhen Zhang , Weian Li , Yuhan Wang , Qi Qi , Kun Huang

Recently the online advertising market has exhibited a gradual shift from second-price auctions to first-price auctions. Although there has been a line of works concerning online bidding strategies in first-price auctions, it still remains…

计算机科学与博弈论 · 计算机科学 2022-05-31 Rui Ai , Chang Wang , Chenchen Li , Jinshan Zhang , Wenhan Huang , Xiaotie Deng

In E-commerce advertising, where product recommendations and product ads are presented to users simultaneously, the traditional setting is to display ads at fixed positions. However, under such a setting, the advertising system loses the…

机器学习 · 计算机科学 2019-09-04 Weixun Wang , Junqi Jin , Jianye Hao , Chunjie Chen , Chuan Yu , Weinan Zhang , Jun Wang , Xiaotian Hao , Yixi Wang , Han Li , Jian Xu , Kun Gai

Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for post-training reasoning models. However, group-based methods such as Group Relative Policy Optimization (GRPO) face a critical dilemma in…

机器学习 · 计算机科学 2026-04-07 Yuning Wu , Ke Wang , Devin Chen , Kai Wei

Recent work on hyperparameters optimization (HPO) has shown the possibility of training certain hyperparameters together with regular parameters. However, these online HPO algorithms still require running evaluation on a set of validation…

机器学习 · 计算机科学 2021-01-19 Jingkang Wang , Mengye Ren , Ilija Bogunovic , Yuwen Xiong , Raquel Urtasun

Reinforcement Learning from Human Feedback (RLHF) is currently the leading approach for aligning large language models with human preferences. Typically, these models rely on extensive offline preference datasets for training. However,…

机器学习 · 计算机科学 2024-12-17 Avinandan Bose , Zhihan Xiong , Aadirupa Saha , Simon Shaolei Du , Maryam Fazel

Running Large Language Models (LLMs) on edge devices is constrained by high compute and memory demands posing a barrier for real-time applications in sectors like healthcare, education, and embedded systems. Current solutions such as…

Solving real-world complex tasks using reinforcement learning (RL) without high-fidelity simulation environments or large amounts of offline data can be quite challenging. Online RL agents trained in imperfect simulation environments can…

Real-Time Bidding (RTB) display advertising is a method for purchasing display advertising inventory in auctions that occur within milliseconds. The performance of RTB campaigns is generally measured with a series of Key Performance…

计算机科学与博弈论 · 计算机科学 2020-07-02 Michael Tashman , Jiayi Xie , John Hoffman , Lee Winikor , Rouzbeh Gerami

Real-time bidding (RTB) has become a new norm in display advertising where a publisher uses auction models to sell online user's page view to advertisers. In RTB, the ad with the highest bid price will be displayed to the user. This ad…

计算机科学与博弈论 · 计算机科学 2018-05-23 Xiang Chen

Algorithms increasingly automate bidding in online auctions, raising concerns about tacit bid suppression and revenue shortfalls. Prior work identifies individual mechanisms behind algorithmic bid suppression, but it remains unclear which…

综合经济学 · 经济学 2026-03-24 Pranjal Rawat

Most machine learning algorithms are configured by one or several hyperparameters that must be carefully chosen and often considerably impact performance. To avoid a time consuming and unreproducible manual trial-and-error process to find…

Considerable research effort has been guided towards algorithmic fairness but real-world adoption of bias reduction techniques is still scarce. Existing methods are either metric- or model-specific, require access to sensitive attributes at…

机器学习 · 计算机科学 2022-07-13 André F. Cruz , Pedro Saleiro , Catarina Belém , Carlos Soares , Pedro Bizarro

Bid shading plays a crucial role in Real-Time Bidding (RTB) by adaptively adjusting the bid to avoid advertisers overspending. Existing mainstream two-stage methods, which first model bid landscapes and then optimize surplus using…

计算机科学与博弈论 · 计算机科学 2026-04-30 Yinqiu Huang , Hao Ma , Wenshuai Chen , Zongwei Wang , Shuli Wang , Yongqiang Zhang , Xue Wei , Yinhua Zhu , Haitao Wang , Xingxing Wang

Real-Time Bidding (RTB) enables advertisers to place competitive bids on impression opportunities instantaneously, striving for cost-effectiveness in a highly competitive landscape. Although RTB has widely benefited from the utilization of…

人工智能 · 计算机科学 2025-02-04 Leng Cai , Junxuan He , Yikai Li , Junjie Liang , Yuanping Lin , Ziming Quan , Yawen Zeng , Jin Xu