English
Related papers

Related papers: Q-Regularized Generative Auto-Bidding: From Subopt…

200 papers

Automated bidding is central to modern digital advertising. Early rule-based methods lacked adaptability, while subsequent Reinforcement Learning approaches modeled bidding as a Markov Decision Process but struggled with long-term…

Artificial Intelligence · Computer Science 2026-05-20 Mingming Zhang , Feiqing Zhuang , Na Li , Shengjie Sun , Xiaowei Chen , Junxiong Zhu , Fei Xiao , Keping Yang , Lixin Zou , Chenliang Li

In the realm of online advertising, advertisers partake in ad auctions to obtain advertising slots, frequently taking advantage of auto-bidding tools provided by demand-side platforms. To improve the automation of these bidding systems, we…

Machine Learning · Computer Science 2025-06-30 Hao Jiang , Yongxiang Tang , Yanxiang Zeng , Pengjia Yuan , Yanhua Cheng , Teng Sha , Xialong Liu , Peng Jiang

Auto-bidding plays a crucial role in facilitating online advertising by automatically providing bids for advertisers. Reinforcement learning (RL) has gained popularity for auto-bidding. However, most current RL auto-bidding methods are…

Machine Learning · Computer Science 2024-10-10 Jiayan Guo , Yusen Huo , Zhilin Zhang , Tianyu Wang , Chuan Yu , Jian Xu , Yan Zhang , Bo Zheng

Auto-bidding, with its strong capability to optimize bidding decisions within dynamic and competitive online environments, has become a pivotal strategy for advertising platforms. Existing approaches typically employ rule-based strategies…

Machine Learning · Computer Science 2025-04-28 Jingtong Gao , Yewen Li , Shuai Mao , Peng Jiang , Nan Jiang , Yejing Wang , Qingpeng Cai , Fei Pan , Peng Jiang , Kun Gai , Bo An , Xiangyu Zhao

Auto-bidding is a critical tool for advertisers to improve advertising performance. Recent progress has demonstrated that AI-Generated Bidding (AIGB), which learns a conditional generative planner from offline data, achieves superior…

Machine Learning · Computer Science 2026-03-04 Zhiyu Mou , Yiqin Lv , Miao Xu , Qi Wang , Yixiu Mao , Jinghao Chen , Qichen Ye , Chao Li , Rongquan Bai , Chuan Yu , Jian Xu , Bo Zheng

Training generally capable agents that thoroughly explore their environment and learn new and diverse skills is a long-term goal of robot learning. Quality Diversity Reinforcement Learning (QD-RL) is an emerging research area that blends…

Machine Learning · Computer Science 2024-01-31 Sumeet Batra , Bryon Tjanaka , Matthew C. Fontaine , Aleksei Petrenko , Stefanos Nikolaidis , Gaurav Sukhatme

We propose a new Q-learning variant, called 2RA Q-learning, that addresses some weaknesses of existing Q-learning methods in a principled manner. One such weakness is an underlying estimation bias which cannot be controlled and often…

Optimization and Control · Mathematics 2024-05-30 Peter Schmitt-Förster , Tobias Sutter

Auto-bidding systems aim to maximize advertiser value over long horizons under budget constraints and ratio targets such as cost-per-acquisition, yet future traffic and auction dynamics are non-stationary and uncertain. Existing approaches…

Artificial Intelligence · Computer Science 2026-05-28 Eunseok Yang , Xingdong Zuo , Kyung-Min Kim

The Q-learning algorithm is known to be affected by the maximization bias, i.e. the systematic overestimation of action values, an important issue that has recently received renewed attention. Double Q-learning has been proposed as an…

Machine Learning · Computer Science 2021-02-03 Rong Zhu , Mattia Rigotti

Auto-bidding is essential in facilitating online advertising by automatically placing bids on behalf of advertisers. Generative auto-bidding, which generates bids based on an adjustable condition using models like transformers and…

Artificial Intelligence · Computer Science 2025-06-04 Yewen Li , Shuai Mao , Jingtong Gao , Nan Jiang , Yunjian Xu , Qingpeng Cai , Fei Pan , Peng Jiang , Bo An

Online advertising has become a core revenue driver for the internet industry, with ad auctions playing a crucial role in ensuring platform revenue and advertiser incentives. Traditional auction mechanisms, like GSP, rely on the independent…

Computer Science and Game Theory · Computer Science 2024-12-17 Ruitao Zhu , Yangsu Liu , Dagui Chen , Zhenjia Ma , Chufeng Shi , Zhenzhe Zheng , Jie Zhang , Jian Xu , Bo Zheng , Fan Wu

This paper proposes a learning algorithm to find a scheduling policy that achieves an optimal delay-power trade-off in communication systems. Reinforcement learning (RL) is used to minimize the expected latency for a given energy constraint…

Systems and Control · Electrical Eng. & Systems 2020-06-11 Yu Zhao , Joohyun Lee , Wei Chen

In the realm of online advertising, automated bidding has become a pivotal tool, enabling advertisers to efficiently capture impression opportunities in real-time. Recently, generative auto-bidding has shown significant promise, offering…

Information Retrieval · Computer Science 2026-02-27 Yulong Gao , Wan Jiang , Mingzhe Cao , Xuepu Wang , Zeyu Pan , Haonan Yang , Ye Liu , Xin Yang

Auto-bidding systems aim to maximize marketing value while satisfying strict efficiency constraints such as Target Cost-Per-Action (CPA). Although Decision Transformers provide powerful sequence modeling capabilities, applying them to this…

Machine Learning · Computer Science 2026-02-10 Binglin Wu , Yingyi Zhang , Xianneng Li , Ruyue Deng , Chuan Yue , Weiru Zhang , Xiaoyi Zeng

Auto-bidding services optimize real-time bidding strategies for advertisers under key performance indicator (KPI) constraints such as target return on investment and budget. However, uncertainties such as model prediction errors and…

Computer Science and Game Theory · Computer Science 2026-04-08 Linghui Meng , Chun Gan , Shengsheng Niu , Chengcheng Zhang , Chenchen Li , Chuan Yang , Yi Mao , Xin Zhu , Jie He , Zhangang Lin , Ching Law

Autonomous driving in multi-agent dynamic traffic scenarios is challenging: the behaviors of road users are uncertain and are hard to model explicitly, and the ego-vehicle should apply complicated negotiation skills with them, such as…

Robotics · Computer Science 2022-06-22 Peide Cai , Hengli Wang , Yuxiang Sun , Ming Liu

This paper explores the application of a reinforcement learning (RL) framework using the Q-Learning algorithm to enhance dynamic pricing strategies in the retail sector. Unlike traditional pricing methods, which often rely on static demand…

Machine Learning · Computer Science 2024-11-28 Mohit Apte , Ketan Kale , Pranav Datar , Pratiksha Deshmukh

Auto-bidding is widely used in advertising systems, serving a diverse range of advertisers. Generative bidding is increasingly gaining traction due to its strong planning capabilities and generalizability. Unlike traditional reinforcement…

Machine Learning · Computer Science 2025-08-26 Yunshan Peng , Wenzheng Shu , Jiahao Sun , Yanxiang Zeng , Jinan Pang , Wentao Bai , Yunke Bai , Xialong Liu , Peng Jiang

We propose Q-learning with Adjoint Matching (QAM), a novel TD-based reinforcement learning (RL) algorithm that tackles a long-standing challenge in continuous-action RL: efficient optimization of an expressive diffusion or flow-matching…

Machine Learning · Computer Science 2026-05-20 Qiyang Li , Sergey Levine

Recent advancements in offline reinforcement learning (RL) have underscored the capabilities of Conditional Sequence Modeling (CSM), a paradigm that learns the action distribution based on history trajectory and target returns for each…

Machine Learning · Computer Science 2024-05-28 Shengchao Hu , Ziqing Fan , Chaoqin Huang , Li Shen , Ya Zhang , Yanfeng Wang , Dacheng Tao
‹ Prev 1 2 3 10 Next ›