中文
相关论文

相关论文: Dynamic Bidding Strategies with Multivariate Feedb…

200 篇论文

The purpose of Inventory Pricing is to bid the right prices to online ad opportunities, which is crucial for a Demand-Side Platform (DSP) to win advertising auctions in Real-Time Bidding (RTB). In the planning stage, advertisers need the…

机器学习 · 计算机科学 2021-10-27 Xu Li , Michelle Ma Zhang , Youjun Tong , Zhenya Wang

Deep reinforcement learning (DRL) agents are often sensitive to visual changes that were unseen in their training environments. To address this problem, we leverage the sequential nature of RL to learn robust representations that encode…

人工智能 · 计算机科学 2022-07-15 Jiameng Fan , Wenchao Li

We present a framework that lets a service provider achieve end-to-end management objectives under varying load. Dynamic control actions are performed by a reinforcement learning (RL) agent. Our work includes experimentation and evaluation…

机器学习 · 计算机科学 2022-10-11 Forough Shahab Samani , Rolf Stadler

Problem definition: Most of the display advertising inventory is sold through real-time auctions. The participants of these auctions are typically bidders (Google, Criteo, RTB House, Trade Desk for instance) who participate on behalf of…

计算机科学与博弈论 · 计算机科学 2023-08-08 Martin Bompaire , Antoine Désir , Benjamin Heymann

Over the past decade, bidding in power markets has attracted widespread attention. Reinforcement Learning (RL) has been widely used for power market bidding as a powerful AI tool to make decisions under real-world uncertainties. However,…

机器学习 · 计算机科学 2024-10-16 Jinyu Liu , Hongye Guo , Yun Li , Qinghu Tang , Fuquan Huang , Tunan Chen , Haiwang Zhong , Qixin Chen

In this paper we investigate the problem of measuring end-to-end Incentive Compatibility (IC) regret given black-box access to an auction mechanism. Our goal is to 1) compute an estimate for IC regret in an auction, 2) provide a measure of…

计算机科学与博弈论 · 计算机科学 2019-06-05 Zhe Feng , Okke Schrijvers , Eric Sodomka

We study an online dynamic pricing problem where the potential demand at each time period $t=1,2,\ldots, T$ is stochastic and dependent on the price. However, a perishable inventory is imposed at the beginning of each time $t$, censoring…

机器学习 · 统计学 2026-01-26 Jianyu Xu , Yining Wang , Xi Chen , Yu-Xiang Wang

Conventional bidding strategies for online display ad auction heavily relies on observed performance indicators such as clicks or conversions. A bidding strategy naively pursuing these easily observable metrics, however, fails to optimize…

机器学习 · 计算机科学 2020-07-10 Daisuke Moriwaki , Yuta Hayakawa , Isshu Munemasa , Yuta Saito , Akira Matsui

Motivated by Carbon Emissions Trading Schemes, Treasury Auctions, Procurement Auctions, and Wholesale Electricity Markets, which all involve the auctioning of homogeneous multiple units, we consider the problem of learning how to bid in…

计算机科学与博弈论 · 计算机科学 2024-11-12 Rigel Galgana , Negin Golrezaei

Deep reinforcement learning (RL) has shown immense potential for learning to control systems through data alone. However, one challenge deep RL faces is that the full state of the system is often not observable. When this is the case, the…

机器学习 · 计算机科学 2023-10-27 Ian Char , Jeff Schneider

A rational behavior of a consumer is analyzed when the user participates in a Peak Time Rebate (PTR) mechanism, which is a demand response (DR) incentive program based on a baseline. A multi-stage stochastic programming is proposed from the…

系统与控制 · 计算机科学 2018-02-23 José Vuelvas , Fredy Ruiz

Effective budget allocation is crucial for optimizing the performance of digital advertising campaigns. However, the development of practical budget allocation algorithms remain limited, primarily due to the lack of public datasets and…

机器学习 · 计算机科学 2025-02-06 Briti Gangopadhyay , Zhao Wang , Alberto Silvio Chiappa , Shingo Takamatsu

When randomness in demand affects the sales of a product, retailers use dynamic pricing strategies to maximize their profits. In this article, we formulate the pricing problem as a continuous-time stochastic optimal control problem and find…

最优化与控制 · 数学 2019-03-13 Asbjørn Nilsen Riseth

In a growing retail electricity market, demand response (DR) is becoming an integral part of the system to enhance economic and operational performances. This is rendered as incentive-based DR (IBDR) in the proposed study. It presents a…

系统与控制 · 电气工程与系统科学 2023-06-02 Vipin Chandra Pandey , Nikhil Gupta , Khaleequr Rehman Niazi , Anil Swarnkar , Tanuj Rawat , Charalambos Konstantinou

There is growing interest in the use of grid-level storage to smooth variations in supply that are likely to arise with increased use of wind and solar energy. Energy arbitrage, the process of buying, storing, and selling electricity to…

最优化与控制 · 数学 2015-09-01 Daniel R. Jiang , Warren B. Powell

Scale-calibrated ranking systems are ubiquitous in real-world applications nowadays, which pursue accurate ranking quality and calibrated probabilistic predictions simultaneously. For instance, in the advertising ranking system, the…

信息检索 · 计算机科学 2024-06-13 Shunyu Zhang , Hu Liu , Wentian Bao , Enyun Yu , Yang Song

Uncertainties in renewable generation and demand dynamics challenge day-ahead scheduling. To enhance renewable penetration and maintain intra-day balance, we develop a multi-agent reinforcement learning framework for self-interested…

多智能体系统 · 计算机科学 2026-04-13 Junhao Ren , Honglin Gao , Lan Zhao , Qiyu Kang , Gaoxi Xiao , Yajuan Sun

We study a finite-horizon restless multi-armed bandit problem with multiple actions, dubbed R(MA)^2B. The state of each arm evolves according to a controlled Markov decision process (MDP), and the reward of pulling an arm depends on both…

机器学习 · 计算机科学 2022-03-25 Guojun Xiong , Jian Li , Rahul Singh

We consider dynamic pricing with covariates under a generalized linear demand model: a seller can dynamically adjust the price of a product over a horizon of $T$ time periods, and at each time period $t$, the demand of the product is…

机器学习 · 计算机科学 2023-11-14 Hanzhao Wang , Kalyan Talluri , Xiaocheng Li

We study Proportional Response Dynamics (PRD) in linear Fisher markets where participants act asynchronously. We model this scenario as a sequential process in which in every step, an adversary selects a subset of the players that will…

计算机科学与博弈论 · 计算机科学 2024-01-17 Yoav Kolumbus , Menahem Levy , Noam Nisan
‹ 上一页 1 8 9 10 下一页 ›