中文
相关论文

相关论文: A Reinforcement Learning Approach in Multi-Phase S…

200 篇论文

Market makers play an important role in providing liquidity to markets by continuously quoting prices at which they are willing to buy and sell, and managing inventory risk. In this paper, we build a multi-agent simulation of a dealer…

交易与市场微观结构 · 定量金融 2019-11-15 Sumitra Ganesh , Nelson Vadori , Mengda Xu , Hua Zheng , Prashant Reddy , Manuela Veloso

Maximizing utility with a budget constraint is the primary goal for advertisers in real-time bidding (RTB) systems. The policy maximizing the utility is referred to as the optimal bidding strategy. Earlier works on optimal bidding strategy…

机器学习 · 计算机科学 2020-04-02 Aritra Ghosh , Saayan Mitra , Somdeb Sarkhel , Viswanathan Swaminathan

We study the effect of persistence of engagement on learning in a stochastic multi-armed bandit setting. In advertising and recommendation systems, repetition effect includes a wear-in period, where the user's propensity to reward the…

机器学习 · 计算机科学 2020-06-19 Priyank Agrawal , Theja Tulabandhula

We introduce a new numerical framework to learn optimal bidding strategies in repeated auctions when the seller uses past bids to optimize her mechanism. Crucially, we do not assume that the bidders know what optimization mechanism is used…

计算机科学与博弈论 · 计算机科学 2021-02-09 Thomas Nedelec , Jules Baudet , Vianney Perchet , Noureddine El Karoui

In this paper, we study how a budget-constrained bidder should learn to bid adaptively in repeated first-price auctions to maximize cumulative payoff. This problem arises from the recent industry-wide shift from second-price auctions to…

计算机科学与博弈论 · 计算机科学 2026-04-14 Yige Wang , Jiashuo Jiang

The proliferation of the Internet has led to the emergence of online advertising, driven by the mechanics of online auctions. In these repeated auctions, software agents participate on behalf of aggregated advertisers to optimize for their…

机器学习 · 计算机科学 2023-06-13 Haozhe Wang , Chao Du , Panyan Fang , Li He , Liang Wang , Bo Zheng

Fairness plays a crucial role in various multi-agent systems (e.g., communication networks, financial markets, etc.). Many multi-agent dynamical interactions can be cast as Markov Decision Processes (MDPs). While existing research has…

机器学习 · 计算机科学 2023-06-02 Peizhong Ju , Arnob Ghosh , Ness B. Shroff

In this work, we investigate the online learning problem of revenue maximization in ad auctions, where the seller needs to learn the click-through rates (CTRs) of each ad candidate and charge the price of the winner through a pay-per-click…

信息检索 · 计算机科学 2024-03-05 Zhe Feng , Christopher Liaw , Zixin Zhou

We present a new recommendation setting for picking out two items from a given set to be highlighted to a user, based on contextual input. These two items are presented to a user who chooses one of them, possibly stochastically, with a bias…

机器学习 · 计算机科学 2016-01-26 Daniel Barsky , Koby Crammer

We consider revenue maximization in online auction/pricing problems. A seller sells an identical item in each period to a new buyer, or a new set of buyers. For the online posted pricing problem, we show regret bounds that scale with the…

计算机科学与博弈论 · 计算机科学 2018-09-13 Sébastien Bubeck , Nikhil R. Devanur , Zhiyi Huang , Rad Niazadeh

We consider reinforcement learning (RL) in episodic Markov decision processes (MDPs) with linear function approximation under drifting environment. Specifically, both the reward and state transition functions can evolve over time but their…

机器学习 · 计算机科学 2024-04-16 Huozhi Zhou , Jinglin Chen , Lav R. Varshney , Ashish Jagmohan

Automated bidding to optimize online advertising with various constraints, e.g. ROI constraints and budget constraints, is widely adopted by advertisers. A key challenge lies in designing algorithms for non-truthful mechanisms with ROI…

计算机科学与博弈论 · 计算机科学 2025-10-21 Yuan Deng , Yilin Li , Wei Tang , Hanrui Zhang

Real-time bidding (RTB) based display advertising has become one of the key technological advances in computational advertising. RTB enables advertisers to buy individual ad impressions via an auction in real-time and facilitates the…

计算机科学与博弈论 · 计算机科学 2018-03-13 Kan Ren , Weinan Zhang , Ke Chang , Yifei Rong , Yong Yu , Jun Wang

The standard framework of online bidding algorithm design assumes that the seller commits himself to faithfully implementing the rules of the adopted auction. However, the seller may attempt to cheat in execution to increase his revenue if…

计算机科学与博弈论 · 计算机科学 2023-11-28 Qian Wang , Xuanzhi Xia , Zongjun Yang , Xiaotie Deng , Yuqing Kong , Zhilin Zhang , Liang Wang , Chuan Yu , Jian Xu , Bo Zheng

In a day-ahead market, energy buyers and sellers submit their bids for a particular future time, including the amount of energy they wish to buy or sell and the price they are prepared to pay or receive. However, the dynamic for forming the…

最优化与控制 · 数学 2024-11-26 Luca Di Persio , Matteo Garbelli , Luca M. Giordano

We study the problem of a buyer (aka auctioneer) who gains stochastic rewards by procuring multiple units of a service or item from a pool of heterogeneous strategic agents. The reward obtained for a single unit from an allocated agent…

计算机科学与博弈论 · 计算机科学 2015-04-30 Satyanath Bhat , Shweta Jain , Sujit Gujar , Y. Narahari

In order to make good decision under uncertainty an agent must learn from observations. To do so, two of the most common frameworks are Contextual Bandits and Markov Decision Processes (MDPs). In this paper, we study whether there exist…

机器学习 · 计算机科学 2019-11-05 Andrea Zanette , Emma Brunskill

We consider the problem of a single seller repeatedly selling a single item to a single buyer (specifically, the buyer has a value drawn fresh from known distribution $D$ in every round). Prior work assumes that the buyer is fully rational…

计算机科学与博弈论 · 计算机科学 2017-11-28 Mark Braverman , Jieming Mao , Jon Schneider , S. Matthew Weinberg

Reinforcement learning has been widely applied in automated bidding. Traditional approaches model bidding as a Markov Decision Process (MDP). Recently, some studies have explored using generative reinforcement learning methods to address…

机器学习 · 计算机科学 2025-07-23 Kaiyuan Li , Pengyu Wang , Yunshan Peng , Pengjia Yuan , Yanxiang Zeng , Rui Xiang , Yanhua Cheng , Xialong Liu , Peng Jiang

First-price auctions have very recently swept the online advertising industry, replacing second-price auctions as the predominant auction mechanism on many platforms. This shift has brought forth important challenges for a bidder: how…

机器学习 · 计算机科学 2025-09-26 Yanjun Han , Zhengyuan Zhou , Aaron Flores , Erik Ordentlich , Tsachy Weissman