中文
相关论文

相关论文: Sequential Selection of Correlated Ads by POMDPs

200 篇论文

Solving partially observable Markov decision processes (POMDPs) with high dimensional and continuous observations, such as camera images, is required for many real life robotics and planning problems. Recent researches suggested machine…

人工智能 · 计算机科学 2025-05-27 Idan Lev-Yehudi , Moran Barenboim , Vadim Indelman

We present a major improvement to the incremental pruning algorithm for solving partially observable Markov decision processes. Our technique targets the cross-sum step of the dynamic programming (DP) update, a key source of complexity in…

人工智能 · 计算机科学 2012-07-19 Zhengzhu Feng , Shlomo Zilberstein

Planning under partial observability is an essential capability of autonomous robots. The Partially Observable Markov Decision Process (POMDP) provides a powerful framework for planning under partial observability problems, capturing the…

机器人学 · 计算机科学 2026-03-11 Marcus Hoerger , Muhammad Sudrajat , Hanna Kurniawati

We consider a finite-state partially observable Markov decision problem (POMDP) with an infinite horizon and a discounted cost, and we propose a new method for computing a cost function approximation that is based on features and…

系统与控制 · 电气工程与系统科学 2025-07-08 Yuchao Li , Kim Hammar , Dimitri Bertsekas

Standard value function approaches to finding policies for Partially Observable Markov Decision Processes (POMDPs) are generally considered to be intractable for large models. The intractability of these algorithms is to a large extent a…

人工智能 · 计算机科学 2011-10-05 N. Roy , G. Gordon , S. Thrun

Ad auctions in sponsored search support ``broad match'' that allows an advertiser to target a large number of queries while bidding only on a limited number. While giving more expressiveness to advertisers, this feature makes it challenging…

计算机科学与博弈论 · 计算机科学 2009-01-26 Eyal Even-dar , Yishay Mansour , Vahab Mirrokni , S. Muthukrishnan , Uri Nadav

The real-time bidding (RTB), aka programmatic buying, has recently become the fastest growing area in online advertising. Instead of bulking buying and inventory-centric buying, RTB mimics stock exchanges and utilises computer algorithms to…

计算机科学与博弈论 · 计算机科学 2013-06-28 Shuai Yuan , Jun Wang , Xiaoxue Zhao

Cascading architecture has been widely adopted in large-scale advertising systems to balance efficiency and effectiveness. In this architecture, the pre-ranking model is expected to be a lightweight approximation of the ranking model, which…

信息检索 · 计算机科学 2023-10-10 Zhishan Zhao , Jingyue Gao , Yu Zhang , Shuguang Han , Siyuan Lou , Xiang-Rong Sheng , Zhe Wang , Han Zhu , Yuning Jiang , Jian Xu , Bo Zheng

Partially Observable Markov Decision Processes (POMDPs) offer an elegant framework to model sequential decision making in uncertain environments. Solving POMDPs online is an active area of research and given the size of real-world problems…

人工智能 · 计算机科学 2018-04-10 Sankalp Arora , Sanjiban Choudhury , Sebastian Scherer

Online portfolio selection is an integral componentof wealth management. The fundamental undertaking is tomaximise returns while minimising risk given investor con-straints. We aim to examine and improve modern strategiesto generate higher…

计算工程、金融与科学 · 计算机科学 2021-09-29 Matthew Kruger , Terence L. van Zyl , Andrew Paskaramoorthy

We consider partially observable Markov decision processes (POMDPs) with a set of target states and positive integer costs associated with every transition. The traditional optimization objective (stochastic shortest path) asks to minimize…

人工智能 · 计算机科学 2016-05-12 Tomáš Brázdil , Krishnendu Chatterjee , Martin Chmelík , Anchit Gupta , Petr Novotný

The synthesis problem for partially observable Markov decision processes (POMDPs) is to compute a policy that satisfies a given specification. Such policies have to take the full execution history of a POMDP into account, rendering the…

人工智能 · 计算机科学 2020-07-20 Leonore Winterer , Ralf Wimmer , Nils Jansen , Bernd Becker

In online display advertising, guaranteed contracts and real-time bidding (RTB) are two major ways to sell impressions for a publisher. For large publishers, simultaneously selling impressions through both guaranteed contracts and in-house…

计算机科学与博弈论 · 计算机科学 2022-03-15 Di Wu , Cheng Chen , Xiujun Chen , Junwei Pan , Xun Yang , Qing Tan , Jian Xu , Kuang-Chih Lee

We consider a distributionally robust Partially Observable Markov Decision Process (DR-POMDP), where the distribution of the transition-observation probabilities is unknown at the beginning of each decision period, but their realizations…

最优化与控制 · 数学 2020-12-09 Hideaki Nakao , Ruiwei Jiang , Siqian Shen

As advertisers increasingly shift their budgets toward digital advertising, accurately forecasting advertising costs becomes essential for optimizing marketing campaign returns. This paper presents a comprehensive study that employs various…

机器学习 · 计算机科学 2024-08-22 Fynn Oldenburg , Qiwei Han , Maximilian Kaiser

A standard objective in partially-observable Markov decision processes (POMDPs) is to find a policy that maximizes the expected discounted-sum payoff. However, such policies may still permit unlikely but highly undesirable outcomes, which…

Direct buy advertisers procure advertising inventory at fixed rates from publishers and ad networks. Such advertisers face the complex task of choosing ads amongst myriad new publisher sites. We offer evidence that advertisers do not excel…

综合经济学 · 经济学 2025-07-04 Carl F. Mela , Jason M. T. Roos , Tulio Sousa

We consider a class of sequential decision-making problems under uncertainty that can encompass various types of supervised learning concepts. These problems have a completely observed state process and a partially observed modulation…

最优化与控制 · 数学 2021-08-24 R. Reid Bishop , Chelsea C. White

In recent years, Optimized Cost Per Click (OCPC) and Optimized Cost Per Mille (OCPM) have emerged as the most widely adopted pricing models in the online advertising industry. However, the existing literature has yet to identify the…

计算机科学与博弈论 · 计算机科学 2024-09-06 Kaichen Zhang , Zixuan Yuan , Hui Xiong

Much recent research in decision theoretic planning has adopted Markov decision processes (MDPs) as the model of choice, and has attempted to make their solution more tractable by exploiting problem structure. One particular algorithm,…

人工智能 · 计算机科学 2013-02-08 Craig Boutilier