中文
相关论文

相关论文: Online Learning for Dynamic Vickrey-Clarke-Groves …

200 篇论文

To address the exponentially increasing data rate demands of end users, necessitates efficient spectrum allocation among co-existing operators in licensed and unlicensed spectrum bands to cater to the temporal and spatial variations of…

计算机科学与博弈论 · 计算机科学 2020-01-22 Indu Yadav , Ankur A. Kulkarni , Abhay Karandikar

Control reserves are power generation or consumption entities that ensure balance of supply and demand of electricity in real-time. In many countries, they are operated through a market mechanism in which entities provide bids. The system…

计算机科学与博弈论 · 计算机科学 2016-11-22 Pier Giuseppe Sessa , Neil Walton , Maryam Kamgarpour

We study the online learning problem of a bidder who participates in repeated auctions. With the goal of maximizing his T-period payoff, the bidder determines the optimal allocation of his budget among his bids for $K$ goods at each period.…

计算机科学与博弈论 · 计算机科学 2017-11-20 Sevi Baltaoglu , Lang Tong , Qing Zhao

We study learning in periodic Markov Decision Process(MDP), a special type of non-stationary MDP where both the state transition probabilities and reward functions vary periodically, under the average reward maximization setting. We…

机器学习 · 计算机科学 2022-07-26 Ayush Aniket , Arpan Chattopadhyay

Online auctions are one of the most fundamental facets of the modern economy and power an industry generating hundreds of billions of dollars a year in revenue. Auction theory has historically focused on the question of designing the best…

计算机科学与博弈论 · 计算机科学 2021-09-23 Thomas Nedelec , Clément Calauzènes , Noureddine El Karoui , Vianney Perchet

This paper addresses the problem of online learning in a dynamic setting. We consider a social network in which each individual observes a private signal about the underlying state of the world and communicates with her neighbors at each…

最优化与控制 · 数学 2013-10-02 Shahin Shahrampour , Alexander Rakhlin , Ali Jadbabaie

This paper investigates reverse auctions that involve continuous values of different types of goods, general nonconvex constraints, and second stage costs. We seek to design the payment rules and conditions under which coalitions of…

计算机科学与博弈论 · 计算机科学 2021-07-14 Orcun Karaca , Pier Giuseppe Sessa , Neil Walton , Maryam Kamgarpour

In display advertising, a small group of sellers and bidders face each other in up to 10 12 auctions a day. In this context, revenue maximisation via monopoly price learning is a high-value problem for sellers. By nature, these auctions are…

机器学习 · 计算机科学 2020-10-21 Lorenzo Croissant , Marc Abeille , Clément Calauzènes

Online auction has been very widespread in the recent years. Platform administrators are working hard to refine their auction mechanisms that will generate high profits while maintaining a fair resource allocation. With the advancement of…

计算机科学与博弈论 · 计算机科学 2021-10-14 Zhanhao Zhang

Sponsored search auctions constitute one of the most successful applications of microeconomic mechanisms. In mechanism design, auctions are usually designed to incentivize advertisers to bid their truthful valuations and to assure both the…

计算机科学与博弈论 · 计算机科学 2014-05-13 Nicola Gatti , Alessandro Lazaric , Marco Rocco , Francesco Trovò

We study resource allocation problems in which a central planner allocates resources among strategic agents with private cost functions in order to minimize a social cost, defined as an aggregate of the agents' costs. This setting poses two…

计算机科学与博弈论 · 计算机科学 2026-03-19 Leo Landolt , Anna Maddux , Andreas Schlaginhaufen , Saurabh Vaishampayan , Maryam Kamgarpour

General purpose intelligent learning agents cycle through (complex,non-MDP) sequences of observations, actions, and rewards. On the other hand, reinforcement learning is well-developed for small finite state Markov Decision Processes…

人工智能 · 计算机科学 2009-12-30 Marcus Hutter

We study learning in periodic Markov Decision Process (MDP), a special type of non-stationary MDP where both the state transition probabilities and reward functions vary periodically, under the average reward maximization setting. We…

机器学习 · 计算机科学 2023-03-20 Ayush Aniket , Arpan Chattopadhyay

We consider an online matching problem with concave returns. This problem is a significant generalization of the Adwords allocation problem and has vast applications in online advertising. In this problem, a sequence of items arrive…

数据结构与算法 · 计算机科学 2015-06-09 Xiao Alison Chen , Zizhuo Wang

In e-commerce advertising, it is crucial to jointly consider various performance metrics, e.g., user experience, advertiser utility, and platform revenue. Traditional auction mechanisms, such as GSP and VCG auctions, can be suboptimal due…

计算机科学与博弈论 · 计算机科学 2021-07-15 Xiangyu Liu , Chuan Yu , Zhilin Zhang , Zhenzhe Zheng , Yu Rong , Hongtao Lv , Da Huo , Yiqing Wang , Dagui Chen , Jian Xu , Fan Wu , Guihai Chen , Xiaoqiang Zhu

Autonomous robots operating in complex, unstructured environments face significant challenges due to latent, unobserved factors that obscure their understanding of both their internal state and the external world. Addressing this challenge…

机器人学 · 计算机科学 2026-04-02 Alejandro Murillo-Gonzalez , Lantao Liu

The Vickrey-Clarke-Groves (VCG) mechanism is infamously revenue non-monotone in combinatorial auctions. I.e., when a buyer increases their value for a bundle of items, the total auction revenue may decrease. Combinatorial auctions exhibit…

理论经济学 · 经济学 2026-02-25 Jason Hartline

We consider the problem of learning optimal reserve price in repeated auctions against non-myopic bidders, who may bid strategically in order to gain in future rounds even if the single-round auctions are truthful. Previous algorithms,…

计算机科学与博弈论 · 计算机科学 2018-05-01 Zhiyi Huang , Jinyan Liu , Xiangning Wang

We study the problem of learning to bid when the bidder's value is dynamic, i.e., when the current value depends on past outcomes. Specifically, we consider a bidder participating in repeated second-price auctions whose value depends on the…

机器学习 · 计算机科学 2026-05-28 Benjamin Heymann , Otmane Sakhi

Online learning algorithms are designed to perform in non-stationary environments, but generally there is no notion of a dynamic state to model constraints on current and future actions as a function of past actions. State-based models are…

机器学习 · 计算机科学 2015-09-01 Peng Guan , Maxim Raginsky , Rebecca Willett