English
Related papers

Related papers: A Game-Theoretic Analysis of the Empirical Revenue…

200 papers

Empirical risk minimization is the main tool for prediction problems, but its extension to relational data remains unsolved. We solve this problem using recent ideas from graph sampling theory to (i) define an empirical risk for relational…

Machine Learning · Statistics 2019-02-25 Victor Veitch , Morgane Austern , Wenda Zhou , David M. Blei , Peter Orbanz

Recently, reinforcement learning with verifiable rewards (RLVR) has been widely used for enhancing the reasoning abilities of large language models (LLMs). A core challenge in RLVR involves managing the exchange between entropy and…

Computation and Language · Computer Science 2025-08-05 Jia Deng , Jie Chen , Zhipeng Chen , Wayne Xin Zhao , Ji-Rong Wen

We consider the sample complexity of revenue maximization for multiple bidders in unrestricted multi-dimensional settings. Specifically, we study the standard model of $n$ additive bidders whose values for $m$ heterogeneous items are drawn…

Computer Science and Game Theory · Computer Science 2021-04-13 Yannai A. Gonczarowski , S. Matthew Weinberg

We study a setting where agents use no-regret learning algorithms to participate in repeated auctions. \citet{kolumbus2022auctions} showed, rather surprisingly, that when bidders participate in second-price auctions using no-regret bidding…

Computer Science and Game Theory · Computer Science 2024-11-15 Gagan Aggarwal , Anupam Gupta , Andres Perlroth , Grigoris Velegkas

Expectation Maximization (EM) is among the most popular algorithms for estimating parameters of statistical models. However, EM, which is an iterative algorithm based on the maximum likelihood principle, is generally only guaranteed to find…

Statistics Theory · Mathematics 2016-08-30 Ji Xu , Daniel Hsu , Arian Maleki

We introduce the Entropy-Driven Uncertainty Process Reward Model (EDU-PRM), a novel entropy-driven training framework for process reward modeling that enables dynamic, uncertainty-aligned segmentation of complex reasoning steps, eliminating…

Machine Learning · Computer Science 2026-03-10 Lang Cao , Renhong Chen , Yingtian Zou , Chao Peng , Huacong Xu , Yuxian Wang , Wu Ning , Qian Chen , Mofan Peng , Zijie Chen , Peishuo Su , Yitong Li

Optimizing risk-averse objectives in discounted MDPs is challenging because most models do not admit direct dynamic programming equations and require complex history-dependent policies. In this paper, we show that the risk-averse {\em total…

Machine Learning · Computer Science 2025-07-15 Xihong Su , Julien Grand-Clément , Marek Petrik

Designing an incentive compatible auction that maximizes expected revenue is a central problem in Auction Design. Theoretical approaches to the problem have hit some limits in the past decades and analytical solutions are known for only a…

Computer Science and Game Theory · Computer Science 2021-10-26 Jad Rahme , Samy Jelassi , Joan Bruna , S. Matthew Weinberg

This paper studies empirical risk minimization (ERM) problems for large-scale datasets and incorporates the idea of adaptive sample size methods to improve the guaranteed convergence bounds for first-order stochastic and deterministic…

Machine Learning · Computer Science 2017-09-05 Aryan Mokhtari , Alejandro Ribeiro

We study the revenue-maximizing mechanism when a buyer's value evolves endogenously because of learning-by-consuming. A seller sells one unit of a divisible good, while the buyer relies on his private, rough valuation to choose his…

Theoretical Economics · Economics 2022-09-07 Huiyi Guo , Wei He , Bin Liu

We study revenue optimization in a repeated auction between a single seller and a single buyer. Traditionally, the design of repeated auctions requires strong modeling assumptions about the bidder behavior, such as it being myopic, infinite…

Computer Science and Game Theory · Computer Science 2019-03-12 Shipra Agrawal , Constantinos Daskalakis , Vahab Mirrokni , Balasubramanian Sivan

We advance a recently flourishing line of work at the intersection of learning theory and computational economics by studying the learnability of two classes of mechanisms prominent in economics, namely menus of lotteries and two-part…

Computer Science and Game Theory · Computer Science 2024-07-17 Maria-Florina Balcan , Hedyeh Beyhaghi

We propose a multi-agent distributed reinforcement learning algorithm that balances between potentially conflicting short-term reward and sparse, delayed long-term reward, and learns with partial information in a dynamic environment. We…

Machine Learning · Computer Science 2022-04-06 Jing Tan , Ramin Khalili , Holger Karl

Mobile payment such as Alipay has been widely used in our daily lives. To further promote the mobile payment activities, it is important to run marketing campaigns under a limited budget by providing incentives such as coupons, commissions…

Social and Information Networks · Computer Science 2020-03-04 Ziqi Liu , Dong Wang , Qianyu Yu , Zhiqiang Zhang , Yue Shen , Jian Ma , Wenliang Zhong , Jinjie Gu , Jun Zhou , Shuang Yang , Yuan Qi

We consider the problem of bid prediction in repeated auctions and evaluate the performance of econometric methods for learning agents using a dataset from a mainstream sponsored search auction marketplace. Sponsored search auctions is a…

Computer Science and Game Theory · Computer Science 2020-11-02 Gali Noti , Vasilis Syrgkanis

In e-commerce advertising, it is crucial to jointly consider various performance metrics, e.g., user experience, advertiser utility, and platform revenue. Traditional auction mechanisms, such as GSP and VCG auctions, can be suboptimal due…

Computer Science and Game Theory · Computer Science 2021-07-15 Xiangyu Liu , Chuan Yu , Zhilin Zhang , Zhenzhe Zheng , Yu Rong , Hongtao Lv , Da Huo , Yiqing Wang , Dagui Chen , Jian Xu , Fan Wu , Guihai Chen , Xiaoqiang Zhu

Academic research in the field of recommender systems mainly focuses on the problem of maximizing the users' utility by trying to identify the most relevant items for each user. However, such items are not necessarily the ones that maximize…

Information Retrieval · Computer Science 2017-07-26 Dietmar Jannach , Gediminas Adomavicius

A large fraction of online advertisement is sold via repeated second price auctions. In these auctions, the reserve price is the main tool for the auctioneer to boost revenues. In this work, we investigate the following question: Can…

Computer Science and Game Theory · Computer Science 2020-02-19 Yash Kanoria , Hamid Nazerzadeh

Empirical Risk Minimization (ERM) algorithms are widely used in a variety of estimation and prediction tasks in signal-processing and machine learning applications. Despite their popularity, a theory that explains their statistical…

Machine Learning · Statistics 2020-07-07 Hossein Taheri , Ramtin Pedarsani , Christos Thrampoulidis

We consider the fundamental scenario where a single item is to be sold to one of two agents. Both agents draw their valuation for the item from the same probability distribution. However, only one of them submits a bid to the mechanism. The…

Computer Science and Game Theory · Computer Science 2025-08-26 Ioannis Caragiannis , Georgios Kalantzis