中文
相关论文

相关论文: A Game-Theoretic Analysis of the Empirical Revenue…

200 篇论文

Empirical risk minimization is the main tool for prediction problems, but its extension to relational data remains unsolved. We solve this problem using recent ideas from graph sampling theory to (i) define an empirical risk for relational…

机器学习 · 统计学 2019-02-25 Victor Veitch , Morgane Austern , Wenda Zhou , David M. Blei , Peter Orbanz

Recently, reinforcement learning with verifiable rewards (RLVR) has been widely used for enhancing the reasoning abilities of large language models (LLMs). A core challenge in RLVR involves managing the exchange between entropy and…

计算与语言 · 计算机科学 2025-08-05 Jia Deng , Jie Chen , Zhipeng Chen , Wayne Xin Zhao , Ji-Rong Wen

We consider the sample complexity of revenue maximization for multiple bidders in unrestricted multi-dimensional settings. Specifically, we study the standard model of $n$ additive bidders whose values for $m$ heterogeneous items are drawn…

计算机科学与博弈论 · 计算机科学 2021-04-13 Yannai A. Gonczarowski , S. Matthew Weinberg

We study a setting where agents use no-regret learning algorithms to participate in repeated auctions. \citet{kolumbus2022auctions} showed, rather surprisingly, that when bidders participate in second-price auctions using no-regret bidding…

计算机科学与博弈论 · 计算机科学 2024-11-15 Gagan Aggarwal , Anupam Gupta , Andres Perlroth , Grigoris Velegkas

Expectation Maximization (EM) is among the most popular algorithms for estimating parameters of statistical models. However, EM, which is an iterative algorithm based on the maximum likelihood principle, is generally only guaranteed to find…

统计理论 · 数学 2016-08-30 Ji Xu , Daniel Hsu , Arian Maleki

We introduce the Entropy-Driven Uncertainty Process Reward Model (EDU-PRM), a novel entropy-driven training framework for process reward modeling that enables dynamic, uncertainty-aligned segmentation of complex reasoning steps, eliminating…

机器学习 · 计算机科学 2026-03-10 Lang Cao , Renhong Chen , Yingtian Zou , Chao Peng , Huacong Xu , Yuxian Wang , Wu Ning , Qian Chen , Mofan Peng , Zijie Chen , Peishuo Su , Yitong Li

Optimizing risk-averse objectives in discounted MDPs is challenging because most models do not admit direct dynamic programming equations and require complex history-dependent policies. In this paper, we show that the risk-averse {\em total…

机器学习 · 计算机科学 2025-07-15 Xihong Su , Julien Grand-Clément , Marek Petrik

Designing an incentive compatible auction that maximizes expected revenue is a central problem in Auction Design. Theoretical approaches to the problem have hit some limits in the past decades and analytical solutions are known for only a…

计算机科学与博弈论 · 计算机科学 2021-10-26 Jad Rahme , Samy Jelassi , Joan Bruna , S. Matthew Weinberg

This paper studies empirical risk minimization (ERM) problems for large-scale datasets and incorporates the idea of adaptive sample size methods to improve the guaranteed convergence bounds for first-order stochastic and deterministic…

机器学习 · 计算机科学 2017-09-05 Aryan Mokhtari , Alejandro Ribeiro

We study the revenue-maximizing mechanism when a buyer's value evolves endogenously because of learning-by-consuming. A seller sells one unit of a divisible good, while the buyer relies on his private, rough valuation to choose his…

理论经济学 · 经济学 2022-09-07 Huiyi Guo , Wei He , Bin Liu

We study revenue optimization in a repeated auction between a single seller and a single buyer. Traditionally, the design of repeated auctions requires strong modeling assumptions about the bidder behavior, such as it being myopic, infinite…

计算机科学与博弈论 · 计算机科学 2019-03-12 Shipra Agrawal , Constantinos Daskalakis , Vahab Mirrokni , Balasubramanian Sivan

We advance a recently flourishing line of work at the intersection of learning theory and computational economics by studying the learnability of two classes of mechanisms prominent in economics, namely menus of lotteries and two-part…

计算机科学与博弈论 · 计算机科学 2024-07-17 Maria-Florina Balcan , Hedyeh Beyhaghi

We propose a multi-agent distributed reinforcement learning algorithm that balances between potentially conflicting short-term reward and sparse, delayed long-term reward, and learns with partial information in a dynamic environment. We…

机器学习 · 计算机科学 2022-04-06 Jing Tan , Ramin Khalili , Holger Karl

Mobile payment such as Alipay has been widely used in our daily lives. To further promote the mobile payment activities, it is important to run marketing campaigns under a limited budget by providing incentives such as coupons, commissions…

社会与信息网络 · 计算机科学 2020-03-04 Ziqi Liu , Dong Wang , Qianyu Yu , Zhiqiang Zhang , Yue Shen , Jian Ma , Wenliang Zhong , Jinjie Gu , Jun Zhou , Shuang Yang , Yuan Qi

We consider the problem of bid prediction in repeated auctions and evaluate the performance of econometric methods for learning agents using a dataset from a mainstream sponsored search auction marketplace. Sponsored search auctions is a…

计算机科学与博弈论 · 计算机科学 2020-11-02 Gali Noti , Vasilis Syrgkanis

In e-commerce advertising, it is crucial to jointly consider various performance metrics, e.g., user experience, advertiser utility, and platform revenue. Traditional auction mechanisms, such as GSP and VCG auctions, can be suboptimal due…

计算机科学与博弈论 · 计算机科学 2021-07-15 Xiangyu Liu , Chuan Yu , Zhilin Zhang , Zhenzhe Zheng , Yu Rong , Hongtao Lv , Da Huo , Yiqing Wang , Dagui Chen , Jian Xu , Fan Wu , Guihai Chen , Xiaoqiang Zhu

Academic research in the field of recommender systems mainly focuses on the problem of maximizing the users' utility by trying to identify the most relevant items for each user. However, such items are not necessarily the ones that maximize…

信息检索 · 计算机科学 2017-07-26 Dietmar Jannach , Gediminas Adomavicius

A large fraction of online advertisement is sold via repeated second price auctions. In these auctions, the reserve price is the main tool for the auctioneer to boost revenues. In this work, we investigate the following question: Can…

计算机科学与博弈论 · 计算机科学 2020-02-19 Yash Kanoria , Hamid Nazerzadeh

Empirical Risk Minimization (ERM) algorithms are widely used in a variety of estimation and prediction tasks in signal-processing and machine learning applications. Despite their popularity, a theory that explains their statistical…

机器学习 · 统计学 2020-07-07 Hossein Taheri , Ramtin Pedarsani , Christos Thrampoulidis

We consider the fundamental scenario where a single item is to be sold to one of two agents. Both agents draw their valuation for the item from the same probability distribution. However, only one of them submits a bid to the mechanism. The…

计算机科学与博弈论 · 计算机科学 2025-08-26 Ioannis Caragiannis , Georgios Kalantzis