中文
相关论文

相关论文: RoxyBot-06: Stochastic Prediction and Optimization…

200 篇论文

Agents that learn to select optimal actions represent a prominent focus of the sequential decision-making literature. In the face of a complex environment or constraints on time and resources, however, aiming to synthesize such an optimal…

机器学习 · 计算机科学 2021-06-23 Dilip Arumugam , Benjamin Van Roy

Automated negotiation is a rising topic in Artificial Intelligence research. Monte Carlo methods have got increasing interest, in particular since they have been used with success on games with high branching factor such as go.In this…

多智能体系统 · 计算机科学 2018-10-17 Cédric Buron , Zahia Guessoum , Sylvain Ductor , Olivier Roussel

We construct prior-free auctions with constant-factor approximation guarantees with ordered bidders, in both unlimited and limited supply settings. We compare the expected revenue of our auctions on a bid vector to the monotone price…

计算机科学与博弈论 · 计算机科学 2012-12-13 Elias Koutsoupias , Stefano Leonardi , Tim Roughgarden

We study revenue maximization in settings where agents' values are interdependent: each agent receives a signal drawn from a correlated distribution and agents' values are functions of all of the signals. We introduce a variant of the…

计算机科学与博弈论 · 计算机科学 2014-08-20 Shuchi Chawla , Hu Fu , Anna Karlin

We introduce a risk-aware multi-objective Traveling Salesperson Problem (TSP) variant, where the robot tour cost and tour reward have to be optimized simultaneously. The robot obtains reward along the edges in the graph. We study the case…

机器人学 · 计算机科学 2021-09-22 Rishab Balasubramanian , Lifeng Zhou , Pratap Tokekar , P. B. Sujit

This paper reports on the first international competition on AI for the traveling salesman problem (TSP) at the International Joint Conference on Artificial Intelligence 2021 (IJCAI-21). The TSP is one of the classical combinatorial…

Autonomous crypto trading systems often spend most of their design effort on finding entries, while exits are left to fixed rules that are rarely tested in a systematic way. This paper examines whether better stop-loss and take-profit…

人工智能 · 计算机科学 2026-05-01 Nathan Li , Aikins Laryea , Yigit Ihlamur

We present a novel model for capturing the behavior of an agent exhibiting sunk-cost bias in a stochastic environment. Agents exhibiting sunk-cost bias take into account the effort they have already spent on an endeavor when they evaluate…

计算机科学与博弈论 · 计算机科学 2021-06-22 Jon Kleinberg , Sigal Oren , Manish Raghavan , Nadav Sklar

We investigate approximately optimal mechanisms in settings where bidders' utility functions are non-linear; specifically, convex, with respect to payments (such settings arise, for instance, in procurement auctions for energy). We provide…

计算机科学与博弈论 · 计算机科学 2017-02-23 Amy Greenwald , Takehiro Oyakawa , Vasilis Syrgkanis

Online auctions have expanded rapidly over the last decade and have become a fascinating new type of business or commercial transaction in this digital era. Here we introduce a master equation for the bidding process that takes place in…

物理与社会 · 物理学 2009-11-11 I. Yang , B. Kahng

Maximizing utility with a budget constraint is the primary goal for advertisers in real-time bidding (RTB) systems. The policy maximizing the utility is referred to as the optimal bidding strategy. Earlier works on optimal bidding strategy…

机器学习 · 计算机科学 2020-04-02 Aritra Ghosh , Saayan Mitra , Somdeb Sarkhel , Viswanathan Swaminathan

We propose a new architecture to approximately learn incentive compatible, revenue-maximizing auctions from sampled valuations. Our architecture uses the Sinkhorn algorithm to perform a differentiable bipartite matching which allows the…

计算机科学与博弈论 · 计算机科学 2021-06-16 Michael J. Curry , Uro Lyi , Tom Goldstein , John Dickerson

We consider a model of a data broker selling information to a single agent to maximize his revenue. The agent has a private valuation of the additional information, and upon receiving the signal from the data broker, the agent can conduct…

理论经济学 · 经济学 2023-08-08 Yingkai Li

We provide a characterization of revenue-optimal dynamic mechanisms in settings where a monopolist sells k items over k periods to a buyer who realizes his value for item i in the beginning of period i. We require that the mechanism…

计算机科学与博弈论 · 计算机科学 2016-07-06 Itai Ashlagi , Constantinos Daskalakis , Nima Haghpanah

Data-driven simulation has become a favorable way to train and test autonomous driving algorithms. The idea of replacing the actual environment with a learned simulator has also been explored in model-based reinforcement learning in the…

机器人学 · 计算机科学 2023-09-29 Zhejun Zhang , Alexander Liniger , Dengxin Dai , Fisher Yu , Luc Van Gool

A recent approach to automated mechanism design, differentiable economics, represents auctions by rich function approximators and optimizes their performance by gradient descent. The ideal auction architecture for differentiable economics…

计算机科学与博弈论 · 计算机科学 2022-02-08 Michael Curry , Tuomas Sandholm , John Dickerson

In the matroid buyback problem, an algorithm observes a sequence of bids and must decide whether to accept each bid at the moment it arrives, subject to a matroid constraint on the set of accepted bids. Decisions to reject bids are…

计算机科学与博弈论 · 计算机科学 2009-11-30 Ashwinkumar B. V. , Robert Kleinberg

We introduce a new Self-Organized Criticality (SOC) model for simulating price evolution in an artificial financial market, based on a multilayer network of traders. The model also implements, in a quite realistic way with respect to…

交易与市场微观结构 · 定量金融 2016-06-30 Alessio Emanuele Biondo , Alessandro Pluchino , Andrea Rapisarda

This paper describe an hybrid agent trained to play in Fantasy Football AI which participated in the Bot Bowl III competition. The agent, MimicBot, is implemented using a specifically designed deep policy network and trained using a…

人工智能 · 计算机科学 2021-08-24 Nicola Pezzotti

In this work, we investigate the online learning problem of revenue maximization in ad auctions, where the seller needs to learn the click-through rates (CTRs) of each ad candidate and charge the price of the winner through a pay-per-click…

信息检索 · 计算机科学 2024-03-05 Zhe Feng , Christopher Liaw , Zixin Zhou