中文
相关论文

相关论文: A Bi-level Decision Framework for Incentive-Based …

200 篇论文

This paper studies a multi-period demand response management problem in the smart grid where multiple utility companies compete among themselves. The user-utility interactions are modeled by a noncooperative game of a Stackelberg type where…

最优化与控制 · 数学 2016-11-17 Khaled Alshehri , Ji Liu , Xudong Chen , Tamer Başar

The rapid progression of sophisticated advance metering infrastructure (AMI), allows us to have a better understanding and data from demand-response (DR) solutions. There are vast amounts of research on the internet of things and its…

信号处理 · 电气工程与系统科学 2019-08-09 Ramin Faraji Fijani , Behrouz Azimian , Ehsan Ghotbi , Xingwu Wang

To address the needs of modeling uncertainty in sensitive machine learning applications, the setup of distributionally robust optimization (DRO) seeks good performance uniformly across a variety of tasks. The recent multi-distribution…

机器学习 · 统计学 2026-01-01 Rafael Hanashiro , Patrick Jaillet

Intelligent wireless networks have long been expected to have self-configuration and self-optimization capabilities to adapt to various environments and demands. In this paper, we develop a novel distributed hierarchical deep reinforcement…

信号处理 · 电气工程与系统科学 2023-12-06 Kaiwen Yu , Chonghao Zhao , Gang Wu , Geoffrey Ye Li

We study the dynamic pricing and replenishment problems under inconsistent decision frequencies. Different from the traditional demand assumption, the discreteness of demand and the parameter within the Poisson distribution as a function of…

机器学习 · 计算机科学 2024-10-29 Yi Zheng , Zehao Li , Peng Jiang , Yijie Peng

The problem of dynamic pricing of electricity in a retail market is considered. A Stackelberg game is used to model interactions between a retailer and its customers; the retailer sets the day-ahead hourly price of electricity and consumers…

最优化与控制 · 数学 2016-03-01 Liyan Jia , Lang Tong

One of the major issues with the integration of renewable energy sources into the power grid is the increased uncertainty and variability that they bring. If this uncertainty is not sufficiently addressed, it will limit the further…

最优化与控制 · 数学 2017-05-15 Joshua Comden , Zhenhua Liu , Yue Zhao

The Stackelberg game depicts a leader-follower relationship wherein decisions are made sequentially, and the Stackelberg equilibrium represents an expected optimal solution when the leader can anticipate the rational response of the…

系统与控制 · 电气工程与系统科学 2024-01-17 Yue Chen , Peng Yi

Automated design of multi-agent interactions with desirable equilibrium outcomes is inherently difficult due to the computational hardness, non-uniqueness, and instability of the resulting equilibria. In this work, we propose the use of…

计算机科学与博弈论 · 计算机科学 2026-03-13 Vinzenz Thoma , Georgios Piliouras , Luke Marris

Demand-side management (DSM) enables distribution system operators (DSOs) to steer electricity consumption through dynamic price signals or incentive mechanisms, thereby leveraging end-users' flexibility potential for delivering grid…

最优化与控制 · 数学 2026-05-04 Silvia Cianchi , Reza Rahimi Baghbadorani , Anibal Sanjab , Sergio Grammatico

Many strategic decision-making problems, such as environment design for warehouse robots, can be naturally formulated as bi-level reinforcement learning (RL), where a leader agent optimizes its objective while a follower solves a Markov…

机器学习 · 计算机科学 2026-04-01 Mikoto Kudo , Takumi Tanabe , Akifumi Wachi , Youhei Akimoto

We design an optimal contract between a demand response aggregator (DRA) and a customer for incentive-based demand response. We consider a setting in which the customer is asked to reduce her consumption by the DRA and she is compensated…

最优化与控制 · 数学 2024-10-30 Donya G. Dobakhshari , Vijay Gupta

The electric vehicle routing problem with time windows (EVRPTW) is a complex optimization problem in sustainable logistics, where routing decisions must minimize total travel distance, fleet size, and battery usage while satisfying strict…

机器学习 · 计算机科学 2026-01-22 Mertcan Daysalilar , Fuat Uyguroglu , Gabriel Nicolosi , Adam Meyers

We study a pessimistic stochastic bilevel program in the context of sequential two-player games, where the leader makes a binary here-and-now decision, and the follower responds a continuous wait-and-see decision after observing the…

最优化与控制 · 数学 2022-06-09 Akshit Goyal , Yiling Zhang , Chuan He

This paper explores an idea of demand-supply balance for smart grids in which consumers are expected to play a significant role. The main objective is to motivate the consumer, by maximizing their benefit both as a seller and a buyer, to…

计算机科学与博弈论 · 计算机科学 2013-04-04 Wayes Tushar , Jian A. Zhang , David B. Smith , Sylvie Thiebaux , H. Vincent Poor

The Pickup and Delivery Problem (PDP) is a fundamental and challenging variant of the Vehicle Routing Problem, characterized by tightly coupled pickup--delivery pairs, precedence constraints, and spatial layouts that often exhibit…

机器学习 · 计算机科学 2026-03-12 Wentao Wang , Lifeng Han , Guangyu Zou

We study operations of a battery energy storage system under a baseline-based demand response (DR) program with an uncertain schedule of DR events. Baseline-based DR programs may provide undesired incentives to inflate baseline consumption…

系统与控制 · 电气工程与系统科学 2019-09-30 Douglas Ellman , Yuanzhang Xiao

This paper studies a class of dynamic Stackelberg games under open-loop information structure with constrained linear agent dynamics and quadratic utility functions. We show two important properties for this class of dynamic Stackelberg…

最优化与控制 · 数学 2016-08-09 Sen Li , Wei Zhang , Jianming Lian , Karanjit Kalsi

Recommendation is crucial in both academia and industry, and various techniques are proposed such as content-based collaborative filtering, matrix factorization, logistic regression, factorization machines, neural networks and multi-armed…

信息检索 · 计算机科学 2019-10-30 Feng Liu , Ruiming Tang , Xutao Li , Weinan Zhang , Yunming Ye , Haokun Chen , Huifeng Guo , Yuzhou Zhang

A rational behavior of a consumer is analyzed when the user participates in a Peak Time Rebate (PTR) mechanism, which is a demand response (DR) incentive program based on a baseline. A multi-stage stochastic programming is proposed from the…

系统与控制 · 计算机科学 2018-02-23 José Vuelvas , Fredy Ruiz