中文
相关论文

相关论文: A Minimal Incentive-based Demand Response Program …

200 篇论文

This paper introduces a novel deep learning-based user-side feedback reduction framework, termed self-nomination. The goal of self-nomination is to reduce the number of users (UEs) feeding back channel state information (CSI) to the base…

信号处理 · 电气工程与系统科学 2025-04-24 Juseong Park , Foad Sohrabi , Jinfeng Du , Jeffrey G. Andrews

In this study, we apply reinforcement learning techniques and propose what we call reinforcement mechanism design to tackle the dynamic pricing problem in sponsored search auctions. In contrast to previous game-theoretical approaches that…

计算机科学与博弈论 · 计算机科学 2017-11-29 Weiran Shen , Binghui Peng , Hanpeng Liu , Michael Zhang , Ruohan Qian , Yan Hong , Zhi Guo , Zongyao Ding , Pengjun Lu , Pingzhong Tang

Reliable demand forecasts are critical for the effective supply chain management. Several endogenous and exogenous variables can influence the dynamics of demand, and hence a single statistical model that only consists of historical sales…

应用统计 · 统计学 2019-09-09 Mahdi Abolghasemi , Ali Eshragh , Jason Hurley , Behnam Fahimnia

The classic Dial-A-Ride Problem (DARP) aims at designing the minimum-cost routing that accommodates a set of user requests under constraints at an operations planning level, where users' preferences and revenue management are often…

最优化与控制 · 数学 2020-11-19 Xiaotong Dong , Joseph YJ Chow , S Travis Waller , David Rey

Unsupervised speech emotion recognition (SER) focuses on addressing the problem of data sparsity and annotation bias of emotional speech. Reinforcement learning (RL) is a promising method which enhances the performance through rule-based or…

音频与语音处理 · 电气工程与系统科学 2026-02-09 Yingying Gao , Shilei Zhang , Runyan Yang , Zihao Cui , Junlan Feng

Direct Preference Optimization (DPO) have emerged as a popular method for aligning Large Language Models (LLMs) with human preferences. While DPO effectively preserves the relative ordering between chosen and rejected responses through…

计算与语言 · 计算机科学 2025-06-05 Lin Sun , Chuang Liu , Peng Liu , Bingyang Li , Weijia Lu , Ning Wu

Demand response has been implemented by distribution system operators to reduce peak demand and mitigate contingency issues on distribution lines and substations. Specifically, the campus based commercial buildings make the major…

最优化与控制 · 数学 2018-12-21 Zheming Liang , Desong Bian , Xiaohu Zhang , Di Shi , Ruisheng Diao , Zhiwei Wang

Preference-based reinforcement learning (PbRL) aligns a robot behavior with human preferences via a reward function learned from binary feedback over agent behaviors. We show that dynamics-aware reward functions improve the sample…

人工智能 · 计算机科学 2024-02-29 Katherine Metcalf , Miguel Sarabia , Natalie Mackraz , Barry-John Theobald

We propose a generic reward shaping approach for improving the rate of convergence in reinforcement learning (RL), called Self Improvement Based REwards, or SIBRE. The approach is designed for use in conjunction with any existing RL…

机器学习 · 计算机科学 2020-12-22 Somjit Nath , Richa Verma , Abhik Ray , Harshad Khadilkar

In the present study, a Particle Swarm Optimization (PSO) based Demand Response (DR) model, using Artificial Neural Network (ANN) to predict load is proposed. The electrical load and climatological data of a residential area in Austin city…

神经与进化计算 · 计算机科学 2022-07-12 Nasrin Bayat

The dominant framework for alignment of large language models (LLM), whether through reinforcement learning from human feedback or direct preference optimisation, is to learn from preference data. This involves building datasets where each…

Demand response (DR) is not only a crucial solution to the demand side management but also a vital means of electricity market in maintaining power grid reliability, sustainability and stability. DR can enable consumers (e.g. data centers)…

计算机科学与博弈论 · 计算机科学 2019-01-11 Jianhai Chen , Deshi Ye , Shouling Ji , Qinming He , Yang Xiang , Zhenguang Liu

Retrieval-Augmented Generation (RAG) has proven its effectiveness in mitigating hallucinations in Large Language Models (LLMs) by retrieving knowledge from external resources. To adapt LLMs for the RAG systems, current approaches use…

计算与语言 · 计算机科学 2025-03-05 Xinze Li , Sen Mei , Zhenghao Liu , Yukun Yan , Shuo Wang , Shi Yu , Zheni Zeng , Hao Chen , Ge Yu , Zhiyuan Liu , Maosong Sun , Chenyan Xiong

In this paper we develop an algorithm for peak load reduction to reduce the impact of increased air conditioner usage in a residential smart grid community. We develop Demand Response Management (DRM) plans that clearly spell out the…

系统与控制 · 计算机科学 2014-08-07 Yawar Ismail Khalid , Naveed Ul Hassan , Chau Yuen , Shisheng Huang

Data-Driven Response Regime Exploration and Identification (DR$^2$EI) is a novel and fully data-driven method for identifying and classifying response regimes of a dynamical system without requiring human intervention. This approach is a…

系统与控制 · 电气工程与系统科学 2023-04-13 Maor Farid

Data generation and labeling are often expensive in robot learning. Preference-based learning is a concept that enables reliable labeling by querying users with preference questions. Active querying methods are commonly employed in…

机器学习 · 计算机科学 2024-02-27 Erdem Bıyık , Nima Anari , Dorsa Sadigh

Within the context of renewable energy communities, this paper focuses on optimal operation of producers equipped with energy storage systems in the presence of demand response. A novel strategy for optimal scheduling of the storage systems…

最优化与控制 · 数学 2025-03-11 Gianni Bianchini , Marco Casini , Milad Gholami

Price elasticity model (PEM) is an appealing and modest model for assessing the potential of flexible demand in DR. It measures the customers demand sensitivity through elasticity in relation to price variation. However, application of PEM…

系统与控制 · 电气工程与系统科学 2021-06-01 Vipin Chandra Pandey , Nikhil Gupta , K. R. Niazi , Anil Swarnkar , Rayees Ahmad Thokar

Recently, there is growing interest and need for dynamic pricing algorithms, especially, in the field of online marketplaces by offering smart pricing options for big online stores. We present an approach to adjust prices based on the…

最优化与控制 · 数学 2021-01-13 David Müller , Yurii Nesterov , Vladimir Shikhman

Residential demand response depends on sustained prosumer participation, yet existing coordination is either fully automated, or limited to one-way dispatch signals and price alerts that offer little possibility for informed…

人工智能 · 计算机科学 2026-03-09 Reda El Makroum , Sebastian Zwickl-Bernhard , Lukas Kranzl , Hans Auer