中文
相关论文

相关论文: Understanding and Shifting Preferences for Battery…

200 篇论文

Fuel optimization of diesel and petrol vehicles within industrial fleets is critical for mitigating costs and reducing emissions. This objective is achievable by acting on fuel-related factors, such as the driving behaviour style. In this…

人工智能 · 计算机科学 2021-07-23 Alberto Barbado , Óscar Corcho

Active inference may be defined as Bayesian modeling of a brain with a biologically plausible model of the agent. Its primary idea relies on the free energy principle and the prior preference of the agent. An agent will choose an action…

机器学习 · 计算机科学 2021-12-14 Jin young Shin , Cheolhyeong Kim , Hyung Ju Hwang

Preference-based reinforcement learning (PbRL) is emerging as a promising approach to teaching robots through human comparative feedback, sidestepping the need for complex reward engineering. However, the substantial volume of feedback…

机器人学 · 计算机科学 2025-01-09 Ruiqi Wang , Dezhong Zhao , Ziqin Yuan , Ike Obi , Byung-Cheol Min

Electric vehicles (EVs) add significant load on the power grid as they become widespread. The characteristics of this extra load follow the patterns of people's driving behaviours. In particular, random parameters such as arrival time and…

计算工程、金融与科学 · 计算机科学 2014-08-12 Farshad Rassaei , Wee-Seng Soh , Kee-Chaing Chua

We state the problem of inverse reinforcement learning in terms of preference elicitation, resulting in a principled (Bayesian) statistical formulation. This generalises previous work on Bayesian inverse reinforcement learning and allows us…

机器学习 · 统计学 2011-06-30 Constantin Rothkopf , Christos Dimitrakakis

Designing incentives for an adapting population is a ubiquitous problem in a wide array of economic applications and beyond. In this work, we study how to design additional rewards to steer multi-agent systems towards desired policies…

机器学习 · 计算机科学 2025-02-11 Jiawei Huang , Vinzenz Thoma , Zebang Shen , Heinrich H. Nax , Niao He

Preference-based reinforcement learning (PbRL) has shown impressive capabilities in training agents without reward engineering. However, a notable limitation of PbRL is its dependency on substantial human feedback. This dependency stems…

机器学习 · 计算机科学 2024-05-30 Fengshuo Bai , Rui Zhao , Hongming Zhang , Sijia Cui , Ying Wen , Yaodong Yang , Bo Xu , Lei Han

Electric Vehicles (EVs), as their penetration increases, are not only challenging the sustainability of the power grid, but also stimulating and promoting its upgrading. Indeed, EVs can actively reinforce the development of the Smart Grid…

计算机科学与博弈论 · 计算机科学 2016-04-18 Wenjing Shuai , Patrick Maillé , Alexander Pelov

We present a novel preference learning framework to capture participant preferences efficiently within limited interaction rounds. It involves three main contributions. First, we develop a variational Bayesian approach to infer the…

机器学习 · 计算机科学 2025-03-20 Yan Wang , Jiapeng Liu , Milosz Kadziński , Xiuwu Liao

Many important behavior changes are frictionful; they require individuals to expend effort over a long period with little immediate gratification. Here, an artificial intelligence (AI) agent can provide personalized interventions to help…

人工智能 · 计算机科学 2024-01-29 Eura Nofshin , Siddharth Swaroop , Weiwei Pan , Susan Murphy , Finale Doshi-Velez

Many proposed methods for explaining machine learning predictions are in fact challenging to understand for nontechnical consumers. This paper builds upon an alternative consumer-driven approach called TED that asks for explanations to be…

机器学习 · 计算机科学 2020-01-17 Michael Hind , Dennis Wei , Yunfeng Zhang

Iterative machine learning algorithms used to power recommender systems often change people's preferences by trying to learn them. Further a recommender can better predict what a user will do by making its users more predictable. Some…

信息检索 · 计算机科学 2022-09-27 Hal Ashton , Matija Franklin

Strategic aggregation of electric vehicle batteries as energy reservoirs can optimize power grid demand, benefiting smart and connected communities, especially large office buildings that offer workplace charging. This involves optimizing…

Bayesian Personalized Ranking (BPR) is a representative pairwise learning method for optimizing recommendation models. It is widely known that the performance of BPR depends largely on the quality of negative sampler. In this paper, we make…

信息检索 · 计算机科学 2018-09-24 Jingtao Ding , Guanghui Yu , Xiangnan He , Yong Li , Depeng Jin

In recent years, the persuasive interventions for inducing sustainable urban mobility behaviours has become a very active research field. This review paper systematically analyses existing approaches and prototype systems and describes and…

Electric vehicles (EVs) have the potential to reduce grid stress through smart charging strategies while simultaneously meeting user demand. This requires accurate forecasts of key charging parameters, such as energy demand and connection…

系统与控制 · 电气工程与系统科学 2025-08-26 Parnian Alikhani , Nico Brinkel , Wouter Schram , Ioannis Lampropoulos , Wilfried van Sark

The growth of Electric Vehicles (EVs) creates a conflict in vehicle-to-building (V2B) settings between building operators, who face high energy costs from uncoordinated charging, and drivers, who prioritize convenience and a full charge. To…

多智能体系统 · 计算机科学 2026-02-18 Rishav Sen , Fangqi Liu , Jose Paolo Talusan , Ava Pettet , Yoshinori Suzue , Mark Bailey , Ayan Mukhopadhyay , Abhishek Dubey

Increasing the adoption of Electric Vehicles (EV) is an integral part of many strategies to address climate change and air pollution. However, Electric Vehicle adoption rates are inhibited by several factors which reduce the confidence of…

最优化与控制 · 数学 2023-06-21 Johnny Tiu , Shankar Ramharack , Patrick Hosein

Evaluations often inform future program implementation decisions. However, the implementation context may differ, sometimes substantially, from the evaluation study context. This difference leads to uncertainty regarding the relevance of…

统计方法学 · 统计学 2023-12-14 Irina Degtiar , Mariel Finucane

Reinforcement Learning from Human Feedback (RLHF) is a powerful paradigm for aligning foundation models to human values and preferences. However, current RLHF techniques cannot account for the naturally occurring differences in individual…

机器学习 · 计算机科学 2024-08-20 Sriyash Poddar , Yanming Wan , Hamish Ivison , Abhishek Gupta , Natasha Jaques