English
Related papers

Related papers: Learning a Reward Function for User-Preferred Appl…

200 papers

Our team is proposing to run a full-scale energy demand response experiment in an office building. Although this is an exciting endeavor which will provide value to the community, collecting training data for the reinforcement learning…

Machine Learning · Computer Science 2021-11-12 Doseok Jang , Lucas Spangher , Manan Khattar , Utkarsha Agwan , Costas Spanos

Electric automation systems offer convenience and efficiency in controlling electrical circuits and devices. Traditionally, these systems rely on predefined commands for control, limiting flexibility and adaptability. In this paper, we…

Machine Learning · Computer Science 2024-03-05 Lochan Basyal , Bijay Gaudel

Recent developments in sequential experimental design look to construct a policy that can efficiently navigate the design space, in a way that maximises the expected information gain. Whilst there is work on achieving tractable policies for…

Machine Learning · Computer Science 2025-08-20 Yasir Zubayr Barlas , Kizito Salako

Reward functions are a common way to specify the objective of a robot. As designing reward functions can be extremely challenging, a more promising approach is to directly learn reward functions from human teachers. Importantly, data from…

We are interested in the design of autonomous robot behaviors that learn the preferences of users over continued interactions, with the goal of efficiently executing navigation behaviors in a way that the user expects. In this paper, we…

Robotics · Computer Science 2020-11-06 Cory Hayes , Matthew Marge

Extreme weather and volatile wholesale electricity markets expose residential consumers to catastrophic financial risks, yet demand response at the distribution level remains an underutilized tool for grid flexibility and energy…

Artificial Intelligence · Computer Science 2026-05-13 Jose E. Aguilar Escamilla , Lingdong Zhou , Xiangqi Zhu , Huazheng Wang

Methods for learning from demonstration (LfD) have shown success in acquiring behavior policies by imitating a user. However, even for a single task, LfD may require numerous demonstrations. For versatile agents that must learn many tasks…

Machine Learning · Computer Science 2022-07-04 Jorge A. Mendez , Shashank Shivkumar , Eric Eaton

The large integration of variable energy resources is expected to shift a large part of the energy exchanges closer to real-time, where more accurate forecasts are available. In this context, the short-term electricity markets and in…

Trading and Market Microstructure · Quantitative Finance 2020-04-14 Ioannis Boukas , Damien Ernst , Thibaut Théate , Adrien Bolland , Alexandre Huynen , Martin Buchwald , Christelle Wynants , Bertrand Cornélusse

Deep learning, as a highly efficient method for metasurface inverse design, commonly use simulation data to train deep neural networks (DNNs) that can map desired functionalities to proper metasurface designs. However, the assumptions and…

Signal Processing · Electrical Eng. & Systems 2023-08-07 Jingxin Zhang , Jiawei Xi , Peixing Li , Ray C. C. Cheung , Alex M. H. Wong , Jensen Li

Smart homes require every device inside them to be connected with each other at all times, which leads to a lot of power wastage on a daily basis. As the devices inside a smart home increase, it becomes difficult for the user to control or…

Artificial Intelligence · Computer Science 2020-09-30 Saurabh Gupta , Siddhant Bhambri , Karan Dhingra , Arun Balaji Buduru , Ponnurangam Kumaraguru

Evidence-based decision-making entails collecting (costly) observations about an underlying phenomenon of interest, and subsequently committing to an (informed) decision on the basis of accumulated evidence. In this setting, active sensing…

Machine Learning · Statistics 2020-06-26 Daniel Jarrett , Mihaela van der Schaar

The combination of deep neural network models and reinforcement learning algorithms can make it possible to learn policies for robotic behaviors that directly read in raw sensory inputs, such as camera images, effectively subsuming both…

Machine Learning · Computer Science 2019-05-17 Avi Singh , Larry Yang , Kristian Hartikainen , Chelsea Finn , Sergey Levine

Residential homes constitute roughly one-fourth of the total energy usage worldwide. Providing appliance-level energy breakdown has been shown to induce positive behavioral changes that can reduce energy consumption by 15%. Existing…

Machine Learning · Computer Science 2019-09-15 Yiling Jia , Nipun Batra , Hongning Wang , Kamin Whitehouse

The evolution of the power grid towards the so-called Smart Grid, where information technologies help improve the efficiency of electricity production, distribution and consumption, allows to use the fine-grained control brought by the…

Optimization and Control · Mathematics 2017-05-03 Nguyen Hoang Son Duong , Patrick Maillé , Ashish Kumar , Laurent Toutain

It is crucial today that economies harness renewable energies and integrate them into the existing grid. Conventionally, energy has been generated based on forecasts of peak and low demands. Renewable energy can neither be produced on…

Signal Processing · Electrical Eng. & Systems 2019-10-02 Alexey Györi , Mathis Niederau , Violett Zeller , Volker Stich

Communications standards are designed via committees of humans holding repeated meetings over months or even years until consensus is achieved. This includes decisions regarding the modulation and coding schemes to be supported over an air…

Machine Learning · Statistics 2021-10-19 Shahrukh Khan Kasi , Sayandev Mukherjee , Lin Cheng , Bernardo A. Huberman

In some agent designs like inverse reinforcement learning an agent needs to learn its own reward function. Learning the reward function and optimising for it are typically two different processes, usually performed at different stages. We…

Artificial Intelligence · Computer Science 2020-04-29 Stuart Armstrong , Jan Leike , Laurent Orseau , Shane Legg

While reinforcement learning algorithms provide automated acquisition of optimal policies, practical application of such methods requires a number of design decisions, such as manually designing reward functions that not only define the…

Machine Learning · Computer Science 2022-12-29 Tim G. J. Rudner , Vitchyr H. Pong , Rowan McAllister , Yarin Gal , Sergey Levine

Reward design is a critical part of the application of reinforcement learning, the performance of which strongly depends on how well the reward signal frames the goal of the designer and how well the signal assesses progress in reaching…

Machine Learning · Computer Science 2022-08-01 Yixiang Wang , Yujing Hu , Feng Wu , Yingfeng Chen

Although reinforcement learning has seen tremendous success recently, this kind of trial-and-error learning can be impractical or inefficient in complex environments. The use of demonstrations, on the other hand, enables agents to benefit…

Machine Learning · Computer Science 2023-03-29 Tongzhou Mu , Hao Su
‹ Prev 1 4 5 6 7 8 10 Next ›