中文
相关论文

相关论文: Multi-Objective Coordination Graphs for the Expect…

200 篇论文

Resource allocation plays a critical role in minimizing cycle time and improving the efficiency of business processes. Recently, Deep Reinforcement Learning (DRL) has emerged as a powerful technique to optimize resource allocation policies…

机器学习 · 计算机科学 2025-09-03 Jeroen Middelhuis , Zaharah Bukhsh , Ivo Adan , Remco Dijkman

Strategic learning studies how decision rules interact with agents who may strategically change their inputs/features to achieve better outcomes. In standard settings, models assume that the decision-maker's sole scope is to learn a…

计算机科学与博弈论 · 计算机科学 2025-10-23 Valia Efthymiou , Ekaterina Fedorova , Chara Podimata

In this work, we consider the problem of Quality-Diversity (QD) optimization with multiple objectives. QD algorithms have been proposed to search for a large collection of both diverse and high-performing solutions instead of a single set…

人工智能 · 计算机科学 2022-06-01 Thomas Pierrot , Guillaume Richard , Karim Beguir , Antoine Cully

Many real-world optimization problems such as engineering design can be eventually modeled as the corresponding multiobjective optimization problems (MOPs) which must be solved to obtain approximate Pareto optimal fronts. Multiobjective…

神经与进化计算 · 计算机科学 2021-11-12 Wang Chen , Jian Chen , Weitian Wu , Xinmin Yang , Hui Li

Many sequential decision-making tasks involve optimizing multiple conflicting objectives, requiring policies that adapt to different user preferences. In multi-objective reinforcement learning (MORL), one widely studied approach} addresses…

机器学习 · 计算机科学 2026-04-28 Ying-Tu Chen , Wei Hung , Bing-Shu Wu , Zhang-Wei Hong , Ping-Chun Hsieh

In practical multi-criterion decision-making, it is cumbersome if a decision maker (DM) is asked to choose among a set of trade-off alternatives covering the whole Pareto-optimal front. This is a paradox in conventional evolutionary…

神经与进化计算 · 计算机科学 2022-04-07 Ke Li , Guiyu Lai , Xin Yao

Real-world engineering systems are typically compared and contrasted using multiple metrics. For practical machine learning systems, performance tuning is often more nuanced than minimizing a single expected loss objective, and it may be…

最优化与控制 · 数学 2016-12-19 Ian Dewancker , Michael McCourt , Samuel Ainsworth

We study the common generalization of Markov decision processes (MDPs) with sets of transition probabilities, known as robust MDPs (RMDPs). A standard goal in RMDPs is to compute a policy that maximizes the expected return under an…

人工智能 · 计算机科学 2025-11-20 Alessandro Abate , Thom Badings , Giuseppe De Giacomo , Francesco Fabiano

In order to coordinate energy interactions among various communities and energy conversions among multi-energy subsystems within the multi-community integrated energy system under uncertain conditions, and achieve overall optimization and…

系统与控制 · 电气工程与系统科学 2024-04-24 Yang Li , Wenjie Ma , Fanjin Bu , Zhen Yang , Bin Wang , Meng Han

As generative agents become increasingly capable, alignment of their behavior with complex human values remains a fundamental challenge. Existing approaches often simplify human intent through reduction to a scalar reward, overlooking the…

机器学习 · 计算机科学 2025-07-30 Kalyan Cherukuri , Aarav Lala

Research on promoting cooperation among autonomous, self-regarding agents has often focused on the bi-objective optimisation problem: minimising the total incentive cost while maximising the frequency of cooperation. However, the optimal…

Distributed energy resources offer a control-based option to improve distribution system reliability by ensuring system states that positively impact component failure rates. This option is an attractive complement to otherwise costly and…

最优化与控制 · 数学 2025-10-27 Gejia Zhang , Robert Mieth

Existing multi-objective reinforcement learning (MORL) algorithms do not account for objectives that arise from players with differing beliefs. Concretely, consider two players with different beliefs and utility functions who may cooperate…

人工智能 · 计算机科学 2017-05-16 Andrew Critch

We address the co-optimization of behind-the-meter (BTM) distributed energy resources (DER), including flexible demands, renewable distributed generation (DG), and battery energy storage systems (BESS) under net energy metering (NEM)…

系统与控制 · 电气工程与系统科学 2025-03-12 Ruixiao Yang , Gulai Shen , Ahmed S. Alahmed , Chuchu Fan

Recent advances in reasoning with large language models (LLMs) have shown the effectiveness of Monte Carlo Tree Search (MCTS) for generating high quality intermediate trajectories, particularly in math and symbolic domains. Inspired by…

人工智能 · 计算机科学 2025-12-23 Bingning Huang , Tu Nguyen , Matthieu Zimmer

As electric vehicles (EV) become more prevalent and advances in electric vehicle electronics continue, vehicle-to-grid (V2G) techniques and large-scale scheduling strategies are increasingly important to promote renewable energy utilization…

系统与控制 · 电气工程与系统科学 2024-07-22 Xin Chen

Independent on-policy policy gradient algorithms are widely used for multi-agent reinforcement learning (MARL) in cooperative and no-conflict games, but they are known to converge sub-optimally when each agent's individual policy gradient…

机器学习 · 计算机科学 2026-05-14 Nicholas E. Corrado , Josiah P. Hanna

Post-training of LLMs with RLHF, and subsequently preference optimization algorithms such as DPO, IPO, etc., made a big difference in improving human alignment. However, all such techniques can only work with a single (human) objective. In…

机器学习 · 计算机科学 2025-05-19 Akhil Agnihotri , Rahul Jain , Deepak Ramachandran , Zheng Wen

We study the interdependence between transportation and power systems considering decentralized renewable generators and electric vehicles (EVs). We formulate the problem in a stochastic multi-agent optimization framework considering the…

最优化与控制 · 数学 2021-01-05 Zhaomiao Guo , Fatima Afifah , Junjian Qi , Sina Baghali

Fair and dynamic energy allocation in community microgrids remains a critical challenge, particularly when serving socio-economically diverse participants. Static optimization and cost-sharing methods often fail to adapt to evolving…

系统与控制 · 电气工程与系统科学 2025-07-28 Abhijan Theja , Mayukha Pal