中文
相关论文

相关论文: Techno-economic optimization of a heat-pipe micror…

200 篇论文

The challenges of the uncertainties in renewable energy generation and the instability of the real-time market limit the effective utilization of clean energy in microgrid communities. Existing peer-to-peer (P2P) and microgrid coordination…

多智能体系统 · 计算机科学 2026-04-06 Junhao Ren , Honglin Gao , Sijie Wang , Lan Zhao , Qiyu Kang , Aniq Ashan , Yajuan Sun , Gaoxi Xiao

This work presents a comparative study of optimization techniques for parameter identification in equivalent electrical models of lithium-ion batteries. The 2RC model is applied to a set of twelve batteries using four publicly available…

Demand forecasting in competitive, uncertain business environments requires models that can integrate multiple evaluation perspectives rather than being restricted to hyperparameter optimization based on a single metric. This traditional…

机器学习 · 计算机科学 2025-12-23 Adolfo González , Víctor Parada

Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for post-training reasoning models. However, group-based methods such as Group Relative Policy Optimization (GRPO) face a critical dilemma in…

机器学习 · 计算机科学 2026-04-07 Yuning Wu , Ke Wang , Devin Chen , Kai Wei

This paper studies the long-term energy management of a microgrid coordinating hybrid hydrogen-battery energy storage. We develop an approximate semi-empirical hydrogen storage model to accurately capture the power-dependent efficiency of…

最优化与控制 · 数学 2025-06-10 Ning Qi , Kaidi Huang , Zhiyuan Fan , Bolun Xu

In this paper, we develop a hybrid prediction framework for accurate electric vehicle (EV) charging time estimation, a capability that is critical for trip planning, user satisfaction, and efficient operation of charging infrastructure. We…

系统与控制 · 电气工程与系统科学 2025-12-09 Praharshitha Aryasomayajula , Ting Bai , Andreas A. Malikopoulos

Hierarchical reinforcement learning (HRL) helps address large-scale and sparse reward issues in reinforcement learning. In HRL, the policy model has an inner representation structured in levels. With this structure, the reinforcement…

人工智能 · 计算机科学 2020-02-07 Wen-Ji Zhou , Yang Yu

In the present study, two-different reduced-order models are proposed for $\text{H}_2\left(\text{X}^1\Sigma_g^+\right)$+$\text{H}\left({}^2\text{S}\right)$ system by leveraging first-principle quasi-classical trajectory simulations and…

化学物理 · 物理学 2026-01-29 Hye Su Jeong , Tae Woong Jeong , Sung Min Jo

Humanoid robots often need to balance competing objectives, such as maximizing speed while minimizing energy consumption. While current reinforcement learning (RL) methods can master complex skills like fall recovery and perceptive…

机器人学 · 计算机科学 2026-03-26 Huanyu Li , Dewei Wang , Xinmiao Wang , Xinzhe Liu , Peng Liu , Chenjia Bai , Xuelong Li

Tree-structured Parzen estimator (TPE) is a versatile hyperparameter optimization (HPO) method supported by popular HPO tools. Since these HPO tools have been developed in line with the trend of deep learning (DL), the problem setups often…

机器学习 · 计算机科学 2025-07-16 Kenshin Abe , Yunzhuo Wang , Shuhei Watanabe

Hyperparameter optimization (HPO) is generally treated as a bi-level optimization problem that involves fitting a (probabilistic) surrogate model to a set of observed hyperparameter responses, e.g. validation loss, and consequently…

机器学习 · 计算机科学 2021-10-18 Hadi S. Jomaa , Jonas Falkner , Lars Schmidt-Thieme

In neural combinatorial optimization (CO), reinforcement learning (RL) can turn a deep neural net into a fast, powerful heuristic solver of NP-hard problems. This approach has a great potential in practical applications because it allows…

机器学习 · 计算机科学 2021-07-14 Yeong-Dae Kwon , Jinho Choo , Byoungjip Kim , Iljoo Yoon , Youngjune Gwon , Seungjai Min

Policy optimization (PO) is a key ingredient for reinforcement learning (RL). For control design, certain constraints are usually enforced on the policies to optimize, accounting for either the stability, robustness, or safety concerns on…

最优化与控制 · 数学 2021-02-16 Kaiqing Zhang , Bin Hu , Tamer Başar

The design of materials structure for optimizing functional properties and potentially, the discovery of novel behaviors is a keystone problem in materials science. In many cases microstructural models underpinning materials functionality…

材料科学 · 物理学 2022-02-23 Rama K. Vasudevan , Erick Orozco , Sergei V. Kalinin

Process rewards have been widely used in deep reinforcement learning to improve training efficiency, reduce variance, and prevent reward hacking. In LLM reasoning, existing works also explore various solutions for learning effective process…

机器学习 · 计算机科学 2026-05-21 Xian Wu , Kaijie Zhu , Ying Zhang , Lun Wang , Wenbo Guo

The use of synthetic fuels is a promising way to reduce emissions significantly. To accelerate cost-effective large-scale synthetic fuel deployment, we optimize a novel 1 MW PtL-plant in terms of PtL-efficiency and fuel production costs.…

最优化与控制 · 数学 2023-11-15 David Huber , Felix Birkelbach , René Hofmann

According to the fundamental theorems of welfare economics, any competitive equilibrium is Pareto efficient. Unfortunately, competitive equilibrium prices only exist under strong assumptions such as perfectly divisible goods and convex…

计算机科学与博弈论 · 计算机科学 2023-05-24 Mete Şeref Ahunbay , Martin Bichler , Johannes Knörr

This paper considers a joint pollution-routing and speed optimization problem (PRP-SO) where fuel costs and $\textit{CO}_2e$ emissions depend on the vehicle speed, arc payloads, and road grades. We present two methods, one approximate and…

最优化与控制 · 数学 2021-05-20 David Lai , Yasel Costa , Emrah Demir , Alexandre Florio , Tom Van Woensel

Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from significant reward hacking, which degrades generative diversity and quality by inducing…

机器学习 · 计算机科学 2026-05-14 Jiaming Li , Chenyu Zhu , Nanxi Yi , Youjun Bao , Li Sun , Quanying Lv , Xiang Fang , Daizong Liu , Jianjun Li , Kun He , Bowen Zhou , Zhiyuan Ma

Production optimization in stress-sensitive unconventional reservoirs is governed by a nonlinear trade-off between pressure-driven flow and stress-induced degradation of fracture conductivity and matrix permeability. While higher drawdown…

机器学习 · 计算机科学 2026-04-02 Mahammad Valiyev , Jodel Cornelio , Behnam Jafarpour