中文
相关论文

相关论文: Behaviorally Grounded Model-Based and Model Free C…

200 篇论文

A data-efficient learning-based control design method is proposed in this paper. It is based on learning a system dynamics model that is then leveraged in a two-level procedure. On the higher level, a simple but powerful optimization…

系统与控制 · 电气工程与系统科学 2026-02-03 Ludvig Svedlund , Constantin Cronrath , Jonas Fredriksson , Bengt Lennartson

Semiconductor supply chains are described by significant demand fluctuation that increases as one moves up the supply chain, the so-called bullwhip effect. To counteract, semiconductor manufacturers aim to optimize capacity utilization, to…

数据库 · 计算机科学 2022-05-17 Nour Ramzy , Soren Auer , Javad Chamanara , Hans Ehm

In this article we quantify the bullwhip effect (the variance amplification in replenishment orders) when demands and lead times are predicted in a simple two-stage supply chain with one supplier and one retailer. In recent research the…

概率论 · 数学 2015-07-30 Zbigniew Michna , Peter Nielsen

Animals and robots exist in a physical world and must coordinate their bodies to achieve behavioral objectives. With recent developments in deep reinforcement learning, it is now possible for scientists and engineers to obtain sensorimotor…

机器人学 · 计算机科学 2024-05-21 Yusheng Jiao , Feng Ling , Sina Heydari , Nicolas Heess , Josh Merel , Eva Kanso

Online planning has proven effective in reinforcement learning (RL) for improving sample efficiency and final performance. However, using planning for environment interaction inevitably introduces a divergence between the collected data and…

机器学习 · 计算机科学 2026-01-16 Guojian Zhan , Likun Wang , Xiangteng Zhang , Jiaxin Gao , Masayoshi Tomizuka , Shengbo Eben Li

Reinforcement Learning is divided in two main paradigms: model-free and model-based. Each of these two paradigms has strengths and limitations, and has been successfully applied to real world domains that are appropriate to its…

机器学习 · 计算机科学 2017-10-19 Somil Bansal , Roberto Calandra , Kurtland Chua , Sergey Levine , Claire Tomlin

We model the behavioral biases of human decision-making in securing interdependent systems and show that such behavioral decision-making leads to a suboptimal pattern of resource allocation compared to non-behavioral (rational)…

密码学与安全 · 计算机科学 2020-11-25 Mustafa Abdallah , Daniel Woods , Parinaz Naghizadeh , Issa Khalil , Timothy Cason , Shreyas Sundaram , Saurabh Bagchi

A self-learning optimal control algorithm for episodic fixed-horizon manufacturing processes with time-discrete control actions is proposed and evaluated on a simulated deep drawing process. The control model is built during consecutive…

系统与控制 · 计算机科学 2020-01-07 Johannes Dornheim , Norbert Link , Peter Gumbsch

Disasters and disruptions such as the COVID-19 pandemic can significantly interrupt supply chains and industries. To control these disruptions, decision-makers must focus on supply chain resiliency. This paper proposes a multi-stage,…

最优化与控制 · 数学 2023-03-06 Hossein Mirzaee , Hamed Samarghandi , Keith Willoughby

Optimal control of complex environments with robotic systems faces two complementary and intertwined challenges: efficient organization of sensory state information and far-sighted action planning. Because the reinforcement learning…

机器学习 · 计算机科学 2026-01-30 Abdullah Akgül , Gulcin Baykal , Manuel Haußmann , Mustafa Mert Çelikok , Melih Kandemir

There have been numerous attempts in explaining the general learning behaviours using model-based and model-free methods. While the model-based control is flexible yet computationally expensive in planning, the model-free control is quick…

机器学习 · 计算机科学 2019-12-12 Krishn Bera , Yash Mandilwar , Bapi Raju

We investigate the use of Reinforcement Learning for the optimal execution of meta-orders, where the objective is to execute incrementally large orders while minimizing implementation shortfall and market impact over an extended period of…

交易与市场微观结构 · 定量金融 2025-11-20 Tomas Espana , Yadh Hafsi , Fabrizio Lillo , Edoardo Vittori

Although recent model-free reinforcement learning algorithms have been shown to be capable of mastering complicated decision-making tasks, the sample complexity of these methods has remained a hurdle to utilizing them in many real-world…

机器学习 · 计算机科学 2020-04-21 Saeed Moazami , Peggy Doerschuk

Learning-based control methods typically assume stationary system dynamics, an assumption often violated in real-world systems due to drift, wear, or changing operating conditions. We study reinforcement learning for control under…

机器学习 · 计算机科学 2026-04-03 Klemens Iten , Bruce Lee , Chenhao Li , Lenart Treven , Andreas Krause , Bhavya Sukhija

We quantify the bullwhip effect (which measures how the variance in replenishment orders is amplified as the orders move up the supply chain) when random demands and random lead times are estimated using the industrially popular moving…

应用统计 · 统计学 2017-02-07 Zbigniew Michna , Stephen M. Disney , Peter Nielsen

The deformable and continuum nature of soft robots promises versatility and adaptability. However, control of modular, multi-limbed soft robots for terrestrial locomotion is challenging due to the complex robot structure, actuator mechanics…

机器人学 · 计算机科学 2016-02-05 Vishesh Vikas , Piyush Grover , Barry Trimmer

Although evidence integration to the boundary model has successfully explained a wide range of behavioral and neural data in decision making under uncertainty, how animals learn and optimize the boundary remains unresolved. Here, we propose…

神经与进化计算 · 计算机科学 2024-08-13 Jamal Esmaily , Rani Moran , Yasser Roudi , Bahador Bahrami

Due to the COVID-19 pandemic, the global supply chain is disrupted at an unprecedented scale under uncertain and unknown trends of labor shortage, high material prices, and changing travel or trade regulations. To stay competitive,…

多智能体系统 · 计算机科学 2022-11-02 Mingjie Bi , Gongyu Chen , Dawn M. Tilbury , Siqian Shen , Kira Barton

Model-free learning has been considered as an efficient tool for designing control mechanisms when the model of the system environment or the interaction between the decision-making entities is not available as a-priori knowledge. With…

网络与互联网体系结构 · 计算机科学 2016-03-10 Wenbo Wang , Andres Kwasinski , Dusit Niyato , Zhu Han

Grid-interactive building control is a challenging and important problem for reducing carbon emissions, increasing energy efficiency, and supporting the electric power grid. Currently researchers and practitioners are confronted with a…

系统与控制 · 电气工程与系统科学 2022-10-20 David Biagioni , Xiangyu Zhang , Christiane Adcock , Michael Sinner , Peter Graf , Jennifer King