中文
相关论文

相关论文: Q-learning Based Optimal False Data Injection Atta…

200 篇论文

Recent advance in deep offline reinforcement learning (RL) has made it possible to train strong robotic agents from offline datasets. However, depending on the quality of the trained agents and the application being considered, it is often…

机器人学 · 计算机科学 2021-11-02 Seunghyun Lee , Younggyo Seo , Kimin Lee , Pieter Abbeel , Jinwoo Shin

Recent research has shown that although Reinforcement Learning (RL) can benefit from expert demonstration, it usually takes considerable efforts to obtain enough demonstration. The efforts prevent training decent RL agents with expert…

机器学习 · 计算机科学 2021-07-09 Si-An Chen , Voot Tangkaratt , Hsuan-Tien Lin , Masashi Sugiyama

Reinforcement learning (RL) is a powerful machine learning technique that enables an intelligent agent to learn an optimal policy that maximizes the cumulative rewards in sequential decision making. Most of methods in the existing…

机器学习 · 统计学 2023-01-06 Chengchun Shi , Zhengling Qi , Jianing Wang , Fan Zhou

Behavior Cloning (BC) has emerged as a highly effective paradigm for robot learning. However, BC lacks a self-guided mechanism for online improvement after demonstrations have been collected. Existing offline-to-online learning methods…

This paper demonstrates that continual relearning of control policies using incremental deep reinforcement learning (RL) can improve policy learning for non-stationary processes. We demonstrate this approach for a data-driven 'smart…

机器学习 · 计算机科学 2020-08-06 Avisek Naug , Marcos Quiñones-Grueiro , Gautam Biswas

Deep Reinforcement Learning (RL) has considerably advanced over the past decade. At the same time, state-of-the-art RL algorithms require a large computational budget in terms of training time to converge. Recent work has started to…

In this work, we present a new model-free and off-policy reinforcement learning (RL) algorithm, that is capable of finding a near-optimal policy with state-action observations from arbitrary behavior policies. Our algorithm, called the…

最优化与控制 · 数学 2025-07-21 Narim Jeong , Donghwan Lee , Niao He

The promise of fault-tolerant quantum computing is challenged by environmental drift that relentlessly degrades the quality of quantum operations. The contemporary solution, halting the entire quantum computation for recalibration, is…

量子物理 · 物理学 2026-03-10 Volodymyr Sivak , Alexis Morvan , Michael Broughton , Rodrigo G. Cortiñas , Johannes Bausch , Andrew W. Senior , Matthew Neeley , Alec Eickbusch , Noah Shutty , Laleh Aghababaie Beni , James S. Spencer , Francisco J. H Heras , Thomas Edlich , Dmitry Abanin , Amira Abbas , Rajeev Acharya , Georg Aigeldinger , Ross Alcaraz , Sayra Alcaraz , Trond I. Andersen , Markus Ansmann , Frank Arute , Kunal Arya , Walt Askew , Nikita Astrakhantsev , Juan Atalaya , Brian Ballard , Joseph C. Bardin , Hector Bates , Andreas Bengtsson , Majid Bigdeli Karimi , Alexander Bilmes , Simon Bilodeau , Felix Borjans , Alexandre Bourassa , Jenna Bovaird , Dylan Bowers , Leon Brill , Peter Brooks , David A. Browne , Brett Buchea , Bob B. Buckley , Tim Burger , Brian Burkett , Nicholas Bushnell , Jamal Busnaina , Anthony Cabrera , Juan Campero , Hung-Shen Chang , Silas Chen , Ben Chiaro , Liang-Ying Chih , Agnetta Y. Cleland , Bryan Cochrane , Matt Cockrell , Josh Cogan , Roberto Collins , Paul Conner , Harold Cook , William Courtney , Alexander L. Crook , Ben Curtin , Martin Damyanov , Sayan Das , Dripto M. Debroy , Sean Demura , Paul Donohoe , Ilya Drozdov , Andrew Dunsworth , Valerie Ehimhen , Aviv Moshe Elbag , Lior Ella , Mahmoud Elzouka , David Enriquez , Catherine Erickson , Vinicius S. Ferreira , Marcos Flores , Leslie Flores Burgos , Ebrahim Forati , Jeremiah Ford , Austin G. Fowler , Brooks Foxen , Masaya Fukami , Alan Wing Lun Fung , Lenny Fuste , Suhas Ganjam , Gonzalo Garcia , Christopher Garrick , Robert Gasca , Helge Gehring , Robert Geiger , Élie Genois , William Giang , Dar Gilboa , James E. Goeders , Edward C. Gonzales , Raja Gosula , Stijn J. de Graaf , Alejandro Grajales Dau , Dietrich Graumann , Joel Grebel , Alex Greene , Jonathan A. Gross , Jose Guerrero , Loïck Le Guevel , Tan Ha , Steve Habegger , Tanner Hadick , Ali Hadjikhani , Michael C. Hamilton , Matthew P. Harrigan , Sean D. Harrington , Jeanne Hartshorn , Stephen Heslin , Paula Heu , Oscar Higgott , Reno Hiltermann , Hsin-Yuan Huang , Mike Hucka , Christopher Hudspeth , Ashley Huff , William J. Huggins , Evan Jeffrey , Shaun Jevons , Zhang Jiang , Xiaoxuan Jin , Chaitali Joshi , Pavol Juhas , Andreas Kabel , Dvir Kafri , Hui Kang , Kiseo Kang , Amir H. Karamlou , Ryan Kaufman , Kostyantyn Kechedzhi , Tanuj Khattar , Mostafa Khezri , Seon Kim , Can M. Knaut , Bryce Kobrin , Fedor Kostritsa , John Mark Kreikebaum , Ryuho Kudo , Ben Kueffler , Arun Kumar , Vladislav D. Kurilovich , Vitali Kutsko , Nathan Lacroix , David Landhuis , Tiano Lange-Dei , Brandon W. Langley , Pavel Laptev , Kim-Ming Lau , Justin Ledford , Joy Lee , Kenny Lee , Brian J. Lester , Wendy Leung , Lily Li , Wing Yan Li , Ming Li , Alexander T. Lill , William P. Livingston , Matthew T. Lloyd , Aditya Locharla , Laura De Lorenzo , Daniel Lundahl , Aaron Lunt , Sid Madhuk , Aniket Maiti , Ashley Maloney , Salvatore Mandrà , Leigh S. Martin , Orion Martin , Eric Mascot , Paul Masih Das , Dmitri Maslov , Melvin Mathews , Cameron Maxfield , Jarrod R. McClean , Matt McEwen , Seneca Meeks , Kevin C. Miao , Zlatko K. Minev , Reza Molavi , Sebastian Molina , Shirin Montazeri , Charles Neill , Michael Newman , Anthony Nguyen , Murray Nguyen , Chia-Hung Ni , Murphy Yuezhen Niu , Logan Oas , Raymond Orosco , Kristoffer Ottosson , Alice Pagano , Agustin Di Paolo , Sherman Peek , David Peterson , Alex Pizzuto , Elias Portoles , Rebecca Potter , Orion Pritchard , Michael Qian , Chris Quintana , Arpit Ranadive , Matthew J. Reagor , Rachel Resnick , David M. Rhodes , Daniel Riley , Gabrielle Roberts , Roberto Rodriguez , Emma Ropes , Lucia B. De Rose , Eliott Rosenberg , Emma Rosenfeld , Dario Rosenstock , Elizabeth Rossi , Pedram Roushan , David A. Rower , Robert Salazar , Kannan Sankaragomathi , Murat Can Sarihan , Kevin J. Satzinger , Max Schaefer , Sebastian Schroeder , Henry F. Schurkus , Aria Shahingohar , Michael J. Shearn , Aaron Shorter , Vladimir Shvarts , Spencer Small , W. Clarke Smith , David A. Sobel , Barrett Spells , Sofia Springer , George Sterling , Jordan Suchard , Aaron Szasz , Alexander Sztein , Madeline Taylor , Jothi Priyanka Thiruraman , Douglas Thor , Dogan Timucin , Eifu Tomita , Alfredo Torres , M. Mert Torunbalci , Hao Tran , Abeer Vaishnav , Justin Vargas , Sergey Vdovichev , Guifre Vidal , Catherine Vollgraff Heidweiller , Meghan Voorhees , Steven Waltman , Jonathan Waltz , Shannon X. Wang , Brayden Ware , James D. Watson , Yonghua Wei , Travis Weidel , Theodore White , Kristi Wong , Bryan W. K. Woo , Christopher J. Wood , Maddy Woodson , Cheng Xing , Z. Jamie Yao , Ping Yeh , Bicheng Ying , Juhwan Yoo , Noureldin Yosri , Elliot Young , Grayson Young , Adam Zalcman , Ran Zhang , Yaxing Zhang , Ningfeng Zhu , Nicholas Zobrist , Zhenjie Zou , Ryan Babbush , Dave Bacon , Sergio Boixo , Yu Chen , Zijun Chen , Michel Devoret , Monica Hansen , Jeremy Hilton , Cody Jones , Julian Kelly , Alexander N. Korotkov , Erik Lucero , Anthony Megrant , Hartmut Neven , William D. Oliver , Ganesh Ramachandran , Vadim Smelyanskiy , Paul V. Klimov

Owing to the openness of wireless channels, wireless communication systems are highly susceptible to malicious jamming. Most existing anti-jamming methods rely on the assumption of accurate sensing and optimize parameters on a single…

信息论 · 计算机科学 2025-11-06 Haoqin Zhao , Zan Li , Jiangbo Si , Rui Huang , Hang Hu , Tony Q. S. Quek , Naofal Al-Dhahir

The 20 Questions (Q20) game is a well known game which encourages deductive reasoning and creativity. In the game, the answerer first thinks of an object such as a famous person or a kind of animal. Then the questioner tries to guess the…

人机交互 · 计算机科学 2026-02-10 Huang Hu , Xianchao Wu , Bingfeng Luo , Chongyang Tao , Can Xu , Wei Wu , Zhan Chen

Voltage control is crucial to large-scale power system reliable operation, as timely reactive power support can help prevent widespread outages. However, there is currently no built in mechanism for power systems to ensure that the voltage…

机器学习 · 计算机科学 2023-05-29 Abhijeet Sahu , Katherine Davis

Traditional power grid systems have become obsolete under more frequent and extreme natural disasters. Reinforcement learning (RL) has been a promising solution for resilience given its successful history of power grid control. However,…

机器学习 · 计算机科学 2022-12-09 Zhenting Zhao , Po-Yen Chen , Yucheng Jin

Reinforcement learning (RL) is one of the most practical ways to learn from real-life use-cases. Motivated from the cognitive methods used by humans makes it a widely acceptable strategy in the field of artificial intelligence. Most of the…

人工智能 · 计算机科学 2026-04-14 Abhishek Sawaika , Samuel Yen-Chi Chen , Udaya Parampalli , Rajkumar Buyya

The primary goal of reinforcement learning is to develop decision-making policies that prioritize optimal performance without considering risk or safety. In contrast, safe reinforcement learning aims to mitigate or avoid unsafe states. This…

机器学习 · 计算机科学 2024-09-13 Zahra Shahrooei , Ali Baheri

This Research proposes a Novel Reinforcement Learning (RL) model to optimise malware forensics investigation during cyber incident response. It aims to improve forensic investigation efficiency by reducing false negatives and adapting…

密码学与安全 · 计算机科学 2025-01-14 Dipo Dunsin , Mohamed Chahine Ghanem , Karim Ouazzane , Vassil Vassilev

This paper presents a Quantum Reinforcement Learning (QRL) solution to the dynamic portfolio optimization problem based on Variational Quantum Circuits. The implemented QRL approaches are quantum analogues of the classical…

机器学习 · 计算机科学 2026-01-29 Vincent Gurgul , Ying Chen , Stefan Lessmann

Applications of reinforcement learning (RL) to stabilization problems of real systems are restricted since an agent needs many experiences to learn an optimal policy and may determine dangerous actions during its exploration. If we know a…

机器学习 · 计算机科学 2021-04-20 Junya Ikemoto , Toshimitsu Ushio

We provide in this paper a concrete method for training a quantum neural network to maximize the relevant information about a property that is transmitted through the network. This is significant because it gives an operationally well…

量子物理 · 物理学 2024-01-23 Ahmet Burak Catli , Nathan Wiebe

Predicting a sequence of actions has been crucial in the success of recent behavior cloning algorithms in robotics. Can similar ideas improve reinforcement learning (RL)? We answer affirmatively by observing that incorporating action…

机器学习 · 计算机科学 2025-11-18 Younggyo Seo , Pieter Abbeel

In heterogeneous networks (HetNets), the overlap of small cells and the macro cell causes severe cross-tier interference. Although there exist some approaches to address this problem, they usually require global channel state information,…

系统与控制 · 电气工程与系统科学 2022-12-16 Kaidi Xu , Nguyen Van Huynh , Geoffrey Ye Li
‹ 上一页 1 8 9 10 下一页 ›