English
Related papers

Related papers: AISYN: AI-driven Reinforcement Learning-Based Logi…

200 papers

This paper studies satisfaction of temporal properties on unknown stochastic processes that have continuous state spaces. We show how reinforcement learning (RL) can be applied for computing policies that are finite-memory and deterministic…

Systems and Control · Electrical Eng. & Systems 2020-09-29 Milad Kazemi , Sadegh Soudjani

Large Language Models (LLMs) have achieved remarkable progress in reasoning, alignment, and task-specific performance. However, ensuring harmlessness in these systems remains a critical challenge, particularly in advanced models like…

Machine Learning · Computer Science 2025-01-29 Manojkumar Parmar , Yuvaraj Govindarajulu

Enhancing the mathematical reasoning of large language models (LLMs) demands high-quality training data, yet conventional methods face critical challenges in scalability, cost, and data reliability. To address these limitations, we propose…

Computation and Language · Computer Science 2025-08-27 Sirui Chen , Changxin Tian , Binbin Hu , Kunlong Chen , Ziqi Liu , Zhiqiang Zhang , Jun Zhou

Reinforcement learning (RL) has emerged as an effective approach for enhancing the reasoning capabilities of large language models (LLMs), especially in scenarios where supervised fine-tuning (SFT) falls short due to limited…

Machine Learning · Computer Science 2026-04-15 Jian Xiong , Jingbo Zhou , Jingyong Ye , Qiang Huang , Dejing Dou

As a key component to intuitive cognition and reasoning solutions in human intelligence, causal knowledge provides great potential for reinforcement learning (RL) agents' interpretability towards decision-making by helping reduce the…

Machine Learning · Computer Science 2025-04-25 Ruichu Cai , Siyang Huang , Jie Qiao , Wei Chen , Yan Zeng , Keli Zhang , Fuchun Sun , Yang Yu , Zhifeng Hao

The problem of automatically generating a computer program from some specification has been studied since the early days of AI. Recently, two competing approaches for automatic program learning have received significant attention: (1)…

Artificial Intelligence · Computer Science 2017-03-23 Jacob Devlin , Jonathan Uesato , Surya Bhupatiraju , Rishabh Singh , Abdel-rahman Mohamed , Pushmeet Kohli

Program synthesis or code generation aims to generate a program that satisfies a problem specification. Recent approaches using large-scale pretrained language models (LMs) have shown promising results, yet they have some critical…

Machine Learning · Computer Science 2022-11-04 Hung Le , Yue Wang , Akhilesh Deepak Gotmare , Silvio Savarese , Steven C. H. Hoi

Joint logical-numerical reasoning remains a major challenge for language models, yet existing datasets rely on fixed rule sets and offer limited control over task complexity, constraining their generalizability for evaluation and training.…

Computation and Language · Computer Science 2025-10-14 Yiwei Liu , Yucheng Li , Xiao Li , Gong Cheng

The progress of AI is bottlenecked by the quality of evaluation, making powerful LLM-as-a-Judge models a core solution. The efficacy of these judges depends on their chain-of-thought reasoning, creating a critical need for methods that can…

Computation and Language · Computer Science 2025-10-14 Chenxi Whitehouse , Tianlu Wang , Ping Yu , Xian Li , Jason Weston , Ilia Kulikov , Swarnadeep Saha

Offline reinforcement learning (RL) defines the task of learning from a fixed batch of data. Due to errors in value estimation from out-of-distribution actions, most offline RL algorithms take the approach of constraining or regularizing…

Machine Learning · Computer Science 2021-12-06 Scott Fujimoto , Shixiang Shane Gu

Reinforcement Learning (RL) has demonstrated significant potential in certain real-world industrial applications, yet its broader deployment remains limited by inherent challenges such as sample inefficiency and unstable learning dynamics.…

Machine Learning · Computer Science 2025-07-03 Tom Maus , Asma Atamna , Tobias Glasmachers

Optimizing quantum circuits is challenging due to the very large search space of functionally equivalent circuits and the necessity of applying transformations that temporarily decrease performance to achieve a final performance…

Quantum Physics · Physics 2023-07-20 Zikun Li , Jinjun Peng , Yixuan Mei , Sina Lin , Yi Wu , Oded Padon , Zhihao Jia

This paper investigates the problem of designing control policies that satisfy high-level specifications described by signal temporal logic (STL) in unknown, stochastic environments. While many existing works concentrate on optimizing the…

Systems and Control · Electrical Eng. & Systems 2024-12-16 Siqi Wang , Shaoyuan Li , Li Yin , Xiang Yin

Neural control is an exciting mystery which we instinctively master. Yet, researchers have a hard time explaining the motor control trajectories. Physiologically accurate biomechanical simulations can, to some extent, mimic live subjects…

Signal Processing · Electrical Eng. & Systems 2019-12-10 Amir H. Abdi , Masoud Malakoutian , Thomas Oxland , Sidney Fels

Reinforcement Learning (RL) trains agents to learn optimal behavior by maximizing reward signals from experience datasets. However, RL training often faces memory limitations, leading to execution latencies and prolonged training times. To…

The promise of fault-tolerant quantum computing is challenged by environmental drift that relentlessly degrades the quality of quantum operations. The contemporary solution, halting the entire quantum computation for recalibration, is…

Quantum Physics · Physics 2026-03-10 Volodymyr Sivak , Alexis Morvan , Michael Broughton , Rodrigo G. Cortiñas , Johannes Bausch , Andrew W. Senior , Matthew Neeley , Alec Eickbusch , Noah Shutty , Laleh Aghababaie Beni , James S. Spencer , Francisco J. H Heras , Thomas Edlich , Dmitry Abanin , Amira Abbas , Rajeev Acharya , Georg Aigeldinger , Ross Alcaraz , Sayra Alcaraz , Trond I. Andersen , Markus Ansmann , Frank Arute , Kunal Arya , Walt Askew , Nikita Astrakhantsev , Juan Atalaya , Brian Ballard , Joseph C. Bardin , Hector Bates , Andreas Bengtsson , Majid Bigdeli Karimi , Alexander Bilmes , Simon Bilodeau , Felix Borjans , Alexandre Bourassa , Jenna Bovaird , Dylan Bowers , Leon Brill , Peter Brooks , David A. Browne , Brett Buchea , Bob B. Buckley , Tim Burger , Brian Burkett , Nicholas Bushnell , Jamal Busnaina , Anthony Cabrera , Juan Campero , Hung-Shen Chang , Silas Chen , Ben Chiaro , Liang-Ying Chih , Agnetta Y. Cleland , Bryan Cochrane , Matt Cockrell , Josh Cogan , Roberto Collins , Paul Conner , Harold Cook , William Courtney , Alexander L. Crook , Ben Curtin , Martin Damyanov , Sayan Das , Dripto M. Debroy , Sean Demura , Paul Donohoe , Ilya Drozdov , Andrew Dunsworth , Valerie Ehimhen , Aviv Moshe Elbag , Lior Ella , Mahmoud Elzouka , David Enriquez , Catherine Erickson , Vinicius S. Ferreira , Marcos Flores , Leslie Flores Burgos , Ebrahim Forati , Jeremiah Ford , Austin G. Fowler , Brooks Foxen , Masaya Fukami , Alan Wing Lun Fung , Lenny Fuste , Suhas Ganjam , Gonzalo Garcia , Christopher Garrick , Robert Gasca , Helge Gehring , Robert Geiger , Élie Genois , William Giang , Dar Gilboa , James E. Goeders , Edward C. Gonzales , Raja Gosula , Stijn J. de Graaf , Alejandro Grajales Dau , Dietrich Graumann , Joel Grebel , Alex Greene , Jonathan A. Gross , Jose Guerrero , Loïck Le Guevel , Tan Ha , Steve Habegger , Tanner Hadick , Ali Hadjikhani , Michael C. Hamilton , Matthew P. Harrigan , Sean D. Harrington , Jeanne Hartshorn , Stephen Heslin , Paula Heu , Oscar Higgott , Reno Hiltermann , Hsin-Yuan Huang , Mike Hucka , Christopher Hudspeth , Ashley Huff , William J. Huggins , Evan Jeffrey , Shaun Jevons , Zhang Jiang , Xiaoxuan Jin , Chaitali Joshi , Pavol Juhas , Andreas Kabel , Dvir Kafri , Hui Kang , Kiseo Kang , Amir H. Karamlou , Ryan Kaufman , Kostyantyn Kechedzhi , Tanuj Khattar , Mostafa Khezri , Seon Kim , Can M. Knaut , Bryce Kobrin , Fedor Kostritsa , John Mark Kreikebaum , Ryuho Kudo , Ben Kueffler , Arun Kumar , Vladislav D. Kurilovich , Vitali Kutsko , Nathan Lacroix , David Landhuis , Tiano Lange-Dei , Brandon W. Langley , Pavel Laptev , Kim-Ming Lau , Justin Ledford , Joy Lee , Kenny Lee , Brian J. Lester , Wendy Leung , Lily Li , Wing Yan Li , Ming Li , Alexander T. Lill , William P. Livingston , Matthew T. Lloyd , Aditya Locharla , Laura De Lorenzo , Daniel Lundahl , Aaron Lunt , Sid Madhuk , Aniket Maiti , Ashley Maloney , Salvatore Mandrà , Leigh S. Martin , Orion Martin , Eric Mascot , Paul Masih Das , Dmitri Maslov , Melvin Mathews , Cameron Maxfield , Jarrod R. McClean , Matt McEwen , Seneca Meeks , Kevin C. Miao , Zlatko K. Minev , Reza Molavi , Sebastian Molina , Shirin Montazeri , Charles Neill , Michael Newman , Anthony Nguyen , Murray Nguyen , Chia-Hung Ni , Murphy Yuezhen Niu , Logan Oas , Raymond Orosco , Kristoffer Ottosson , Alice Pagano , Agustin Di Paolo , Sherman Peek , David Peterson , Alex Pizzuto , Elias Portoles , Rebecca Potter , Orion Pritchard , Michael Qian , Chris Quintana , Arpit Ranadive , Matthew J. Reagor , Rachel Resnick , David M. Rhodes , Daniel Riley , Gabrielle Roberts , Roberto Rodriguez , Emma Ropes , Lucia B. De Rose , Eliott Rosenberg , Emma Rosenfeld , Dario Rosenstock , Elizabeth Rossi , Pedram Roushan , David A. Rower , Robert Salazar , Kannan Sankaragomathi , Murat Can Sarihan , Kevin J. Satzinger , Max Schaefer , Sebastian Schroeder , Henry F. Schurkus , Aria Shahingohar , Michael J. Shearn , Aaron Shorter , Vladimir Shvarts , Spencer Small , W. Clarke Smith , David A. Sobel , Barrett Spells , Sofia Springer , George Sterling , Jordan Suchard , Aaron Szasz , Alexander Sztein , Madeline Taylor , Jothi Priyanka Thiruraman , Douglas Thor , Dogan Timucin , Eifu Tomita , Alfredo Torres , M. Mert Torunbalci , Hao Tran , Abeer Vaishnav , Justin Vargas , Sergey Vdovichev , Guifre Vidal , Catherine Vollgraff Heidweiller , Meghan Voorhees , Steven Waltman , Jonathan Waltz , Shannon X. Wang , Brayden Ware , James D. Watson , Yonghua Wei , Travis Weidel , Theodore White , Kristi Wong , Bryan W. K. Woo , Christopher J. Wood , Maddy Woodson , Cheng Xing , Z. Jamie Yao , Ping Yeh , Bicheng Ying , Juhwan Yoo , Noureldin Yosri , Elliot Young , Grayson Young , Adam Zalcman , Ran Zhang , Yaxing Zhang , Ningfeng Zhu , Nicholas Zobrist , Zhenjie Zou , Ryan Babbush , Dave Bacon , Sergio Boixo , Yu Chen , Zijun Chen , Michel Devoret , Monica Hansen , Jeremy Hilton , Cody Jones , Julian Kelly , Alexander N. Korotkov , Erik Lucero , Anthony Megrant , Hartmut Neven , William D. Oliver , Ganesh Ramachandran , Vadim Smelyanskiy , Paul V. Klimov

While reinforcement learning (RL) can empower autonomous agents by enabling self-improvement through interaction, its practical adoption remains challenging due to costly rollouts, limited task diversity, unreliable reward signals, and…

Program Synthesis is the task of generating a program from a provided specification. Traditionally, this has been treated as a search problem by the programming languages (PL) community and more recently as a supervised learning problem by…

Artificial Intelligence · Computer Science 2018-06-11 Riley Simmons-Edler , Anders Miltner , Sebastian Seung

Logic synthesis is a challenging and widely-researched combinatorial optimization problem during integrated circuit (IC) design. It transforms a high-level description of hardware in a programming language like Verilog into an optimized…

Machine Learning · Computer Science 2021-10-25 Animesh Basak Chowdhury , Benjamin Tan , Ramesh Karri , Siddharth Garg

Large Language Models (LLMs), when enhanced through reasoning-oriented post-training, evolve into powerful Large Reasoning Models (LRMs). Tool-Integrated Reasoning (TIR) further extends their capabilities by incorporating external tools,…

Computation and Language · Computer Science 2025-07-30 Yifan Wei , Xiaoyan Yu , Yixuan Weng , Tengfei Pan , Angsheng Li , Li Du