English
Related papers

Related papers: Implicitly Restarted Lanczos Enables Chemically-Ac…

200 papers

Inverse reinforcement learning (IRL) recovers reward functions from observed behavior, yet traditional methods assume a single stationary reward that cannot capture goal switching within an episode. Recent multi-intention IRL methods…

Machine Learning · Computer Science 2026-05-27 Wenyuan Sheng , Hao Zhu , Joschka Boedecker

The key approaches for machine learning, especially learning in unknown probabilistic environments are new representations and computation mechanisms. In this paper, a novel quantum reinforcement learning (QRL) method is proposed by…

Quantum Physics · Physics 2008-10-22 Daoyi Dong , Chunlin Chen , Hanxiong Li , Tzyh-Jong Tarn

Post-Training Quantization (PTQ) is essential for deploying Large Language Models (LLMs) on memory-constrained devices, yet it renders models static and difficult to fine-tune. Standard fine-tuning paradigms, including Reinforcement…

Machine Learning · Computer Science 2026-02-04 Yinggan Xu , Risto Miikkulainen , Xin Qiu

Neural Quantum States (NQS) are now among the most accurate methods for studying strongly correlated many-fermion systems, outperforming existing many-body approaches for large systems. However, NQS calculations remain extremely…

Strongly Correlated Electrons · Physics 2026-04-29 Yuntian Gu , Zeyao Han , Wenrui Li , Zhiyu Xiao , Tao Xiang , Mingpu Qin , Liwei Wang , Dingshun Lv

Reinforcement learning with verifiable rewards (RLVR) has become a trending paradigm for training reasoning large language models (LLMs). However, due to the autoregressive decoding nature of LLMs, the rollout process becomes the efficiency…

Machine Learning · Computer Science 2026-02-17 Yuhang Li , Reena Elangovan , Xin Dong , Priyadarshini Panda , Brucek Khailany

Multiple rotation averaging (MRA) is a fundamental optimization problem in 3D vision and robotics that aims to recover globally consistent absolute rotations from noisy relative measurements. Established classical methods, such as L1-IRLS…

Computer Vision and Pattern Recognition · Computer Science 2026-02-11 Shuteng Wang , Natacha Kuete Meli , Michael Möller , Vladislav Golyanik

Reinforcement learning (RL) has transformed sequential decision making, yet traditional algorithms like Deep Q-Networks (DQNs) and Proximal Policy Optimization (PPO) often struggle with efficient exploration, stability, and adaptability in…

Machine Learning · Computer Science 2025-05-13 Umberto Gonçalves de Sousa

The gloabal objective of inverse Reinforcement Learning (IRL) is to estimate the unknown cost function of some MDP base on observed trajectories generated by (approximate) optimal policies. The classical approach consists in tuning this…

Machine Learning · Computer Science 2021-05-26 Firas Jarboui , Vianney Perchet

The preparation of quantum states is essential in the realm of quantum information processing, and the development of efficient methodologies can significantly alleviate the strain on quantum resources. Within the framework of deep…

Quantum Physics · Physics 2024-07-24 Zhao-Wei Wang , Zhao-Ming Wang

Due to the large state space of the two-qubit system, and the adoption of ladder reward function in the existing quantum state preparation methods, the convergence speed is slow and it is difficult to prepare the desired target quantum…

Quantum Physics · Physics 2024-05-14 Wenjie Liu , Jing Xu , Bosi Wang

We present BenchRL-QAS, a unified benchmarking framework for reinforcement learning (RL) in quantum architecture search (QAS) across a spectrum of variational quantum algorithm tasks on 2- to 8-qubit systems. Our study systematically…

Quantum Physics · Physics 2025-09-30 Azhar Ikhtiarudin , Aditi Das , Param Thakkar , Akash Kundu

Recent studies show that the reasoning capabilities of Large Language Models (LLMs) can be improved by applying Reinforcement Learning (RL) to question-answering (QA) tasks in areas such as math and coding. With a long context length, LLMs…

Computation and Language · Computer Science 2025-10-17 Stephen Chung , Wenyu Du , Jie Fu

In this paper, the problem of associating reconfigurable intelligent surfaces (RISs) to virtual reality (VR) users is studied for a wireless VR network. In particular, this problem is considered within a cellular network that employs…

Information Theory · Computer Science 2020-02-24 Christina Chaccour , Mehdi Naderi Soorki , Walid Saad , Mehdi Bennis , Petar Popovski

Variational quantum algorithms (VQAs) are leading strategies for using near-term quantum devices, with a well-studied bottleneck being their trainability. Standard expectation-value objectives with expressive circuits frequently encounter…

Quantum Physics · Physics 2026-05-05 Yixian Qiu , Josep Lumbreras , Xiufan Li , Patrick Rebentrost

This paper studies the problem of identifying low-order linear systems via Hankel nuclear norm regularization. Hankel regularization encourages the low-rankness of the Hankel matrix, which maps to the low-orderness of the system. We provide…

Machine Learning · Statistics 2022-04-01 Yue Sun , Samet Oymak , Maryam Fazel

Inverse reinforcement learning (IRL) is the problem of finding a reward function that generates a given optimal policy for a given Markov Decision Process. This paper looks at an algorithmic-independent geometric analysis of the IRL problem…

Machine Learning · Computer Science 2021-02-19 Abi Komanduru , Jean Honorio

We develop an efficient data-driven and model-free unsupervised learning algorithm for achieving fully passive intelligent reflective surface (IRS)-assisted optimal short/long-term beamforming in wireless communication networks. The…

Signal Processing · Electrical Eng. & Systems 2024-11-01 Spyridon Pougkakiotis , Hassaan Hashmi , Dionysis Kalogerias

This letter investigates the reconfigurable intelligent surface (RIS)-assisted multiple-input single-output (MISO) wireless system, where both half-duplex (HD) and full-duplex (FD) operating modes are considered together, for the first time…

Signal Processing · Electrical Eng. & Systems 2021-10-12 Alice Faisal , Ibrahim Al-Nahhal , Octavia A. Dobre , Telex M. N. Ngatched

The growing availability and usage of low precision foating point formats has attracts many interests of developing lower or mixed precision algorithms for scientific computing problems. In this paper we investigate the possibility of…

Numerical Analysis · Mathematics 2024-02-14 Haibo Li

Aligning large language models (LLMs) with human preferences usually requires fine-tuning methods such as RLHF and DPO. These methods directly optimize the model parameters, so they cannot be used in test-time to improve model performance,…

Machine Learning · Computer Science 2025-07-04 Xinnan Zhang , Chenliang Li , Siliang Zeng , Jiaxiang Li , Zhongruo Wang , Kaixiang Lin , Songtao Lu , Alfredo Garcia , Mingyi Hong