English
Related papers

Related papers: Conformal Bootstrap with Reinforcement Learning

200 papers

Reinforcement finetuning (RFT) is a key technique for aligning Large Language Models (LLMs) with human preferences and enhancing reasoning, yet its effectiveness is highly sensitive to which tasks are explored during training. Uniform task…

Artificial Intelligence · Computer Science 2026-02-02 Qianli Shen , Daoyuan Chen , Yilun Huang , Zhenqing Ling , Yaliang Li , Bolin Ding , Jingren Zhou

Reinforcement learning (RL) is often credited with improving language model reasoning and generalization at the expense of degrading memorized knowledge. We challenge this narrative by observing that RL-enhanced models consistently…

Computation and Language · Computer Science 2025-11-11 Renfei Zhang , Manasa Kaniselvan , Niloofar Mireshghallah

The computational cost of stiff chemical kinetics remains a dominant bottleneck in reacting-flow simulation, yet hybrid integration strategies are typically driven by hand-tuned heuristics or supervised predictors that make myopic decisions…

Machine Learning · Computer Science 2026-04-02 Eloghosa Ikponmwoba , Opeoluwa Owoyele

Finite-size effects limit the accuracy with which conformal data can be extracted from lattice simulations of critical systems. While action improvement suppresses some corrections to scaling, it does not address operator-dependent effects…

Strongly Correlated Electrons · Physics 2026-05-29 Lior Oppenheim , Snir Gazit , Zohar Ringel

Mobile robots are increasingly being employed for performing complex tasks in dynamic environments. Reinforcement learning (RL) methods are recognized to be promising for specifying such tasks in a relatively simple manner. However, the…

Artificial Intelligence · Computer Science 2017-11-08 Angel Martínez-Tenor , Juan Antonio Fernández-Madrigal , Ana Cruz-Martín , Javier González-Jiménez

Reinforcement learning (RL) is a fundamental framework for sequential decision-making, in which an agent learns an optimal policy through interactions with an unknown environment. In settings with function approximation, many existing RL…

Machine Learning · Computer Science 2026-05-05 Ruiquan Huang , Donghao Li , Yingbin Liang , Jing Yang

We demonstrate that the Ising model on a general triangular graph with 3 distinct couplings $K_1,K_2,K_3$ corresponds to an affine transformed conformal field theory (CFT). Full conformal invariance of the $c= 1/2$ minimal CFT is restored…

High Energy Physics - Theory · Physics 2023-08-02 Richard C. Brower , Evan K. Owen

Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergence guarantees under minimal assumptions on the problem data, they can exhibit the slow…

Optimization and Control · Mathematics 2026-05-18 Jeremy Bertoncini , Alberto De Marchi , Matthias Gerdts , Simon Gottschalk

We use the embedding formalism to construct conformal fields in $D$ dimensions, by restricting Lorentz-invariant ensembles of homogeneous neural networks in $(D+2)$ dimensions to the projective null cone. Conformal correlators may be…

High Energy Physics - Theory · Physics 2025-10-07 James Halverson , Joydeep Naskar , Jiahua Tian

In this thesis we study two-dimensional conformal field theories with Virasoro algebra symmetry, following the conformal bootstrap approach. Under the assumption that degenerate fields exist, we provide an extension of the analytic…

High Energy Physics - Theory · Physics 2019-02-06 Santiago Migliaccio

Transformers can acquire Chain-of-Thought (CoT) capabilities to solve complex reasoning tasks through fine-tuning. Reinforcement learning (RL) and supervised fine-tuning (SFT) are two primary approaches to this end. In this work, we…

Machine Learning · Computer Science 2026-05-27 Bochen Lyu , Yiyang Jia , Xiaohao Cai , Zhanxing Zhu

We present a systematic exploration of conformal field theories (CFTs) constrained by duality-inspired fusion rules using the conformal bootstrap. We classify the operator spectrum into three sectors: $[\sigma]$, $[\epsilon]$, and $[1]$.…

High Energy Physics - Theory · Physics 2026-05-20 Yu Nakayama , Toshiki Onagi

Adapting large language models to multiple tasks can cause cross-skill interference, where improvements for one skill degrade another. While methods such as LoRA impose orthogonality constraints at the weight level, they do not fully…

Computation and Language · Computer Science 2025-04-29 Andy Zhou

We use the conformal bootstrap to perform a precision study of 3d maximally supersymmetric ($\mathcal{N}=8$) SCFTs that describe the IR physics on $N$ coincident M2-branes placed either in flat space or at a $\C^4/\Z_2$ singularity. First,…

High Energy Physics - Theory · Physics 2018-08-01 Nathan B. Agmon , Shai M. Chester , Silviu S. Pufu

Geosteering, a key component of drilling operations, traditionally involves manual interpretation of various data sources such as well-log data. This introduces subjective biases and inconsistent procedures. Academic attempts to solve…

Machine Learning · Computer Science 2025-04-08 Ressi Bonti Muhammad , Apoorv Srivastava , Sergey Alyaev , Reidar Brumer Bratvold , Daniel M. Tartakovsky

We explain how the axioms of Conformal Field Theory are used to make predictions about critical exponents of continuous phase transitions in three dimensions, via a procedure called the conformal bootstrap. The method assumes conformal…

Mathematical Physics · Physics 2020-11-09 Slava Rychkov

Reinforcement learning (RL) has proven to be well-performed and general-purpose in the inventory control (IC). However, further improvement of RL algorithms in the IC domain is impeded due to two limitations of online experience. First,…

Machine Learning · Computer Science 2025-02-18 Zifan Liu , Xinran Li , Shibo Chen , Gen Li , Jiashuo Jiang , Jun Zhang

This paper studies satisfaction of temporal properties on unknown stochastic processes that have continuous state spaces. We show how reinforcement learning (RL) can be applied for computing policies that are finite-memory and deterministic…

Systems and Control · Electrical Eng. & Systems 2020-09-29 Milad Kazemi , Sadegh Soudjani

The dimensional reductions in the branched polymer and the random field Ising model (RFIM) are discussed by a conformal bootstrap method. The small size minors are applied for the evaluations of the scale dimensions of these two models and…

Disordered Systems and Neural Networks · Physics 2019-06-27 Shinobu Hikami

Online matching problems arise in many complex systems, from cloud services and online marketplaces to organ exchange networks, where timely, principled decisions are critical for maintaining high system performance. Traditional heuristics…

Machine Learning · Statistics 2025-10-09 Chiara Mignacco , Matthieu Jonckheere , Gilles Stoltz
‹ Prev 1 8 9 10 Next ›