中文
相关论文

相关论文: Logarithmic light cone, slow entanglement growth, …

200 篇论文

The evolution of Large Language Model (LLM) reasoning is bottlenecked by the scarcity of high-quality process data. While self-alignment via endogenous rewards offers a solution, mining valid supervision faces three challenges: (1) Label…

人工智能 · 计算机科学 2026-05-26 Yanyu Chen , Jiyue Jiang , Dianzhi Yu , Zheng Wu , Jiahong Liu , Jiaming Han , Xiao Guo , Jinhu Qi , Yu Li , Yifei Zhang , Irwin King

Large Language Models (LLMs) face significant challenges in long-context processing, including quadratic computational costs, information forgetting, and the context fragmentation inherent in retrieval-augmented generation (RAG). We propose…

计算与语言 · 计算机科学 2026-02-10 Zhuoen Chen , Dongfang Li , Meishan Zhang , Baotian Hu , Min Zhang

We have recently shown that the logarithmic growth of the entanglement entropy following a quantum quench in a many-body localized (MBL) phase is accompanied by a slow growth of the number entropy, $S_N\sim\ln\ln t$. Here we provide an…

无序系统与神经网络 · 物理学 2021-01-22 Maximilian Kiefer-Emmanouilidis , Razmik Unanyan , Michael Fleischhauer , Jesko Sirker

Converging evidence suggests that human systems of semantic categories achieve near-optimal compression via the Information Bottleneck (IB) complexity-accuracy tradeoff. Large language models (LLMs) are not trained for this objective, which…

计算与语言 · 计算机科学 2026-03-16 Nathaniel Imel , Noga Zaslavsky

Length generalization (LG) is a challenging problem in learning to reason. It refers to the phenomenon that when trained on reasoning problems of smaller lengths or sizes, the resulting model struggles with problems of larger sizes or…

人工智能 · 计算机科学 2024-04-02 Changnan Xiao , Bing Liu

Many-body localization (MBL) is understood theoretically through the existence of an extensive number of local integrals of motion (LIOMs). These conserved quantities are related to the microscopic quantum degrees of freedom that are…

无序系统与神经网络 · 物理学 2025-12-11 Ben Craps , Oleg Evnin , Dmitry Kovrizhin , Gabriele Pascuzzi

Large Language Models (LLMs) exhibit impressive capabilities yet suffer from sensitivity to slight input context variations, hampering reliability. Conventional metrics like accuracy and perplexity fail to assess local prediction…

计算与语言 · 计算机科学 2026-02-12 Deyuan Liu , Zecheng Wang , Zhanyue Qin , Zhiying Tu , Dianhui Chu , Dianbo Sui

Large language models (LLMs) have shown their power in different areas. Attention computation, as an important subroutine of LLMs, has also attracted interests in theory. Recently the static computation and dynamic maintenance of attention…

数据结构与算法 · 计算机科学 2023-04-11 Yichuan Deng , Sridhar Mahadevan , Zhao Song

Designing large coupling memory quasi-cyclic spatially-coupled LDPC (QC-SC-LDPC) codes with low error floors requires eliminating specific harmful substructures (e.g., short cycles) induced by edge spreading and lifting. Building on our…

信息论 · 计算机科学 2026-01-21 Lei Huang

Token-based time series large language models (TS-LLMs) have emerged as a promising direction for time series analysis and reasoning. However, prior studies largely overlook the inherent continuity and ordinality of time series tokens,…

机器学习 · 计算机科学 2026-05-29 Musheng Li , Ziying Zhang , Cheng jin , Yuantao Gu

We prove global existence and uniqueness of dynamics on the quasi-local algebra $\mathcal{A}$ of a quantum lattice system for spatially growing derivations $\mathcal{L}_\Phi = \sum_x [ \Phi_x , \cdot ]$. Existing results assume that the…

数学物理 · 物理学 2025-12-10 Stefan Teufel , Marius Wesle , Tom Wessel

Large Language Models are increasingly proposed as cognitive components for robotic systems, yet their opaque decision processes make it difficult to explain success or failure in closed-loop embodied tasks. Following an empirical AI…

人工智能 · 计算机科学 2026-05-20 Oussama Zenkri , Oliver Brock

Large language model (LLM) performance on reasoning problems typically does not generalize out of distribution. Previous work has claimed that this can be mitigated with chain of thought prompting-a method of demonstrating solution…

人工智能 · 计算机科学 2025-03-13 Kaya Stechly , Karthik Valmeekam , Subbarao Kambhampati

A rich line of work has been addressing the computational complexity of locally checkable labelings (LCLs), illustrating the landscape of possible complexities. In this paper, we study the landscape of LCL complexities under bandwidth…

数据结构与算法 · 计算机科学 2021-05-18 Alkida Balliu , Keren Censor-Hillel , Yannic Maus , Dennis Olivetti , Jukka Suomela

Generalized Linear Bandits (GLBs), a natural extension of the stochastic linear bandits, has been popular and successful in recent years. However, existing GLBs scale poorly with the number of rounds and the number of arms, limiting their…

机器学习 · 统计学 2017-10-24 Kwang-Sung Jun , Aniruddha Bhargava , Robert Nowak , Rebecca Willett

The connection between entanglement dynamics and non-equilibrium statistics in isolated many-body quantum systems has been established both theoretically and experimentally. Many-Body Localization (MBL), a phenomenon where interacting…

量子物理 · 物理学 2025-07-04 Peyman Azodi , Herschel A. Rabitz

A key signature of MBL (many-body localization) is the slow rate at which information spreads. It is shown that the infinite random Heisenberg XXZ spin-$\frac12$ chain exhibits slow propagation of information (logarithmic light cone) in any…

数学物理 · 物理学 2026-03-10 Alexander Elgart , Abel Klein

Compressing long chain-of-thought (CoT) from large language models (LLMs) is an emerging strategy to improve the reasoning efficiency of LLMs. Despite its promising benefits, existing studies equally compress all thoughts within a long CoT,…

计算与语言 · 计算机科学 2025-05-27 Yansong Ning , Wei Li , Jun Fang , Naiqiang Tan , Hao Liu

Large language models (LLMs) have shown strong knowledge reserves and task-solving capabilities, but still face the challenge of severe hallucination, hindering their practical application. Though scientific theories and rules can…

计算与语言 · 计算机科学 2026-04-09 Maotian Ma , Zheni Zeng , Zhenghao Liu , Yukun Yan

Despite their impressive capabilities, aligned large language models (LLMs) often generate outputs that lack diversity. What drives this consistency in the generation? We investigate this phenomenon through the lens of probability…

计算与语言 · 计算机科学 2026-03-04 Chenghao Yang , Sida Li , Ari Holtzman