English
Related papers

Related papers: Linear Depth Increase of Lambda Terms along Leftmo…

200 papers

Confidence estimation is crucial for reflecting the reliability of large language models (LLMs), particularly in the widely used closed-source models. Utilizing data augmentation for confidence estimation is viable, but discussions focus on…

Machine Learning · Computer Science 2025-06-16 Rui Wang , Renyu Zhu , Minmin Lin , Runze Wu , Tangjie Lv , Changjie Fan , Haobo Wang

The optimized linear $\delta$-expansion is applied to the $\lambda \phi^4$ theory at high temperature. Using the imaginary time formalism the thermal mass is evaluated perturbatively up to order $\delta^2$. A variational procedure…

High Energy Physics - Phenomenology · Physics 2014-11-17 Marcus B. Pinto , Rudnei O. Ramos

The increasing size and complexity of pre-trained language models have demonstrated superior performance in many applications, but they usually require large training datasets to be adequately trained. Insufficient training sets could…

Computation and Language · Computer Science 2025-02-03 Yaping Chai , Haoran Xie , Joe S. Qin

In a previous paper an asymptotic expansion for lambda_d in powers of 1/d was developed. The results of computer computations for some terms in the expansion, as well as various quantities associated to the expansion, are herein presented.…

Mathematical Physics · Physics 2008-05-30 Paul Federbush

We use the optimized perturbation theory, or linear delta expansion, to evaluate the critical exponents in the critical 3d O(N) invariant scalar field model. Regarding the implementation procedure, this is the first successful attempt to…

Other Condensed Matter · Physics 2009-11-10 Marcus Benghi Pinto , Rudnei O. Ramos , Paulo J. Sena

Self-Refinement refers to a model's ability to revise its own responses to produce improved outputs. This capability can also serve as a fundamental mechanism for Self-Improvement, for example, by reconstructing datasets with refined…

Computation and Language · Computer Science 2025-10-28 Yongcheng Zeng , Xinyu Cui , Xuanfa Jin , Qirui Mi , Guoqing Liu , Zexu Sun , Mengyue Yang , Dong Li , Weiyu Ma , Ning Yang , Jian Zhao , Jianye Hao , Haifeng Zhang , Jun Wang

This paper is devoted to the design of efficient primal-dual algorithm (PDA) for solving convex optimization problems with known saddle-point structure. We present a new PDA with larger acceptable range of parameters and correction, which…

Optimization and Control · Mathematics 2019-12-04 Xiaokai Chang , Sanyang Liu

For a real number $0<\lambda<2$, we introduce a transformation $T_\lambda$ naturally associated to expansion in $\lambda$-continued fraction, for which we also give a geometrical interpretation. The symbolic coding of the orbits of…

Probability · Mathematics 2011-04-04 Elise Janvresse , Benoît Rittaud , Thierry De La Rue

We introduce a calculus of extensional resource terms. These are resource terms \`a la Ehrhard-Regnier, but in infinitely eta-long form. The calculus still retains a finite syntax and dynamics: in particular, we prove strong confluence and…

Logic in Computer Science · Computer Science 2026-04-22 Lison Blondeau-Patissier , Pierre Clairambault , Lionel Vaux Auclair

The basis of the $\{\beta\}$-expansion for the perturbative series evaluated in the $\overline{MS}$ scheme for the renormalization group invariant quantities is summarized.Comparison with a similar representation,used within the…

High Energy Physics - Phenomenology · Physics 2015-01-20 A. L. Kataev , S. V. Mikhailov

The training of deep residual neural networks (ResNets) with backpropagation has a memory cost that increases linearly with respect to the depth of the network. A way to circumvent this issue is to use reversible architectures. In this…

Machine Learning · Computer Science 2021-07-23 Michael E. Sander , Pierre Ablin , Mathieu Blondel , Gabriel Peyré

Scalable sequence models, such as Transformer variants and structured state-space models, often trade expressivity power for sequence-level parallelism, which enables efficient training. Here we examine the bounds on error and how error…

Machine Learning · Computer Science 2026-03-09 Gyuryang Heo , Timothy Ngotiaoco , Kazuki Irie , Samuel J. Gershman , Bernardo Sabatini

In recent months, substantial progress has been made in complex reasoning of Large Language Models, particularly through the application of test-time scaling. Notable examples include o1/o3/o4 series and DeepSeek-R1. When responding to a…

Artificial Intelligence · Computer Science 2025-08-28 Xifeng Yao , Chengyuan Ma , Dongyu Lang , Yinhao Ni , Zhiwei Xu , Huarui Xie , Zihao Chen , Guang Shen , Dandan Tu , Yi Bai , Changzheng Zhang

The bisimulation proof method can be enhanced by employing `bisimulations up-to' techniques. A comprehensive theory of such enhancements has been developed for first-order (i.e., CCS-like) labelled transition systems (LTSs) and…

Logic in Computer Science · Computer Science 2023-06-22 Jean-Marie Madiot , Damien Pous , Davide Sangiorgi

We study the typical growth rate of the number of words of length n which can be extended to beta-expansions of x. In the general case we give a lower bound for the growth rate, while in the case that the Bernoulli convolution associated to…

Dynamical Systems · Mathematics 2012-03-27 Tom Kempton

This paper introduces a simple and scalable approach to improve the data efficiency of large language model (LLM) training by augmenting existing text data with thinking trajectories. The compute for pre-training LLMs has been growing at an…

Computation and Language · Computer Science 2025-10-20 Liang Wang , Nan Yang , Shaohan Huang , Li Dong , Furu Wei

We positively answer the question A.1.6 in J. Klop's "Ustica Notes": "Is there a recursive normalizing one-step reduction strategy for micro $\lambda$-calculus?" Micro $\lambda$-calculus refers to an implementation of the $\lambda$-calculus…

Logic in Computer Science · Computer Science 2014-05-02 Anton Salikhmetov

Consider $\beta > 1$ and $\lfloor \beta \rfloor$ its integer part. It is widely known that any real number $\alpha \in \Bigl[0, \frac{\lfloor \beta \rfloor}{\beta - 1}\Bigr]$ can be represented in base $\beta$ using a development in series…

Dynamical Systems · Mathematics 2022-05-11 Victor Vargas

The linear-algebraic lambda-calculus and the algebraic lambda-calculus are untyped lambda-calculi extended with arbitrary linear combinations of terms. The former presents the axioms of linear algebra in the form of a rewrite system, while…

Logic in Computer Science · Computer Science 2012-03-29 Pablo Buiras , Alejandro Díaz-Caro , Mauro Jaskelioff

The evolving sophistication and intricacies of Large Language Models (LLMs) yield unprecedented advancements, yet they simultaneously demand considerable computational resources and incur significant costs. To alleviate these challenges,…

Computation and Language · Computer Science 2023-10-03 Hongye Jin , Xiaotian Han , Jingfeng Yang , Zhimeng Jiang , Chia-Yuan Chang , Xia Hu