English

Reinforced Large Language Model is a formal theorem prover

Artificial Intelligence 2025-02-14 v1

Abstract

To take advantage of Large Language Model in theorem formalization and proof, we propose a reinforcement learning framework to iteratively optimize the pretrained LLM by rolling out next tactics and comparing them with the expected ones. The experiment results show that it helps to achieve a higher accuracy compared with directly fine-tuned LLM.

Keywords

Cite

@article{arxiv.2502.08908,
  title  = {Reinforced Large Language Model is a formal theorem prover},
  author = {Zhiling Luo},
  journal= {arXiv preprint arXiv:2502.08908},
  year   = {2025}
}
R2 v1 2026-06-28T21:42:28.573Z