arXiv · 2502.08908
Reinforced Large Language Model is a formal theorem prover
Abstract
To take advantage of Large Language Model in theorem formalization and proof, we propose a reinforcement learning framework to iteratively optimize the pretrained LLM by rolling out next tactics and comparing them with the expected ones. The experiment results show that it helps to achieve a higher accuracy compared with directly fine-tuned LLM.
Explore related subjects
Keep this discovery
Zhiling Luo. 2025-02-13. Reinforced Large Language Model is a formal theorem prover. https://arxiv.org/abs/2502.08908
Cite the original work for its findings. Save a collection to share your selection of sources.