2000 character limit reached
Reinforced Large Language Model is a formal theorem prover
Published 13 Feb 2025 in cs.AI | (2502.08908v1)
Abstract: To take advantage of LLM in theorem formalization and proof, we propose a reinforcement learning framework to iteratively optimize the pretrained LLM by rolling out next tactics and comparing them with the expected ones. The experiment results show that it helps to achieve a higher accuracy compared with directly fine-tuned LLM.
Paper Prompts
Sign up for free to create and run prompts on this paper using GPT-5.
Top Community Prompts
Collections
Sign up for free to add this paper to one or more collections.