LeanConjecturer: Automatic Generation Of Mathematical Conjectures For Theorem Proving

arXiv:2506.22005v1 Announce Type: new
Abstract: We introduce LeanConjecturer, a pipeline for automatically generating university-level mathematical conjectures in Lean 4 using Large Language Models (LLMs). Our hybrid approach combines rule-based context extraction with LLM-based theorem statement generation, addressing the data scarcity challenge in formal theorem proving. Through iterative generation and evaluation, LeanConjecturer produced 12,289 conjectures from 40 Mathlib seed files, with 3,776 identified as syntactically valid and non-trivial, that is, cannot be proven by \texttt{aesop} tactic. We demonstrate the utility of these generated conjectures for reinforcement learning through Group Relative Policy Optimization (GRPO), showing that targeted training on domain-specific conjectures can enhance theorem proving capabilities. Our approach generates 103.25 novel conjectures per seed file on average, providing a scalable solution for creating training data for theorem proving systems. Our system successfully verified several non-trivial theorems in topology, including properties of semi-open, alpha-open, and pre-open sets, demonstrating its potential for mathematical discovery beyond simple variations of existing results.

Source link

What's Hot

NASA, IBM’s AI model set to unlock Sun mysteries

Cisco ties storage networking gear to IBM z17 mainframe

RynnEC: Bringing MLLMs into Embodied World – Takara TLDR

LeanConjecturer: Automatic Generation of Mathematical Conjectures for Theorem Proving

LTLCrit: A Temporal Logic-based LLM Critic for Safe and Efficient Embodied Agents

From Imitation to Innovation: The Emergence of AI Unique Artistic Styles and the Challenge of Copyright Protection

VerifyLLM: LLM-Based Pre-Execution Task Plan Verification for Robots

Tanya Bonakdar Gallery to Close Los Angeles Space

Ancient Silver Coins Suggest New History of Trading in Southeast Asia

Sasan Ghandehari Sues Christie’s Over Picasso Once Owned by a Criminal

Ancient Roman Villa in Sicily Reveals Mosaic of Flip-Flops

NASA, IBM’s AI model set to unlock Sun mysteries

Cisco ties storage networking gear to IBM z17 mainframe

RynnEC: Bringing MLLMs into Embodied World – Takara TLDR

What's Hot

LeanConjecturer: Automatic Generation of Mathematical Conjectures for Theorem Proving

Related Posts

Subscribe to Updates