SAFE: Enhancing Mathematical Reasoning in Large Language Models via Retrospective Step-aware Formal VerificationPublished in ACL 2025, 2025SAFE uses Lean 4 proofs to identify hallucinations in natural-language mathematical reasoning at the step level.Share on Twitter Facebook LinkedIn Previous Next