Current LLMs struggle with complex, unsolved mathematical proofs that require deep logical reasoning. While synthetic data improves basic arithmetic, these models often hallucinate steps in high-level theory. Researchers now focus on integrating formal verification to stop these errors. This shift determines if AI can actually discover new math or just mimic existing textbooks.