Unsolved mathematical conjectures now serve as the ultimate benchmark for LLM reasoning. Current models struggle with proofs requiring deep logical leaps rather than pattern matching. This gap highlights a critical ceiling in synthetic reasoning. Researchers must move beyond training on existing datasets to achieve genuine mathematical discovery and autonomous problem solving.