Paper 2608.12325v1 argues that generative AI lacks a concrete operational definition for reasoning. This ambiguity makes current evaluation metrics unverifiable and hinders progress toward trustworthy systems. The authors propose returning to symbolic AI logic to establish construct validity. Practitioners should scrutinize whether current benchmarks measure actual reasoning or mere pattern matching.