Tests on Coconut and CODI models show they barely use hidden reasoning steps for logical tasks. Forcing early termination of these steps rarely changes the final response. High performance stems from specific training data rather than inference-time thinking. This suggests latent reasoning may be less critical for logic than previously assumed.