Claude Opus 5 hits 30.2 percent on the ARC-AGI-3 benchmark, nearly four times the score of GPT-5.6 Sol. Anthropic claims the model matches Fable 5 performance while cutting token costs by half. This price drop makes high-end reasoning more accessible. Developers can now deploy flagship-grade coding and knowledge work at a fraction of previous costs.