The Beijing-based lab Z.ai released GLM 5.2, an open-weights model with API costs at $4.40 per million output tokens. This pricing is less than a fifth of many frontier competitors. Engineers now route simple tasks to this cheaper model while reserving high-reasoning models like Anthropic’s Fable for complex problems to optimize spend.