The Beijing-based lab Z.ai released GLM 5.2, an open-weights model with API costs of $4.40 per million output tokens. This pricing is less than a fifth of many frontier competitors. Engineers now route simple tasks to this cheaper model while reserving high-reasoning models for complex problems. This tiered approach optimizes developer spend without sacrificing code quality.