Three GPT-5.6 variants—Sol, Terra, and Luna—now support cross-Region inference across 25 AWS Regions. This system routes requests to destination regions via inference profiles to bypass local capacity bottlenecks. It functions primarily as a throughput mechanism. Practitioners can now draw from a broader compute pool to reduce latency and avoid regional outages.