Google launched an LLM router to optimize model selection based on task complexity. Meanwhile, Anthropic partnered with Volta to scale specialized AI deployments. These updates reflect a push toward efficiency over raw power. Practitioners can now route queries to smaller models to reduce latency and compute costs.