Three distinct components—models, writers, and routers—now drive specific AI workflow optimizations. This approach separates the generation of content from the logic used to select the best model for a task. Ben's Bites highlights how this modularity reduces latency. Practitioners can now optimize cost by routing simple queries to smaller models.