Google is scaling its custom TPU hardware to power Gemini models, reducing reliance on external vendors. Meanwhile, AMD challenges Nvidia's dominance by optimizing its Instinct accelerators for large-scale inference. These infrastructure shifts lower operational costs for developers. The move signals a broader industry trend toward vertically integrated AI silicon to escape high GPU premiums.