Engineers on Lobste.rs are debating the efficiency of specialized AI accelerators versus general-purpose GPUs. The discussion focuses on memory bandwidth bottlenecks and the viability of RISC-V for custom silicon. These technical critiques highlight a growing frustration with current hardware constraints. Practitioners should monitor these architectural shifts to optimize future model inference costs.