Engineers on Lobste.rs are debating the efficiency of specialized AI accelerators versus general-purpose GPUs. The discussion centers on memory bandwidth bottlenecks and the viability of RISC-V for custom silicon. These technical trade-offs determine how developers optimize large-scale inference. Most contributors view current hardware gains as incremental rather than structural.