The AFM 3 Core Advanced model powers Siri's real-time, on-device speech synthesis. Apple uses a decoupled temporal depth diffusion transformer to convert semantic tokens into high-fidelity audio. This architecture fits within the strict memory limits of the Apple Matrix Coprocessor. It enables rich, configurable voice output without relying on cloud compute.