The vLLM high-throughput inference engine now supports Baidu Kunlun accelerators. This integration expands the library's hardware compatibility beyond NVIDIA and AMD GPUs. Developers can now deploy large language models on Baidu's proprietary silicon. It is an incremental update that reduces vendor lock-in for users of Chinese AI infrastructure.