A new system called KVBoost uses a dual-hash keying scheme to reuse key-value cache chunks regardless of their position in a prompt. This removes the requirement for shared content to be a contiguous leading prefix. Developers can now reduce prefill latency for prompts with fragmented shared content across HuggingFace decoder models.