Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms Paper • 2608.04074 • Published 11 days ago • 2