Spend Bits Where Queries Look: KV Cache Vector Quantization with Attention-Preserving Transforms Paper • 2608.04074 • Published 9 days ago • 1