What is KV cache?

A temporary store of results an AI model has already computed while processing a conversation. Being able to hand this off directly to another model, instead of recalculating everything from scratch, is what makes switching between models faster.

Briefings mentioning this
2026-08-252026-08-22

Terms seen alongside this one