Systems session 3/3 · Quiz & review

KV-cache sizing and optimization

What happens to KV memory when active context doubles?

Show answer

Under the simplified linear estimate, it approximately doubles.

Is provider prompt caching identical to a request’s runtime KV cache?

Show answer

No. Provider prefix reuse has provider-specific persistence and eligibility rules.

What evidence closes this lesson?

Show answer

A measured result with assumptions, exact revisions, and a quality check.