Product feature · models-context
Prompt cachingCommon product term
Reuse eligible repeated prompt content under documented caching behavior.
Terminology basis: Common product term. Anthropic — How Claude Code uses prompt caching
Prompt caching: 1 supported, 1 partial, 0 unsupported, 29 unreviewed across 31 cataloged products.
Current evidence by product
Can my agent use Prompt caching?
Read across for the answer. 2 of 31 current product columns have reviewed evidence; unreviewed does not mean unsupported.
Web
9 productsDesktop
13 productsCLI
9 productsDefinition and scope
What this capability means
This row concerns model-request prompt or context caching: eligible repeated prefixes receive documented processing, latency, or billing reuse. It does not include browser caches, downloaded-file caches, embedding indexes, retrieval caches, build caches, or a conversation merely retaining its history.
Evidence should identify whether reuse is automatic or explicit, minimum eligible prefix size, exact-prefix requirements, supported models and regions, cache lifetime, isolation boundary, invalidation behavior, and read-versus-write pricing. A model API feature does not prove that a hosted chat or coding harness preserves stable prefixes or passes cache controls through.
Traceable compatibility
Assertion ledger
Documentation evidence only. No runtime conformance test is implied.
- Target
- current Claude Code documentation · dated-documentation
- Environment
- local-default
- Observed
- 2026-08-28
- runtimeautomatic exact-prefix caching covers stable request layers; switching models, reconnecting MCP servers, compaction, and upgrades can invalidate all or part of the prefix
- policycache infrastructure and retention depend on the authentication and serving provider
- Anthropic — How Claude Code uses prompt cachingdocumented · reviewed 2026-08-28
- Target
- current Gemini CLI documentation · dated-documentation
- Environment
- local-default
- Observed
- 2026-08-28
- runtimeautomatic caching reuses previous system instructions and context for Gemini API-key and Vertex AI users
- policyOAuth users through Google Personal or Enterprise Code Assist do not receive cached-content creation
- Google — Gemini CLI token cachingdocumented · reviewed 2026-08-28
- 1. Evidence checked 2026-08-28: Claude Code documents automatic prefix-based prompt caching for its system prompt, project context, conversation history, and tool results, with explicit invalidation behavior.
- 2. Evidence checked 2026-08-28: Gemini CLI documents automatic token caching for Gemini API-key and Vertex AI authentication, while OAuth through Code Assist does not support cached-content creation.
No sourced issues are attached to this row.
- Methodology note
- Anthropic — How Claude Code uses prompt caching documented · Anthropic · reviewed 2026-08-28
- Google — Gemini CLI token caching documented · Google · reviewed 2026-08-28
Found a wrong or incomplete result?
Name the exact product, describe what happened, add the date, and link public evidence when available. A community report starts a review. It does not change the support state by itself.