Uncached baseline
$3780.00exact-derivedAll prefix, suffix, and output terms.
[2]Portfolio 13 · prefix economics
Keep cache writes, reads, token-time storage, uncached suffixes, outputs, threshold eligibility, TTL expiry, and unstable-prefix misses visible as independent terms.
Decision this answers: At what reuse level does the stable prefix become cheaper, and is it actually eligible under the selected provider rule and TTL?
Uncached baseline
$3780.00exact-derivedAll prefix, suffix, and output terms.
[2]Cached scenario
$2118.70exact-derivedWrites, reads, storage, suffix, and output.
[1][2]Scenario saving
$1661.30exact-derivedCaching is eligible and cheaper: true.
[1]Break-even requests
2exact-derivedSuppressed as zero when caching never becomes cheaper.
[1]The tables below are the accessible source of truth; bar lengths never carry identity alone.
| Candidate | Write | Read | Storage | Eligibility | Operational model | Evidence |
|---|---|---|---|---|---|---|
| Anthropic | Explicit write | Explicit read | TTL-dependent by model | Threshold + stable prefix + TTL | Explicit cache control | published list[1] |
| Gemini | Model-specific | Cached token rate | Token-time storage | Provider rule | Explicit context cache | published list[2] |
| OpenAI | Automatic | Cached input rate | Provider-managed | Provider rule | Automatic caching | published list[3] |
Use the official surface to confirm native prices, counting rules, limits, and product eligibility.
Provider docs define eligibility, TTL, and cache meters. This page reconciles the complete request family and suppresses a break-even when caching never becomes cheaper.
Every result-affecting reference is visible here without JavaScript and is retained in the JSON export.
prompt cachinggemini-3.6-flash standard cachingautomatic prompt caching