Prompt caching live for GPT-OSS 20B; 50% cost savings on cached tokens
Automatic prompt caching now live for openai/gpt-oss-20b: 50% cost savings on cached input tokens ($0.037/M vs $0.075/M), lower latency, and automatic prefix matching. Zero setup required.
Fetched August 11, 2026

