releases.sh

Prompt caching live for GPT-OSS 20B; 50% cost savings on cached tokens

September 25, 2025ChangelogView original ↗
1 featureThis release1 featureNew capabilitiesAI-tallied from the release notes

Automatic prompt caching now live for openai/gpt-oss-20b: 50% cost savings on cached input tokens ($0.037/M vs $0.075/M), lower latency, and automatic prefix matching. Zero setup required.

Fetched August 11, 2026

Prompt caching live for GPT-OSS 20B; 50% cost savings on… — releases.sh