Releases Index

API

OpenAI platform API and developer tools

$npx @buildinternet/releases get api
Mon
Wed
Fri
OctNovDecJanFebMarAprMayJunJulAugSepOct
Less
More
Releases92Avg Interval24hAvg Cadence30/mo

Responses API users can set service_tier: "ultrafast" on gpt-6.1-sol to reduce time between generated output tokens. Available to all API users subject to rate limits, with global processing and US and EU data residency.

Read more →

The chat-latest snapshot now points to the latest model available in ChatGPT for Plus, Pro, Business, and Enterprise users, with the underlying snapshot updated regularly. Production API usage should use the GPT-6 model family; chat-latest is for testing the latest chat improvements.

Read more →

Ultrafast mode for GPT-6 Astra is available in the Responses API via service_tier: "ultrafast", reducing the time between generated output tokens. It is subject to rate limits, restricted to global processing and US data residency, and EU and other regional inference residency aren't supported.

Read more →
gpt-6.1-sol

GPT-6.1 Sol (gpt-6.1-sol) is available for complex coding and professional work at a lower cost than GPT-6 Astra, priced at $2 input, $0.10 cached input, $2.50 cache write, and $10 output per 1M tokens for prompts up to 272K input tokens. It also supports Multi-agent in beta, letting the model delegate work to subagents in a Responses API request.

Read more →

An image encoding bug that degraded image understanding in GPT-6 Sol and GPT-6 Luna is fixed, improving results on visual tasks in the API and Codex, including computer use. OpenAI recommends rerunning evaluations and retrying workflows that use image inputs.

Read more →

Administrators can now restrict API key creation at the organization and project levels to service-account keys only, user-owned project keys only, or disable new key creation entirely, with organization restrictions taking precedence over project settings. Existing API keys are unaffected.

Read more →

Project API keys can now be created with expiration dates, and administrators can enforce a maximum key lifetime at the organization or project level in Platform settings so newly created keys expire within the configured limit.

Read more →

Prompt Cache Diagnostics is now generally available in the Responses API for GPT-5.6 and later supported models. It lets developers compare cache reuse against a previous response, identify reasons for cache misses, and follow troubleshooting guidance to improve cache reuse.

Read more →

GPT-Rosalind (gpt-rosalind-research) is now generally available through the trusted-access program for approved internal life sciences research. Standard pricing is $5 per 1M input tokens, $0.50 per 1M cached input tokens, and $25 per 1M output tokens, with billing beginning October 5, 2026.

Read more →

GPT Image 2.5 Sunburst and GPT Image 2.5 Flare are available for image generation and editing through the Image API and the Responses API image generation tool. Both support new xhigh and max quality settings and use GPT Image 2 token rates.

Read more →

Prompt Cache Diagnostics is now generally available in the Responses API for GPT-5.6 and later supported models. It lets developers compare cache reuse against a previous response, identify reasons for cache misses, and follow troubleshooting guidance to improve cache reuse.

Read more →

GPT Image 2.5 Sunburst and Flare are available for image generation and editing through the Image API and the Responses API image generation tool. Both support the new xhigh and max quality settings and use GPT Image 2 token rates.

Read more →

New Responses API controls for long-running GPT-6 Astra work let the model keep running while your app executes function or custom tools and return results as available, accept mid-turn steering over WebSockets while a response is in progress, and change reasoning effort mid-conversation while preserving the cached prompt prefix.

Read more →

GPT-6 Astra is released for reasoning, coding, computer use, research, and document creation, carrying complex tasks from an initial request to a finished result. It does not support the none reasoning effort level, custom temperature or top_p values, or log probabilities, and tool calling requires the Responses API.

Read more →
Latest
Oct 8, 2026
Category
Tags