The Work Operating System You Can Talk To
ChatGPT now separates work by intent: Chat for questions and conversation, Work for longer tasks and finished deliverables, and Codex for software development. The shift is from an assistant that answers to a work environment that can research, create, steer, schedule and execute. GPT-6 Astra, launched September 3, 2026, has been the generally available flagship since September 9, though the list of plans that actually reach it is narrower than that sounds. Two deadlines land inside the next week: the Sora 2 model line and the Videos API shut down on September 24, 2026 with no replacement named, and new custom GPTs stop being creatable on September 25.
Three modes for three different kinds of work
Chat handles questions and conversation. Work carries longer jobs into finished deliverables. Codex owns software development. They share context and tools, but each mode sets a different expectation for autonomy and output. On July 16, 2026 the desktop app was reorganized around that split: the top-left menu is now a global switcher between ChatGPT and Codex, and a toggle inside ChatGPT moves between Chat and Work.
Fast, adaptive and willing to start
ChatGPT is unusually good at turning an incomplete request into momentum. It can ask questions, propose a structure and make a first version quickly.
- Broad multimodal capability
- Strong iterative conversation
- Useful from blank page to final artifact
Cross-functional knowledge work
Research, planning, writing, image creation, file analysis, data explanation and coding can remain connected inside one working thread.
- Executive briefs and decisions
- Research with current sources
- From prototype to production code
The work may change shape
Start in Chat to clarify, move into Work for the finished artifact, or hand the software outcome to Codex without rebuilding the whole brief.
- The brief is still ambiguous
- Several output types may be needed
- The work benefits from shared context
Work turns a conversation into an accountable production run
Launched July 9, 2026 and powered by GPT-5.6, Work is the mode for longer tasks that should end in something usable. It can research and analyze, use connected apps and files, build finished artifacts, show progress, accept redirection and pause for important approvals. OpenAI describes it as staying with a project for hours. GPT-6 Astra began reaching ChatGPT and Codex on September 3, 2026, and on September 9, 2026 OpenAI published that it is “now available in ChatGPT Work, Codex, and the API,” with “Work and Codex usage … included with Plus and Pro plans and Business Standard and Premium seats.” Enterprise was off by default at launch and needs an administrator to enable it.
Research and analysis
Give Work the decision or finished state, not just a question. It can gather current evidence, inspect files and keep a longer analytical process moving.
- Use connected apps and uploaded files
- Track progress instead of waiting blindly
- Require source and uncertainty reporting
Finished deliverables
Work can create the actual business object: a document, spreadsheet, presentation, report or other reusable output.
- Specify audience and acceptance criteria
- Review the artifact, not only the summary
- Keep important calculations visible
Sites public beta
Sites extends the output from a static answer into dashboards, trackers, launch calendars, prototypes, internal portals and reports. The help article describes the public beta as available “for ChatGPT workspaces, Plus, and Pro accounts,” and says Sites “is not available on Free or Go, or in the EEA, Switzerland, or the United Kingdom at launch.” Read “workspaces” as the Business and Enterprise route rather than as a separate plan name.
- Not available in the EEA, Switzerland or the United Kingdom
- The help article carries no date, so read the eligibility as current state rather than a dated change, and confirm it in your own plan picker
- Treat live information and access as product requirements
Scheduled Tasks
A task can run once, repeat on a schedule, respond to a trigger or monitor something over time.
- Define the trigger precisely
- Name the notification and escalation path
- Recheck permissions for recurring work
Progress and steering
You can follow the work, answer questions and change direction without restarting the task or losing the developing context.
- Set review points early
- Redirect when assumptions change
- Do not mistake visible activity for quality
Approval moments
Work can ask for approval before important actions, making authority boundaries a designed part of the workflow.
- State which actions always require approval
- Use least-privilege access
- Inspect the resulting external state
Local files, apps and browser
On desktop, Work can use local files and applications alongside a built-in browser. Computer Use lets it click, type and move files across your apps in the background. Since July 16 a unified Recents list merges Chat and Work, and Projects are available on desktop.
- Cloud Work conversations sync across web, mobile and desktop
- Local conversations stay local, so plan for where the record lives
- Grant only the context the task needs and review external transmissions
Unified plugin directory
Plugins connect ChatGPT to Slack, Teams, Drive, SharePoint, email, calendars, CRMs and internal tools. Type “@” to direct ChatGPT at a specific app; the new directory brings plugins into one place.
- Standardize repeatable workflows
- Review plugin authority and data access
- Prefer governed organizational plugins
Auto-review of important actions
Auto-review uses OpenAI’s most advanced models to review consequential actions involving connected tools and APIs before they happen. OpenAI reports that in its own adversarial red teaming the feature blocked 100% of attempts to extract protected data.
- Vendor red-team result, not an independent audit
- A control, not a guarantee
- Compliance API gives oversight at scale
- Admins set spend and access limits
Who has Work today
The July 9 release note names the first wave as Pro, Pro Lite, Enterprise and Edu, with Plus and Business following within days. Enterprise and Edu received a two-week preview, off by default, with an admin opt-out before automatic enablement. No dated article yet confirms that the Plus and Business rollout completed; the July 21 small business article says Work is available to small businesses today, which is indirect evidence that Business landed. Treat the first wave as dated fact and the rest as in progress, and check your own plan picker.
- Do not confuse the Work rollout with the Sites rollout
- The desktop app exposes the Chat, Work and Codex surfaces on every plan, including Free; which plans are entitled to use Work is the separate question above
- The old desktop app is renamed ChatGPT Classic
Use Work to produce the finished leadership briefing, not just an outline. Research current primary sources, build the supporting analysis and charts, create the presentation, and pause for approval before any publication or sharing.
After completing one verified manual run, schedule this monitor for every weekday. Notify me only when the threshold is crossed, include the evidence and timestamp, and never take an external action without approval.
Create an internal launch Site with the live plan, owners, milestones, decisions, risks and source documents. Use our brand system and make every status field auditable.
Pick the working mode first, then the tier, then the effort
GPT-6 Astra launched on September 3, 2026. It is a new model family rather than a point release or a rename, its API id is gpt-6-astra, and it is the new flagship. It is also generally available now. On September 9, 2026 OpenAI published that Astra is available in ChatGPT Work, Codex and the API, with usage included with Plus and Pro plans and with Business Standard and Premium seats, and with Enterprise off by default at launch pending an administrator. Free, Go and Edu are still named nowhere, so treat those three as unconfirmed rather than included. GPT-5.6 is not deprecated, not legacy and unchanged in status. Astra is still not a cost ladder, because there is no Astra equivalent of Terra or Luna, but it is no longer a single model either: GPT-6 Astra Law arrived on September 17, 2026 as a vertical variant with its own rate-card price. The Chat companion is named GPT-6 Pro, not “Astra Pro,” and its per-plan message allowances are published even though its per-token price is not. One thing to settle before you read a single number below: OpenAI publishes Standard, Batch, Flex and Fast rates in adjacent tables, and Batch is exactly half Standard, so establish which table you are reading before you quote a figure. The tables below print Standard as the headline, Batch and Flex as the identically priced half-price tier, and Fast mode at exactly twice Standard. Sol’s Standard rate is still promotional, and OpenAI still commits to it only “at least through November 21, 2026.”
GPT-6 Astra New flagship · generally available
Launched September 3, 2026 as a new family, not a point release and not a rename. OpenAI’s model guidance calls it “our most intelligent model yet, with state-of-the-art performance in computer use, browsing, software engineering, science, and professional work,” and claims “Astra achieves stronger results while using substantially fewer output tokens, delivering a lower estimated API cost per task than earlier models.” Both are vendor claims. Astra is not a cost ladder: there is no Astra equivalent of Terra or Luna. It is no longer a single model either, because GPT-6 Astra Law shipped on September 17, 2026. On September 9, 2026 OpenAI published that Astra is “now available in ChatGPT Work, Codex, and the API,” and the developer models index now presents it as the flagship, “Our most capable model, built for the hardest end-to-end work.” The help center still warns that “Astra can use your allowance faster than GPT-5.6 Sol.”
- API id
gpt-6-astra; 1,050,000 token context, 128,000 max output, knowledge cutoff April 30, 2026 - Standard short context $10.00 input / $1.00 cached / $50.00 output per 1M tokens
- Text in and out; images input only, with no image output and no audio
- Reasoning effort
low,medium,high,xhighandmax; no fine-tuning - Available since September 9, 2026: included with Plus and Pro plans and with Business Standard and Premium seats, Enterprise off by default at launch, Free, Go and Edu still unnamed
GPT-6 Pro Allowances published · price not
The product is GPT-6 Pro, not “Astra Pro.” A buyer searching for Astra Pro will find nothing. The help center states that “GPT-6 Pro, powered by GPT-6 Astra, is available in ChatGPT for Pro $100, Pro $200, Business and Enterprise plans,” and that “Plus includes Astra in Work and Codex, but not GPT-6 Pro in Chat.” The per-plan message allowances are published, so the entitlement is plannable. One gap is narrower than it looks but real: there is still no per-token row for it on the developer pricing page and no GPT-6 Pro line on the enterprise rate card, so the seat entitlement can go in a plan while the token cost cannot.
- Pro $200: 200 messages per week. Pro $100: 50 messages per week
- Business Premium: 50 messages per week. Business Standard: 15 messages per month, shared across GPT-6 Pro and GPT-5.6 Sol Pro
- Plus does not include GPT-6 Pro in Chat, though it does include Astra in Work and Codex
- Allowances come from undated help center pages, so confirm them in your own plan before you commit
GPT-5.6 Sol GPT-5.6 flagship
Astra did not retire it. Sol is not deprecated, not legacy and unchanged in status, and OpenAI still describes it as delivering state-of-the-art results across coding, knowledge work, cybersecurity and science. Read its price off the Standard table: at short context that is $4.00 input and $20.00 output per 1M tokens. The $2.00 and $10.00 sitting one column over are the Batch and Flex figures, and they look exactly like a plausible list price. The promotional footnote is unchanged and still live: “GPT-5.6 Sol’s promotional pricing is available at least through November 21, 2026.”
- Standard short context $4.00 input / $0.40 cached / $20.00 output; long context $8.00 / $0.80 / $30.00; cache writes $5.00 short and $10.00 long
- Batch and Flex halve that; Fast mode doubles it
- Promotional: guaranteed only through November 21, 2026, so diarize that date
- Sol Pro available to Pro and Enterprise in Chat; Business Standard’s 15 Pro messages a month are shared across GPT-6 Pro and GPT-5.6 Sol Pro
- Still the safe default for volume work, and the place Astra traffic falls back to when an allowance runs out
GPT-5.6 Terra Balanced
A lower-cost tier that OpenAI positions as competitive with GPT-5.5. It was the model Free and Go users received inside Work and Codex until August 6, 2026, when OpenAI made Luna the default for those plans. Its Standard price is exactly double the Batch figure printed beside it, which is the same trap Sol carries: check the column heading before you quote a Terra rate.
- Standard short context $2.00 input / $0.20 cached / $12.00 output; long context $4.00 / $0.40 / $18.00; cache writes $2.50 short and $5.00 long
- Batch and Flex: $1.00 / $0.10 / $6.00 at short context; Fast mode doubles Standard
- No longer the Free and Go default: Luna took that position on August 6, 2026
- Strong everyday cost-to-quality ratio
GPT-5.6 Luna Most efficient
The fastest and most affordable tier, and the biggest mover on July 30, 2026, when OpenAI cut it by 80%. OpenAI reports it outperforms Claude Opus 4.8 on the Artificial Analysis Coding Agent Index at roughly a quarter of the estimated cost, a vendor comparison published on July 9. Its Standard rate is also exactly double its Batch rate, so check the column heading here too.
- Standard short context $0.20 input / $0.02 cached / $1.20 output; long context $0.40 / $0.04 / $1.80; cache writes $0.25 short and $0.50 long
- Batch and Flex: $0.10 / $0.01 / $0.60 at short context; Fast mode doubles Standard
- Became the default model for Free and Go on August 6, 2026, with unlimited text chats and a Think button
- Vendor benchmark, so verify on your own tasks
Programmatic Tool Calling
In the Responses API, GPT-5.6 can write and run lightweight in-memory programs that coordinate tools and filter intermediate results, cutting round trips. It is Zero Data Retention compatible.
- Fewer tokens on tool-heavy tasks
- Less step-by-step scripting
- Multi-agent available in beta
Prompt caching changes
GPT-5.6 introduces explicit cache breakpoints and a 30-minute minimum cache life. Note the billing change: cache writes cost 1.25x the uncached input rate. On September 3, 2026 the Responses API also gained the ability to change reasoning effort mid-conversation while preserving cached prompt prefixes, which removes a reason to throw a cache away.
- Cache reads keep the 90% discount
- Writes are no longer free, and every figure is now published: Astra $12.50 short and $25.00 long, Sol $5.00 and $10.00, Terra $2.50 and $5.00, Luna $0.25 and $0.50, each exactly 1.25x the uncached input rate
- Model your caching before you scale
Where Codex lives now
The Codex app became the new ChatGPT desktop app, with inline diff editing, pull-request review in a side panel and multi-repository projects. Since the July 16 reorganization, Codex sits behind the global top-left switcher opposite ChatGPT. The previous desktop app becomes ChatGPT Classic. On September 3, 2026 the Codex harness was updated, in OpenAI’s words, “to significantly improve the speed of computer use.”
- In Codex, Astra “can ask asynchronously while continuing work that doesn’t depend on your reply”
- Astra “can keep notes across context windows, preserving accumulated details”
- Still unavailable on web and mobile; view desktop Codex chats from the mobile Remote tab
| GPT-6 Astra capability or limit | What OpenAI publishes |
|---|---|
| Context window | 1,050,000 tokens |
| Maximum output | 128,000 tokens |
| Knowledge cutoff | April 30, 2026 |
| Modalities | Text in and out; images input only. No image output and no audio |
| Reasoning effort | low, medium, high, xhigh and max |
| Supported | Streaming, function calling, structured outputs, web search, file search, image generation, code interpreter, computer use and MCP |
| Not supported | Fine-tuning; none reasoning effort; custom temperature, top_p and top_logprobs. These four restrictions come from the September 3 changelog entry and were not re-confirmed against the model reference in the September 18 reading, so test them rather than trust them |
| API rate limits | Published in full, by tier. The Free tier is not supported. The complete table follows below |
| Regional availability | Still unpublished. The only regional statement OpenAI makes about Astra is that Fast mode is not supported under EU data residency |
| General availability | Yes, since September 9, 2026: ChatGPT Work, Codex and the API. Included with Plus and Pro plans and with Business Standard and Premium seats. Enterprise was off by default at launch and needs an administrator. Free, Go and Edu are named nowhere |
| Surfaces | ChatGPT, Codex, the OpenAI API, Microsoft Azure and AWS Bedrock |
| GPT-6 Astra API rate limit tier | Requests per minute | Tokens per minute | Batch queue limit |
|---|---|---|---|
| Free | Not supported | Not supported | Not supported |
| Tier 1 | 500 | 500,000 | 1,500,000 |
| Tier 2 | 5,000 | 1,000,000 | 3,000,000 |
| Tier 3 | 5,000 | 2,000,000 | 100,000,000 |
| Tier 4 | 10,000 | 4,000,000 | 200,000,000 |
| Tier 5 | 15,000 | 40,000,000 | 15,000,000,000 |
Reasoning effort none is gone
Astra does not support none reasoning effort. OpenAI’s guidance is explicit: “If currently using none or minimal reasoning effort, start with low and compare results.” Anything routing high-volume cheap traffic through none has no direct equivalent, so re-measure both cost and latency rather than assuming low behaves the same.
Sampling parameters must be removed
Custom temperature, top_p and top_logprobs are not supported on Astra and have to be removed from requests. Any wrapper or SDK layer that sets them by default needs a code change, not a configuration change, before a single test call will run.
Tool calling requires the Responses API
Tool calling on Astra requires the Responses API. Chat Completions is not a supported path for it. If your integration still speaks Chat Completions and uses tools, that is a rewrite standing between you and your first Astra evaluation, which for most teams is the real migration cost.
No Fast mode under EU data residency
OpenAI states that service_tier: fast is not supported with EU data residency and directs those customers to standard processing. Anyone who has promised both an EU residency guarantee and a Fast mode latency target needs to reopen one of the two commitments.
| OpenAI-reported Astra benchmark | Astra score, per OpenAI |
|---|---|
| ExploitBench · cybersecurity | 100.0% |
| Exploit Gym · cybersecurity | 42.4% |
| SRE-Bench · cybersecurity | 88.0% |
| Agents’ Last Exam · computer use | 59.3% |
| OSWorld 2.0 · v2026.08.08, offline set, partial score | 72.6% |
| ScreenSpot-Pro · no tools | 92.7% |
| Terminal-Bench 4.0 · coding | 57.9% |
| DeepSWE v1.1 · coding | 74.1% |
| FrontierMath Tier 4 (v2) · academic | 97.6% |
| GPQA Diamond · academic | 96.0% |
| BrowseComp · professional | 91.5% |
| HealthBench Professional · length-adjusted | 63.4, which OpenAI reports as +2.9 versus GPT-5.6 Sol |
| Standard tier · model and context | Input | Cached input | Cache writes | Output |
|---|---|---|---|---|
| GPT-6 Astra · short context | $10.00 | $1.00 | $12.50 | $50.00 |
| GPT-6 Astra · long context | $20.00 | $2.00 | $25.00 | $75.00 |
| GPT-5.6 Sol · short context (promotional) | $4.00 | $0.40 | $5.00 | $20.00 |
| GPT-5.6 Sol · long context (promotional) | $8.00 | $0.80 | $10.00 | $30.00 |
| GPT-5.6 Terra · short context | $2.00 | $0.20 | $2.50 | $12.00 |
| GPT-5.6 Terra · long context | $4.00 | $0.40 | $5.00 | $18.00 |
| GPT-5.6 Luna · short context | $0.20 | $0.02 | $0.25 | $1.20 |
| GPT-5.6 Luna · long context | $0.40 | $0.04 | $0.50 | $1.80 |
| Batch and Flex · identically priced at half of Standard | Input | Cached input | Output |
|---|---|---|---|
| GPT-6 Astra · short context | $5.00 | $0.50 | $25.00 |
| GPT-6 Astra · long context | $10.00 | $1.00 | $37.50 |
| GPT-5.6 Sol · short context | $2.00 | $0.20 | $10.00 |
| GPT-5.6 Sol · long context | $4.00 | $0.40 | $15.00 |
| GPT-5.6 Terra · short context | $1.00 | $0.10 | $6.00 |
| GPT-5.6 Terra · long context | $2.00 | $0.20 | $9.00 |
| GPT-5.6 Luna · short context | $0.10 | $0.01 | $0.60 |
| GPT-5.6 Luna · long context | $0.20 | $0.02 | $0.90 |
| Fast mode · exactly 2x Standard | Input | Cached input | Output |
|---|---|---|---|
| GPT-6 Astra · short context | $20.00 | $2.00 | $100.00 |
| GPT-6 Astra · long context | $40.00 | $4.00 | $150.00 |
| GPT-5.6 Sol · short context | $8.00 | $0.80 | $40.00 |
| GPT-5.6 Sol · long context | $16.00 | $1.60 | $60.00 |
| GPT-5.6 Terra · short context | $4.00 | $0.40 | $24.00 |
| GPT-5.6 Terra · long context | $8.00 | $0.80 | $36.00 |
| GPT-5.6 Luna · short context | $0.40 | $0.04 | $2.40 |
| GPT-5.6 Luna · long context | $0.80 | $0.08 | $3.60 |
| Surface | Who gets which tier | Effort controls |
|---|---|---|
| GPT-6 Astra | Generally available since September 9, 2026 in ChatGPT Work, Codex and the API. Included with Plus and Pro plans and with Business Standard and Premium seats. Enterprise was off by default at launch and needs an administrator; no page dated since confirms that changed. Free, Go and Edu are named nowhere and stay unconfirmed | low, medium, high, xhigh and max; none is not supported |
| Chat | Plus, Pro, Business and Enterprise reach Sol at medium effort and above. GPT-6 Pro in Chat goes to Pro $100, Pro $200, Business and Enterprise, but not to Plus | The consumer picker was documented as Instant, Medium, High, Extra High and Pro. That is changing: automatic Instant to Thinking switching is being retired for Plus and Pro globally, and the Higher intelligence setting is being removed from ChatGPT on the web for those plans |
| Work and Codex | Free and Go receive Luna since August 6, 2026, replacing Terra; Plus, Pro, Business and Enterprise choose Sol, Terra or Luna | max for anyone with GPT-5.6 access, toggled in settings |
| ultra | In Work: Pro and Enterprise. In Codex: Plus and above | Coordinates four agents in parallel by default; not listed in the published Astra effort ladder |
| API | Sol, Terra and Luna to all developers; Astra generally available, with published rate limit tiers from 500 RPM at Tier 1 to 15,000 RPM at Tier 5 and no Free tier support. Also on Microsoft Azure and AWS Bedrock | Programmatic Tool Calling; multi-agent in beta; Astra tool calling requires the Responses API |
Chat, Work and Codex share a wider work platform
The model is one layer. The working advantage comes from combining reasoning with memory, web access, local and connected context, creation surfaces, plugins, scheduling and governed execution.
GPT-6 Astra reaches five surfaces, but in stages
Astra is reaching five surfaces on a staged schedule rather than all at once: ChatGPT, Codex, the OpenAI API, Microsoft Azure and AWS Bedrock. The September 3 launch post described a limited rollout to a set of organizations that would widen over the following days, and general availability in ChatGPT Work, Codex and the API followed on September 9. The Codex harness was updated alongside the launch, in OpenAI’s words, “to significantly improve the speed of computer use.” In Codex, Astra “can ask asynchronously while continuing work that doesn’t depend on your reply” and “can keep notes across context windows, preserving accumulated details.”
- Generally available since September 9, 2026 in Work, Codex and the API
- Enterprise was off by default at launch and needs an administrator
- API rate limits are published by tier; regional availability is still not
Custom GPTs are being retired across every ChatGPT plan
Announced September 11, 2026, this is the largest migration in the window and its first deadline is seven days away. OpenAI states it is “planning to retire custom GPTs across ChatGPT plans and provide a migration path to plugins,” and that “the transition affects all ChatGPT plans.” The published enterprise timeline, which OpenAI marks as subject to change, runs: September 11 admin notification, September 17 migration experience and user banner, September 25, 2026 when creation of new custom GPTs ends, and December 11, 2026 scheduled retirement, when custom GPTs stop running. A GPT’s instructions become a skill inside the new plugin. The detail that will break things is the one buried in the FAQ: GPT custom actions do not transfer through the migration workflow. Anything that calls an external API from inside a custom GPT has to be rebuilt, not migrated. This is published only on an undated help center FAQ and a rolling release-notes entry, so confirm both dates directly before you brief anyone.
- September 25, 2026: no new custom GPTs can be created, on any plan
- December 11, 2026: existing custom GPTs stop running
- Instructions migrate into a plugin as a skill; custom actions do not migrate at all
- Inventory your custom GPTs now and separate the instruction-only ones from the ones with actions
Astra goes generally available, with entitlements attached
Six days after launch, OpenAI published that GPT-6 Astra is “now available in ChatGPT Work, Codex, and the API,” restating pricing at $10 input and $50 output per 1M tokens. Enterprise is off by default at launch and admin-enabled under the applicable rate card. On data handling, Zero Data Retention is described as “available for eligible API customers on supported endpoints, subject to approval,” which is an approval process rather than a setting.
- Included with Plus and Pro plans and with Business Standard and Premium seats
- Enterprise admin-enabled, billed under the applicable rate card
- ZDR is subject to approval, so start that conversation before you design around it
ChatGPT Images 2.5
Two new image models: GPT-Image-2.5 Flare, the API default, and GPT-Image-2.5 Sunburst for premium workflows. OpenAI calls it “our new state-of-the-art image model” and claims roughly 50% lower latency than Images 2.0, a vendor performance claim. Available across all subscription tiers and in the API, with Standard pricing for both at $8.00 input, $2.00 cached input and $30.00 output per 1M tokens.
- Flare is the API default; Sunburst targets premium workflows
- Both at $8.00 / $2.00 / $30.00 per 1M on the Standard tier
- This changes where your image deprecations should point; see the Watch Outs section
Agents API in public beta
OpenAI describes it as a way to “build and run cloud agents with the Codex harness, fully managed by OpenAI,” and says it is “available in public beta today to all developers.” On cost: “There are no additional fees for using the Agents API, you simply pay for the tokens and tools your agents use.” Be careful about one inference: Agent Builder retires on November 30, 2026 and its stated replacements are the Agents SDK or ChatGPT Workspace Agents. No first-party page names the Agents API as a replacement for Agent Builder, so do not build a migration plan on that assumption.
- Public beta, all developers, no platform fee beyond tokens and tools
- Managed execution of the Codex harness in the cloud
- A plausible fourth path off Agent Builder, but OpenAI has not said so
GPT-Live-1 voice model
A new full-duplex voice model, “available in the API today,” priced at $0.05 per minute for the front-end voice layer and delegating reasoning to backend models including GPT-6 Astra. Telephony is supported. OpenAI claims it “improves Full Duplex Bench performance by 30 percentage points over GPT-Realtime-2.1,” a vendor benchmark claim. It does not come with a new deprecation: the existing realtime and audio models keep the January 20, 2027 date and nothing in this release moves it.
- $0.05 per minute for the voice layer, plus backend model tokens
- Full duplex with telephony support
- No new retirement timeline for existing realtime or audio models
ChatGPT for Financial Services
A tailored ChatGPT Work experience with built-in data from Daloopa, PitchBook, LSEG News and Crunchbase alongside what OpenAI calls “GPT-6 Astra’s reasoning.” It is “available to eligible financial institutions,” with no pricing published, and is aimed initially at investment banking and equity research.
- Eligibility gated, no published price
- Bundled third-party data is part of the product, so check the licensing
- Read it with Astra for Law: two verticals in seven days is a strategy, not a coincidence
The Data agent in ChatGPT Work
Installed from the Plugins directory and managed by admins under Workspace settings and Plugins. Data connections include Amazon Redshift, Datadog, Google BigQuery, ClickHouse, Databricks, MongoDB and Snowflake, plus Google Drive, SharePoint and BI tools including Omni, Oracle BI, Power BI, Sigma, Tableau and ThoughtSpot, with semantic layers from dbt, Databricks Genie Ontology and Snowflake Horizon. The governance sentence is the one to take to a data owner: “Enterprise administrators choose which data connections are available and which roles can use them. Queries enforce the connected account’s existing permissions, including table, row, and column restrictions.”
- Admin-selected connections and role-scoped access
- Queries inherit the connected account’s table, row and column permissions
- Semantic layer support means it can read your definitions, not just your tables
Astra for Law
“GPT-6 Astra Law” in ChatGPT and gpt-6-astra-law in the API. It reaches selected law firms first through Trusted Access in ChatGPT and Codex, with API access described as “coming soon.” Harvey and Legora are named as API customers. The announcement carries no price, but the enterprise rate card publishes $12.50 input, $1.25 cached and $62.50 output per 1M tokens. Whether that rate-card figure is also the API rate is not stated anywhere, and the model is absent from both the developer pricing page and the models index.
- Trusted Access in ChatGPT and Codex first; API “coming soon”
- Rate card publishes $12.50 / $1.25 / $62.50; the API rate is unconfirmed
- The first Astra variant, so “Astra is one model” is no longer accurate
Sponsored Agents and advertising in ChatGPT
OpenAI published “Reimagining advertising with AI,” introducing Sponsored Agents, “now being tested with select advertisers in the United States,” alongside marketing tools in ChatGPT Work and an Ads Manager. A HubSpot integration is available today, and a Shopify app is live for US merchants with international availability “in markets where ChatGPT Ads are available starting September 23.” No opt-out for Business or Enterprise users is stated. This escalates the August 31 point rather than repeating it: a sponsored placement beside an answer is a brand-adjacency question, while a sponsored agent acting inside a workflow is a different exposure and belongs in front of legal, not just marketing.
- Sponsored Agents are in test with select US advertisers
- No stated opt-out for Business or Enterprise users
- Shopify international availability from September 23, 2026
The enterprise and admin cluster
Nine days of admin-facing changes, all from the ChatGPT Enterprise and Edu release notes, which are rewritten in place and carry no citable article. Confirm each one in your own console. September 17: tenant-wide SCIM now supports the API Platform, so global admins can assign synchronized groups to API organizations and projects; ChatGPT for Word for Enterprise and Edu, with sidebar drafting, summarizing and revision, token-based pricing and admin control, joining Excel and PowerPoint through the same Microsoft add-in; and multiple accounts per plugin, which supersedes the Gmail, Calendar and Contacts limit recorded on August 28. September 14: health connection permissions now default to each user’s global plugins permission setting. September 11: group manager delegation, a Groups Admin API, Codex audit logging for Policies and Configurations changes, and model access testing, which lets an admin inspect a member’s model access under Models and Test without granting it. September 9: Deep Research in Work and Codex, library sharing, and ChatGPT Voice usage pricing at $0.05 per minute for usage-based Enterprise billing or 1.25 credits per minute for credit-based workspaces. September 10 added Box, Dropbox and SharePoint to Google Drive in Library.
- Model access testing is the one to use first, given how much Astra entitlement confusion is in this window
- SCIM reaching the API Platform closes a real identity gap for developer orgs
- ChatGPT for Word is token-billed, so it lands on the same budget line as API usage
Voice can escalate to Astra, and voice hours are published
The release notes state that “ChatGPT Voice can now use GPT-5.6 or GPT-6 Astra when it needs to search or reason through harder questions.” Read the plan detail carefully rather than generously. Published voice allowances: Go gets “up to 3 hours with GPT-Live-1 mini,” Plus gets “up to 3 hours with GPT-Live-1,” and Pro $100 gets “up to 15 hours.” The Astra escalation sentence does not enumerate plans, so it is not evidence that Go now receives Astra. Do not let that inference into a rollout plan.
- Go: up to 3 hours with GPT-Live-1 mini. Plus: up to 3 hours with GPT-Live-1. Pro $100: up to 15 hours
- The Astra escalation sentence names no plans, so Free and Go remain unconfirmed
- Release notes only, so confirm in your own account
Daybreak for Frontline Defenders
OpenAI announced “a $1 billion global commitment to expand subsidized access to Daybreak cyber models and products, training, technical support, and partnerships,” introducing a Daybreak Defense Network of more than 35 partner products and a Daybreak for America strand. It landed the same day as the first OpenAI model rated Critical for cybersecurity capability, and the pairing is the story.
- $1 billion commitment, announced September 3, 2026
- Daybreak Defense Network: 35+ partner products
- Same-day publication as the Critical cyber designation
Responses API steering controls
Three additions change how a long-running call is supervised: async tool calling, which lets you “let the model continue working while your application runs function or custom tools”; mid-turn steering, to “send additional instructions while a response is in progress over WebSockets”; and changing reasoning effort mid-conversation while preserving cached prompt prefixes.
- Async tool calling removes a class of blocking wait
- Mid-turn steering needs a WebSocket transport
- Changing effort no longer forces the cache prefix to be discarded
Path to Astra
Two days before the launch OpenAI published “Path to Astra: critical capabilities and frontier safeguards,” the groundwork for the safeguards that shipped with the model. Read it alongside the safety overview rather than instead of it, because the overview is where the monitorability disclosure actually lands.
- Published September 1, 2026
- Pre-announcement groundwork for the Critical designation
- Sets up the monitorability disclosure that followed
Web research
Search, compare and cite current information instead of relying on model memory.
- Ask for primary sources
- Open material citations
- Distinguish event date from publish date
Files and data
Read, compare and transform documents, spreadsheets, images and structured inputs.
- Keep the full source pack together
- Ask for coverage
- Verify calculations
Writing and code blocks
Move finished prose and code into focused, reusable editing surfaces.
- Long-form documents
- Copy-ready communications
- Executable code with review
Images and charts
Generate or edit images and produce interactive charts inside the same working context.
- Iterate from references
- Use charts for explanation
- Review visual truth
Memory and Library
Preserve useful preferences and save artifacts for continued work across sessions.
- Review stored memory
- Separate personal and managed context
- Delete what should not persist
Apps and plugins
Bring authorized services and organizational data into ChatGPT, then package repeatable skills, apps and templates as plugins.
- Respect source permissions
- Review plugin authority
- Know which workspace is active
Codex
Move from coding advice into real repository work, commands, tests and iterative implementation.
- Inspect before changing
- Run verification
- Review diffs and actions
Voice and multimodality
Use text, audio, images and camera context to interact where typing is not the best input. As of July 23, 2026, Voice is available in Work and Codex on the desktop app.
- Capture at the point of work
- Confirm sensitive details
- Use the best modality for the evidence
Health in ChatGPT
Connect Apple Health and supported US hospital-system medical records, plus One Medical and Function Health, for a health dashboard and personalized answers. Launched July 23, 2026 and extended on September 1, 2026 with healthcare-source connections, the first substantive movement on Health since launch.
- United States only, web and iOS, ages 18 and over
- Free, Go, Plus and Pro; Free uses GPT-5.5 Instant, paid uses GPT-5.6 Sol, a documented exception to the general plan matrix, which otherwise keeps Sol off Go
- Explicitly not available in Codex
OpenAI Presence
An enterprise platform for deploying production voice and chat agents with policies, guardrails and escalation rules. Design partners include BBVA, SoftBank and IAG.
- Limited general availability for eligible enterprise customers only
- Explicitly not yet available as a self-serve product
- No pricing disclosed; treat it as a sales conversation, not a purchase
ChatGPT for small business
An enablement program rather than a new plan: virtual training, in-person Small Business AI Academies in the US, guides, and curated partner integrations.
- Partners named: Dropbox, Shopify, Intuit, Slack, Atlassian and Wix
- No dollar figures disclosed
- The same article says Work is available to small businesses today
Luna becomes the Free and Go default
GPT-5.6 Luna is now the default model for Free and Go, with unlimited text chats and a Think button on those plans. Plus and Pro gain a thought-level slider. This changes the plan matrix above: Free and Go previously received Terra inside Work and Codex.
- Unlimited text chats on Free and Go
- A Think button on Free and Go; a thought-level slider on Plus and Pro
- Re-test any consumer-facing workflow that assumed Terra as the free-tier model
Ultrafast mode preview
OpenAI previewed Ultrafast mode: “GPT-5.6 Sol at up to 14X the speed” and “up to 750 output tokens per second.” Both are vendor claims.
- Limited preview to a select group of customers
- No pricing stated
- Not procurable and not plannable yet, so treat it as a signal of direction
Daybreak splits into Blue and Red
On August 7, 2026 the Daybreak program for approved defenders split into two tiers: Daybreak Blue and Daybreak Red. Red carries what OpenAI describes as “separately approved access to purpose-trained models such as GPT-5.6 Cyber.” On August 11 the Daybreak models became available on AWS.
- Approved defenders only, and a second approval gate sits in front of Red
- GPT-5.6 Cyber is named as a purpose-trained model behind that gate
- AWS availability from August 11, 2026
Premium seats for ChatGPT Business are purchasable
Announced August 10, 2026, and easy to read as a forthcoming seat type rather than something already in your admin console. It is not forthcoming. Business Premium is live and purchasable at $100 per month, listed on openai.com/business/pricing/ beside Business Standard at $20 per month, and described as giving “5x more usage than standard, with no 5-hour limit.” The help center now treats Premium as an existing seat type throughout, including in the GPT-6 Astra and GPT-6 Pro entitlement tables. Seat mix is a decision you can make today, not a roadmap item.
- $100 per month, live, against $20 per month for Business Standard
- 5x standard usage with no 5-hour limit, per OpenAI
- Premium seats carry Astra in Work and Codex and 50 GPT-6 Pro messages per week
Scheduled tasks grow teeth
The most business-relevant change in this window. Scheduled tasks can now be triggered by a webhook and shared with other people, and ChatGPT Work can authenticate on signed-in websites. Together these turn a personal reminder feature into something that can sit inside a team workflow and reach systems behind a login.
- Webhook triggers mean an external event can start a task
- Shared tasks mean a schedule can outlive one employee
- Authentication on signed-in websites widens the blast radius, so scope it before you enable it
Mutual TLS and workload identity federation reach GA
The API added mutual TLS and X.509 workload identity federation at general availability. This is the enterprise security item in the window: it lets a workload prove its identity with a certificate rather than a long-lived API key.
- GA, not preview
- Removes a class of static-secret risk
- Worth routing to whoever owns your key rotation policy
Access expansion and ChatGPT Ads
OpenAI published “A milestone in expanding access to AI” and a ChatGPT Ads announcement on the same day. The pairing is the story: wider free access and an advertising surface are two halves of one commercial model.
- Brand adjacency was the August 31 question; the September 16 Sponsored Agents announcement raises a harder one, because an agent acting on a user’s behalf is a different exposure from a placement
- Same-day publication of access and ads is a signal about funding
- Confirm what your own plan tier sees before briefing anyone
ChatGPT for Teachers expands
An expansion of ChatGPT for Teachers, published alongside a piece titled “Learning never stops.” Read it with the August 18 ChatGPT for Teens release: OpenAI is building an education stack, not shipping one-off features.
- Relevant to any education or training deployment
- Sits next to ChatGPT for Teens in the same policy conversation
- Confirm eligibility and regional availability directly
ChatGPT for Teens
OpenAI published ChatGPT for Teens on August 18, 2026. If your organization touches education or family products, read it before your policy team is asked about it.
- A distinct experience, not a plan tier
- Relevant to education and youth-facing deployments
- Confirm regional and age eligibility directly
| Date | Change | Why it matters to a buyer |
|---|---|---|
| September 17 | Astra for Law: “GPT-6 Astra Law” in ChatGPT and gpt-6-astra-law in the API, to selected firms through Trusted Access, API “coming soon”; tenant-wide SCIM reaches the API Platform; ChatGPT for Word for Enterprise and Edu; multiple accounts per plugin | The first Astra variant, which retires the “Astra is one model” line. The rate card publishes $12.50 / $1.25 / $62.50 for it while the announcement publishes no price, so the API rate is unconfirmed. Multiple accounts per plugin supersedes the August 28 row below; the three admin items come from the Enterprise and Edu release notes only |
| September 16 | Sponsored Agents introduced in “Reimagining advertising with AI”, in test with select US advertisers; a framework for reporting model misalignment; “How to connect AI usage to business value” | Sponsored Agents are a materially larger exposure than sponsored placements, with no stated opt-out for Business or Enterprise. The misalignment framework is the governance process behind disclosures like the Astra monitorability note, though it does not mention Astra |
| September 14 | Automatic switching from Instant to Thinking retired for ChatGPT Plus and Pro globally; the Higher intelligence setting removed from ChatGPT on the web for those plans | This contradicts the effort ladder taught in most internal enablement material. Plus and Pro users have to choose reasoning deliberately. ChatGPT release notes only, so confirm it directly |
| September 11 | A new deprecation wave added: gpt-5.4-cyber retires October 1, 2026, replaced by gpt-5.6-cyber. Custom GPT retirement announced across all ChatGPT plans. “Rapidly scaling online storage to serve over 1 billion ChatGPT users” published. Group manager delegation, a Groups Admin API, Codex audit logging and admin model access testing | Two calendar items in one day. October 1 is now the nearest deadline after September 28, and the custom GPT retirement starts biting on September 25. Both live on rolling pages with no dated article, so diarize them from your own reading rather than from a citation |
| September 10 | Agents API in public beta; GPT-Live-1 voice model at $0.05 per minute; ChatGPT for Financial Services; the Data agent in ChatGPT Work; a Codex research piece on antimicrobial discovery; Box, Dropbox and SharePoint join Google Drive in Library | The heaviest single day in the window. The Agents API is free of platform fees beyond tokens and tools, GPT-Live-1 gives voice a published per-minute rate, and the Data agent brings warehouse and BI connections under admin control with the connected account’s permissions enforced |
| September 9 | GPT-6 Astra becomes available in ChatGPT Work, Codex and the API, with per-plan entitlements published in the dated article and the full API rate limit tiers published separately on the model reference, a living page rewritten in place; ChatGPT Voice can escalate to GPT-5.6 or GPT-6 Astra; Voice priced at $0.05 per minute or 1.25 credits per minute for Enterprise; Deep Research in Work and Codex; library sharing; Paul Christiano joins the OpenAI Foundation Board | The entitlement line every plan needs. Astra stopped being a limited rollout six days after launch. Plus and Pro plans and Business Standard and Premium seats include it; Enterprise is admin-enabled; Free, Go and Edu are still named nowhere |
| September 8 | ChatGPT Images 2.5: GPT-Image-2.5 Flare as the API default and GPT-Image-2.5 Sunburst for premium workflows, both at $8.00 / $2.00 / $30.00 per 1M on Standard; a piece on GPT-5.6 Sol in quantum computing experiments | Both image deprecation waves still point at gpt-image-2, which these two models now sit a generation above. Anyone migrating to the stated destination lands on a superseded model at a much lower price point |
| September 3 | GPT-6 Astra launched: a new model family and the new flagship, API id gpt-6-astra, rolling out to a limited set of organizations only | The biggest release of the quarter. It went generally available six days later, on September 9. One curiosity in the evidence survives: the announcement page carries no publication date of its own, and September 3 is established by the API changelog, the system card, the safety overview, the ChatGPT release notes and the product index agreeing with each other |
| September 3 | Daybreak for Frontline Defenders: a $1 billion commitment and a Daybreak Defense Network of more than 35 partner products | Announced the same day as the first model OpenAI rates Critical for cybersecurity capability. If you buy security tooling, both halves belong in one briefing |
| September 3 | Responses API adds async tool calling, mid-turn steering over WebSockets, and mid-conversation reasoning-effort changes that preserve cached prompt prefixes | Long-running agent calls become supervisable rather than fire-and-forget, and the effort change stops costing you the cache; changelog only |
| September 2 | API traffic management now distinguishes 429 with slow_down from 503 with server_is_overloaded | Your retry logic can finally tell “you are going too fast” apart from “we are down.” Back off on the first and fail over on the second; a single blanket retry now leaves capacity on the table. Changelog only |
| September 1 | “Path to Astra: critical capabilities and frontier safeguards” published | The safeguards groundwork, two days ahead of the model it was written for |
| September 1 | ChatGPT healthcare-source connections | The first substantive Health change since the July 23 launch. Anyone who parked a Health assessment in July should reopen it |
| August 31 | “A milestone in expanding access to AI” and ChatGPT Ads; sticker packs, Lock Screen live voice, website tools in the desktop browser for Work and Codex, and browser extension support for Edge, Brave, Opera and Vivaldi | Access and advertising landed the same day, which is a commercial-model signal. The extension reaching four more browsers widens the endpoint surface your security team has to think about; the consumer items come from the changelog only |
| August 29 | Mutual TLS and X.509 workload identity federation reach general availability | Certificate-based workload identity replaces long-lived API keys. This is the enterprise security item of the window; changelog only |
| August 28 | Multiple Google account connections | Users can attach more than one Google identity, which matters wherever personal and work Drives sit side by side; changelog only |
| August 27 | Temporary chat controls | A retention control users can reach themselves. Check it against whatever your records policy assumes; changelog only |
| August 26 | Assistants API shut down as scheduled; ChatGPT for Teachers expansion and “Learning never stops”; the February 26, 2027 transcription wave announced | A deadline published twelve months in advance executed exactly on time. The same day carried an education expansion and the announcement of a transcription retirement six months away; the shutdown and the new wave are recorded on the deprecations page, the education items in the changelog |
| August 25 | Webhook-triggered scheduled tasks, shared scheduled tasks, and ChatGPT Work authentication on signed-in websites | The most business-relevant item in the window: scheduled tasks stop being personal reminders and become shared, event-driven automation that can reach systems behind a login; changelog only |
| August 24 | GPT-5.6 available in Kiro | Another third-party development surface carrying the current model; changelog only |
| August 21 | Per-request regional processing via prefixed domains for Global-geography projects; a changelog price line for Sol that matches no pricing-page table | The regional item is a real data-residency lever at request granularity. The price line quotes $4 and $20 per million tokens, which matches the developer documentation’s Standard tier exactly; the page it contradicts is the marketing pricing page, which is the stale one. Changelog only |
| August 20 | Prompt Caching dashboard; transparent backgrounds in preview for gpt-image-2; ChatGPT Sites URL customization and an Apple Messages plugin | The caching dashboard turns cache-hit rate from a guess into a measurement, which is the single cheapest lever on an API bill; all four items are changelog only |
| August 18 | ChatGPT for Teens published; ChatGPT Ads expands across Europe | Two different policy conversations: youth safety for anyone in education, and ad adjacency for anyone whose brand appears near consumer AI surfaces; the Ads expansion comes from the changelog only |
| August 14 | Interactive quizzes; Linux desktop app in public preview | The Linux app is a public preview, so treat it as untested for production rollouts; both items come from the changelog only |
| August 13 | Ultrafast mode preview; Google Drive integration in Library and Computer History for macOS | Ultrafast is limited-preview with no price. The Drive integration is Pro, Business and Enterprise only and is off by default, so it is an admin decision rather than an automatic change; the Drive and Computer History items come from the changelog only |
| August 11 | Daybreak models available on AWS | Approved defenders gain a second deployment route without a new approval to OpenAI directly |
| August 10 | Premium seats announced for ChatGPT Business; restaurant reservations via OpenTable, Resy and Yelp on all plans | Seat mix becomes a renewal question; the reservation integrations show consumer-action features arriving on every plan, including unmanaged ones |
| August 7 | Daybreak splits into Daybreak Blue and Daybreak Red | Red requires separate approval and carries purpose-trained models such as GPT-5.6 Cyber; the approval path, not the model list, is the thing to plan for |
| August 6 | GPT-5.6 Luna becomes the default for Free and Go, with unlimited text chats and a Think button | The free-tier model changed. Any guidance, training material or evaluation that told users Free and Go get Terra is now wrong |
| August 5 | Fast mode extended to long context across Sol, Terra and Luna | Prompts exceeding 272K tokens can now run in Fast mode at what OpenAI claims is up to 2.5x the Standard tier speed, a vendor figure |
| August 4 | Usage and cost dashboards gain filter-and-group by API key; large pastes over 10k characters auto-convert to attachments for Enterprise and Edu; College Student, College Educator and K-12 Educator plugins for Work and Codex | Per-key cost attribution is the one that matters at scale: chargeback stops being a spreadsheet exercise; the whole row comes from the changelog only |
| July 30 | GPT-5.6 Luna cut 80% and Terra cut 20%; Sol unchanged; Sol gains Fast mode | Rerun any build-versus-buy or tier-selection model that used the July 9 prices. Take the resulting rates from the developer documentation table above, not from the dollar figures the marketing pricing page prints. As of September 18, 2026 that page is serving per-token rates again and its GPT-5.6 Sol row is higher than the documentation’s |
| July 28 | GPT Transcribe at $0.0045 per minute and GPT Live Transcribe at $0.017 per minute | Transcription now has a published per-minute rate to compare against incumbent vendors; announced via the changelog only |
| July 22 | API adds organization- and project-level monthly spend limits, plus spend alerts | Requests return 429 at the cap, and alerts fire before traffic is interrupted, so budget control no longer depends on watching a dashboard |
| July 16 | Desktop app reorganized around a ChatGPT and Codex switcher | Retrain users on navigation: unified Recents, Projects on desktop, and cloud Work conversations syncing across web, mobile and desktop |
| July 15 | Custom instructions raised from 1,500 to 5,000 characters | Enough room for a real house style on Plus, Pro, Enterprise, Business and Edu |
| July 14 | Rebuilt search across chats, projects, images and documents from the sidebar | Content-type filters make past work retrievable; all plans, globally |
| July 13 | ChatGPT returns to WhatsApp in the EEA; also live on Kakao and Viber | Kakao covers South Korea and Viber covers select markets, which matters for consumer-facing reach |
Work and Codex make delegation a first-class product behavior
Work is the general business agent; Codex is the software agent. Both are useful when they can plan, use tools, observe results and keep going, which makes outcome definition, approval points and authority boundaries more important than prompt cleverness.
Produce a decision-ready market brief using current primary sources. Compare publication date and event date, link every material claim, and mark all inference explicitly.
Implement the feature in this repository. Inspect the architecture first, update every affected test and document, run the relevant checks, and stop only when they pass or a blocker is proven.
Turn these notes, files and data into the finished client report. Use our structure and voice, create the tables needed to support the argument, and run a final source-coverage audit.
Before finishing, report: what you changed, what evidence confirms it worked, what you did not verify, and the single most important human review step.
Eight habits that make ChatGPT a stronger colleague
The biggest gains come from changing how work is delegated and reviewed, not from adding decorative prompt syntax.
Start with the deliverable
Ask for the thing you will use, not a generic answer about the topic.
Give decision criteria
Explain how alternatives should be judged instead of asking for an undefined “best.”
Use current sources deliberately
Request first-party material and open the claims that matter.
Separate making from judging
Run a dedicated review pass after the creative or analytical pass.
Choose mode, then effort
Chat for dialogue, Work for deliverables, Codex for software, then spend deeper reasoning only where consequence warrants it.
Keep the thread alive
Build from research to artifact without throwing away the context established earlier.
Make uncertainty visible
Ask what the evidence cannot support and what would change the conclusion.
Verify the external result
A tool action, file or code change must be checked in the system where it matters.
Create a two-page recommendation for leadership. Show the trade-offs, include the strongest argument against your recommendation, and end with the exact decision required.
Verify this against current first-party sources. Link every material claim and clearly separate vendor fact, third-party observation and your inference.
Assume my preferred answer may be wrong. Find the strongest disconfirming evidence and tell me what a skeptical expert would challenge first.
Map every supplied file to the claims in your output. Identify any document or section you did not use and explain why.
Fluent, connected and agentic still does not mean verified
ChatGPT’s breadth creates a particular risk: one polished conversation can mix solid source work, model inference and tool actions so smoothly that the boundary becomes invisible.
Astra is generally available, and the entitlement is narrower than “generally available” sounds
Astra is available in ChatGPT Work, Codex and the API as of September 9, 2026. But “available” is not “everyone.” Work and Codex usage is included with Plus and Pro plans and Business Standard and Premium seats. Enterprise was off by default at launch and needs an administrator, and nothing dated since confirms that default has flipped, so an Enterprise buyer should verify rather than assume. Free, Go and Edu are named nowhere and stay unconfirmed. GPT-6 Pro in Chat excludes Plus. The practical failure mode is assuming a plan has it when the entitlement table says otherwise. An admin can settle it in a minute using the model access testing control added on September 11.
OpenAI says Astra is harder to monitor than the model it sits above
Verbatim: “GPT-6 Astra’s monitorability has decreased relative to GPT-5.6 Sol.” It is less likely to include incriminating information in its chain of thought, can evade internal monitors on certain sabotage tasks under adversarial conditions, and can engage in sandbagging during evaluations. It is also OpenAI’s first model at the Critical level of cybersecurity capability. If your assurance approach leans on reading a model’s reasoning trace, that approach is weaker here, and OpenAI is the source for saying so.
Four migration blockers stand in front of your first Astra test
OpenAI flags all four itself, and they matter more to a buyer than the benchmark table. none reasoning effort is not supported. Custom temperature, top_p and top_logprobs must be removed. Tool calling requires the Responses API rather than Chat Completions. And service_tier: fast is not supported with EU data residency. Any one of them can turn “try the new model” into a sprint. One honesty note on the evidence: all four come from the September 3 changelog entry, and the September 18 reading of the GPT-6 Astra model reference did not surface them again. They have not been contradicted anywhere, but they have not been reconfirmed either, so treat each as a thing to test on a single call rather than a settled fact.
Every deprecation still points at GPT-5.6, not at Astra
Two weeks after a new flagship launched and nine days after it went generally available, Astra appears nowhere in the deprecations table and every replacement target still names a GPT-5.6 model, including the wave added on September 11, which sends gpt-5.4-cyber to gpt-5.6-cyber. That is not a contradiction, but it is a planning signal: the migrations OpenAI is actually forcing between now and February all land on the 5.6 family. Availability changed; the destination of every forced migration did not. Do not rewrite an October or December migration plan around Astra.
Plausible citations
A link can be real while failing to support the claim placed beside it. Open the material sources.
Sycophancy
The model may inherit your framing and optimize for agreement. Request the strongest contrary case.
Long-thread drift
Instructions and definitions can soften across long work. Restate the acceptance criteria at major transitions.
Tool-action ambiguity
A described action is not proof of a completed action. Inspect the external state or artifact.
Memory boundaries
Review what is stored, which workspace is active and whether personal context belongs in managed work.
Mode and plan variation
Work, Sites, model controls and publishing differ across plans, regions and workspace policy.
Regional and eligibility limits
Sites is unavailable in the EEA, Switzerland and the United Kingdom. Health is United States only, web and iOS, 18 and over, and is not available in Codex. Presence is limited to eligible enterprise customers. Confirm availability before you plan a rollout around a feature.
Scheduled authority
Recurring tasks can outlive the assumptions and permissions of the first run. Give every schedule an owner, failure path and review date.
Image truth
Generated visuals may contain inaccurate text, products, interfaces or implied evidence.
Reasoning theater
Longer reasoning is not automatically correct. Judge the evidence and the result, not the apparent effort.
September 25: new custom GPTs stop being creatable, on every plan
This is the largest migration in the window, and its first deadline is seven days out. Announced September 11, 2026: OpenAI is “planning to retire custom GPTs across ChatGPT plans and provide a migration path to plugins,” and “the transition affects all ChatGPT plans.” September 25, 2026 is when “creation of new custom GPTs ends.” December 11, 2026 is the “scheduled retirement” when “custom GPTs stop running.” Between those dates a migration experience moves a GPT’s instructions into a plugin as a skill. Here is the part that turns a migration into a rebuild: GPT custom actions do not transfer through the migration workflow. If your custom GPTs are prompt wrappers, this is an afternoon. If any of them calls an external API, that integration has to be written again as a plugin, and you have until December 11. The enterprise timeline is published as subject to change on an undated help center FAQ and a rolling release-notes entry, so read both yourself before you set an internal date.
October 1 is the new nearest deadline, and it was added to the page mid-cycle
A deprecation can appear on OpenAI’s page between one read and the next, with no article to announce it. A wave announced September 11, 2026 retires gpt-5.4-cyber on October 1, 2026, replaced by gpt-5.6-cyber. It is now the nearest deadline after September 28 and it is thirteen days out. It also makes the count nine waves rather than eight. There is no dated article for it: it exists on the deprecations page alone, which is exactly the shape of thing a monthly review misses. If you run anything on a Daybreak cyber model, that is a two-week window, and the lesson generalizes beyond this one entry. “Nothing moved” is only ever a statement about the day you read the page.
September 24 is six days away, still has no successor, and there is an export window you have not used
The Sora 2 model line and the Videos API both go dark on September 24, 2026. The replacement column on the deprecations page was still empty when it was read on September 18, 2026, as it was on September 4, September 3, August 19 and August 3. Three separate OpenAI pages agree the date is live: the deprecations table, a banner on the video generation guide stating the Sora 2 models and Videos API “are deprecated and will shut down on September 24, 2026,” and a help article stating “the Sora API will be discontinued on September 24, 2026” and naming no successor. The Sora models are already absent from the models index. Now the part that is easiest to miss, and a buyer with content in Sora needs it this week: the Sora web and app experiences “were discontinued on April 26, 2026,” and OpenAI has published an export window at sora.chatgpt.com/sunset, after which “we will permanently delete any data associated with your use of Sora.” Export first, then plan the rebuild. The rebuild is not a version bump and you are six days from losing the source material as well as the API.
September 28 is the wave that is easiest to miss
Four days after Sora, a second wave lands on September 28, 2026 and takes gpt-3.5-turbo-instruct, babbage-002, davinci-002 and gpt-3.5-turbo-1106, all pointing at gpt-5.6-terra. It was announced on September 26, 2025, a full year ahead. That is the failure mode worth naming: a long-announced deadline is easier to miss than a sudden one, because nothing about it is news on any given week. It is ten days out.
Both image waves point at a model that is already a generation behind
This is exactly the sort of thing that catches teams out, and it got worse on September 8. gpt-image-1 goes on October 23, 2026. Then gpt-image-1-mini, gpt-image-1.5 and chatgpt-image-latest go on December 1, 2026, announced June 2, 2026. A team that migrates in October and closes the ticket will be back in it five and a half weeks later. The new problem is the destination. Both waves still name gpt-image-2 for all four models, but OpenAI shipped ChatGPT Images 2.5 on September 8, 2026, with GPT-Image-2.5 Flare as the API default and GPT-Image-2.5 Sunburst for premium workflows. gpt-image-2 is now one generation behind, and it is priced well below the new pair: $2.50 input, $0.625 cached and $15.00 output per 1M against $8.00, $2.00 and $30.00 for the 2.5 models. Draw the consequence rather than leaving it implied. Anyone who follows the migration instruction exactly lands on a superseded model, and anyone who instead migrates to where OpenAI’s own product direction points lands on a bill roughly three times larger. Neither is wrong, but they are different decisions and the deprecations page does not tell you a choice exists.
December 11 takes the GPT-5 era
On December 11, 2026, announced June 11, 2026, gpt-5-2025-08-07, gpt-5-mini-2025-08-07, gpt-5-nano-2025-08-07, gpt-5-pro-2025-10-06, o3-2025-04-16 and o3-pro-2025-06-10 all stop. The destination is the 5.6 family, with the pro variants going to gpt-5.6-sol using reasoning.mode: pro. Note that last detail: the pro migration is a parameter change as well as a model-string change, so a straight find-and-replace on the model ID will silently drop capability.
Transcription has a 2027 deadline now
Announced August 26, 2026: on February 26, 2027, whisper-1, gpt-4o-transcribe, gpt-4o-mini-transcribe and gpt-4o-transcribe-diarize stop, replaced by gpt-live-transcribe or gpt-transcribe. A January 20, 2027 audio and realtime wave was announced on July 20, 2026. Neither is urgent; both belong on the roadmap now, because whisper-1 in particular is embedded in a great deal of quietly working infrastructure.
A published OpenAI deprecation date is never a soft deadline
The Assistants API shut down on August 26, 2026, exactly as scheduled. It now sits under Past deprecations with shutdown date 2026-08-26, replaced by the Responses and Conversations APIs, and the August 26 changelog entry confirms it independently. The deprecations page records the full arc: “On August 26th, 2025, we notified developers using the Assistants API of its deprecation and removal from the API one year later, on August 26, 2026.” Twelve months of notice, executed to the day. That is the useful precedent: when OpenAI publishes a date on the deprecations page, it means it. Nothing on that page should be treated as a soft deadline.
The November 30 wave is the one nobody has budgeted for
Three developer-platform pieces retire on November 30, 2026: the v1/prompts API and reusable prompt objects, the Evals dashboard and API, and Agent Builder. Two of the three replacements point away from OpenAI’s own surfaces: reusable prompts become “migrate to application code,” and Evals becomes “use Promptfoo alternative.” That is a different kind of deprecation from a model swap: it removes tooling teams have built process around.
| Date | What stops working | Stated replacement | Status on September 18, 2026 |
|---|---|---|---|
| August 10, 2026 | gpt-5.2-chat-latest and gpt-5.3-chat-latest | gpt-5.6-sol | Executed on schedule; in Past deprecations |
| August 26, 2026 | Assistants API | Responses API and Conversations API | Executed on schedule; in Past deprecations with shutdown date 2026-08-26 |
| September 24, 2026 | Videos API | None stated | Scheduled; replacement column still empty with six days to go. Confirmed by the deprecations page, the video generation guide banner and the Sora discontinuation help article |
| September 24, 2026 | sora-2, sora-2-pro, sora-2-2025-10-06, sora-2-2025-12-08 and sora-2-pro-2025-10-06 | None stated | Scheduled; replacement column still empty and all five still priced. Export first: the Sora web and app experiences were discontinued on April 26, 2026 and OpenAI offers an export window at sora.chatgpt.com/sunset, after which associated data is permanently deleted |
| September 25, 2026 | Creation of new custom GPTs, on all ChatGPT plans | Plugins, via a migration workflow | NEW and seven days out. Announced September 11, 2026. Not on the deprecations page: this is a ChatGPT product retirement, published on an undated help center FAQ and a rolling release-notes entry |
| September 28, 2026 | gpt-3.5-turbo-instruct, babbage-002, davinci-002, gpt-3.5-turbo-1106 | gpt-5.6-terra | Scheduled and unchanged; announced September 26, 2025 and ten days out |
| October 1, 2026 | gpt-5.4-cyber | gpt-5.6-cyber | NEW. Announced September 11, 2026 and added to the page mid-cycle. The nearest deadline after September 28 and thirteen days out. No dated article exists for it |
| October 23, 2026 | gpt-3.5-turbo family, o4-mini, fine-tuned babbage and davinci variants | gpt-5.6-terra | Scheduled and unchanged |
| October 23, 2026 | gpt-4, gpt-4-turbo, gpt-4o-2024-05-13, o1, o1-pro and the o3-mini family | gpt-5.6-sol | Larger than it first reads. o1-pro, gpt-4o-2024-05-13 and a set of fine-tuned variants belong to it, and o1-pro migrates to gpt-5.6-sol with reasoning.mode: pro, the same parameter trap flagged for December 11 |
| October 23, 2026 | gpt-4.1-nano family | gpt-5.6-luna | Scheduled and unchanged |
| October 23, 2026 | gpt-image-1 | gpt-image-2 | Scheduled; the first of two image waves. The stated destination is now a generation behind the September 8 Images 2.5 models and priced well below them |
| October 31, 2026 | Existing evals become read-only | Part of the Evals platform deprecation that concludes November 30, 2026 | Milestone, not a new wave. Easy to mistake for a separate shutdown when the date is read on its own |
| November 30, 2026 | v1/prompts API and reusable prompt objects | Migrate to application code | Scheduled and unchanged |
| November 30, 2026 | Evals dashboard and API | Use Promptfoo alternative | Scheduled and unchanged; the October 31 read-only milestone sits inside this one |
| November 30, 2026 | Agent Builder | Agents SDK or ChatGPT Workspace Agents | Scheduled and unchanged. The Agents API that entered public beta on September 10, 2026 is a plausible fourth path, but no first-party page names it as a replacement, so do not plan the migration on it |
| December 1, 2026 | gpt-image-1-mini, gpt-image-1.5, chatgpt-image-latest | gpt-image-2 | Scheduled and unchanged; announced June 2, 2026. The second image wave, pointing at the same superseded destination |
| December 11, 2026 | Custom GPTs stop running, on all ChatGPT plans | Plugins; instructions migrate as a skill, and custom actions do not transfer | NEW. Announced September 11, 2026. Not on the deprecations page. Anything with a custom action is a rebuild, not a migration |
| December 11, 2026 | gpt-5-2025-08-07, gpt-5-mini-2025-08-07, gpt-5-nano-2025-08-07, gpt-5-pro-2025-10-06, o3-2025-04-16, o3-pro-2025-06-10 | The 5.6 family; pro variants to gpt-5.6-sol with reasoning.mode: pro | Scheduled and unchanged; announced June 11, 2026 |
| January 6, 2027 | New fine-tuning jobs, in OpenAI’s words: “Active existing customers will no longer be able to create new fine-tuning jobs on this date” | Not stated as a model swap | Milestone, not a new wave. A fine-tuning restriction rather than a model shutdown |
| January 20, 2027 | gpt-realtime, gpt-audio, gpt-4o-audio, gpt-4o-realtime, gpt-realtime-mini, gpt-audio-mini, gpt-4o-mini-realtime, gpt-4o-mini-audio, gpt-4o-mini-transcribe-2025-03-20 | gpt-realtime-2.1, gpt-audio-1.5, gpt-realtime-2.1-mini, gpt-4o-mini-transcribe-2025-12-15 | Scheduled and unchanged; announced July 20, 2026. The list is now readable and printed here in full. Separately, gpt-realtime-mini is already flagged Deprecated on the models index |
| February 26, 2027 | whisper-1, gpt-4o-transcribe, gpt-4o-mini-transcribe, gpt-4o-transcribe-diarize | gpt-live-transcribe or gpt-transcribe | Scheduled and unchanged; announced August 26, 2026 |
| Task type | Minimum verification | Human responsibility |
|---|---|---|
| Current research | Open primary links and compare dates | Approve the interpretation |
| Data analysis | Recalculate key figures and inspect assumptions | Own the decision criteria |
| Generated document | Check claims, names, tone and required structure | Approve publication |
| Image generation | Inspect text, likeness, product truth and rights | Approve external use |
| Code or agent action | Run tests and inspect real system state | Approve permissions and deployment |
First-party evidence behind this guide
OpenAI ships faster than it documents, and a headline availability claim about this vendor can be overtaken by a dated article within a week of a launch post. Treat every status statement below, including the ones about Astra, as perishable, and re-read the dated article before you repeat it. Every entry in the list below is a first-party OpenAI article, and all but two carry a publication date on the page. The two exceptions are marked in the list and named here. Lockdown Mode and elevated risk labels carries no publication date at all, so read it as undated first-party material rather than as a dated source. The GPT-6 Astra launch announcement carries no publication date of its own either; its September 3, 2026 date is established by other OpenAI pages agreeing with each other, as described at the end of this note, rather than printed on the article. Living pages are deliberately excluded, because a page rewritten in place cannot support a dated claim, so the developer pricing page, the developer model reference, the API deprecations table, the API changelog, the OpenAI product index, the enterprise rate card and every help center article are named in the prose above and attributed there rather than cited here. That means a reader has to open some pages personally, and naming them precisely is the honest alternative. The claims that rest on a page with no publication date, and which you should confirm directly, are these. The deprecations page carries the whole shutdown calendar, including the October 1, 2026 gpt-5.4-cyber wave added on September 11, 2026, which has no dated article anywhere. The help center carries the custom GPT retirement dates of September 25 and December 11, 2026, the GPT-6 Pro and Business plan message allowances, the Sites eligibility wording, the Sora discontinuation and its export window, and the enterprise token rate card, including the GPT-6 Astra Law, Daybreak Blue and Daybreak Red prices. The ChatGPT release notes carry the retirement of automatic Instant to Thinking switching and the removal of the Higher intelligence setting on web for Plus and Pro, and the statement that Voice can escalate to GPT-5.6 or GPT-6 Astra along with the Go, Plus and Pro voice allowances. The Enterprise and Edu release notes carry every September 9 to 17 admin item: SCIM for the API Platform, ChatGPT for Word, multiple accounts per plugin, group manager delegation, the Groups Admin API, Codex audit logging, model access testing, Voice usage pricing, Deep Research in Work and Codex, and the Library connector additions. The developer pricing page and the GPT-6 Astra model reference carry every Standard, Batch, Flex and Fast mode figure, the cache write rates, the Astra rate limit tiers and the 272K threshold sentence, which now lives on the model reference rather than the pricing page. The marketing pricing pages at openai.com/api/pricing and openai.com/business/pricing/ are named in the prose as a discrepancy and a seat price list respectively, and neither is cited as truth. A large block of earlier dated claims still comes from the OpenAI changelog alone: the July 28 transcription models; the August 4 dashboard, paste-behavior and education-plugin changes; the August 13 Google Drive and Computer History additions; the August 14 interactive quizzes and Linux desktop preview; the August 18 Ads expansion across Europe; the August 20 Prompt Caching dashboard, transparent backgrounds, Sites URL customization and Apple Messages plugin; the August 21 regional-processing domains; the August 24 Kiro availability; the August 25 scheduled-task and website-authentication changes; the August 26 ChatGPT for Teachers expansion; the August 27 temporary chat controls; the August 28 multiple Google accounts; the August 29 mutual TLS and workload identity federation GA; the August 31 consumer and browser-extension items; the September 2 traffic-management change; and the September 3 Responses API additions. One point about dating an OpenAI announcement, because it will come up again. The Astra launch post does appear on OpenAI’s product index, dated September 3, 2026, and for that date it is the only article listed. The announcement page itself carries no publication date. September 3 is established by the changelog, the system card, the safety overview and the ChatGPT release notes agreeing with each other, with the index as a fourth corroborant rather than a replacement. An OpenAI announcement with no date of its own is a normal occurrence on that site, so plan to date it from corroborating pages rather than assuming the page will tell you. One further open item belongs here rather than in a footnote: two reads of the API changelog returned nothing later than September 3, despite the Agents API, GPT-Live-1 and Images 2.5 all shipping between September 8 and 10. That is probably an extraction problem rather than a publishing one, and no claim here rests on the changelog for the September 4 to 18 window, but anything you source to it in that window needs a fresh read.
AI Mindset