AIThis post was created with the assistance of artificial intelligence (AI).

Sep 30, 2026 · @Thorsten Three

Buying for a business?Offer from Amazon

Get business pricing on tech for your team

  • Business-only prices and quantity discounts
  • Tax-exempt purchasing
  • Multiple users, one account, clear invoices
As an affiliate, we earn on qualifying purchases.

OpenAI used its DevDay on 29 September 2026 to move from chatbot to agent platform. It announced more than 20 products and features, and pitched ChatGPT as a shared surface for humans and agents that reaches 1.2 billion weekly users. Five announcements matter most if you build or sell software:

  1. Dots are always-on agents powered by GPT-6 Astra. Each has its own cloud computer and browser, and can reach over 4,000 apps through plugins.
  2. Plugins got a major upgrade: plugin extensions that live inside ChatGPT, a new creator and submission flow, better discovery, Sites that host plugins, and MCP events that trigger automations.
  3. The Decisions API hands narrow, repetitive choices to Luna, with answers drawn from a finite list you define. It is in limited preview.
  4. The Agents API now supports computer use, so agents can operate software through its interface. OpenAI runs the Codex harness for you.
  5. Sign in with ChatGPT lets users bring their existing plan to 16 launch partners, which makes a free core app with paid upsells workable.

OpenAI DevDay 2026 at a glance

29 September 2026: more than 20 announcements across ChatGPT, Codex, the API and enterprise.

The five launches that matter, plus Sol

Six of the 20-plus announcements decide what builders do next.
DotsAlways-on agents on GPT-6 Astra with their own cloud computerrolling out
PluginsExtensions, Plugin Creator, discovery, Sites and MCP eventslive now
Decisions APILuna answers from a finite list you definelimited preview
Agents API + computer useManaged Codex harness that can drive a browserpublic beta
Sign in with ChatGPTUsers bring their plan to 16 partner appslive now
GPT-6.1 SolNear-Astra intelligence at one-fifth of the priceWork, Codex, API

The rest of the slate

Green means live, amber means preview or beta, blue means coming.

Collaboration

ChatGPT SpacelivePagesliveCollaborative slidesnext few weeksTeams and Team Taskslive@ChatGPT in Slack and TeamsliveMeetings pluginbeta, macOSShareable profileslive

Codex

Codex CloudliveCLI refreshvoice, /agentsCode ReviewliveCodex Security Cloudlive

Plans, enterprise, speed

Pro 500$500, 25× PlusUltrafastAstra now, Sol soonPrivate IntelligencePrivate Inference this fallBedrock Managed Agentsruns in AWSOpenAI Marketplace32 partners, apply
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

The model news is GPT-6.1 Sol: near-Astra intelligence at one-fifth of Astra’s token prices. I cover it in its own section, along with Ultrafast, Pro 500 and the rest of the slate.

Screenshot

Dots: OpenAI’s personal agent platform

Dots are OpenAI’s biggest consumer bet. A dot is a persistent agent that works toward your goals around the clock, learns your preferences and standards over time, and keeps several projects moving at once. You name your first dot and make it your own, and OpenAI says it envisions teams of dots working together later.

QuestionAnswer
What runs itGPT-6 Astra, with its own cloud computer and browser
What it connects toOver 4,000 apps through plugins; your laptop only if you grant permission
Where you reach itChatGPT on desktop, web and mobile, plus Slack and Teams. Voice calls work today, texting is coming
Who gets itPro and Business Premium in eligible markets now. Enterprise, Edu and Healthcare can use a beta once an admin enables it; it is off by default
What it costsThe first dot is included in the plan. Chatting with it does not count toward usage limits, but tasks it starts in Codex or ChatGPT Work do. Later you can add dots or buy more speed and monthly output per dot

What a dot does in practice, per OpenAI’s own examples: it watches customer feedback and returns tested pull requests with videos, revises a launch when scope changes, reruns a scientist’s analysis as new data arrives, and updates a sales proposal as requirements shift. One early tester’s dot noticed he had not invoiced a publication, prepared the invoice and sent it after his approval.

Dots: the personal agent platform

OpenAI’s always-on agent, how it works and how it is controlled.

How a dot works

A persistent agent with its own computer, reachable where you already work.
Diagram of a dot: you talk to it through ChatGPT, Slack, Teams and voice; it runs on GPT-6 Astra with its own cloud computer and browser and connects to over 4,000 apps through plugins; five controls keep you in chargeYouChatGPT: desktop,web and mobileSlack and Teamsvoice callstexting: comingaskprogressA dotPowered by GPT-6 AstraIts own cloud computerIts own browserMemory: learns your goals,standards and preferencesWorks 24/7, several projects at once4,000+ appsthrough pluginsYour laptop only ifyou grant permissionOther devices it canconnect toHow you stay in controlCustom Rulesallow, requireapproval or blockactionsActivity Viewfollow backgroundwork, redirectAuto-reviewchecks account-affecting actionsRead-only researchbackground toolscannot send orchangeMonitoringcan pause or stop adotSome things always stay with you, such as changing a password. Passwords are used without being shown to the model.

Who gets one, and what dots do

Pro and Business Premium
rolling out now in eligible markets. First dot included; chatting does not use plan limits, tasks started in Codex or Work do
live now
Enterprise, Edu, Healthcare
beta once a workspace admin enables it. Off by default
preview
Specialist dots
enterprise pilots: procurement, invoices, email marketing, support, contracts. Microsoft Agent 365 integration planned
coming

What OpenAI says dots do today

Feedback to fixes
Watches customer feedback, returns tested pull requests with videos
Launch changes
Revises launch materials when scope changes
Analysis
Reruns analyses as new data arrives
Proposals
Updates a sales proposal as requirements shift
Content
Drafts clips, show notes and posts for approval
Read the fine print: these are OpenAI’s own examples, and OpenAI warns that dots can make mistakes. Review consequential work.
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

Control is the design constraint. Each dot works on its own cloud computer, so yours stays separate. Dots can use saved passwords without exposing them to the model. Background “proactive research” is limited to read-only tools. Custom Rules let you allow, require approval for or block specific actions, an Activity View shows background work, and monitoring can pause or stop a dot. Certain sensitive tasks, such as changing a password, always stay with you. OpenAI still warns that dots can make mistakes.

Specialist dots are the enterprise version: agents with their own identity, credentials and IT-provisioned hardware, taking on defined responsibilities. OpenAI is starting with pilots in procurement, invoice processing, email marketing, customer support and commercial contracting, and plans a Microsoft Agent 365 integration for governance.

OpenAI doubles down on plugins

Plugins are now the way into ChatGPT, Codex and dots. A plugin bundles reusable skills, connections to external services through MCP servers, and optional UI, and you publish it once to a shared directory. Dots reach over 4,000 apps through this ecosystem, so plugin distribution decides who gets work from agents.

AnnouncementWhat it means for a builder
Plugin extensionsYour plugin gets a home in the ChatGPT sidebar, interactive panels beside the conversation, and viewers for your file types. Figma and Adobe showed integrations
Plugin Creator and new submission flowEasier building, review tracking, clearer feedback and simpler updates to a live plugin
Better discoveryImproved ranking and recommendations in the directory and inside conversations. Users choose which plugins to use and approve each one’s access
Sites can host pluginsTeammates use the same app with their own connected data and permissions (Business, Enterprise, Healthcare, Edu)
MCP eventsSupport for the proposed MCP Events specification lets a plugin start an automation when something happens in a connected app, such as a new task on a project board
OpenAI MarketplaceEligible enterprise customers can spend part of their OpenAI commitment on 32 launch partners, including Figma, Adobe, Salesforce, ServiceNow, Harvey, CrowdStrike and Baseten

The developer docs also list conversion specs for restaurant reservations, quote requests and product checkout, and a guide for submitting a Claude Code plugin, which suggests OpenAI wants to accept skills and plugins built for other agent ecosystems.

OpenAI doubles down on plugins

Plugins are the storefront for agents.

How an agent hires a business

Distribution now runs through plugins. Agents pick a capability by the outcome it delivers.
Flow: customer asks for an outcome, ChatGPT understands the job, your plugin is selected, your business does the work, result is deliveredCustomerasks for an outcome“I need this permitfixed”ChatGPT or dotunderstands the joband picks acapabilityYour plugingets selectedfor the workYour businessdoes the work,fixes and filesResultdelivered:permit approvedA worked example: an agent hiring a permit expediter. The customer never needs to know your name.

What a plugin is

Publish once, and it reaches ChatGPT, Codex and dots.
Skills
reusable instructions and scripts
MCP server
tools, data and services
Optional UI
panels and viewers in ChatGPT

What changed at DevDay

Plugin extensions
Sidebar entry, interactive panels beside the conversation, file viewers. Figma and Adobe showed integrations
Plugin Creator and submission
Easier building, review tracking, clearer feedback, simpler updates
Discovery
Better ranking and recommendations in the directory and in conversations
Sites host plugins
Teammates use the same app with their own data and permissions
MCP events
Plugins start automations when something happens in a connected app
Enterprise Marketplace
32 partners such as Figma, Salesforce, Harvey and CrowdStrike

The metric that decides who gets hired

If you do not know why your plugin was selected or ignored, you cannot improve it. Measure each step.
Funnel: intent, activation, completionIntentagent chose your plugin from the requestActivationthe plugin actually started the jobCompletionthe outcome was delivered
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

The strategic point, in the commentary slides you shared: an agent picks a business by outcome, not by brand. The customer asks for a result, ChatGPT understands the job, your plugin is selected, and your business does the work. You can be hired before the customer knows your name, so the question becomes why your plugin was selected or ignored, and that means measuring intent, activation and completion.

The Decisions API: a decision model from OpenAI

The Decisions API applies GPT-6 Luna to a narrow job. You define a set of questions with a finite list of allowed answers, send text or images as context, and get back a selection your app can use to classify content, route requests or choose an agent’s next action. It is in limited preview, with a broad release promised “in the coming days”. OpenAI published no price and no accuracy figure.

Trade press read it as a response to TypeSafe’s Jev, which is a fair inference but not something OpenAI said. What is clear is that the category I have been writing about, models that return typed decisions instead of prose, now has a second major vendor.

Decisions API (OpenAI)Jev (TypeSafe)
Built onGPT-6 LunaPurpose-built decision model
OutputA choice from your predefined answersnoul, choice and score answers with confidence
InputText or imagesText or JSON state
SpeedOpenAI claims 150 ms, against 1.6 s for a Luna prompt doing the same jobVendor: 70 to 500 ms. My fleet: 0.3 to 0.9 s
PriceNot published. Luna’s list price is $0.10 in and $0.50 out per 1M tokens$0.042 per 1M input tokens, output free
AvailabilityLimited preview from 29 SeptemberPublic since 15 September
Accuracy evidenceNone publishedMy own production measurements, plus TypeSafe’s benchmarks

The speed numbers are not comparable: one is a vendor claim for the model, the other is what I measure end to end.

The Decisions API vs Jev

OpenAI enters the decision-model category.

OpenAI now sells decisions too

The Decisions API points GPT-6 Luna at a fixed set of questions with finite answers. Reported as a response to TypeSafe’s Jev, which OpenAI did not say itself.
Status: limited preview, broad release “in the coming days”. No price and no accuracy figure published.

Speed: read with care

One number is a vendor claim for the model, the other is what I measure end to end, so this is not like for like.
Luna with a prompt
OpenAI: same classification job
1.6 s
Decisions API
OpenAI claim, on Luna
150 ms
Jev, my fleet
measured end to end
0.3–0.9 s
Jev, vendor claim
TypeSafe documentation
70–500 ms
prompt to a general modelDecisions APIJev

Side by side

Decisions API (OpenAI)Jev (TypeSafe)
Built onGPT-6 LunaPurpose-built decision model
AnswersA choice from your listnoul, choice and score, each with confidence
InputText or imagesText or JSON state
PriceNot published (Luna: $0.10 in, $0.50 out per 1M)$0.042 per 1M input, output free
AvailabilityLimited preview, 29 SepPublic since 15 Sep
Accuracy evidenceNone publishedMy production data, plus vendor benchmarks

What I will do when it opens

My method does not change with the vendor: fit test first, then shadow-test 300 to 500 past decisions per confidence band, and check that confidence is calibrated.
A cascade: your rule, then a decision model, then a frontier LLM, then a humanYour rulefree, transparentDecision modelDecisions API or Jev, clear casesFrontier LLMthe gray zoneHumanthe rest
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

What to do about it. Nothing in my method changes. The four-condition fit test and the shadow test decide whether a use case is worth wiring, whichever vendor answers. When the Decisions API opens, replay the same 300 to 500 past decisions through it, compare per confidence band, and check whether its confidence is calibrated, meaning that high-confidence answers really are right about as often as claimed. My 24 use cases are vendor-neutral, and a cascade of rule, decision model, LLM and human works with either.

The Agents API now has computer use

The Agents API entered public beta on 10 September as a managed service that runs the same Codex harness OpenAI uses internally. At DevDay it gained computer use, so an agent can operate software through its interface instead of only calling APIs. OpenAI runs the harness and, if you choose, the sandbox, and charges no extra fee beyond the tokens and tools your agents use.

CapabilityWhat you get
Managed harnessLong-running sessions with automatic context compaction, tool search, programmatic tool calling, and MCP, function and web-search tools
Multi-agentA main agent delegates to parallel subagents, each with its own context
EnvironmentsOpenAI-hosted sandboxes, or your own infrastructure, or partners such as Cloudflare, Modal, Vercel, E2B, Daytona and Oracle
Computer useThe agent drives a browser in an OpenAI-hosted desktop, with optional screenshots in the session output
ElsewhereAlso available in Codex and ChatGPT Work on Pro 500 and Enterprise, and inside AWS through Bedrock Managed Agents

How computer use is controlled. The browser needs the user’s approval before it reaches each new website origin, including public ones. Your application handles sign-in through a dedicated event that supports email, passwords and verification codes but not passkeys or QR codes, and submitted credentials stay out of the model’s input. Website content is treated as untrusted and cannot grant permission.

Agents API with computer use

A managed harness, a browser and the controls around it.

Computer use, inside a managed agent

The Agents API entered public beta on 10 September. DevDay added computer use, so agents can operate software through its interface.
Architecture: your app talks to the Agents API, which runs a sandbox and an OpenAI-hosted browser; the browser needs approval for each websiteYour appstarts the sessionfollows eventshandles approvalsand sign-inAgents APIManaged Codex harnessContext compactionTool search + tool callingSubagents in parallelMCP, functions, web searchSandboxOpenAI-hosted, your own infrastructure,or partners: Cloudflare, Modal, Vercel, E2B …Browser (computer use)OpenAI-hosted desktop. The agent seesthe page and operates it like a personWebsitesAny site the user approvesapprove each siteOrigin approvalEach new website needs the user’sapproval, public ones includedSign-in eventYour app handles it: email, password,codes. No passkeys or QR. Credentialsstay out of the model inputUntrusted pagesWebsite content cannot grant permissionor override the user
The limit that matters: origin approval is not confirmation before each action. For purchases or destructive changes, restrict the hosted browser to resources that cannot do harm, or run a browser you control.

Early customer results

Published by OpenAI in its launch post. Treat them as claims to reproduce.
60%
lower cost per case
SafetyKit
86%
fewer failed agent responses
Hypha
0.71 → 0.85
evaluation score, 4× lower latency
Ciridae
No extra fee: you pay for the tokens and tools your agents use. Also available in Codex and ChatGPT Work on Pro 500 and Enterprise, and inside AWS through Bedrock Managed Agents.
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

One limit matters for anything consequential: origin approval does not enforce confirmation before individual actions such as purchases or destructive changes. OpenAI’s own guidance is to restrict the hosted browser to resources that cannot do harm, or to run a browser you control.

Early results are vendor-published. OpenAI’s launch post quotes customers reporting a 60% lower cost per case (SafetyKit), 86% fewer failed agent responses after separating harness from sandbox (Hypha), and an evaluation score that rose from 0.71 to 0.85 with a 4x latency reduction from subagents (Ciridae). Treat those as claims to reproduce, not benchmarks.

Sign in with ChatGPT: bring your own plan

Sign in with ChatGPT now does two separate things. It is an identity login that anyone can use, available globally. It is also a way for Plus and Pro subscribers to let a partner app draw on the usage already included in their ChatGPT plan, so the app does not have to pay for inference on that user.

Sign in onlyUse my ChatGPT plan
Who can use itAnyone with a ChatGPT accountPlus and Pro subscribers
What the app getsIdentity, and faster plugin connectionEligible AI requests billed to the user’s plan
What it draws onNothingThe ChatGPT Work and Codex usage in the plan
Who sets limitsNot applicableThe user, as a weekly cap per app, shown under Settings > Usage
Partners at launch5, including Airtable, Canva, GitLab, HubSpot and Supabase16, including Devin, Notion, Vercel, T3, OpenClaw, Dactyl, Amp, Warp and OpenCode; Lovable is coming soon

The guardrails are explicit. Apps do not get access to the user’s conversations or memories. Plan usage does not reserve capacity or raise the user’s overall limits, and when the whole plan runs out, apps stop unless the user has allowed them to use credits.

Sign in with ChatGPT: bring your own plan

Users bring their ChatGPT plan, so a free core app with paid upsells works.

Two permissions, one sign-in

Signing in and using your plan are separate choices.
Flow: the user signs in, separately grants use of the ChatGPT plan, and the partner app draws on it within a weekly capUserPlus or Pro planwith Work and Codexusage included1Sign inidentity login,anyone, global2Use my planseparate permission,Plus and Pro only3Partner appeligible AI requestsdraw on the planUser caps each appweekly limit, as % of planPlan runs out?app stops, unless creditsThe app never sees the user’s conversations or memories. Plan usage does not raise the user’s overall limits.

Why this changes the business model

The commentary slides’ example: a user pays $10 a month and burns $8 of inference. If users bring their own plan, you carry no model bill, so a free core app with paid upsells works.
$2 left of $10
model bill $8
left to run the business
Before: the app pays for inference
illustrative
$10 left of $10
left to run the business
After: the user’s plan pays
before other costs

Sell workflow, not tokens

Products too niche to become OpenAI features, but valuable to a small audience.
Free core app leads to paid upsell; six things worth paying forCore appfree to try, no model billfor youPaid upselladvanced features people pay forthe workflow, not the tokensProprietary informationA specialist workflowAccess to outside systemsCollaboration and auditabilityA network of humans or suppliersA measurable business resultWhat is worth paying for once tokens are out of the price:
An agent that cleans architectural CAD files
A researcher monitoring one scientific field
A Shopify catalog cleanup desktop app
A video tool built around one editing workflow
A contract-review tool for one franchise agreement

Who takes part at launch

16 partners can use a plan; 5 more offer sign-in only. Lovable is coming soon.
DevinNotionVercelT3OpenClawDactylAmpWarpOpenCodeConductor
Airtable sign-in onlyCanva sign-in onlyGitLab sign-in onlyHubSpot sign-in onlySupabase sign-in only
Named examples only; OpenAI lists 16 plan partners in total.
Limits: Free and Go users cannot bring a plan. It draws on Work and Codex usage, not general chat. Users can cap you, and developers apply to join.
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

Why this changes the business model. Until now, a niche AI app often had to charge enough to cover its own model bill. The commentary slides you shared put it this way: if a user pays $10 a month while consuming $8 of inference, the business barely works, but if users bring their ChatGPT plan, a company can focus on a narrow workflow without carrying the model bill. The consequence is that a free core app with paid upsells becomes workable, because the free tier no longer costs you inference. You sell workflow instead of tokens.

The same slides list what is worth paying for once tokens are out of the price: proprietary information, a specialist workflow, access to outside systems, collaboration and auditability, a network of humans or suppliers, and a measurable business result. Their examples are deliberately narrow: an agent that cleans architectural CAD files, a researcher that monitors one scientific field, a Shopify catalog cleanup desktop app, a video tool built around one repeatable editing workflow, and a contract-review tool for one type of franchise agreement. The opportunity, in their words, is products too niche to become OpenAI features but valuable enough for a small audience to pay for.

Limits to plan around. Only Plus and Pro users can bring a plan, so Free and Go users cannot. The plan usage is the Work and Codex allowance, not general chat. Each user can cap your app, so heavy use can be throttled by the person you are serving. And OpenAI launched with a limited partner set, so developers apply to take part.

GPT-6.1 Sol, Ultrafast and the new plans

GPT-6.1 Sol is the model story of the day. OpenAI calls it near-Astra intelligence for coding, computer use and professional work, at one-fifth of Astra’s standard token prices.

ModelInput per 1MCached input per 1MOutput per 1M
GPT-6 Astra$10$1$50
GPT-6.1 Sol$2$0.10$10
GPT-6 Luna$0.10$0.01$0.50

Sol is available to Plus, Pro, Business, Enterprise and Edu users in ChatGPT Work and Codex, and through the API as gpt-6.1-sol. It is not yet in regular Chat.

GPT-6.1 Sol

OpenAI’s new workhorse, in price and performance.

Near-Astra intelligence at one-fifth of the price

Prices per 1M tokens.
GPT-6 Astra
Input $10
Output $50
Cached input $1
GPT-6.1 Sol
Input $2
Output $10
Cached input $0.10
GPT-6 Luna
Input $0.10
Output $0.50
Cached input $0.01

The independent view

Artificial Analysis Intelligence Index v4.3.2, cost per task. Sol sits 1 to 2 points under Astra and Fable, and 5 under Opus 5.5 at xhigh, at a fraction of the cost.
GPT-6.1 Sol medium
$0.21index 48
GPT-6.1 Sol high
$0.32index 50
GPT-6.1 Sol xhigh
$0.39index 51
Opus 5.5 high
$1.82index 54
GPT-6 Astra max
$3.26index 53
Opus 5.5 xhigh
$3.46index 56
Fable 5.1 max
$7.63index 53
Sol first-token time is 57 to 69 s at high and xhigh, so it is not an interactive model at those settings.

What OpenAI claims

Vendor-reported. Competitor scores come from public reports.
~1/5
Cost of Astra
DeepSWE v1.1: matches Astra on complex software-engineering tasks.
−2.1
Points from Astra
OSWorld 2.0 offline at max effort, at roughly one-seventh of the cost per task.
+2.2
Points over Opus 5.5
AutomationBench at medium effort, at about a third of the cost.

Cost per task, Terminal-Bench Science (max effort)

GPT-6.1 Sol
$5.47
Opus 5.5
$23.21
GPT-6 Astra
$23.80
Astra still scores highest at 68.1% and OpenAI recommends it for the hardest scientific tasks.
Where I use it: Opus 5.5 at high or xhigh stays my main model. Sol digs into details and reviews, because a second opinion at $0.32 to $0.39 per task is affordable on every change. Available to Plus, Pro, Business, Enterprise and Edu in Work and Codex, not yet in Chat, and via the API as gpt-6.1-sol.
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

What OpenAI claims (vendor-reported; competitor scores come from public reports):

  • On DeepSWE v1.1 it matches Astra at roughly one-fifth of the cost.
  • On GDP.pdf it scores above Opus 5.5 with fallbacks at less than half the cost per task, and near Astra at about one-fifth.
  • On AutomationBench it beats Opus 5.5 at medium effort by 2.2 points at about a third of the cost.
  • On OSWorld 2.0 it comes within 2.1 points of Astra at roughly one-seventh of the cost per task.
  • On Terminal-Bench Science it costs $5.47 per task against $23.21 for Opus 5.5 and $23.80 for Astra, though Astra still scores highest at 68.1%.
  • Factual errors on hard prompts fall from 11.4% to 7.7% at low effort.

The independent view. On the Artificial Analysis Intelligence Index (v4.3.2), Sol scores 48 at medium for $0.21 per task, 50 at high for $0.32 and 51 at xhigh for $0.39. That is 1 to 2 points under Astra and Fable 5.1 (both 53 at max) and 5 under Opus 5.5 at xhigh (56), at a fraction of the cost. It is also concise, but its first token takes 57 to 69 seconds at high and xhigh.

Where I use it: Opus 5.5 at high or xhigh stays my main model. Sol is the model I use to dig into details and to review, because at $0.32 to $0.39 per task a second opinion on every change is affordable.

Speed and plans

  • Ultrafast is a paid speed tier with up to 8× faster generation (300 tokens per second) in Codex and up to 6× in the API. Astra Ultrafast is live in the API and on Pro 500 and Enterprise. Sol Ultrafast is promised within days.
  • Pro 500 costs $500 a month, gives 25× the Plus allowance and includes Ultrafast.
  • Pro 200 reopened, with less usage. The Next Web reports that from 30 October its Work and Codex allowance falls from 20× to 10× Plus, existing subscribers keep current limits until 29 October and get a one-time $2,500 credit, and the five-hour limit will not return. An earlier BGR report said existing subscribers would keep 20× for an unspecified period, so check the help page.
  • Latency: OpenAI reports 45% lower API time to first token and over 30% faster tool calls and workflows.

Everything else announced or shared

OpenAI counted more than 20 announcements. The rest of the slate falls into four groups: team collaboration, Codex, enterprise and security, and distribution.

GroupAnnouncementDetails and availability
CollaborationChatGPT SpaceA shared home where teammates, ChatGPT and dots keep files and project context. Pro, Business and Enterprise on desktop and web
CollaborationPagesEditable documents built for people and agents together, including charts and interactive tools
CollaborationCollaborative slidesMulti-editor decks with comments, exportable to PowerPoint or Google Slides. Coming in the next few weeks
CollaborationTeams and Team TasksShare pages, slides and plugins, and assign recurring work. Business and Enterprise
Collaboration@ChatGPT in Slack and TeamsMention it in a channel, thread or DM; teammates need no individual license. Business and Enterprise
CollaborationMeetings pluginNotes and action items saved to Space; audio is deleted once notes are ready. Beta on macOS for Pro and Business
CollaborationShareable profilesOne page showcasing your Sites and plugins
CodexCodex CloudTasks run while your laptop is closed, from any device, with reusable team environments
CodexCLI refreshTwo-way voice, an /agents view for tracking several tasks, better session and worktree handling
CodexCode ReviewSummaries, diffs and questions about changes; automatic first-pass reviews in the cloud
CodexCodex Security CloudScheduled or on-demand repository scans that investigate findings, remove duplicates and prepare fixes
EnterprisePrivate IntelligenceZero Data Retention with Private Safety Processing now; a Private Inference preview with confidential computing this fall
EnterpriseBedrock Managed AgentsOpenAI-powered agents that run entirely inside AWS
EnterpriseOpenAI MarketplaceSpend part of an OpenAI commitment on 32 partners such as Figma, Salesforce, Harvey and CrowdStrike
DistributionPlugins, Sites, MCP events, Sign in with ChatGPTCovered above

Plans, speed and dates to watch

Pro 500, Ultrafast and what changes next.

The usage ladder

Work and Codex allowance, as a multiple of Plus.
Plus
the baseline
1× Plus
Pro 100
stays at 5×
5× Plus
Pro 200
reopened; 10× from 30 Oct
10× Pluswas 20× until 29 Oct
Pro 500
$500 a month, includes Ultrafast
25× Plus
Pro 200 change as reported by The Next Web. An earlier BGR report said existing subscribers would keep 20× for an unspecified period, so check OpenAI’s help page. Existing Pro 200 subscribers also get a one-time $2,500 credit, and the five-hour limit will not return.

Ultrafast

A paid speed tier for work where speed matters most.
300
tokens per second
Up to 8× faster in Codex and 6× in the API.
Astra Ultrafast
Live in the API and on Pro 500 and Enterprise.
live now
Sol Ultrafast
Promised within days.
coming

Dates to watch

29 SepDevDay: dots, Sol, plugins, Agents API, Sign in with ChatGPT, Pro 500
Coming daysDecisions API broad release; GPT-6.1 Sol Ultrafast
Next few weeksCollaborative slides
Soon, no dateTexting your dot; dots for more users; more dots per account
29 OctExisting Pro 200 subscribers keep current limits until this date
30 OctPro 200 Work and Codex allowance drops to 10× Plus; GPT-6 Pro chat messages fall from 200 to 100 a week
This fallPrivate Inference preview with confidential computing
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

A few other things surfaced around the keynote. OpenAI said the ChatGPT surface reaches 1.2 billion weekly users. It briefly mentioned an OpenClaw enterprise harness without detail and announced a worldwide usage reset at the end. Sam Altman teased AI hardware as “something that’s worth waiting for” and an agent safety platform similar to Nvidia’s, according to Axios. Axios also noted the launch came the same week OpenAI said, citing the New York Times, that it would not release its newest flagship model, GPT-6.1 Astra, because of security concerns.

What it means for builders

The pattern across all five key announcements is the same: the model will own the decision, so build what happens before and after it. Models keep getting cheaper, and the Decisions API now sells the decision itself as a product.

Where to build the moat

What the DevDay launches mean for builders.

Own two of three

The model will own the decision. Own what happens before and after it.
Four steps: trigger, decision, action, feedback. Own two of trigger, action and feedback1. TriggerAn invoice becomes dueInventory runs lowIt is tax timeOwn the event2. DecisionWhat should happen next?Models get cheaperCommodity3. ActionCorrect and submitOwn the ability4. FeedbackApproved or rejected?Own the outcomeOwn 2 of 3: trigger, action, feedback.If all you own is the decision, a better model can replace you.

Where this week’s launches land

My mapping of the DevDay announcements onto the four steps.
Trigger
MCP events, dots’ proactive research, Team Tasks
Decision
Decisions API, GPT-6.1 Sol, Jev
Action
Agents API with computer use, plugins, Codex Cloud
Feedback
Plugin conversion data, approvals, outcome history

Two product ideas

From the commentary slides, matched to this week’s launches.
Real-world work API
Agents send structured jobs to vetted specialists, such as permit expediters or customs brokers.
Money: Transaction fee
Moat: Supply and quality history
Agent distribution
Learn why your plugin was selected or ignored. Test: intent, then activation, then completion.
Money: Agency work becomes software
Moat: Intent-to-conversion data

How to get there

Start by doing the work yourself, then turn what you learn into rules and software.
Ladder: service first, then exceptions, then rules, then softwareService firstExceptionsRulesSoftware
From “OpenAI DevDay 2026: Everything Announced and What It Means” on thorstenmeyerai.com. Sources: OpenAI DevDay recap and product pages, OpenAI developer docs, The New Stack, The Next Web, Axios, Artificial Analysis. Checked 30 September 2026. OpenAI figures are OpenAI’s own claims.

The commentary slides you shared turn this into a rule: own two of three, trigger, action and feedback. If all you own is the decision, a better model can replace you.

StepWhat happensWhat to own
1. TriggerAn invoice becomes due, inventory runs low, it is tax timeThe event that starts work
2. DecisionWhat should happen next?Nothing: models get cheaper
3. ActionCorrect and submitThe ability to do the work
4. FeedbackApproved or rejected?The outcome, and the data on it

The same slides sketch two product ideas that fit this week’s launches:

Real-world work APIAgent distribution
IdeaAgents send structured jobs to vetted specialists, such as permit expediters or customs brokersLearn why your plugin was selected or ignored
TestCan a job arrive as structured input and end as a verified result?Intent, then activation, then completion
MoneyTransaction feeAgency work becomes software
MoatSupply and quality historyIntent-to-conversion data

They add a path for both: service first, then exceptions, then rules, then software.

My own reading of the week:

  1. Treat plugins as your storefront. Build an outcome-shaped plugin and MCP server, then measure selection, activation and completion. Dots make agents your biggest new channel.
  2. Try Sign in with ChatGPT if you sell a narrow workflow to Plus and Pro users, and design the free tier around the fact that users can cap you.
  3. Watch the Decisions API when it opens, and shadow-test it against Jev and your current model on the same past decisions.
  4. Pilot computer use only with limits. Approval is per origin, not per action, so keep consequential steps outside the hosted browser.
  5. Keep a two-model stack. Opus 5.5 builds, GPT-6.1 Sol reviews, and the price of a second opinion is now small.

Dates to watch: the Decisions API broad release (“in the coming days”), Sol Ultrafast, collaborative slides (next few weeks), the Pro 200 change on 30 October, and Private Inference this fall.

Sources

Checked on 30 September 2026. Where OpenAI is the source, the figures are its own claims.

FALL

Fall Picks

As an affiliate, we earn on qualifying purchases.

You May Also Like

Bosch Faces a Major Shake‑Up: What the Cost‑Cutting Drive Means for Workers and Why Artificial Intelligence Matters

AIThis post was created with the assistance of artificial intelligence (AI).A Tale…

AI Productivity Paradox: If AI Is So Powerful, Why Isn’t Productivity Booming?

AIThis post was created with the assistance of artificial intelligence (AI).Despite widespread…

GPT-6 Sol and Luna: Cheaper Intelligence Changes the Economics of Work — but the Review Bill Remains

AIThis post was created with the assistance of artificial intelligence (AI).OpenAI’s new…

The City That Watches Itself: The Living Digital Twin, and the God’s-Eye View We’re Building

AIThis post was created with the assistance of artificial intelligence (AI).Somewhere in…