Mark RadarMARK RADAR
About
EN
Sign in

OpenAI's GPT-5.5 Delivers Major Performance Gains, Expands Codex's General-Purpose Computing Capabilities

3 reports · First detected 2026-05-01 · Last active 2026-05-09

OpenAI launched GPT-5.5 on April 23, 2026, and expanded Codex's role beyond software development to research, data analysis, document and spreadsheet creation, and software operation. The shift means AI agents can do more than assist with coding: they can complete longer, general-purpose computer workflows across multiple tools. The model became available through the API on April 24.

The UK AI Security Institute published an evaluation on April 30, 2026, finding that GPT-5.5 performed comparably to Anthropic's Claude Mythos Preview across 95 cybersecurity tasks spanning four difficulty levels. It also became the second model to complete an end-to-end simulated enterprise-network attack, a test estimated to take a human 20 hours. API pricing is $5 per million input tokens and $30 per million output tokens.

All Coverage

3 original reports
NEWS.SMOL.AI 2026-05-08
not much happened today
NEWS.SMOL.AI 2026-04-30
not much happened today

The Backstory

The history behind this event
OpenAI Suspected of Secretly Testing GPT-5.6 as ChatGPT Quality and Reasoning Improve Sharply2026-06-23 · 1 reports · similarity 0.80

OpenAI has not publicly confirmed GPT-5.6, but ChatGPT users this week have widely reported more coherent responses and deeper reasoning, along with longer wait times. This has fueled speculation that the platform is conducting A/B testing with a subset of accounts. If true, OpenAI is testing the performance and computing costs of its next-generation model, with potential implications for competition in the generative AI market.

Recent hands-on tests indicate that the suspected new model excels at complex tasks such as generating 3D games and simulating robots, with some users describing it as better than Fable 5. Market rumors suggest GPT-5.6 could be released as early as June 25. However, OpenAI has yet to disclose the scale of the testing, model parameters, pricing or a firm timetable, and the claims still await official confirmation.

OpenAI Launches GPT-5.5, Touting Agentic Computing and 12M-Token Context Window2026-04-24 · 8 reports · similarity 0.83

Large language models are evolving from single-turn question-answering tools into agentic systems that can independently plan, use tools and complete multistep tasks. OpenAI is positioning GPT-5.5 as the primary workhorse for ChatGPT and Codex, targeting software development, data analysis, document creation and scientific research. The move shows that competition among AI models has expanded into enterprise workflows and research output.

OpenAI released GPT-5.5 on April 23, 2026, and made it available through the API on April 24. Official specifications list a 1 million-token context window for the API and 400,000 tokens for Codex, not 12 million. The model topped the Artificial Analysis Intelligence Index and scored 82.7% on Terminal-Bench 2.0. Pricing is $5 per million input tokens and $30 per million output tokens.

OpenAI Launches Six-Week Codex Update Sprint to Counter Anthropic’s Claude Code2026-03-17 · 4 reports · similarity 0.82

OpenAI and Anthropic are competing for the AI software development tools market, with Codex going head-to-head with Claude Code. OpenAI President Greg Brockman has instructed engineering teams to make AI agents central to their daily development work. The aim is to shorten coding, testing and deployment cycles, deepen the model’s adoption in enterprise software workflows and strengthen the ChatGPT ecosystem.

OpenAI recently began a six-week product sprint that includes rapid releases of the Codex desktop app, GPT-5.4 and security-agent features. It also plans to combine ChatGPT, Codex and the Atlas browser into a desktop “super app.” The company has not disclosed the investment involved or exact release dates. The central goal of the update push is to counter Anthropic’s Claude Code.

OpenAI Launches GPT-5.4 With Native Computer Use and Enhanced AI Agent Capabilities2026-03-06 · 4 reports · similarity 0.80

Since launching the GPT-5 series, OpenAI has continued to develop large language models beyond conversational tools into AI agents capable of executing multistep tasks. Released on March 5, 2026, GPT-5.4 is the company’s first general-purpose model with native computer-use capabilities. It can issue mouse and keyboard commands based on screenshots and complete tasks across websites, spreadsheets and documents, lowering the integration barrier for automating professional workflows.

OpenAI made GPT-5.4 Thinking available in ChatGPT, the API and Codex on the same day. API pricing is $2.50 per million input tokens and $15 per million output tokens. The model achieved an 83.0% win-or-tie rate on the GDPval professional-work benchmark. In Tool Search testing across 250 tasks, it reduced token usage by 47% without sacrificing accuracy.

Mark Radar|MARK RADAR

If you search news on Google, you can set Mark Radar as a preferred source—our coverage will show up more often in your results. Set as preferred source on Google →

All times are in Taipei time (GMT+8)