OpenAI logo

OpenAI

GPT-5.5

FrontierAvailable in FlowHunt

GPT-5.5 is a top-tier frontier AI model, developed by OpenAI and released in April 2026. It is a proprietary closed-weight model, available through the OpenAI API. The 1M-token context window is large enough to hold entire codebases, lengthy technical documents, or extended agentic sessions without truncation.

On the FlowHunt AI Leaderboard it is one of 24 tracked models, evaluated across 6 benchmarks. Standout scores: 82.7% on Terminal-Bench (#1 of 8); 88.7% on SWE-bench Verified (#2 of 13); 93.6% on GPQA Diamond (#2 of 14).

GPT-5.5 was released by OpenAI on April 23, 2026 as an iterative refinement of GPT-5. At launch it matched Claude Opus 4.6 on GPQA Diamond at 93.6% and recorded the highest Terminal-Bench score of any model measured — 82.7% — reflecting particular strength in autonomous system administration, shell scripting, and DevOps automation. OpenAI positioned GPT-5.5 as the primary general-purpose model in its API, and it became the default model powering ChatGPT for most users. It also recorded 68.4% on SWE-bench Pro and 88.7% on SWE-bench Verified, placing it among the top three coding models globally.

Note: Highest SWE-bench Verified before Fable 5. Best OpenAI Terminal-Bench score.

Released
April 2026
Context
1M tokens
Weights
Closed
Tier
Frontier
Provider
OpenAI

GPT-5.5 Benchmark Rankings

Rank among all 24 models tracked on this leaderboard. SR = self-reported · 3P / AA / C = independently verified.

BenchmarkScoreRankSourceWhat it measures
SWE-bench Verified88.7%2 of 13SRReal GitHub issue resolution (500 verified issues)
SWE-bench Pro58.6%7 of 9SRMulti-language, standardised scaffold
GPQA Diamond93.6%2 of 14SRGraduate-level Google-proof science Q&A
Terminal-Bench82.7%1 of 8SRAgentic Linux terminal task completion
HLE41.4%3 of 6SRHumanity's Last Exam (50+ STEM disciplines)
AA Intelligence Index60.2 pts2 of 9AAArtificial Analysis composite score (independent)

Detailed Benchmark Scores

SWE-V SR
88.7%
Real GitHub issue resolution (500 verified issues)
SWE-Pro SR
58.6%
Multi-language, standardised scaffold
GPQA ◇ SR
93.6%
Graduate-level Google-proof science Q&A
Terminal SR
82.7%
Agentic Linux terminal task completion
HLE SR
41.4%
Humanity's Last Exam (50+ STEM disciplines)
AA Idx AA
60.2 pts
Artificial Analysis composite score (independent)

GPT-5.5 vs. Frontier Peers

Head-to-head benchmark comparison with other Frontier-tier models. Higher is better for all metrics.

BenchmarkGPT-5.5Claude Opus 4.8Claude Opus 4.7
SWE-V88.7%88.6%87.6%
SWE-Pro58.6%69.2%64.3%
GPQA ◇93.6%93.6%94.2%
Terminal82.7%74.6%66.1%
HLE41.4%
AA Idx60.2 pts61.4 pts57.3 pts

About OpenAI

OpenAI Founded 2015 · San Francisco, CA
Website →

OpenAI was founded in December 2015 as a nonprofit AI research laboratory by Sam Altman, Elon Musk, Greg Brockman, Ilya Sutskever, John Schulman, and Wojciech Zaremba, with the mission of developing artificial general intelligence (AGI) that benefits all of humanity. In 2019 it created a capped-profit subsidiary, OpenAI LP, to raise institutional capital and scale model training. The company is known for the GPT series of language models (GPT-1 through GPT-5.5), the DALL-E image generation series, Whisper speech recognition, and Sora video generation. OpenAI operates the ChatGPT consumer application, a developer API, and enterprise services through Microsoft Azure. As of 2026 it maintains a close strategic partnership with Microsoft, which is the company's largest investor and cloud provider.

Use Cases

What to Use GPT-5.5 For

Software Engineering
Scientific & Technical Reasoning
Agentic CLI Automation
Long-Document & Codebase Analysis

Strengths & Limitations

Strengths

  • Strong software engineering — 88% SWE-bench Verified
  • Graduate-level science reasoning — 93% GPQA Diamond
  • Agentic Linux / CLI automation — 82% Terminal-Bench
  • Composite intelligence — 60 pts AA Intelligence Index (independent)
  • Very long documents and large codebases (1M context)

Limitations

  • Closed-weight — cannot be self-hosted or fine-tuned on private data

Browse the Full AI Model Leaderboard

Frequently asked questions

Learn more

GPT-5 — Benchmarks, Capabilities, and Use Cases
GPT-5 — Benchmarks, Capabilities, and Use Cases

GPT-5 — Benchmarks, Capabilities, and Use Cases

GPT-5 is a proprietary AI model from OpenAI, classified as Frontier tier with a 400K tokens context window. It is available directly in FlowHunt. Explore benchm...

3 min read
GPT-5.4 — Benchmarks, Capabilities, and Use Cases
GPT-5.4 — Benchmarks, Capabilities, and Use Cases

GPT-5.4 — Benchmarks, Capabilities, and Use Cases

GPT-5.4 is a proprietary AI model from OpenAI, classified as Frontier tier with a 1M tokens context window. It is available directly in FlowHunt. Explore benchm...

3 min read
GPT-4.1 — Benchmarks, Capabilities, and Use Cases
GPT-4.1 — Benchmarks, Capabilities, and Use Cases

GPT-4.1 — Benchmarks, Capabilities, and Use Cases

GPT-4.1 is a proprietary AI model from OpenAI, classified as Advanced tier with a 1M tokens context window. It is available directly in FlowHunt. Explore benchm...

3 min read