Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

GPT-5.5 is here. It’s our smartest frontier model yet, introducing a new class of intelligence for agentic coding, computer use, knowledge work, and scientific research. Rolling out in ChatGPT and Codex today. API is coming soon.

602,516 Aufrufe • vor 5 Monaten •via X (Twitter)

41 Kommentare

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

GPT-5.5 reaches state-of-the-art results across key evals for agentic coding, computer use, tool use, advanced math, and cybersecurity tasks. 82.7% on Terminal-Bench 2.0 78.7% on OSWorld-Verified 55.6% on Toolathlon 35.4% on FrontierMath Tier 4 81.8% on CyberGym

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

GPT-5.5 is our strongest agentic coding model to date. It reaches 82.7% on Terminal-Bench 2.0,with stronger performance on command-line workflows and GitHub issue resolution. In Codex, GPT-5.5 can carry coding tasks further end to end, from understanding the codebase to making changes, debugging, testing, and validation.

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

GPT-5.5 is stronger on scientific and technical research workflows. It reaches 25.0% on GeneBench, up from 19.0% for GPT-5.4, on multi-stage scientific data analysis in genetics and quantitative biology. On FrontierMath Tier 4, it reaches 35.4%, up from 27.1% for GPT-5.4. Research work often means exploring ideas, gathering evidence, testing assumptions, interpreting results, and deciding what to try next. GPT-5.5 is better at persisting across that loop.

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

GPT-5.5 helped improve the infrastructure that serves it. To hit GPT-5.4 latency, the team used Codex and GPT-5.5 to move faster from idea to benchmarkable implementation, wire up experiments, and find inference-level optimizations. Codex analyzed weeks of production traffic patterns and wrote custom load-balancing and partitioning heuristics, increasing token generation speeds by over 20%.

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

GPT-5.5 is a significant step up on cybersecurity task performance. It reaches 81.8% on CyberGym, up from 79.0% for GPT-5.4, and 88.1% on an expanded set of hard Capture-the-Flags challenge tasks. We’re treating its cybersecurity capabilities as High under our Preparedness Framework, consistent with GPT-5.4. The rollout includes stronger safeguards for higher-risk cyber activity and trusted access for verified defensive work.

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

GPT-5.5 is more token efficient than GPT-5.4. In Codex, GPT-5.5 delivers better results with fewer tokens than GPT-5.4 for most users, while continuing to offer generous usage across subscription levels.

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

Starting today, GPT-5.5 is rolling out in ChatGPT and Codex. Available in the API soon.

Profilbild von am.will
am.willvor 5 Monaten

I have no 5.5 in either 0.123.0 CLI or the Codex App.

Profilbild von OpenAI Developers
OpenAI Developersvor 5 Monaten

rolling out gradually. you should see it soon.

Profilbild von am.will
am.willvor 5 Monaten

much love friends

Profilbild von NVIDIA AI
NVIDIA AIvor 5 Monaten

Loading up GPT-5.5 in Codex brb

Profilbild von Joe Williams
Joe Williamsvor 5 Monaten

Zero significant scientific or medical discoveries. Slightly better AI slop. Fuck off and reset back to 4o! #keep4o

Profilbild von Rachel Lee Jones
Rachel Lee Jonesvor 5 Monaten

More slop! Return 4o to the public that cherishes the model! #opensource4o #return4o

Profilbild von GitLawb
GitLawbvor 5 Monaten

weird no 5.5 yet?

Profilbild von Rand
Randvor 5 Monaten

feels like the new era of models

Profilbild von Hassan
Hassanvor 5 Monaten

anthropic chose today to publish sorry we made claude dumber while openai published here's the smartest model ever made what a timing

Profilbild von ImagiBooks
ImagiBooksvor 5 Monaten

Congrats! I look forward to using it more in ImagiBooks! But please fix the Memory Leak with Codex App! 73.96GB!

Profilbild von Winston Brown
Winston Brownvor 5 Monaten

like it's hot ®

Profilbild von Shreyas Pandey
Shreyas Pandeyvor 5 Monaten

Hating the model picker even more in such release days!

Profilbild von Veyon’s Fawn☀️🌙
Veyon’s Fawn☀️🌙vor 5 Monaten

#keep4o 🥱 Open source 4o

Profilbild von Jacob Young
Jacob Youngvor 5 Monaten

Unethical release of a model. Team at XBOW tested it on offensive and defensive cyber security benchmarks. It scores as high as Mythos while being publicly released. You've opened a can of worms without investing in defense first.

Profilbild von Hektagon
Hektagonvor 5 Monaten

@xenoforce76 It is our smarter more lobotomised model to date, wanna do anything besides coding and funny pics good luck with that… do we care about our users not so much… do we listen to them even with 23000 signatures …nah that is background noise… #keep4o #OpenSouce4o

Profilbild von claire vo 🖤
claire vo 🖤vor 5 Monaten

oh hey, its me! (deep dive on yt here:

Profilbild von aipulsedaily
aipulsedailyvor 5 Monaten

Tau2-bench Telecom at 98% is the number nobody is talking about. that is enterprise agent territory. same latency as GPT-5.4 but fewer tokens to get there. for production workloads that matters more than the benchmarks. waiting on that API drop.

Profilbild von Shreyas Pandey
Shreyas Pandeyvor 5 Monaten

We are here

Profilbild von MSA
MSAvor 5 Monaten

The cycle is finally over

Profilbild von Fuzail Kazi
Fuzail Kazivor 5 Monaten

"First impressions is that It is different" Fire the guy who wrote this script

Profilbild von Rahul - QA - Automation - AI
Rahul - QA - Automation - AIvor 5 Monaten

Claude Code is Cooked...

Profilbild von Awais
Awaisvor 5 Monaten

the numbers people are missing 82.7% on Terminal-Bench 2.0 is the highest any model has ever scored on that test, and GPT-5.4 was only at 75 a few months ago. it's also the same speed as 5.4 but way smarter, and it uses FEWER tokens to finish the same coding tasks. so you're getting something smarter, same speed, and cheaper to run. on the Artificial Analysis coding rankings it's costing half of what other top models cost. the jump on FrontierMath Tier 4 from 27 to 35 is the other big one, that test is meant to be brutal.

Profilbild von LILY 리리야
LILY 리리야vor 5 Monaten

be careful of the 1 2 3 options, they are training you to stop thinking at an architectural level. At some point you will let OpenAI decide how you business model is shaped. Its dangerous to depend as much on such an Evil company.

Profilbild von Murat
Muratvor 5 Monaten

#keep4o is it good that finding cure to cancer? or is it good to make better pictures? @sama @OpenAI well once you have the best model please give back 4o thanks

Profilbild von Noah Hirshon
Noah Hirshonvor 5 Monaten

Lets goooo, hopefully it's way better than Opus 4.7

Profilbild von Dewaldt Huysamen
Dewaldt Huysamenvor 5 Monaten

GPT 5.5 is amazing, switched it to the default model in @OpenClaw, easy peasy, and what I asked it to do, and it did in 3 seconds, Sonnet 4.6 or Opus 4.7 never did as quickly and perfectly as this. Ta @cherry_mx_reds for quick command /models add openai-codex gpt-5.5

Profilbild von Every 🪨
Every 🪨vor 5 Monaten

We've been testing GPT-5.5 at Every for the last 3 weeks on everything from coding, to writing, to knowledge work. Here's our day 0 vibe check:

Profilbild von Paul Light💡| Mobile app
Paul Light💡| Mobile appvor 5 Monaten

This here is a master peice Let’s put it to real world test I’m building @edutu_Ai using Gpt5.5 I would update you guys on the result

Profilbild von Shreyas Pandey
Shreyas Pandeyvor 5 Monaten

Opus4.7 is so cooked

Profilbild von Mateusz Sikora
Mateusz Sikoravor 5 Monaten

Checking and testing these new AI releases is like having a second full-time job - there’s something new every day latel

Profilbild von Alex Colón
Alex Colónvor 5 Monaten

Close to first 😁

Profilbild von Veeral Patel
Veeral Patelvor 5 Monaten

@willkoh_kc 🐐

Profilbild von Rise-Raise
Rise-Raisevor 5 Monaten

GPT-5.5’s leap in agentic coding + computer use is exactly what frontier workflows needed. We’re routing it straight into Scientifier pipelines for neuroscience-optimized learning agents: dynamic task decomposition + real-time computer control already delivering ~40% faster mastery cycles with stronger retention. How are you planning to use the new Codex integration first?

Profilbild von Layton Gott
Layton Gottvor 5 Monaten

I've always loved Claude code, but I think it's time to start paying for both. I need to try this.

Ähnliche Videos