正在加载视频...

视频加载失败

Over the past six months, we shipped 30+ models, features, and upgraded tools for the API. Our changelog has been busy. Here’s what you may have missed for the API: New models • GPT-5.5 • GPT-5.4 mini • GPT-5.4 nano • GPT-Realtime-2 • GPT-Realtime-Whisper • GPT-Realtime-Translate • GPT-Image-2 Agent...

160,100 次观看 • 3 个月前 •via X (Twitter)

0 条评论

暂无评论

原始帖子的评论将显示在这里

相关视频

The week in OpenAI and Anthropic news (Week 19, 2026) OpenAI rolled out GPT-5.5 Instant as the new ChatGPT default model with memory sources and ChatGPT for Excel and Google Sheets globally, launched three new realtime voice models in the API (GPT-Realtime-2, GPT-Realtime-Translate, GPT-Realtime-Whisper) and an OpenAI CLI, introduced Trusted Contact safety feature, GPT-5.5-Cyber for defenders, B2B Signals report, ChatGPT Futures Class of 2026, EMEA youth safety blueprint, privacy in model training explainer, expanded ads pilot, published engineering posts on low-latency voice, MRC supercomputer networking, running Codex safely internally, and investigating accidental chain-of-thought grading during reinforcement learning, plus discovered ChatGPT Personal Wiki and dropped a goblin-themed merch line that sold out Anthropic hosted Code with Claude developer conference in San Francisco, announced a new enterprise AI services company with Blackstone, Hellman & Friedman, and Goldman Sachs, signed a SpaceX compute partnership and raised Claude Code and API usage limits, made Claude for Excel, PowerPoint, and Word generally available with Claude for Outlook in beta, launched Workload Identity Federation, financial services agent templates, dreaming, outcomes, and multiagent orchestration in Managed Agents, shipped 60+ Claude Code reliability fixes, published research on agentic misalignment training, sandbagging mitigation, model spec midtraining, and Natural Language Autoencoders, donated Petri to Meridian Labs, introduced The Anthropic Institute research agenda, plus discovered Orbit proactive assistant for Cowork and /radio command in Claude Code, and more

Tibor Blaho

12,602 次观看 • 4 个月前

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 次观看 • 3 个月前