正在加载视频...

视频加载失败

here's how things have progressed since codex gpt-5 last year. this takes fable about 11 minutes. identical simple prompt, just asking for a minimal but playable roguelite deckbuilder.

26,752 次观看 • 3 个月前 •via X (Twitter)

6 条评论

Tenobrus 的头像
Tenobrus3 个月前

obviously it looks much better, albeit still minimal in many ways. also has more cards, relics, elites, branching paths, rest sites, a boss fight, enemies have actual abilities, meta-progression, etc etc.

Tumbles 的头像
Tumbles10 天前

lol it gave you adrenaline on floor one and an energy relic on floor 2, SNORE so much for agi this shit aint balanced

Habanero 的头像
Habanero3 个月前

It’s a very honest comparison. But obviously the more interesting one would be Codex (5.5 + imagen) vs just GPT5. Basically a product comparison instead of a pure model comparison. Codex vs Codex. App vs cli.

Seventh 的头像
Seventh3 个月前

how much $$$?

Tenobrus 的头像
Tenobrus3 个月前

$16

Matar Paneer 的头像
Matar Paneer9 天前

Since when was making prototypes comparable to any real world use?

相关视频

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 次观看 • 2 个月前