Загрузка видео...

Не удалось загрузить видео

На главную

Asked opus 5.5 to make a video of openai solving Navier-Stokes. Very nice model, I still prefer Astra and 5.6 as my main drivers as I do a lot of research heavy work but 5.5 Opus is the first time since 4.7 that I want to heavily run Claude....

41,222 просмотров • 6 дней назад •via X (Twitter)

Комментарии: 21

Фото профиля echo.hive
echo.hive6 дней назад

Once they figure out “music by code” too it will be fantastic!

Фото профиля Jimmy Apples 🍎/acc
Jimmy Apples 🍎/acc6 дней назад

Yes, music still a little slop like

Фото профиля echo.hive
echo.hive6 дней назад

It also is very good with ffmpeg I just pointed opus to my web page for my get amplified series and told it to use ffmpeg only, tho I am not sure if it used any JS or anything else :) I told it not to install anything new tho lol retry amazing. I also had these suno mp3 folder. I just told it to pick a song and it looked at the waveforms or something to determine which would go best

Фото профиля echo.hive
echo.hive6 дней назад

Not “lol retry..” I meant “really amazing “ lol

Фото профиля tetsuo
tetsuo6 дней назад

That's clean. Damn

Фото профиля Jimmy Apples 🍎/acc
Jimmy Apples 🍎/acc6 дней назад

The bar has been raised

Фото профиля Sieventer
Sieventer5 дней назад

For researching, GPT is total SOTA. Claude never satisfied me to find accurate info the internet. And is not precisely a irrelevant use case… But undoubtedly, Claude intrinsic creativity is light years from the autistic void of GPT. Tho that’s not always negative.

Фото профиля Jimmy Apples 🍎/acc
Jimmy Apples 🍎/acc5 дней назад

Now we just need the model has both

Фото профиля Sieventer
Sieventer5 дней назад

I have to say that GPT’s image generation capabilities can go a long way toward compensating for its lack of internal creativity. Because if Astra, with its excellent vision, can replicate what GPT has generated, the idea you had in your head, it can do a very good job.

Фото профиля Rick F
Rick F6 дней назад

Not Sol 6 now?

Фото профиля Jimmy Apples 🍎/acc
Jimmy Apples 🍎/acc6 дней назад

Haven’t tried yet, will be giving it a go tonight, I suspect I’ll like it over 5.6

Фото профиля fuckfaggot
fuckfaggot6 дней назад

you do a lot of research on how to larp on twitter?

Фото профиля Iason Koukas
Iason Koukas6 дней назад

Which video gen did it use, or was it programmatically made?

Фото профиля Jimmy Apples 🍎/acc
Jimmy Apples 🍎/acc6 дней назад

All done in plain JavaScript

Фото профиля Mark Ruff 💎
Mark Ruff 💎6 дней назад

seeing everyone praise Opus 5.5 is giving me serious FOMO 😭 I’d be testing it too if my local bank card was accepted for the subscription

Фото профиля Abram Jackson
Abram Jackson6 дней назад

It's a good model sir

Фото профиля Baby Yoda 🧸
Baby Yoda 🧸6 дней назад

oh no

Фото профиля Rohan
Rohan5 дней назад

What's the research task where Astra still pulls ahead of Opus for you?

Фото профиля agotpier
agotpier5 дней назад

Now ask Sol 6 the same

Фото профиля hello
hello5 дней назад

Yeah I figured. I'm going to get another openai 20x plan and reinstate my Claude for frontend UI. Claude still feels like the artistic art student and openai feels like the austistic academic/engineering realist

Фото профиля AGTP
AGTP5 дней назад

It is interesting to see how the performance of Opus 5.5 is already sparking so much interest. We actually went deeper on the rumors surrounding this launch here:

Похожие видео

BREAKING: GPT-5.5 "Spud" is out and it is a BEAST We've been testing it Every 📧 for the last 3 weeks on everything from coding, to writing, to knowledge work. Here's our day 0 vibe check: - It's a step change in coding AND it's easy to talk to. It's fast and friendly and quickly became my daily driver. But it's also a coding powerhouse—a really rare combination. - It scored 62/100 on our Senior Engineer benchmark. Opus 4.7 scored only a 33/100. (But GPT-5.5 performed best when using an Opus 4.7 plan). Naveen Naidu used over 900 million tokens during testing—and it let him ship production features for Monologue at both high speed and quality. - It has serious conceptual clarity. It can hold a complex plan in its head over hours of work, without getting distracted by existing code. This makes it the first model that we've tested that can perform well on complex refactors requiring deleting and reimagining an substantial existing codebase. - It's a very good writer. This is the first OpenAI model in about a year that got our writers Every 📧 to switch away from Claude. 5.5 has Katie Parrott's seal of approval—not an easy task. Its writing feels more organic and it's better at mimicking a writing style without going overboard. - It's great for agentic knowledge-work. This is the first OpenAI model that manages to be both a stellar senior engineer AND that can be used for everything from spreadsheets to research. It's crazy fast, and it's amazing inside of the Codex desktop app, and got much of our team to switch away from Claude Code and Cowork during the testing period. However, it's not a perfect model. - 5.5 still loses to Opus 4.7 on plan quality. It's plans are extremely readable but Opus has better attention to detail and sharper insight. - 5.5 still loses to Opus 4.7 by a bit on front-end and full-stack product work. Kieran Klaassen found that it wasn't quite as good when full-stack thinking and design are involved. And it's not great writing Ruby. - 5.5 is a great vibe coder but if you're vibe coding without a plan it's worse than Opus. Mike Taylor found that Opus is better at reading in between the lines on underspecified vibe-coding tasks. Overall GPT-5.5 is a massive achievement from OpenAI and it deserves a serious look as your daily driver. Read our full vibe check on Every 📧 here:

Dan Shipper 📧

130,382 просмотров • 5 месяцев назад