Loading video...

Video Failed to Load

Go Home

i asked opus 5.5 to explain why personal benchmarks are so important (this is one shot)

17,669 views • 4 days ago •via X (Twitter)

9 Comments

Jonny Miller's profile picture
Jonny Miller4 days ago

incredible

黄月英's profile picture
黄月英4 days ago

So good! 🤯

adam brotman's profile picture
adam brotman4 days ago

This is amazing. How long did it work on it? Effort max? I’m trying something similar. It’s been working for 2 hours on the video. But that might be normal timing for something so high fidelity…

John's profile picture
John4 days ago

A benchmark you wrote yourself is the only one the model has not already seen.

JoRoan Lazaro's profile picture
JoRoan Lazaro3 days ago

i've been experimenting this year with how to shape models' explanations so I comprehend them better and faster, despite a habit of multi-tasking and always being under deadline. slowing it down into a spoken video explainer feels useful (maybe because I watch it full screen for focus?)

Ian Lucas's profile picture
Ian Lucas4 days ago

Old World: Idiocracy New World: Idiographicy

Persona's profile picture
Persona4 days ago

When the Yellow soil of China appeared, I chuckled, it is about open source models anyway, cool.

Treynor Tetik's profile picture
Treynor Tetik3 days ago

I've been sleeping on this for too long. Think I'm going to figure out something this week for a benchmark. I think some sort of slide deck benchmark. I do AI enablement and make a lot of decks.

Jonathan Malkin 🦊's profile picture
Jonathan Malkin 🦊3 days ago

Running Opus 5.5 headless at low effort for research briefs: about 15 seconds and twelve cents a run. Low effort is the real feature. You only escalate when the call is genuinely ambiguous.

Related Videos