Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

i asked opus 5.5 to explain why personal benchmarks are so important (this is one shot)

17,669 Aufrufe • vor 4 Tagen •via X (Twitter)

9 Kommentare

Profilbild von Jonny Miller
Jonny Millervor 4 Tagen

incredible

Profilbild von 黄月英
黄月英vor 4 Tagen

So good! 🤯

Profilbild von adam brotman
adam brotmanvor 4 Tagen

This is amazing. How long did it work on it? Effort max? I’m trying something similar. It’s been working for 2 hours on the video. But that might be normal timing for something so high fidelity…

Profilbild von John
Johnvor 4 Tagen

A benchmark you wrote yourself is the only one the model has not already seen.

Profilbild von JoRoan Lazaro
JoRoan Lazarovor 4 Tagen

i've been experimenting this year with how to shape models' explanations so I comprehend them better and faster, despite a habit of multi-tasking and always being under deadline. slowing it down into a spoken video explainer feels useful (maybe because I watch it full screen for focus?)

Profilbild von Ian Lucas
Ian Lucasvor 4 Tagen

Old World: Idiocracy New World: Idiographicy

Profilbild von Persona
Personavor 4 Tagen

When the Yellow soil of China appeared, I chuckled, it is about open source models anyway, cool.

Profilbild von Treynor Tetik
Treynor Tetikvor 4 Tagen

I've been sleeping on this for too long. Think I'm going to figure out something this week for a benchmark. I think some sort of slide deck benchmark. I do AI enablement and make a lot of decks.

Profilbild von Jonathan Malkin 🦊
Jonathan Malkin 🦊vor 4 Tagen

Running Opus 5.5 headless at low effort for research briefs: about 15 seconds and twelve cents a run. Low effort is the real feature. You only escalate when the call is genuinely ambiguous.

Ähnliche Videos