Loading video...
Video Failed to Load
stop picking one model, a spider merged OPUS 5.5 + SONNET 5.5 + JEV + OPENAI DOTS into one business system one command splits a business into 6 jobs: plan, code, decide, run 24/7, review and ship every job has 4 candidates. the spider scans them one by one,... show more
11,588 views • 4 days ago •via X (Twitter)
7 Comments

This is actually very useful. Having four candidates where the spider scans them one by one, and then Jeff scores them, I think, reduces latency and also makes the quality of decisions better.

6 jobs across 4 models?

Include GLiDE among the candidates for that “decide” job. We compared it with Jev across 155,390 Decision Index requests: 64.81 vs 57.91. I’d run your job-specific test with GLiDE before locking the decision slot to Jev.

This is a strong argument for using the right model for each job instead of forcing one model to do everything.

This is like staffing a relay team by each leg: match models to the job, then make handoffs and a shared scoring rubric the thing you test before trusting the whole system.

other people scores are good only as an example and your own check decides

wrote my own take on the dots guide earlier today mostly about the morning review part that's where it actually breaks or works for me
