Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Both Claude Opus 5 and GPT-5.6 Sol have been performing exceptionally well on our leaderboards, ranking in the top 3 across almost all of our non-agentic tasks. We put both models through the same 3D design and website challenges to see where each excels the most. Watch the full...

30,487 görüntüleme • 11 gün önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

HERMES AGENT NOW RUNS CLAUDE OPUS 5. NEAR FABLE 5 INTELLIGENCE. HALF THE PRICE. SELF-VERIFIES ITS OWN WORK. AVAILABLE TODAY VIA NOUS PORTAL (20% OFF ALL MODELS). Anthropic shipped Opus 5 on July 24, 2026. same $5/$25 per million tokens as Opus 4.8. but the benchmarks tell a different story. WHAT CHANGED FROM OPUS 4.8: FrontierBench v0.1: Opus 5: 43.3%. Opus 4.8: 18.7%. 2.3x jump on the same test. ARC-AGI-3: Opus 5: 30.2%. 3x better than the next closest model. beat Fable 5 on 8 out of 13 benchmarks. at half the cost ($5/$25 vs $10/$50). same price as Opus 4.8. twice the intelligence. no reason to stay on 4.8. THE SPECS: model ID: claude-opus-5 context: 1M tokens (default and maximum) max output: 128K tokens thinking: on by default effort toggle: low / medium / high per request fast mode: $10/$50, 2.5x faster knowledge cutoff: May 2026 minimum cacheable prompt: 512 tokens (was 1,024) SELF-VERIFICATION (the biggest change): Opus 5 checks its own work automatically. Anthropic says: delete your verification prompts. "include a final verification step" now causes OVER-verification because the model already does it. for Hermes /goal tasks this is a direct upgrade. the judge checks evidence. the model also checks evidence. double layer of verification without extra tokens. EFFORT TOGGLE: low: fast, cheap, routine work. medium: balanced, daily tasks. high: full reasoning, complex problems. set per request. not a global switch. matches Hermes /reasoning command: /reasoning low (routine) /reasoning high (complex) Opus 5 effort toggle + Hermes reasoning control = precise cost management per turn. WHERE OPUS 5 FITS IN HERMES: DAILY DRIVER (replaces Opus 4.8): same price. 2.3x better benchmarks. set as your main model: Desktop app / Dashboard: Models → claude-opus-5 CHIEF OF STAFF: synthesis across multiple agents. reads Kanban, prioritizes, routes tasks. self-verification catches routing errors before they cascade. COMPLEX CODING: SOTA on agentic coding benchmarks. FrontierBench 43.3% = best public model for coding. set as coder profile model. /GOAL TASKS: self-verification + completion contracts = the model proves its work AND double-checks the proof. long-horizon goals finish correctly more often. MoA AGGREGATOR: strongest synthesis model at $5/$25. pair with GPT-5.6 and Grok 4.5 as references. Opus 5 aggregates. best quality at mid-range price. presets: max-quality: reference_models: - provider: openai-codex model: gpt-5.6-sol - provider: xai model: grok-4.5 aggregator: provider: anthropic model: claude-opus-5 COMPUTER USE: near-Fable 5 quality for browser automation. at half the token cost per session. computer_use tasks burn lots of vision tokens. Opus 5 halves that bill vs Fable 5. WHAT TO KEEP OPUS 5 AWAY FROM: cron monitoring: too expensive. use DeepSeek or no_agent mode. sub-agent grunt work: use GPT-5.6 Luna ($1/$6) or DeepSeek. auxiliary tasks: use Gemini Flash. routine web extraction: use a cheap model. Opus 5 is for the turns where quality compounds. planning, synthesis, verification, complex reasoning. budget models handle everything else. NOUS PORTAL: 20% OFF ALL MODELS Nous Portal currently runs a 20% discount on all models including Opus 5. $5/$25 official → $4/$20 through Nous Portal. the cheapest way to run Opus 5 right now. hermes setup --portal select claude-opus-5 as your model. discount applies automatically. Opus 5 replaces Opus 4.8 everywhere. same price. better at everything. no tradeoff. straight upgrade. hermes update /model claude-opus-5

YanXbt

16,744 görüntüleme • 20 gün önce

June 4th, 1994 our lives forever changed. We said, “I do!”. With those two words, we said, yes, to all the highs, the lows, and everything in between. God has blessed us with four absolutely amazing children who are now amazing adults, with their own best friends/significant others (that they’re doing life with), we have three incredible grandsons, and a beautiful granddaughter on the way. We’ve lived where we both grew up (on the East Coast), and have now been out here in San Diego for just over 11 years. We’ve gotten jobs (and lost jobs), we’ve had more times than we can count where we couldn’t make ends meet, even though both you and I were working two, and sometimes three jobs at a time, and we’ve been blessed in ways that we could’ve never dreamed of. We’ve watched both my parents pass on, and are now dealing with the overwhelmingly difficult challenge of seeing your parents struggle with their own health in ways that no one should have to go through. Through it all (even in the midst of the chaos), we’ve been blessed to be by each other‘s sides! I thank God for you every day, Jillian! I love our adventures together (the big ones where we fly to somewhere we’ve never been before, and the little ones where we hop in the car with no agenda, and just drive). I love when we find ourselves in deeper conversation, laughter, and tears of joy then ever expected, and in the moments of silence, where no words are even spoken, but when we’re together, just being where our feet are. As the world (as we know it), keeps getting crazier and crazier, let’s continue to keep Christ in the center of all we do, keep leaning on and lifting each other up when it’s needed, and keep living the lives that we have been so incredibly blessed to live together. I love you with all my heart Jillian. Happy 32nd (heading into our 33rd year), Anniversary.

Coach Hines 🇺🇸

10,530 görüntüleme • 2 ay önce