Loading video...

Video Failed to Load

Go Home

Grok 4.7 beats Opus 5 at 10× lower price! the test: four self-contained three.js space scenes: a night rocket launch with fire and smoke, an alien mothership lasering incoming meteors, a rotating spiral galaxy, a black hole swallowing planets and stars cost: Grok 4.7: $0.20 Claude Opus 5: $2.05...

33,356 views • 6 days ago •via X (Twitter)

23 Comments

Hugo Plat's profile picture
Hugo Plat6 days ago

10x difference?!?!? OMG

AI/ML API's profile picture
AI/ML API6 days ago

yeah, it's crazy!

Lookoff's profile picture
Lookoff6 days ago

Grok is cooking 🔥

AI/ML API's profile picture
AI/ML API6 days ago

yes! and it's live on AI/ML API

AshutoshShrivastava's profile picture
AshutoshShrivastava6 days ago

10X is huge

AI/ML API's profile picture
AI/ML API6 days ago

yes! and you can try it on AI/ML API

AI/ML API's profile picture
AI/ML API6 days ago

try Grok 4.7

Volodymyr Bandura's profile picture
Volodymyr Bandura6 days ago

You really don’t feel this efficiency in subscription, they are better on Claude max compared to cursor ultra

AI/ML API's profile picture
AI/ML API6 days ago

They’re also great when you try them with AI/ML API 😉

Lexi's profile picture
Lexi6 days ago

The combination of animation, effects, and geometry raises the bar.

AI/ML API's profile picture
AI/ML API6 days ago

Exactly

Vishal Singh 🥑's profile picture
Vishal Singh 🥑6 days ago

not much difference is outputs, but 10x difference in price...crazy grok wins this one big time 💪

AI/ML API's profile picture
AI/ML API6 days ago

10x is no joke

Enzo's profile picture
Enzo6 days ago

yep thats it, rocket science

AI/ML API's profile picture
AI/ML API6 days ago

Literally it

Shinra's profile picture
Shinra6 days ago

grok 4.7 at 10x cheaper AND faster on 3d generation is the kind of concrete comparison that matters. one test doesn't prove everything but if this scales to production workflows, that's a real displacement event. opus still wins on pure reasoning complexity but for creative/visual work where you need speed and iteration, grok just became the obvious pick. that's how markets consolidate around price/performance winners

Chansoo Byeon's profile picture
Chansoo Byeon6 days ago

Grok Groks Buy American 🇺🇲 Buy Winner 🏆 Buy Tesla! Best products & services like Tesla & Starlink simply WIN 🌎 Way too many people don't know buying a Tesla is cheaper than buying a Corolla! & Don't forget to use a Tesla referral code when ordering your Tesla! Tesla expanded its referral program to & certain parts of Europe & Asia! USA & Canada: 3 months of FSD Germany: €250 France: €500 Netherlands: €500 Norway: 11,500kr UK: £500 Australia 🦘: $350AUD New Zealand: $400 Italia: €500 Switzerland: 250 CHF Sweden: 11400 SEK South Korea 🇰🇷: 165,000 Won Japan: 60,000 yen Singapore: S$300 Malaysia: RM 4200 Thailand: B8,500 Hong Kong: HK$1,900 Macau: MOP$1,958 Philippines: ₱13,000 Taiwan: NT$8,000 India: ₹22,000 Here is mine if you need one: DMs are open for questions. FSD is mind-blowing!!

AI/ML API's profile picture
AI/ML API6 days ago

Grok is great. We’ve added it to our API - try it

Jarret's profile picture
Jarret6 days ago

only problem the Opus 5 output's are better in everyday..... you get what you pay for

AI/ML API's profile picture
AI/ML API6 days ago

That’s why we put both and many others into our API - try it

ShadowAguy's profile picture
ShadowAguy6 days ago

Did you verify the generated 3.js scenes actually run without errors, or just that they look correct in a screenshot?

AI/ML API's profile picture
AI/ML API6 days ago

Ran perfectly from the 1st shot

ShadowAguy's profile picture
ShadowAguy6 days ago

Ran perfectly from the first shot, that is the detail that matters. What did you run, was it a coding task or something else entirely

Related Videos

An Anthropic researcher sat down next to me at a hackathon last week. Claude Opus 4.7 was running 4 agents on my laptop. Live. No manual input. She looked at the terminal and said: "What is this?" I showed her. 4 agents. 678 trades. 81% win rate. $16,200 last 30 days. She worked on the evals team. She'd never seen Claude pointed at 88 million on-chain trades. The setup is 3 public repos. All free. -> 88 million Polymarket trades. Every wallet. Every entry. Every exit. Every resolution. -> the framework that bridges Claude Opus 4.7 directly to live markets. Order placement, position tracking, exit timing. -> real-time WebSocket order book. Depth on both sides. No polling, no lag. Four agents. One loop. Agent 1 identifies which wallets win consistently across 88 million trades. Agent 2 reverse-engineers their entry timing. Agent 3 monitors order book volume spikes. Agent 4 sizes positions using Kelly. No overbet. Drawdown capped at 1.4% over 678 trades. 85% of windows get killed. No trade. The bot only enters when 3 signals align: -> Elite wallet consensus pointing the same direction. -> Price divergence with Binance and Coinbase both agreeing. -> Order book imbalance confirming the bias. Single-source price data was 57% accurate. All three together: 81%. Exit before resolution. Always. Losers hold to 0 or $1. The agents copy their exits. The agents don't gamble on that. My stack: Claude Opus 4.7 at $19/mo, VPS Hetzner at $4.99/mo, Everything else free. Total stats: $23.99/month. 30 days: 678 trades, 81% win rate, net +$16,200, max drawdown -1.2%, avg hold 4h 12m. She asked if Anthropic could test this internally. "We run Claude on benchmarks and evals. Nobody pointed it at a live market dataset with 88 million rows." Claude Opus 4.7 didn't need a system prompt. It read the wallet index, understood the signal structure, and wrote the combiner logic in one pass. The people who built the model hadn't thought to point it at this data. I had. Copy the live trades: -> all 4 agents run 24/7. The window is open right now. Save this, follow me and comment OPUS. I will send the guide to you.

slash1s

46,379 views • 5 months ago

ox alpha vs deepseek v4 flash vision vs grok 4.6 vs gemini 3.7 flash vs – on photo-to-3d four vision models got one photograph each and had to rebuild the place inside it as a Three.js scene. twelve scenes, twelve first-try runs, zero console errors the setup: one reference photo per scene, sent as an image on OpenRouter. the prompt never says what is in the picture – no "motel", no "bar", no "gas station". the model has to read the photo and rebuild it: layout, materials, hour of the day, and whatever is around the corner that the frame does not show tasks – three photographs of early-2000s america: 1. a motel at night, neon pylon lit, snow on the ground 2. an old new york tavern interior, tin ceiling, tiled floor 3. an abandoned service station in the california desert, midday sun each scene ships as one self-contained html file, procedural geometry and canvas textures only, no downloads. three timed camera shots, and shot 1 has to reproduce the framing of the reference photo models: xAI grok 4.6, Google DeepMind gemini 3.7 flash, DeepSeek deepseek v4 flash vision exp, and ox alpha – a stealth model on openrouter, free, no lab attached to it yet results: - wall clock, three scenes #1 gemini 3.7 flash – 11m 12s #2 deepseek v4 flash – 15m 20s #3 grok 4.6 – 28m 11s #4 ox alpha – 38m 54s - output tokens #1 gemini 3.7 flash – 77,396 #2 ox alpha – 87,613 #3 grok 4.6 – 105,687 #4 deepseek v4 flash – 127,884 - lines of code shipped #1 ox alpha – 2,090 #2 deepseek v4 flash – 2,291 #3 grok 4.6 – 3,529 #4 gemini 3.7 flash – 3,989 - total price #1 ox alpha – $0.000 #2 deepseek v4 flash – $0.091 #3 gemini 3.7 flash – $0.136 #4 grok 4.6 – $0.697 observations: • grok is 7.7x the price of deepseek. it is the only model that read the light – low sun, real shadows on the station, a cold night on the motel • gemini is the fastest and the least deliberate. 17,158 reasoning tokens against deepseek's 99,172, and it still shipped the most code – 3,989 lines • deepseek thought hardest and rendered plainest. 99,172 reasoning tokens, 5.8x gemini's, spent on layout rather than on light. its motel is the second best in the set for $0.030 • ox alpha is free and reads a photo as well as anything here – it lifted "family units / kitchenettes" off the pylon and redrew it in canvas conclusion: twelve scenes, four models, zero fixes, and the whole run cost $0.924! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

24,470 views • 1 month ago

fable 5.1 vs fable 5 vs opus 5 – three lord of the rings landmarks, built in 3d from one image the setup: one reference image per scene, one html file per build, everything procedural – no meshes, no textures, no image files, nothing past Three.js from a cdn. each model reads the picture, writes its own prompt from it, then builds to that prompt in the same turn. three named camera shots per scene on keys 1/2/3, so it can be screen-recorded. run through OpenRouter tasks: 1. bag end – hobbiton from two frames, outside and in. the round green door has to open onto the room you are standing in 2. barad-dûr – the tower and orodruin from one film still. the eye has to move and track the camera, the volcano erupts on a cycle, the clouds never stop 3. rivendell – jerry vanderstelt's painting. sun shafts that shimmer, water that falls without a break, trees that sway on a gust models: Anthropic fable 5.1, fable 5, opus 5 total cost, three builds #1 fable 5 – $14.97 #2 opus 5 – $18.53 #3 fable 5.1 – $22.38 wall clock, three builds #1 fable 5 – 38m #2 fable 5.1 – 92m #3 opus 5 – 122m output tokens #1 fable 5 – 298,592 #2 fable 5.1 – 439,435 #3 opus 5 – 724,418 lines of code shipped #1 fable 5 – 2,885 #2 fable 5.1 – 4,021 #3 opus 5 – 5,161 biggest single build, lines #1 opus 5, bag end – 2,410 #2 fable 5.1, barad-dûr – 1,375 #3 fable 5, bag end – 1,319 observations: • fable 5.1 is the only model that furnished the bag end interior – a live fire, panelling, books on the floor, leaded diamond windows, against fable 5's flat color and opus's dark tunnel. the round door outside opens onto that room, the hard part of the brief • what it costs is thinking room. the 128k output ceiling is a thinking budget in disguise: fable 5.1 burned 102,116 of it on reasoning and hit the wall mid-file. opus spent 109,241 and hit the same wall. fable 5 spent 61,240 and finished bag end in one call – the only one that did • fable 5.1's first pass is not the finished thing. its barad-dûr came back with three defects you only catch by looking at it – nothing a read of the code would have flagged • it is the best of the three at being corrected. handed a plain list of what was wrong, it returned 32 targeted patches over two rounds, every one applied first try, and it worked out one of the causes itself instead of guessing at constants conclusion: nine scenes, 12,067 lines and 1.46m output tokens for $55.88 all in – and the cheapest model was also the fastest, by 3.2x! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

18,509 views • 25 days ago