Загрузка видео...

Не удалось загрузить видео

На главную

🚨 codex, claude and grok users!! I have FINALLY tested this for YOU!!! the $20 subscriptions! Codex VS Claude VS Grok (Watch video for results) Claude Pro: - Opus 5.5 (max) - 10% Weekly limit = 115 Million tokens - 2 hr work - 5/10 Quality Codex Plus: -...

81,439 просмотров • 4 дней назад •via X (Twitter)

Комментарии: 96

Фото профиля shownotover
shownotover4 дней назад

Here are the usage bars AFTER the test

Фото профиля Axel 🇬🇧 Dev | Software Engineer | AI Claude GPT
Axel 🇬🇧 Dev | Software Engineer | AI Claude GPT4 дней назад

Thanks for sharing your experience. I'm a ChatGPT plus subscriber but I think I'm about to buy a Claude Pro subscription, I have Opus 5.5 FOMO.

Фото профиля shownotover
shownotover4 дней назад

Hahaha... Definitly try Opus 5.5 ... its a good model (especially better than GPT 5.6 Sol) Let me know your experience after that.

Фото профиля Rajantha
Rajantha4 дней назад

Wow....Claude is wayy better than I thought. Opus 5.5 is pretty good at high and token efficient. Should get a lot of mileage that way.

Фото профиля shownotover
shownotover4 дней назад

Its not token efficient... Opus 5.5 is token hungry... but the limits are generous to cover up for that... I think for 1 Billion tokens per month for Opus 5.5 ... you'll be a able to do A LOT of work ngl

Фото профиля Thomas Klein
Thomas Klein4 дней назад

@RajanthaR I’m already at 2 billion in the past 5 days with deep seek Am I missing something? You can get anything done with 1 billion tokens?

Фото профиля Rajantha
Rajantha3 дней назад

Deepseek V4.1 Flash is really fast and it also uses so many tokens. In token count, my top model is also DSV4.1 followed by Luna followed by Grok. DSV4.1 is on par with Fable 5 and reliable. Doesn't come close to Fable 5.1 and more reliable than Opus 5. Opus 5.5 beats Fable 5.1 and it beats Astra in terms of consistency and reliability. Astra sometimes so really dumb things. Astra can also run around in circles. Only models that beat DeepSeek V4.1 Flash that is reliable is Fable 5.1 and Opus 5.5 and with the cost difference, they are definitely not worth it at API cost. Claude $20 Gives you $4000 ish of API spend. It's the most cost effective plan when you compare with Claude max plans. You can get 200+ TPS from Deepseek, 300+ TPS with CoreWeave and 400+ from Makora at 2x price. Can't go wrong with that.

Фото профиля Thomas Klein
Thomas Klein3 дней назад

@shownotover It’s wild how reliable DSV4.1 is it’s literally god tier imo best model since 4.6 when it came out. Now it’s the best imo aside from fable and astra and supposedly 5.5 but idk if I even need 5.5 tbh

Фото профиля Rajantha
Rajantha3 дней назад

Yeah. It's wild how Deepseek Flash leaped from preview to V0731 to V4.1 in a span of few months. Unless you're doing something new, I don't think you need to switch. The 20x increase in the max plan is only for the 5hr window btw. In a lawsuit it says that 20x weekly usage is 1.5x of the 10x plan. And 10x plan weekly is like 6.5x of $20 plan. Who knows how next Deepseek will be. Might catch up to Opus 5.5. You can just try the $20 plan and see how far the model can be pushed. Plus you get $100 of Cloud usage. That will probably burn at API rates so might be better spent with what the next Sonnet model will be.

Фото профиля Silas He
Silas He4 дней назад

So what are you saying—that Anthropic is somehow giving Max 20x users the equivalent of $80,000 in API usage, or that Claude Max subscribers are all idiots paying more money for usage that doesn’t even scale with the price?

Фото профиля shownotover
shownotover4 дней назад

No no 🙂‍↔️😭 I got this much usage somehow on a $20 plan 20x max plan is only 4-5x more usage than Pro plan anyways... Mostly you get $1100-1400 worth of usage on the $20 plan monthly Idk how I was able to juice out 115 million tokens out of Opus 5.5

Фото профиля Silas He
Silas He4 дней назад

Opus 5.5 is great—I’ve used it in Cursor—but I’m never subscribing to Claude again. I’ll never forget how a single conversation in Chat, not Code, burned through 36% of my five-hour usage limit when I was using Opus 4.7. I couldn’t even use the chat comfortably. Even though GPT-6 Sol was a disappointing release, I still really like its Chat mode. Even back when I was on the Plus plan, I never once ran into a usage limit there.

Фото профиля shownotover
shownotover4 дней назад

😂🤣 It was the same time I cancelled my Claude sub Opus 4.8 was launched and it ate my 5 hr limit in 20 mins... I canceled and moved to codex But now, idk how and why... its somehow fixed. But again, Good luck with Cursor... I have never used it... maybe in the future

Фото профиля Antho
Antho4 дней назад

Per your results Sol produced better quality, used less weekly allowance, took less time, but you still recommend Claude Code because of the token numbers? Looks like your recommendation focuses on the only number that isn't meaningful here

Фото профиля shownotover
shownotover4 дней назад

Why do I recommend Claude? Because: - Codex gives you only $300 API usage VS $1100 monthly - 115 Million tokens vs 29 Million is a difference The thing is, If I don't one-shot project and use with care, I can use Claude a LOT more than Codex But again, if you like Codex... please choose that... you will always curse me if anything goes wrong on Claude 🥹

Фото профиля Antho
Antho4 дней назад

I guess it depends where you consider the value lives, in raw API prices vs what it actually produced. It is hard (probably impossible) to tell whether the figures you observed translate across many use cases though. Thanks for your tests and reports

Фото профиля Furkan Ceyran
Furkan Ceyran4 дней назад

I got Claude Pro just to try it out, and I can comfortably run 4 mid-sized tasks within a single 5-hour window. My dad is on GPT Plus, while I’m on the 20x plan. And I can honestly say, even with him using Luna, he burns through his limits so easily that I genuinely feel bad for him.

Фото профиля shownotover
shownotover4 дней назад

🙂 Bro... you are on a 20x plan... what do you expect? That's like a $360 vs $8000 battle If you were both on $20 plans... I could have said that I feel bad for him too (even though I am still feeling bad for him) Get your dad a Claude Pro ... will work enough for him (And one opencode Go sub for Deepseek v4.1)

Фото профиля Furkan Ceyran
Furkan Ceyran4 дней назад

Nah, I wasn’t comparing my 20x plan to his Plus 😭 I meant that I can comfortably run around 4 mid tasks on Claude Pro’s $20 plan, while my dad burns through his GPT Plus limits surprisingly fast even though he mostly uses Luna. The 20x part was just context about what I personally use. And yeah, you’re right. I’m definitely considering getting him Claude Pro now 🙂

Фото профиля shownotover
shownotover4 дней назад

Oh... my bad 🥲 Good luck with coding bro... Also use these plugins... improve quality, token efficiency and saves usage limits: - Graphify - Ponytail - Obsidian - Codegraph Good luck!

Фото профиля Anikesh Paul
Anikesh Paul3 дней назад

@furkan_ceyran Hey, can you explain what the plugins do other than ponytail? Thanks.

Фото профиля shownotover
shownotover3 дней назад

@furkan_ceyran They save context, take notes, build memory and ofc save tokens instead of rereading everything

Фото профиля Compilethings
Compilethings4 дней назад

I told you so. Then you went and tested it, Great work, wonderful info for everyone!

Фото профиля shownotover
shownotover4 дней назад

Your welcome bro... anything else you want... I can cook it up too (for now) haha

Фото профиля Compilethings
Compilethings4 дней назад

You must not have believed the numbers I showed you so you went and proved it. I like that in people, keep it up!

Фото профиля Donald Epstein
Donald Epstein4 дней назад

Your conclusion to go with Claude is stupid. Your experiment proves that Codex is the superior choice: its faster, its cheaper, it uses less of your limit, its higher quality. Even if you were paying API usage, Codex is better. You are a Claude cuck.

Фото профиля shownotover
shownotover4 дней назад

Bro... its your choice... I gave my opinion... Choice is yours... sure go with Codex. Good luck

Фото профиля ミ Α Ω 彡
ミ Α Ω 彡4 дней назад

Great job. Opus 5.5 is doing great for me. I'm using chatgpt to help organize it though but opus is the coder

Фото профиля shownotover
shownotover4 дней назад

Cool bro... Good luck coding

Фото профиля Marlon Chocce
Marlon Chocce3 дней назад

No entiendo, me dices que gpt me entrega mejor calidad y gasta menos mi suscripción, pero me recomiendas Claude? Es absurdo, tu test solo demuestra lo ineficiente que es Claude. Has analizado 3 variables (calidad, tiempo y tokens) y usaste la peor para recomendar.

Фото профиля shownotover
shownotover3 дней назад

1. Claude usage is a LOT, you can revise your work 5 times rather than one-shotting on codex 2. Opus 5.5 at max overthinks and messed up at the end... Med Opus 5.5 >>> GPT-6 Sol 3. Because of stupid overthinking, it took so much time testing... But again, its your choice... You can still choose codex.

Фото профиля Blazerdan
Blazerdan4 дней назад

are you sure you used Opus 5.5? or why does it says "automode on" for claude

Фото профиля shownotover
shownotover4 дней назад

I definitly used Opus 5.5 because after all the tests... I asked Mimo v2.6 to make a report which can be seen at 11th sec of the video... it didn't give me "Oh and other models were used too" ... it simply said Opus 5.5

Фото профиля Homelander
Homelander4 дней назад

Thanks man. Let's hope Claude does not nerf it's limits.

Фото профиля shownotover
shownotover4 дней назад

I really hope too... But it will happen in future and I'll update all you guys

Фото профиля Yeltsin Valero | Valero Dev
Yeltsin Valero | Valero Dev3 дней назад

What jumps out: Claude used 115M tokens for a 5/10, Codex 29M for a 7/10. So the "$4,000 of API value" measures how many tokens Opus burns, not how much work gets done. How much of those 115M were cache reads? That changes the math a lot.

Фото профиля shownotover
shownotover3 дней назад

99.8% was cache read. Total cost was $90 Tbh, still feels the best value

Фото профиля H.Ali
H.Ali4 дней назад

good analysis bro. Sol did the job pretty well. right?

Фото профиля shownotover
shownotover4 дней назад

I mean, it all depends on how you look at it... For me, even though it messed up... usage is too much for me to overlook... Codex did a good job 7/10 but the limits are nerfed so... your choice

Фото профиля JustEd
JustEd3 дней назад

What the result actually shows is that COT-6 Sol and Grok xhigh are better than Opus 5.5 and that you don't need the max efford to get a similar resulte than 5.5 max. They did a better job quality and faster.

Фото профиля shownotover
shownotover3 дней назад

What about the 5x tokens you got on Claude? You can do 5 times over if you DO NOT one-shot things... I am stupid, but you aren't and you'll get a lot more usage ngl

Фото профиля JustEd
JustEd3 дней назад

It is simple. Open AI gives less tokens because their models use less and Anthropicss more cuz their model uses more. Usage apparently is proportional to token consumsion by every model.

Фото профиля Kaio Ferreira | Game Developer
Kaio Ferreira | Game Developer4 дней назад

@thsottiaux Where are those “great limits”? Or do you still see Plus users as casual users?

Фото профиля shownotover
shownotover4 дней назад

@thsottiaux He sees all the plus users as "targets" and somehow make them get Pro 5x plan 🫠🫠

Фото профиля Kaio Ferreira | Game Developer
Kaio Ferreira | Game Developer4 дней назад

@thsottiaux I already paid for the 5x Pro plan. It actually lasted a decent amount of time, but still not even 3 days. Brutal, especially for me as a Brazilian. It’s extremely expensive here.

Фото профиля shownotover
shownotover4 дней назад

@thsottiaux Codex never worked for me in Sep (was working good till Aug) same with SuperGrok (I was able to use it for 3 days straight 24/7 in hermes agent but now can't even do 10 hrs) And... Claude... that's surprising to me

Фото профиля Sinn
Sinn4 дней назад

Since users in China cannot access Claude, which subscription do you recommend?

Фото профиля shownotover
shownotover4 дней назад

That'll need workaround... Buy Codex (I know its nerfed but best you got) and then connect with Hermes agent and use Deepseek API as "implementor" So Codex thinks and Deepseek does.

Фото профиля Sinn
Sinn4 дней назад

Thanks for the reply. If possible, I’d love for you to review Grok 4.7 in Cursor; I believe it offers a higher usage limit than SuperGrok.

Фото профиля shownotover
shownotover4 дней назад

Sure, I will but currently out of budget... I will test it next month... have to cancel SuperGrok and Codex

Фото профиля Ethan
Ethan3 дней назад

However, cladue is limited, so codex is still relatively friendly.

Фото профиля shownotover
shownotover3 дней назад

Go with Deepseek API $2 minimum... That'll give you at least 300 Million tokens... No need to pay for Codex $20 for Luna

Фото профиля Adarsh Subham
Adarsh Subham4 дней назад

Honestly speaking codex has always felt nerfed to me since I started using it idk after astra tho, for me claude code is the perfect use getting the max out of it rather than just getting something bland

Фото профиля shownotover
shownotover4 дней назад

I agree with you... Before Opus 4.8 which was 1.6 months ago... Codex was good... now its nerfed af

Фото профиля Adarsh Subham
Adarsh Subham4 дней назад

Glad I didnt follow the hype and didnt get chatgpt plus when Astra came out stuck to Opus 5 knew they are gonna release something fast to take over🤣

Фото профиля shownotover
shownotover4 дней назад

Claude has been on top for 2 years now... Total dominance

Фото профиля Adarsh Subham
Adarsh Subham4 дней назад

Exactly feel a bit behind as I actually started using claude code since July before that I was on codex for a bit and antigravity

Фото профиля Darren
Darren4 дней назад

wait that's not what I experieced at that time when grok build just launch , Elon cut that hard?

Фото профиля shownotover
shownotover4 дней назад

Yeah!!! Don't buy grok for now

Фото профиля Gio
Gio4 дней назад

How much did claude pay you?

Фото профиля shownotover
shownotover4 дней назад

I wish they had paid me ☹️

Фото профиля Divine 〽️achine
Divine 〽️achine3 дней назад

5/10 quality on opus 5.5?

Фото профиля shownotover
shownotover3 дней назад

yeah... check the video from the end

Фото профиля Allahoum zeyd
Allahoum zeyd3 дней назад

@thsottiaux is this right?

Фото профиля shownotover
shownotover3 дней назад

@thsottiaux He won't answer haha

Фото профиля Jules
Jules3 дней назад

Can you test again 10 bucks Opencode go?

Фото профиля shownotover
shownotover3 дней назад

Opencode Go is very transparent ngl... we don't need that do we?

Фото профиля Sahil Panhotra | Indie Builder
Sahil Panhotra | Indie Builder4 дней назад

Great bro 😅 even though limits are nerfed i think codex is still good coz it uses less tokens and has lower pricing so api pricing wise codex still wins

Фото профиля shownotover
shownotover4 дней назад

Maybe... I think Claude is giving a LOT of tokens here... But whatever you choose, it's best for you

Фото профиля hixonchan
hixonchan4 дней назад

首富的api 居然是最贵最垃圾的

Фото профиля shownotover
shownotover4 дней назад

What do you mean by that ? 😭

Фото профиля Pavel Chinyaev
Pavel Chinyaev4 дней назад

true, the only thing is Claude’s quality is 10/10

Фото профиля shownotover
shownotover4 дней назад

well, it wasn't able to 1-shot the project with the same prompt so... its kinda bad

Фото профиля David Zhang | Algoryn
David Zhang | Algoryn4 дней назад

Thanks for sharing,

Фото профиля shownotover
shownotover4 дней назад

No probs... Let me know if I can do it some other way if it helps

Фото профиля Xeno
Xeno3 дней назад

What about using only mobile app?

Фото профиля shownotover
shownotover3 дней назад

No... its on Linux I have recorded the screen

Фото профиля Nikita Terekhin
Nikita Terekhin4 дней назад

Качество клода 5/10, аххаха, вот клоун, а.

Фото профиля shownotover
shownotover4 дней назад

I am saying that the quality isn't as good as others but there are so many tokens and so much work that you can repeat the process one or two times and you will double the quality, tbh It's not like you can one-shot an entire project. 🥲

Фото профиля Crazy
Crazy3 дней назад

Does claude have 5 hour limits?

Фото профиля Thomas Klein
Thomas Klein4 дней назад

Thank you for this

Фото профиля shownotover
shownotover4 дней назад

No problem

Фото профиля Thomas Klein
Thomas Klein4 дней назад

Also, you should mention that the groc subscription also has grok bot so there’s a lot more usage than just grok build…

Фото профиля shownotover
shownotover4 дней назад

Grok bot washed the limits in 4 prompts... those seprate limits are a joke (SuperGrok $30 plan)

Фото профиля Thomas Klein
Thomas Klein4 дней назад

Oh on superfrok plus it’s not bad

Фото профиля shownotover
shownotover4 дней назад

Yeah but if the budget is $20-30 you can't do anything can you?

Фото профиля Thomas Klein
Thomas Klein4 дней назад

Nope lol also, I don’t know who is expecting in 2026 to pay $20 a month and expect to get anything done. If it’s truly for work and making you money, then $100 a month is nothing.

Фото профиля shownotover
shownotover4 дней назад

I agree with you bro. SuperGrok Plus is not something to buy right now because the limits are so much nerfed. If I had $100 then I would definitely get Claude Max.

Фото профиля Sunjin woo (Dreamer)
Sunjin woo (Dreamer)3 дней назад

Lol I felt this too, That's why I'm wondering why 5.6 SOL is taking more usage than I used to work with Claude. So rn Claude ain't expensive it's GPT

Фото профиля shownotover
shownotover3 дней назад

Exactly!

Фото профиля Lucas
Lucas4 дней назад

20刀的claude周订阅有10亿?

Фото профиля shownotover
shownotover4 дней назад

YES! 1 Billion tokens is what I got I have also conducted two recent tests in which I tested Claude Code and in each test I got at least 50 million tokens for 10-13% weekly usage... You can expect around 500 million to 900 million tokens per week

Фото профиля Lucas
Lucas4 дней назад

谢谢,非常有用的数据

Фото профиля AnimeVibeX
AnimeVibeX4 дней назад

Thanks, i will buy Claude Today

Фото профиля shownotover
shownotover4 дней назад

Good luck bro! Lemme know if you need anything... I have plenty of info

Фото профиля AnimeVibeX
AnimeVibeX3 дней назад

I'll do that (: I'll be working a lot on my Unity game with Opus 5.

Похожие видео

I tested Claude Code on a fresh account - 1,500 lines of HTML cost me 50% of my window. Full video and summary is here.. I just ran a recorded test on Claude Code with a fresh account (Pro, not Max - my main account was 20x Max) , and the result is honestly insane. The task was trivial: create 3 simple demo HTML pages, around 500 lines each. Roughly 1,500 lines of code total. Nothing massive. Nothing enterprise-grade. Nothing that should meaningfully stress a premium coding product. And yet Claude Code burned through 40% of my 5-hour window almost immediately. I ran the exact same test with Codex, and it consumed only 2%. Then it got even worse: after the session ended, I did absolutely nothing for 15 minutes, and Claude still ate another 10%. Total: 50% of the 5-hour window gone for a tiny HTML demo. My weekly usage had already started at 2% before I even really used it, and after this tiny test it jumped to 8%. Now let us be generous and assume this entire run used around 30k tokens total. If 30k tokens represents 10% of weekly usage, that implies around 300k tokens per week. That is roughly 1.2M-1.3M tokens per month, and even if you round up aggressively, you are still in the 1.5M token range. Using the Sonnet 4.6 pricing you list: $3 per 1M input tokens $15 per 1M output tokens How exactly is this supposed to make sense for a paid coding product? Because from the user side, this no longer looks like "premium usage protection." It looks like a quota system that is either wildly inefficient, badly broken, or being accounted in a way users are not being told about. And that is before I even get to my main account: my $200 Max plan now dies in a single day. Just a few months ago, similar or heavier usage would last me about a week. So no, I do not buy the "maybe you just used it more" excuse anymore. Something is clearly broken in Claude Code. Either token accounting is broken, context handling is broken, background consumption is broken, or all three. Alex Albert is this really the experience you want users to pay for? Just watch the video. I tried to be very transparent and clear for your team! I was fan of Claude but just disappointed! And if you want, send me the detailed token accounting for this session and let us inspect it together publicly. Because from where I am standing, this is no longer a small pricing annoyance. It looks like something seriously wrong is happening, and users deserve a real explanation.

Hayrettin Tüzel

26,854 просмотров • 6 месяцев назад

I just compared Claude Code vs Codex vs Cursor CLI The task was to build a Next.js app with Tailwind 4 and shadcn components to collect customer feedback and showcase it with a widget. I gave all three the same prompt and let them go for 30 minutes to see what they came up with. Claude Code with Opus 4.1 Even though I told it to set up the app in the existing project folder, it tried to create a directory for it. After I interrupted and told it not to do that, it built a demo form and landing page with no errors. I had to ask it to make the demo interactive so users could submit a testimonial and preview it. The landing page looked like AI and was pretty basic, but it worked and it was done in a fraction of the time of the others. Total tokens used: 33k Codex with GPT-5 At the end of the 30 minutes I just could not get Codex to produce a working app. It got stuck in a loop of not being able to set up Tailwind 4 and despite many, MANY, attempts, I ended up with a "failed to compile" error. Total tokens used: 102k Cursor Agent with GPT-5 This was the slowest agent by far and a couple of times I actually thought it got stuck in a loop and was close to Ctrl+C'ing to cancel it. The TUI is really nice though, especially how it shows diffs and it did eventually build a working app (after one or two slight errors that needed fixing) The demo was interactive and it had a very minimal design that looked bare but also a lot less like an "AI generated" app than the Opus 4.1 design. It also wasn't too chatty and just did what it needed to do! Code quality was on a par with Opus 4.1, but it did use 5.5x as many tokens to get there. Still cheaper than Opus on a direct comparison but not when you factor in a Claude Code Max subscription. Total tokens: 188k I'll be able to do a proper comparison and record some videos when I'm back from holiday but for now, Opus is still the more capable model out of the box and Claude Code is the more complete CLI product. It will be interesting to see how Cursor evolve their CLI though with commands and subagents because I think with GPT-5 they have a real shot at providing competition for Claude Code if they can optimise output to get similar quality with less tokens. Jump to 0:40 in the video to see the two apps. Which do you think is which? ;)

Ian Nuttall

195,173 просмотров • 1 год назад

Three skills I use every day in Claude Code and Codex to solve my hardest problems: 1️⃣ /agent-watchdog When I have one agent like Codex working on a task and I don't fully trust it's going to do everything right, I'll open up another one like Claude Code and tell it to watchdog the Codex thread. You can copy the Codex deep link into Claude Code and it'll look at the prompt you sent, watch the Codex thread until it's done, then compare the Codex solution to how it was planning to solve it and automatically fix anything that Codex missed. It can also test the work of the other agent end-to-end. Similar to the idea of OpenRouter's new Fusion feature, I've definitely found that two models thinking through a problem and checking each other's work can be wildly more impactful than just one. 2️⃣ /plan-arbiter Similar ideas as /agent-watchdog - but with this one you have both make plans, compare plans, negotiate the differences, and make a final plan to execute. I find Claude Code is better at writing plans, but Codex is faster and cheaper to execute on them. Then I usually have Claude Code watchdog the Codex work and fix anything that was missed. 3️⃣ /read-the-damn-docs One thing that drives me crazy with coding agents is they're so reluctant to look up docs. They'll just guess and guess and guess at the right API surface for things, or the right solution to an integration of two things. Once I explicitly tell it to look up the docs, it says "Oh, I see the answer," and it fixes the problem. So I made the /read-the-damn-docs skill. Add it and your agents will know when and how to do efficient web searches to look up docs for the types of problems you really should look up docs for. All of these are totally open source over on my GitHub. If you try them, let me know your feedback. Will link to them below:

Steve (Builder.io)

43,089 просмотров • 3 месяцев назад