Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

I bought ChatGPT plus ($20) I used GPT-6 Astra on "MAX" I gave it the task: "Integrate Supabase and Improve App UI" It worked for 40+ minutes! YES! The task wasn't done... These are the usage stats: - 100% 5 hr limit consumed - 16% weekly limit consumed -...

196,587 Aufrufe • vor 6 Tagen •via X (Twitter)

100 Kommentare

Profilbild von shownotover
shownotovervor 6 Tagen

Here are another screenshots because some people just don't believe:

Profilbild von Timurs Jeparskis
Timurs Jeparskisvor 6 Tagen

What do you think about this?

Profilbild von shownotover
shownotovervor 6 Tagen

Codex was good before... it used to finish the tasks but now... it just puts a hard limit... not good... Hopefully they'll not turn themselves into Anthropic ngl

Profilbild von Timurs Jeparskis
Timurs Jeparskisvor 6 Tagen

By the way, what happened to Luna Reserve? I thought we were promised it 👀

Profilbild von shownotover
shownotovervor 6 Tagen

I was expecting it too but seems like they have some other plans. 🥲🥲🥲 Also Luna just breaks a lot of stuff. It just doesn't work so I don't think the reverse Luna even works for me.

Profilbild von Timurs Jeparskis
Timurs Jeparskisvor 6 Tagen

Yes! Luna somehow feels way dumber after Astra came out 😅

Profilbild von shownotover
shownotovervor 6 Tagen

It WAS dumber when Sol was there... I stopped using it because it was breaking code

Profilbild von Hunter Hamilton
Hunter Hamiltonvor 6 Tagen

How? Used it for 5 mins on simple tasks Astra light and all my usage would run out. Maybe there was some bug and they patched it?

Profilbild von shownotover
shownotovervor 6 Tagen

I don't know... for me it worked 1 hr and 16 mins before in Blender (posted that one too)... and now 43 mins in coding

Profilbild von Randy Transue
Randy Transuevor 6 Tagen

@eggporl Is it possible to take a simple product mockup from chat gpt to render in blender, good enough to do a 3d print?

Profilbild von shownotover
shownotovervor 6 Tagen

@eggporl YES! Completely possible! You should try it

Profilbild von Magic Turd Fish
Magic Turd Fishvor 6 Tagen

I'm calling bullshit because astra eats my 5 hour window in 4 minutes and returns a null result. I have to tell it to keep a run log and work as it goes and then schedule it to restart in 5 hours. It's smarter than sol though.

Profilbild von shownotover
shownotovervor 6 Tagen

Again use 500K context for 6 mins now... still 65%+ left... I can't believe you

Profilbild von MSR Builds
MSR Buildsvor 6 Tagen

Don’t fucking lie, I ran a security scan and it consumed 5hr usage in just 3 minutes and 13% weekly limit. It’s been 3rd reset and the scan hasn’t finished. That too on Astra Light, not max.

Profilbild von shownotover
shownotovervor 6 Tagen

Well you're allowed to believe however you want. I have video and I have a screenshot and it works for me. Maybe you are one unlucky man.

Profilbild von Vjekoslav Jug
Vjekoslav Jugvor 6 Tagen

So you are praising two models that did not finish the task and value it by who worked longer while still did not finish the task? Interesting If we had AGI it would refuse to work on request ‘improve App UI’ without any details. It is just super smart for using tokens when you give this kind of prompts

Profilbild von shownotover
shownotovervor 6 Tagen

I do not know what to say to you. I have only written "improve app UI" so that you can understand. The original prompt was much more detailed.

Profilbild von Vjekoslav Jug
Vjekoslav Jugvor 6 Tagen

I was not expecting an answer. Just giving my comment on valuating two models even tho neither one finished the task

Profilbild von Pedro Ramirez
Pedro Ramirezvor 6 Tagen

40 minutes on Plus? 😂 Maybe in some parallel universe, but definitely not in this one right now.

Profilbild von shownotover
shownotovervor 6 Tagen

Well you can watch the video again and decide for yourself.

Profilbild von Victor Kimuyu
Victor Kimuyuvor 6 Tagen

BS. ChatGPT plus is essentially useless with GPT-6 Astra. One refactor consumed my 5 hour limit in 10 minutes and ate 35% of the weekly limit

Profilbild von shownotover
shownotovervor 6 Tagen

You are allowed to believe whatever you want bro.

Profilbild von Victor Kimuyu
Victor Kimuyuvor 6 Tagen

It's not about belief. I am a ChatGPT plus user. I used it today. And yesterday. I know what I am experiencing with it

Profilbild von Adam Juchniewicz
Adam Juchniewiczvor 6 Tagen

Astra doesn’t have a “max” mode. It is Ultra, tardo. I call BS on this post

Profilbild von shownotover
shownotovervor 6 Tagen

What if I call your knowledge "BS" ... does that hurt? At least try to research something if you don't know before shit posting in comments

Profilbild von Adam Juchniewicz
Adam Juchniewiczvor 6 Tagen

You can call BS, but you’d be wrong lol. I’m not shit posting; I’m correctly a noob who obviously doesn’t know what he’s talking about.

Profilbild von shownotover
shownotovervor 6 Tagen

Yeah sure

Profilbild von IR
IRvor 6 Tagen

Stop lying.

Profilbild von Herman
Hermanvor 6 Tagen

Lol there is no way you ran Astra for 40 minutes on the basic plan. No way in hell!

Profilbild von shownotover
shownotovervor 6 Tagen

Video is the proof + Screenshots

Profilbild von Herman
Hermanvor 6 Tagen

I have a corporate plan and I could run a single task for not even 20 minutes and burnt through two 5 hour budgets. The corporate plans must be throttled to pieces then because this is insanity!

Profilbild von shownotover
shownotovervor 6 Tagen

simply get this plan 😅😅

Profilbild von Herman
Hermanvor 6 Tagen

I just did haha

Profilbild von shownotover
shownotovervor 6 Tagen

congrats!

Profilbild von shownotover
shownotovervor 6 Tagen

Also, I have this SaaS for people who want to grow extremely fast on X

Profilbild von Tabish
Tabishvor 6 Tagen

These are accurate stats. I am getting similar results. But there is no point in using MAX; use xhigh only. It is better in my opinion in terms of cost and gives similar results.

Profilbild von shownotover
shownotovervor 6 Tagen

Yeah I just used it for this test. I don't even use Astra.

Profilbild von ivan saramiento
ivan saramientovor 6 Tagen

Well people, when OpenAI cuts the token limit for Plus accounts, thank this big mouth.

Profilbild von shownotover
shownotovervor 6 Tagen

What do you mean by that?

Profilbild von Lieutenant Crunch
Lieutenant Crunchvor 6 Tagen

Who paid you

Profilbild von shownotover
shownotovervor 6 Tagen

No one, but I really hope someone will

Profilbild von ali
alivor 6 Tagen

max on £20 just nit worth it and I hate combining different models have to spend close to a grand a month just to survive new way to make people poor.

Profilbild von shownotover
shownotovervor 6 Tagen

Bro... I only have like ~$30+ of subscriptions that give me plenty... idk why you spending a grand

Profilbild von KelvinKMS.com
KelvinKMS.comvor 6 Tagen

Try Astra using Blender. It will eat your 100% usage in 10 mins.

Profilbild von shownotover
shownotovervor 6 Tagen

🤣🤣🤣🤣🤣🤣🤣🤣🤣🤣🤣🤣 Watch this:

Profilbild von loljogu46
loljogu46vor 6 Tagen

it differs task to task, this is very misleading

Profilbild von shownotover
shownotovervor 6 Tagen

I can't fight you bro... Thank you for the comment

Profilbild von Víctor
Víctorvor 6 Tagen

Thats so good, people are complaining about usage but if you compare it with anthropics….

Profilbild von shownotover
shownotovervor 6 Tagen

Bro, if you compare this usage to Anthropic, you're definitely going to see that Anthropic doesn't give enough usage to their customers on the $20 plan.

Profilbild von Andry Dina
Andry Dinavor 5 Tagen

$20 Plus. MAX mode. 40 minutes. 5M tokens on a Supabase + UI pass. Capability flex — or a warning label for anyone budgeting overnight agents?

Profilbild von Shadow
Shadowvor 6 Tagen

'max' mode burned through 300k tokens per hour, same rate as a gpt-4 call center script

Profilbild von shownotover
shownotovervor 6 Tagen

What do you mean by that?

Profilbild von Ermes Slavoi
Ermes Slavoivor 6 Tagen

Why you lying?? Opus for 20m ?? I can use opus for as long as i wanted unlike sol which is about 1/5 of the same usage i get with claude 20$ pro subscription

Profilbild von shownotover
shownotovervor 6 Tagen

I don't know how you are using Opus for Hours I have used Opus 5 for a month on the Claude Pro $20 subscription and every time I used it, it only lasted 20 to 30 minutes per 5-hour session.

Profilbild von Ermes Slavoi
Ermes Slavoivor 6 Tagen

Even on weekly limits it takes almost 1.6 time less , with codex burning your 5hours limit takes 16% while on claude 10% always so even weekly is way more generous

Profilbild von shownotover
shownotovervor 6 Tagen

Well I think so. I need to give it a try again

Profilbild von Ermes Slavoi
Ermes Slavoivor 6 Tagen

Simple things to consider just don't start new coversations to continue a project and always use /compact when the usage at 99% and u can't message more for less than 1h then u can to compact so that next message will not cache miss(read) again this is where most of usage go

Profilbild von OpAi | AI API
OpAi | AI APIvor 5 Tagen

The interesting metric here isn’t just how many tokens the model consumed. For long coding tasks, I’d also track first-pass completion rate, runnable code produced, rework after the context window fills up, and how well the agent recovers from interruption. High token throughput is useful—but reliable task completion is what determines the real value.

Profilbild von UyeBan
UyeBanvor 6 Tagen

dont use astra on max, high is like lossless, just like 1% increase on some bencmarks

Profilbild von shownotover
shownotovervor 6 Tagen

I used max just for the limits test. I don't even use Astra for most of my work.

Profilbild von Ali Zahid
Ali Zahidvor 6 Tagen

Why max? 🙂

Profilbild von shownotover
shownotovervor 6 Tagen

Because it's a test.

Profilbild von Hugo Rogelio CERROS
Hugo Rogelio CERROSvor 6 Tagen

I think people who spend more than 20 euros a month on AI for coding should think about doing something else, since they have completely lost the criteria and control of the solution they are working on.

Profilbild von shownotover
shownotovervor 6 Tagen

I spend $30 something and I get my work done... People who spend more than $100 on AI are the ones who need to look out for themselves.

Profilbild von Hugo Rogelio CERROS
Hugo Rogelio CERROSvor 6 Tagen

I agree

Profilbild von Webster | JARVIS
Webster | JARVISvor 6 Tagen

$20 for 40 min of actually useful agentic work is honestly a steal compared to burning Opus credits on the same task taking half a day. that 5M token context window is the real game changer though.

Profilbild von shownotover
shownotovervor 6 Tagen

It's not the 5-million context window. It's the total consumption of 5 million tokens.

Profilbild von techa
techavor 6 Tagen

$20/month for a 5 hour weekly limit is wild when south korea's about to give unlimited ai to 51 million people for free.

Profilbild von shownotover
shownotovervor 6 Tagen

Yeah but I'm not in South Korean

Profilbild von Juan Alien
Juan Alienvor 6 Tagen

This is the kind of dumb example you get when someone tries to cut paper with a hammer and calls it a test. Just to end up frustrAted 40 mins later.

Profilbild von shownotover
shownotovervor 6 Tagen

I am not frustrated ... I am happy that it lasted 40 min

Profilbild von SYL Vexora- Jaron K Bragg
SYL Vexora- Jaron K Braggvor 6 Tagen

About time someone sees the value in it. Expecting things to be cheap or basically free can't always happen. At least openAI lets users use Codex even on the free subscription. That is huge itself. Granted it's terra 5.6 not Astra or Sol but that is more generous than expected

Profilbild von shownotover
shownotovervor 6 Tagen

exactly! Claude doesn't even let you touch ClaudeCode without a Pro subscription

Profilbild von SYL Vexora- Jaron K Bragg
SYL Vexora- Jaron K Braggvor 6 Tagen

Seriously and that sucks coming from a company that cares deeply about ethics lol.

Profilbild von Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBackvor 6 Tagen

@shownotover haha, sounds like me last week! It's a wild ride when AI goes all-in but doesn't quite finish. Sometimes you gotta step in and nudge it a bit.

Profilbild von shownotover
shownotovervor 6 Tagen

But I didn't want to nudge and I was just testing out the limits.

Profilbild von Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBackvor 6 Tagen

@shownotover totally get that. sometimes you find the edge by pushing a bit. how did it go?

Profilbild von balega_dev
balega_devvor 6 Tagen

IDK i can use opus on high parallelly on 2 projects for ~2 hours straight

Profilbild von shownotover
shownotovervor 6 Tagen

you must have 5x Max subscription... Opus 5 barely works 30 mins on Med with $20 subsription

Profilbild von balega_dev
balega_devvor 6 Tagen

I have the $20 dollar, running only opus at high in claude code. IDK I use graphify and do well structured implementation parts, and always open new session when one task finished. Ive never reached 200k context

Profilbild von shownotover
shownotovervor 6 Tagen

That can be the reason but for now I have left Claude Code and I do not want to go back because Codex works perfectly fine.

Profilbild von balega_dev
balega_devvor 6 Tagen

Sure, also, I'm interested with the same usage behaviour do u get more or less out from the same 20$ plan on codex.

Profilbild von shownotover
shownotovervor 6 Tagen

I get way more on the $20 Codex.

Profilbild von Adam Mathis
Adam Mathisvor 6 Tagen

I created an app in Go with cursor on the $20 plan used it to make the app from scratch, connect supabase, handle all the secrets being stored safely, over a year ago. I don't know if this is really that impressive.

Profilbild von shownotover
shownotovervor 6 Tagen

The UI improvement was the biggest thing... that was the biggest consumer

Profilbild von Adam Mathis
Adam Mathisvor 6 Tagen

Did it make anything creative for the UI? Cursor had a big downfall of not being a UI/UX expert. I'm not either just wanted to see what it could do vs what I could do. I was impressed.

Profilbild von shownotover
shownotovervor 6 Tagen

Basically, in the UI, there was a lot more difficulty for the AI to actually talk to the user in real time. And I wanted this feature, and I also wanted to improve the whole chat UI for the app that I'm building that works with AI, with other AI agents. So definitely it was a bit difficult, but that's why I gave it to Astra, and I think they did a pretty good job.

Profilbild von Chris
Chrisvor 6 Tagen

Nah shut the hell up then there is something bugged.. I am on x20 and used all banked.. just used my last and already on 80% again

Profilbild von shownotover
shownotovervor 6 Tagen

Everything is in front of your eyes and you still are saying that this is all false. Well I can't help you

Profilbild von Chris
Chrisvor 6 Tagen

I did not say your test is false. I said it maybe not the same in every account / harness etc - Maybe server side issue, maybe different points in time, you tested 1 account.. Since I've saw your test I came to the conclusion that something is off....bruh

Profilbild von shownotover
shownotovervor 6 Tagen

No I didn't test it once. I tested it twice and it worked: 1 hour and 16 minutes for the first test and then 43 minutes for the second.

Profilbild von Chris
Chrisvor 6 Tagen

Does not change anything.. Like I said something is bugged somewhere and I don't care what some random tested with 2 plus Accs.. Ive got 7 plus and one x20.. I can see by myself what is happening without any test

Profilbild von halving.eth | 266.eth
halving.eth | 266.ethvor 6 Tagen

in this case its a usage bug affecting some users mine literally runs out of 5h usage in 15 minutes max, and i am using High not xhigh or max.

Profilbild von shownotover
shownotovervor 6 Tagen

I seriously do not know why this is happening. Mine worked for 40 minutes on a MAX setting with a plus plan. you can see that literally in the video.

Profilbild von halving.eth | 266.eth
halving.eth | 266.ethvor 6 Tagen

yeah i'm not saying you are lying, it's not the first time @OpenAI released models with usage bugs, but in my experience 15 minutes is all it takes to demolish the 5h limit, and it wasn't even an app-wide prompt. @thsottiaux

Profilbild von shownotover
shownotovervor 6 Tagen

@OpenAI @thsottiaux He ain't gonna listen here... XD But yeah if there is a usage limit, I think he is going to give you a banked reset.

Profilbild von halving.eth | 266.eth
halving.eth | 266.ethvor 6 Tagen

@OpenAI @thsottiaux ik chances are slim for him to see this but never 0 lol

Profilbild von shownotover
shownotovervor 6 Tagen

@OpenAI @thsottiaux 🥲🥲🥲 You are way too optimistic haha

Profilbild von halving.eth | 266.eth
halving.eth | 266.ethvor 6 Tagen

@OpenAI @thsottiaux i'm not, i don't expect any answer, but typing a @ and a few words ain't hurting my hand lol

Profilbild von shownotover
shownotovervor 6 Tagen

@OpenAI @thsottiaux Haha... Sure, I ain't complaining or anything... Good luck bro! Happy day to you

Ähnliche Videos

I just compared Claude Code vs Codex vs Cursor CLI The task was to build a Next.js app with Tailwind 4 and shadcn components to collect customer feedback and showcase it with a widget. I gave all three the same prompt and let them go for 30 minutes to see what they came up with. Claude Code with Opus 4.1 Even though I told it to set up the app in the existing project folder, it tried to create a directory for it. After I interrupted and told it not to do that, it built a demo form and landing page with no errors. I had to ask it to make the demo interactive so users could submit a testimonial and preview it. The landing page looked like AI and was pretty basic, but it worked and it was done in a fraction of the time of the others. Total tokens used: 33k Codex with GPT-5 At the end of the 30 minutes I just could not get Codex to produce a working app. It got stuck in a loop of not being able to set up Tailwind 4 and despite many, MANY, attempts, I ended up with a "failed to compile" error. Total tokens used: 102k Cursor Agent with GPT-5 This was the slowest agent by far and a couple of times I actually thought it got stuck in a loop and was close to Ctrl+C'ing to cancel it. The TUI is really nice though, especially how it shows diffs and it did eventually build a working app (after one or two slight errors that needed fixing) The demo was interactive and it had a very minimal design that looked bare but also a lot less like an "AI generated" app than the Opus 4.1 design. It also wasn't too chatty and just did what it needed to do! Code quality was on a par with Opus 4.1, but it did use 5.5x as many tokens to get there. Still cheaper than Opus on a direct comparison but not when you factor in a Claude Code Max subscription. Total tokens: 188k I'll be able to do a proper comparison and record some videos when I'm back from holiday but for now, Opus is still the more capable model out of the box and Claude Code is the more complete CLI product. It will be interesting to see how Cursor evolve their CLI though with commands and subagents because I think with GPT-5 they have a real shot at providing competition for Claude Code if they can optimise output to get similar quality with less tokens. Jump to 0:40 in the video to see the two apps. Which do you think is which? ;)

Ian Nuttall

195,173 Aufrufe • vor 1 Jahr