Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

CODEX SKILL TO BRUTALLY TEST ANY STARTUP IDEA! Most startup ideas sound good. This Codex skill tells you why they probably won’t work. Just give Codex your idea and it pressure-tests it for you -> finds the core assumption -> exposes fatal flaws -> checks if the problem is...

513,769 Aufrufe • vor 4 Monaten •via X (Twitter)

74 Kommentare

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

this is the repo👇

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

matbe

Profilbild von Nishan
Nishanvor 4 Monaten

Bookmarked it. This is actually a good way to stop wating time on something which has low probabilities of working out. What i already do is use my manual frameworks to filter out ideas. This would be an interesting addition to that.

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

@nishancodes Thank you, Nishan. Yes, I saw it exactly the same way, which is why I created it. If you use it, let me know how it goes!

Profilbild von Vzdohi
Vzdohivor 4 Monaten

@nishancodes Are u a founder yourself?

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

@nishancodes yes

Profilbild von fed
fedvor 4 Monaten

the problem is that codex (or any ai) will most likely make up the “fatal flaws” just because skill says it

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

nahhh

Profilbild von Zach Roseman
Zach Rosemanvor 4 Monaten

Best way to get other people to give you their best ideas 🤣

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

no?

Profilbild von Zach Roseman
Zach Rosemanvor 4 Monaten

ha just kidding - but that would be an epic move

Profilbild von Andi
Andivor 4 Monaten

maybe time to build a codex skill marketplace :)

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

why not..

Profilbild von funkycol 🇺🇸
funkycol 🇺🇸vor 4 Monaten

Is there a back door that sends you all the good ideas ?

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

nope

Profilbild von funkycol 🇺🇸
funkycol 🇺🇸vor 4 Monaten

That’s good. I’ll check it out.

Profilbild von Far
Farvor 4 Monaten

installing that npx command just to get roasted by an open source script finally a tool that admits my startup idea is garbage before i

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

AHHAHAHAHA if u tty it let me know what you think!

Profilbild von Jason Shuman
Jason Shumanvor 4 Monaten

Using it to come up with ideas is the real hack

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

reallll

Profilbild von Jason Shuman
Jason Shumanvor 4 Monaten

Very

Profilbild von Antoine
Antoinevor 4 Monaten

I love these pressure tests. So much more useful than the classic llm bias « yes sir you’re totally right »

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

I was tired of him always saying I was right, so why not strain my mind for better ideas?

Profilbild von Mo Dawod
Mo Dawodvor 4 Monaten

I can confirm that it works with claude code as well :)

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

really good!

Profilbild von Mo Dawod
Mo Dawodvor 4 Monaten

Honest note on the pressure-test skill: it's designed to find the narrowest defensible wedge for pre-seed. That's useful. But it doesn't tell you how to keep the bigger vision while pitching the wedge - that's a structural blind spot of the skill, not a fact about your company.

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

Okay, I'll try to improve. Thanks!

Profilbild von Mo Dawod
Mo Dawodvor 4 Monaten

Thank you! this is a great skill - lots of amazing potential that will emerge from here

Profilbild von Eric Forgy
Eric Forgyvor 4 Monaten

I love asking our Overlords to roast us. It burns, but in a good way 😅

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

AHAHHAHAH

Profilbild von TobiasJames.ΞTH ᵍᵐ
TobiasJames.ΞTH ᵍᵐvor 4 Monaten

This is really good. I got an 8/10. Codex: "Why not higher yet? Because this is still doctrine unless it becomes product..." I wont list the what Codex bulleted for now 😊

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

Aoh nice

Profilbild von Arman Babakhani
Arman Babakhanivor 4 Monaten

Interesting. But it seems unlikely that general prompting will give the LLM more info to do proper analysis. It might tell you things that were already available out there but in the new ideas era, a lot of startups are proposing ideas that haven’t been tested before. So, interesting to see how often LLMs will give good analysis for a brand new idea

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

real

Profilbild von Emanuele Di Pietro
Emanuele Di Pietrovor 4 Monaten

cooked

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

thanks broda

Profilbild von Emanuele Di Pietro
Emanuele Di Pietrovor 4 Monaten

strunz rispunnm ngopp a whatsapp

Profilbild von Ivan Andrescov
Ivan Andrescovvor 4 Monaten

That's the Mom Test applied to an LLM’s interpretation of founders’ fantasies. A triple layer of epistemic noise. It generates highly confident prescriptive outputs with zero structural verification. This is not a pressure test of the idea — it’s a pressure test of its narrative.

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

ok!

Profilbild von Ivan Andrescov
Ivan Andrescovvor 4 Monaten

I would give more details of my analysis, but sorry—lately many people have been releasing products based on my ideas and passing them off as their own. GPT often presents wishful thinking as fact.

Profilbild von KP 🇺🇸🇦🇺
KP 🇺🇸🇦🇺vor 4 Monaten

Bro this is just their way to scan for new feature and product ideas, you guys don’t get it yet

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

real

Profilbild von Sebastián
Sebastiánvor 4 Monaten

Awesome! This is great

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

thabks bro

Profilbild von Olivia Bennett
Olivia Bennettvor 4 Monaten

This is actually very useful - saves time by showing what’s wrong before you start. Every startup idea should go through this kind of real test.

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

💪🏻

Profilbild von J A Z I I
J A Z I Ivor 4 Monaten

you da life saver man

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

AHHAAH

Profilbild von AR
ARvor 4 Monaten

Does it work for Claude?

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

maybe idk

Profilbild von Anshul Soni
Anshul Sonivor 4 Monaten

gotta give this a try

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

let me know what u thinks looks so cool

Profilbild von Anshul Soni
Anshul Sonivor 4 Monaten

yes

Profilbild von Lumus 👽
Lumus 👽vor 4 Monaten

half of startup twitter needs this immediately

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

real

Profilbild von CODIFY
CODIFYvor 4 Monaten

Thanks for sharing this!

Profilbild von Abdulmuiz Adeyemo
Abdulmuiz Adeyemovor 4 Monaten

Bro Just killed my startup idea with a codex skill 😂

Profilbild von 0xmaddy | Tech Adrenaline
0xmaddy | Tech Adrenalinevor 4 Monaten

Codex as a devil's advocate is underrated. Best founders I've seen already know their fatal flaw—this just forces honesty faster.

Profilbild von Alex Dochioiu
Alex Dochioiuvor 4 Monaten

Too much obsession on the idea and too little on execution. Look at literally all the AI models out there (Claude, Kimi, Deepseek, etc). None of them started with "I have this crazy idea". They all started with "I can do this better/cheaper than OpenAI".

Profilbild von Mild Mannered Maniac
Mild Mannered Maniacvor 4 Monaten

I'll take that, thank you very much!

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

let me know what u think?

Profilbild von Mild Mannered Maniac
Mild Mannered Maniacvor 4 Monaten

Ok yeah got a couple big angles here in Houston I’ll pretend I’ve finally learned my lesson (I have not)

Profilbild von Saad Ali
Saad Alivor 4 Monaten

Wow

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

do u like it?

Profilbild von Pato
Patovor 4 Monaten

codex finding the fatal flaw is useful but founders usually already know the flaw and hope no one notices. the real test is whether you can explain why it won't kill you

Profilbild von Jack Su
Jack Suvor 4 Monaten

Thanks for the sharing. Will help?

Profilbild von Sebastian Buzdugan
Sebastian Buzduganvor 4 Monaten

cool demo but without real market data it just hallucinates confidence, not insight

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

mmh...I know but in the end it's just a basic assessment

Profilbild von Dan
Danvor 4 Monaten

This is sweet. Building something similar to this that goes beyond the skill, would be interested if you had any thoughts on where this still falls short? An issue I find is that everything is infinitely critique-able and the model will find flaws that exist only in prose, not reality, after enough loops.

Profilbild von Supersocks
Supersocksvor 4 Monaten

@llmgram

Profilbild von Kwisatz Sazerac
Kwisatz Sazeracvor 4 Monaten

Does this work in a retroactive way, in terms of projects already in flight not just in ideation?

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

nope

Profilbild von AlMeqdad Yaseen
AlMeqdad Yaseenvor 4 Monaten

How about another codex skill for investing, generating leads, etc., so you fully found a startup and codex would be the co-founder.

Profilbild von Kappaemme
Kappaemmevor 4 Monaten

👀👀

Ähnliche Videos

EVERYTHING YOU NEED TO KNOW ABOUT CHATGPT'S "LOVABLE KILLER" CODEX SITES (in 25 mins): TLDR; the coolest part is that apps you build can update themselves autonomously 1. Codex Sites is not Replit or Lovable or Bolt. Those are great for one-prompting a full app. Codex Sites is for building apps that the agent keeps improving without you touching them. 2. Your personal website can update its own stats. Your internal dashboard can refresh its own data. Your product can add features while you sleep. The app is alive. 3. Start by invoking at-sites. Use realistic sample data. Always say "save for review, do not deploy." This unlocks building a real product, not a homepage. 4. Add persistent storage so the app remembers everything between visits. Without this it resets every time. Ask Codex to show you the data model before it builds. 5. Create safe actions. These are the specific things the agent is allowed to do to your app: add data, update cards, move things, score things. You define the boundaries. The agent operates within them. 6. Build skills so any future Codex chat knows how to interact with your app. The skill is basically a manual for the agent. Without it, every new chat starts from zero. 7. Save gate like a video game. Codex doesn't auto-save. Create checkpoints before you deploy so you can roll back if something breaks. 8. Close the autonomous loop. This is the magic. Once memory, safe actions, and skills are set up, the agent can update your app from any chat, any context, without you switching tabs. 9. Use the plugins most people are sleeping on. Figma, Canva, HeyGen for avatar videos, Game Studio for interactive experiences, FAL for image generation, Hugging Face for open source models. Worth adding a few. 10. The big picture: we went from building apps to raising apps. You set up the structure, the guardrails, and the skills. The agent does the rest. That's autonomous product building and it's here right now. Tbh, Codex sites isn't perfect. Still a lot to be desired like domains, db, authentication etc. But it's a glimpse into this idea that apps can be updated/improved upon automonously. And Codex Sites is REALLY good if you live in Codex everyday. Which more and more of are. And that's really cool. Will be interesting to see how Lovable, Bolt, Replit etc react to this. full tutorial on The Startup Ideas Podcast (SIP) 🧃 where you get your pods watch share with a friend i'm rooting for you What do you think of Codex and Codex sites?

GREG ISENBERG

69,219 Aufrufe • vor 3 Monaten

Three skills I use every day in Claude Code and Codex to solve my hardest problems: 1️⃣ /agent-watchdog When I have one agent like Codex working on a task and I don't fully trust it's going to do everything right, I'll open up another one like Claude Code and tell it to watchdog the Codex thread. You can copy the Codex deep link into Claude Code and it'll look at the prompt you sent, watch the Codex thread until it's done, then compare the Codex solution to how it was planning to solve it and automatically fix anything that Codex missed. It can also test the work of the other agent end-to-end. Similar to the idea of OpenRouter's new Fusion feature, I've definitely found that two models thinking through a problem and checking each other's work can be wildly more impactful than just one. 2️⃣ /plan-arbiter Similar ideas as /agent-watchdog - but with this one you have both make plans, compare plans, negotiate the differences, and make a final plan to execute. I find Claude Code is better at writing plans, but Codex is faster and cheaper to execute on them. Then I usually have Claude Code watchdog the Codex work and fix anything that was missed. 3️⃣ /read-the-damn-docs One thing that drives me crazy with coding agents is they're so reluctant to look up docs. They'll just guess and guess and guess at the right API surface for things, or the right solution to an integration of two things. Once I explicitly tell it to look up the docs, it says "Oh, I see the answer," and it fixes the problem. So I made the /read-the-damn-docs skill. Add it and your agents will know when and how to do efficient web searches to look up docs for the types of problems you really should look up docs for. All of these are totally open source over on my GitHub. If you try them, let me know your feedback. Will link to them below:

Steve (Builder.io)

43,089 Aufrufe • vor 3 Monaten