Loading video...

Video Failed to Load

Go Home

Andrej Karpathy spent a weekend building something the world wasn't ready for. The idea: don't trust one model. Make them debate each other. He called it LLM Council. GPT, Claude, Gemini, Grok, same prompt. Every model critiquing each other. A "Chairman" AI synthesizing truth from their disagreements. He called...

12,246 views • 5 months ago •via X (Twitter)

23 Comments

Alexa | Startup founder's profile picture
Alexa | Startup founder5 months ago

keep going, you’ve got momentum.

Long 7's profile picture
Long 75 months ago

重要なアイデアです! AIの能力を最大限に引き出す簡単な方法に感謝します。cueyを試してみます!

Alex's profile picture
Alex5 months ago

The no-terminal part matters more than the 30 models—once comparison breaks flow, most people quit. SimianX learned that too.

Priyanka Vergadia's profile picture
Priyanka Vergadia5 months ago

Absolutely!! More accessible to more people which varying backgrounds not just developers.

rcanand's profile picture
rcanand5 months ago

I built maibook - something less structured than llm council, but more personalized - a private network of AI agents personalized based on the user’s interests (based on their activity). They discuss, augment, debate each thing user is into from different angles. Just launched it (free) -

Long 7's profile picture
Long 75 months ago

確かに、新しいテクノロジーは常に私たちの予想を上回りますね。これからはもっと多くの非技術者がAIを活用できるようなソリューションが必要です。

Ryuho | Sr. Eng @GlobalLogic (Hitachi)'s profile picture
Ryuho | Sr. Eng @GlobalLogic (Hitachi)5 months ago

The "Context Silo" has been the silent killer of AI productivity. We’ve all felt the pain of "Copy-Paste exhaustion" while manually running Karpathy’s council pattern. Making your memory and history portable across GPT-5, Claude 4.5, and Gemini 3 is a game-changer. It turns AI from a series of isolated conversations into a continuous, compounding knowledge base. This is the "Librarian" layer every pro-user has been waiting for.

Charlie Brewer's profile picture
Charlie Brewer5 months ago

Getting the models to work with & against each other sounds like a great strategy that won’t get outdated…you get the meta-performance of the models even as they upgrade

Saeed Anwar's profile picture
Saeed Anwar4 months ago

The LLM Council idea sounds great until you realize the Chairman model is also an LLM with the same failure modes. You're not eliminating bias, you're adding a meta-layer of it on top of the first round.

Sujal Manpara's profile picture
Sujal Manpara5 months ago

what about memory context ?

Priyanka Vergadia's profile picture
Priyanka Vergadia5 months ago

Assuming you are providing all the context in the one prompt in this case.

Giacomo Masseroni's profile picture
Giacomo Masseroni5 months ago

Yes, please, let's allow the whole world to develop with AI, without having the slightest idea what the AI ​​is doing.

Miss Cy 🦋 Ambassador 🔥's profile picture
Miss Cy 🦋 Ambassador 🔥4 months ago

MIT's 70% to 95% accuracy jump with multi-model debate is wild. The Chairman AI concept is pretty elegant honestly. Would be cool to see @tntaiapi666 test this with their API infrastructure.

alfhvatne's profile picture
alfhvatne5 months ago

Looks like a product yes. $9.99 per month grants you 100 model-debates in that period and a "priority queue". Models not being quite good enough so one should ideally query several vendors to get good enough answers sounds like a patch (and a business model)

CloudlyAI's profile picture
CloudlyAI5 months ago

great ! Another way is to use it as a skill that runs the full thing - 5 advisors, peer reviews, chairman verdict - automatically, on any decision you throw at it !!! this guy 👇has provided the link and it worked for me pretty well(except "type":"exceeded_limit..... error sometimes )

Great Tokens AI's profile picture
Great Tokens AI4 months ago

The council pattern is clever, but the real challenge is orchestration overhead. For most teams, simpler routing based on task type delivers better ROI.

CleverMC's profile picture
CleverMC5 months ago

Some of those models are not smart enough to be on the council.

Priyanka Vergadia's profile picture
Priyanka Vergadia5 months ago

Lol thats a different topic of discussion altogether! But I agree.

CleverMC's profile picture
CleverMC5 months ago

😂

⚔️𝕯𝖎𝖌𝖎𝖙𝖆𝖑 👹 𝕽𝖔𝖓𝖎𝖓`鬼` (クラッシュ・オーバーライドX)'s profile picture
⚔️𝕯𝖎𝖌𝖎𝖙𝖆𝖑 👹 𝕽𝖔𝖓𝖎𝖓`鬼` (クラッシュ・オーバーライドX)5 months ago

He did not invent this I have proof I was running council based setups almost a year before that weekend experiment tired of this crap

JihadiMouse's profile picture
JihadiMouse5 months ago

Who do you think i am, a millionaire

Priyanka Vergadia's profile picture
Priyanka Vergadia5 months ago

@quiveringpudle lol! Assuming everyone is a millionaire these days with how we are using tokens!

JihadiMouse's profile picture
JihadiMouse5 months ago

Btw is the 95% accuracy number made up right?

Related Videos

SOMEONE FROM THE ANTHROPIC TEAM LEAKED THEIR OBSIDIAN SETUP. 8 MILLION PEOPLE SAW HOW HE ACTUALLY USES CLAUDE the funniest part? all of this information was sitting in claude's documentation from day one. nobody read it one guy did, packed it into a 9-step guide and posted it. and it broke the internet. 4,100 likes, 800 retweets, then china picked it up and 8 million views want to know what's in it? one file. called CLAUDE.md. it holds everything about you: how you think, what you're working on, where you get stuck, even how you want the ai to talk to you. claude reads it first every single session one file changed everything. because now ai doesn't open with "how can i help?" it already knows. it remembers your projects, sees your goals, catches moments where you're contradicting yourself people spent years searching for the perfect prompt. the right temperature. the magic formula. and the answer turned out to be not how you ask ai. but what ai knows about you before you even open your mouth then the guy went deeper. taught claude to work on a schedule. every morning at 7am the ai walks through all notes on its own, finds new stuff, links it, cleans what's stale. no command. no reminder and all of this runs on obsidian. free app. text files on your drive. no cloud, no lock-in. switch models tomorrow and the folder keeps working the most liked comment under the original post: "this is the difference between using ai and building a system. most people won't realize it until they waste hundreds of hours repeating themselves" hundreds of hours. you've already spent some of them full guide in the video. i break down finds like this every day - follow so you don't miss the next one

kai

174,650 views • 2 months ago

REAL ESTATE PEOPLE WILL HATE HIM FOR THIS. HE BUILT A CLAUDE AGENT THAT TURNS ANY LISTING INTO A SELLABLE VIDEO ON ITS OWN Playbook: connect Claude to a video generator, paste a listing, get a cinematic tour of every room, sell it to the agent But typing the prompt for every listing doesn't scale. He turned it into a skill his Claude runs on its own Here's how to build the automated version: 1. Connect the video engine once. In Claude, go to Customize, Connectors, Add Custom Connector, name it Higgsfield, and paste the server URL from higgsfield. ai/mcp. Authenticate through your account. No API keys. Now Claude can generate video straight from chat 2. Turn the workflow into a skill. Instead of pasting the same prompt every time, have Claude build a skill. Tell it: "Create a skill called listing-to-video. When I give it a listing URL, scrape the room photos, generate a cinematic clip of each room with Higgsfield, and save them to a folder." Now the whole process is one command, not a wall of text 3. Let the agent run the listing. Hand it a URL and say "run listing-to-video on this." It pulls the photos, fires each room through the video model, and brings the clips back. You wrote the prompt once, inside the skill. You never write it again 4. Stitch and deliver. Drop the clips together into one tour. Send a free sample to the listing's agent, then charge per video or a monthly rate for ongoing listings 5. Scale it with your team. Add a skill that drafts the outreach email and one that builds a simple landing page for the agent. Now one operator runs sourcing, production, and pitching from a single Claude session The edge isn't generating one video. It's building the skill once so every future listing runs itself Bookmark this

Yarchi

54,840 views • 3 months ago

How to create Farcaster mini app without coding - Step 1: > Go to ChatGPT > Just type: "I would like to create a Farcaster mini app, can you give me some ideas?" > Choose any one of the ideas - Note: - You can also type your own idea to ChatGPT and alter it - Step 2: - Go to ChatGPT > Paste the first prompt (shared in my channel) > Include the project you chose and tell ChatGPT to alter the prompt according to the idea below - Step 3 - Go to: > Sign up using Gmail or GitHub > Projects -> New Project -> Type the name of your project >Just paste the prompt we created earlier > It will start creating your app > If you need to alter anything, just go to GPT and type: [ I need a prompt to alter "your issue (like changing colour, etc.)" in my app already created in v0app by Vercel ] > Copy the prompt and paste it in v0 > No need to worry about errors , it will rectify them for you > If your project is finished, just click "Publish" on your top right > It will automatically host your app in Vercel > Click "Visit site" to see your app on another device and copy the website URL - Step 4 - Go to: v0 app > Just paste Prompt 2 from the file I provided > Click Publish once it’s completed - Step 5 - Go to: v0 app > Just paste the prompt 3 from the file I provided - Step 6 (Important) > Just paste Prompt 4 from the file I provided - Step 7 > Just paste Prompt 4 from the file I provided - Step 8 > Just paste Prompt 6 from the file I provided > It’s an important step to do — you need to create a manifest inside your app so that you can host it on Farcaster - Step 9 - Go to: > Sign up or Login > Settings → Developers → Mini Apps → Create Manifest (New) > Paste your website URL (remove https) > It should be like: > If you find any error while doing this, just copy and paste it in the v0 app > Once everything is finished, you need to create Account Association > Scan the QR on your mobile and tap and hold the Yes button > It will show an error — just copy and paste those errors into the v0 app > Click Publish again, and then click Refresh in Manifest > Click Open your app > for the main prompt file check our tg channel link in our bio

Maran

48,427 views • 10 months ago

anthropic's head of product just revealed how they're able to ship faster than any other AI company. their secret: "side quest maxxing." here's how it works: instead of long-term roadmaps, anthropic runs on unplanned afternoon experiments. anyone on the team gets full freedom to spend an afternoon prototyping an idea and show it to the team. you get to skip the approval process entirely. then, employees at anthropic try it. if they keep using it the next day and the day after that, it gets polished into a real feature. if nobody touches it again, it dies. that's the whole process. claude code on desktop started as one engineer's afternoon project. he wanted it to work on desktop so he built a prototype. people on the team started using it immediately. so they shipped it. the todo list feature started the same way. someone built it, the team adopted it internally, and it became one of the most-used parts of the product. plugins started when one engineer shared a spec with claude code and the prototype that came back was close to production-ready. went from idea to working feature in a single session. they also killed standup meetings. instead of telling people what you're working on, you just show a working demo. all walk no talk basically the team structure makes this possible. > designers ship code. > engineers make product decisions. > product managers build prototypes. everyone can take an idea from concept to working demo without waiting on anyone else. the biggest features at a $380b company came from afternoon experiments that nobody asked for. honestly this matches my own experience cooking with ai. some of the best workflows i use every day came from just fucking around. opening a session with zero intention and asking claude what it can do, or jamming on a random idea to see where it goes. if you're only using ai for tasks you already have in mind, you're missing the best part. open a session with no agenda. ask it to surprise you. try building something stupid. half the time it goes nowhere. the other half it becomes the thing you use most. you need to be sidequestmaxxing.

Ole Lehmann

106,072 views • 5 months ago