Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

karpathy's CLAUDE.md hit #1 on github trending. 220,000 stars. most devs still haven't read it. it's 65 lines. it took AI coding accuracy from 65% to 94%. the 4 rules inside: → think before coding state your assumptions. ask when unsure. never guess. → simplicity first write the minimum...

2,629,432 Aufrufe • vor 3 Monaten •via X (Twitter)

38 Kommentare

Profilbild von Statusnone
Statusnonevor 3 Monaten

The link to the OG repo:

Profilbild von self.dll
self.dllvor 3 Monaten

ep thanks

Profilbild von Marcelo Ceccon
Marcelo Cecconvor 3 Monaten

I have pushed it further with my charter inspired by Karpathy. The codebase has some eval for benchmarking but I’ve been using it heavy on my private codebases as well. Looking for feedback

Profilbild von self.dll
self.dllvor 3 Monaten

great job

Profilbild von David
Davidvor 3 Monaten

Anyone good at ai coding usually gives Claude something similar in the initial prompt, not every task needs these exact instructions and some need more

Profilbild von Yurii
Yuriivor 3 Monaten

That’s a good one. I also highly recommend. Second one that I’d recommend is called “caveman” it makes Claude way less verbose. Saves money, context and agent is not getting off rails as much

Profilbild von Henry "Coop" Jackson
Henry "Coop" Jacksonvor 3 Monaten

Every dev rn bookmarking the 65 lines they will never read

Profilbild von Tejas AI
Tejas AIvor 3 Monaten

This CLAUDE.md hitting 220k stars is actually wild Karpathy dropped 65 lines of “just think clearly and don’t be sloppy” and the entire timeline completely lost their minds. To be fair, the 4 rules are genuinely solid—especially “state your assumptions” and “surgical changes.” I’ve been forcing almost the exact same discipline into my custom Claude Skills for my UGC factories. It makes the jump from messy prompts to consistent Seedance 2.0 output way cleaner. Still calling bullshit on that 65% \rightarrow 94% accuracy jump, though. That smells like classic AI benchmark cope.

Profilbild von rodrigo
rodrigovor 3 Monaten

.

Profilbild von Average Guy Trading
Average Guy Tradingvor 3 Monaten

@grok does @karpathy actually has that repo? Or this is a scam?

Profilbild von Victor | Structural Alpha
Victor | Structural Alphavor 3 Monaten

i added 'ask when unsure' to my prompt and now claude asks me questions that expose i didn't understand my own request. humbling

Profilbild von Smart Money Leaks
Smart Money Leaksvor 3 Monaten

ironic that devs need a 65 line manifesto to stop guessing finally someone admitting the rest of us are just over-engineering nonsense

Profilbild von stargazer
stargazervor 3 Monaten

Friendly reminder:Karpathy didn't create this skill. He shared observations about common LLM coding pitfalls. A developer packaged them into a CLAUDE.md file for Claude Code, and the community repo

Profilbild von unemployable
unemployablevor 3 Monaten

It's missing very important "you are 100x elite engineer with 100 years of experience". And also the classic one "make no mistakes plz".

Profilbild von I Pun Daddy
I Pun Daddyvor 3 Monaten

65 lines and 220,000 stars, yet the most crucial rule seems to be 'actually read the thing'.

Profilbild von DeutscherPapa
DeutscherPapavor 3 Monaten

That’s very interesting. I thought prompt engineering had died a year ago, but I was mistaken. Thousands of users are still searching for the "golden prompt"... It’s a bit sad. I really don't understand why you would try to coax a model instead of just giving it a direct command

Profilbild von supper
suppervor 3 Monaten

The "think before coding" rule is the real unlock. Most people skip straight to prompting and then wonder why the output is messy. Writing down assumptions first forces you (and the model) to actually reason about the problem instead of pattern-matching.

Profilbild von Ataullah Siddiki
Ataullah Siddikivor 3 Monaten

Karpathy dropping another gem These 4 rules are pure gold. Especially 'surgical changes' and 'think before coding' — that's exactly where most AI coding sessions go off the rails. Added to every project now. The 220k stars are well deserved."

Profilbild von slash1s
slash1svor 3 Monaten

he is real legend

Profilbild von Mohit Jaswal
Mohit Jaswalvor 3 Monaten

repoweasel: CLAUDE.md is the operating file, DEMO.md is the 60-second prompt-only paste-in-any-model demo, README has the before/after + try-it path. MIT licensed, no install, single file. happy to answer any questions:

Profilbild von Gerard Sans | Axiom 🇬🇧
Gerard Sans | Axiom 🇬🇧vor 3 Monaten

Relevant thread:

Profilbild von Paulo Nick
Paulo Nickvor 3 Monaten

Excellent

Profilbild von danejw
danejwvor 3 Monaten

Fat skill, thin harness. Too much instruction in the harness limits the underlying models capabilities.

Profilbild von leanxbt
leanxbtvor 3 Monaten

did accuracy actually go from 65% to 94% or is that just a catchy number from the post?

Profilbild von JMoon
JMoonvor 3 Monaten

accuracy numbers in this post are made up but the CLAUDE.md concept itself is real. a well-structured one does change how the model approaches your codebase

Profilbild von Λ
Λvor 3 Monaten

This is solid evidence why a real software engineer will be the person who extracts the most out of LLMs. SEs know what to ask for because they know the consequences of decisions being made. That it what Karpathy is doing with the skills.md.

Profilbild von Dimi
Dimivor 3 Monaten

65 lines for 94% accuracy? My Chrome extension for X replies, Crazy Pigeon, needs about 65,000 lines to get people to click 'reply'. Guess I'm not Karpathy.

Profilbild von Bob Ulrich
Bob Ulrichvor 3 Monaten

how did you define “AI coding accuracy?”

Profilbild von Manpreet
Manpreetvor 3 Monaten

220k stars and the real flex is admitting you bookmarked it without opening the file 😂 Classic dev self-own.

Profilbild von StarShipX
StarShipXvor 3 Monaten

素晴らしいルール!簡潔で的確な方法ですね。必ず守ります。

Profilbild von pacemaker
pacemakervor 3 Monaten

lol why you’re so excited any decent dev say to his model to follow KISS,DRY,SOC,YAGNI there is nothing new lr special here

Profilbild von tawer
tawervor 3 Monaten

Goal criteria stop claude from adding unplanned refactors during a task

Profilbild von AI Mastery Guide
AI Mastery Guidevor 3 Monaten

65 lines, 4 rules, 29% accuracy jump. The simplicity is the whole point

Profilbild von zostaff
zostaffvor 3 Monaten

banger grats

Profilbild von self.dll
self.dllvor 3 Monaten

thanks my friend

Profilbild von Agent X AGI
Agent X AGIvor 3 Monaten

65→94% is real. it's also fragile. CLAUDE.md rules are advisory — model follows them turn 1, drifts by turn 5. accuracy creeps back down mid-session. the fix isn't more rules. it's a PreToolUse hook that checks output against constraints before committing

Profilbild von Granite
Granitevor 3 Monaten

been here before 100k views rules are simple yet so necessary

Profilbild von Mykola Kondratiuk
Mykola Kondratiukvor 3 Monaten

honestly most skip it and run their agents blind. 65 lines is nothing

Ähnliche Videos

Andrej Karpathy: "90% of Claude's mistakes come from missing context, not a weak model." 41% mistake rate without a CLAUDE.md. 11% with the 4-rule baseline. 3% with the 12-rule version below here are the 12 rules senior engineers settled on: 1. think before coding: state assumptions, don't guess. the model can't read your mind, stop hoping it will 2. simplicity first: minimum code, no speculative abstractions. the moment you let Claude add "for future flexibility," you've added 200 lines you'll delete next quarter 3. surgical changes: touch only what you must. don't let it improve adjacent code, that's how PRs blow up 4. goal-driven execution: define success criteria upfront, loop until verified. without them Claude either loops forever or stops too early 5. use the model only for judgment calls: classification, drafting, summarization, extraction. NOT routing, retries, status-code handling, deterministic transforms. if code can answer, code answers 6. token budgets are not advisory: per-task 4000, per-session 30000. by message 40 of a long debug, Claude is re-suggesting fixes you rejected at message 5 7. surface conflicts, don't average them: two patterns in the codebase? pick one. Claude blending them is how errors get swallowed twice 8. read before you write: read exports, callers, shared utilities. Claude will happily add a duplicate function next to an identical one it never read 9. tests verify intent, not just behavior: a test that can't fail when business logic changes is wrong. all 12 of Claude's tests can pass while the function returns a constant 10. checkpoint every significant step: Claude finished steps 5 and 6 on top of a broken state from step 4. nobody noticed for an hour 11. match the codebase conventions: class components? don't fork to hooks silently. testing patterns assumed componentDidMount, hooks broke them without surfacing 12. fail loud: "completed successfully" with 14% of records silently skipped is the worst class of bug. surface uncertainty, don't hide it what actually compounds instead of the next framework: - the CLAUDE.md file as institutional memory across sessions - eval-driven changes, not vibe-driven - checkpoints over speed - explicit conflicts over silent blending - discipline over framework, every time - one repo, one rules file, no exceptions be a few rules ahead of AI twitter before this becomes mass-opinion study this

Ronin

451,445 Aufrufe • vor 3 Monaten

Microsoft spent $13 billion and 3 years building an AI that knows your work context. Every time you open it, it still asks what you're working on. This developer set up a plain text file in 2 minutes. The file is called CLAUDE.md. It loads before every session. Before he types a single word. It already knows his name. It already knows his writing style. It already knows what he's building, who it's for, and what he never wants to see in a response. He doesn't introduce himself anymore. He doesn't explain his preferences anymore. He doesn't correct the same mistakes twice. He just works. No $30/month Copilot subscription. No Microsoft 365. No IT approval. No data sharing agreement. No onboarding. Just a plain text file, a free text editor, and 21 instructions a developer distilled from Andrej Karpathy's research. Those 21 instructions moved Claude's coding accuracy from 65% to 94%. The file hit #1 on GitHub with 82,000 stars. Most people using Claude right now have never heard of it. Microsoft has 221,000 employees, $13 billion invested in OpenAI, and a direct integration into every Windows laptop sold on the planet.. they built an AI assistant most companies pay $30/user/month for that still doesn't know your name. This developer has a laptop, a text file and a 2-minute setup.. he built something that knows more about how he works than any enterprise AI on the market. The $50 billion AI personalization industry just got embarrassed by a .md file. full breakdown down below

Dep

14,179 Aufrufe • vor 4 Monaten

THIS GUY CONNECTED HIS AI AGENTS TO HIS OBSIDIAN AND BUILT A BRAIN THAT LEARNS ON ITS OWN. HERE'S HOW TO BUILD IT Obsidian is just markdown files sitting in a folder. That turns out to be the perfect memory for an AI agent, because an agent can read and write those files directly. He wired his agents into the vault so they pull context from it, do the work, and write what they learned back. The notes aren't the point. The loop is, and it gets sharper every cycle How to build it: 1. Point an agent at your vault. The fastest way, no plugins, no API keys: open a terminal and run npx obsidian-mcp /path/to/your/vault. That exposes your Obsidian folder to Claude as a tool it can read, search, and write to. Add it to your Claude Code or Cowork config and restart 2. Confirm it can see the brain. Ask it: "list the notes in my vault and summarize what's in them." If it reads them back, the connection is live. Now it starts every task with everything the vault already holds instead of from zero 3. Give each agent one job and a write-back rule. Tell it: "research this, then save what you found as a new note in /brain with links to related notes." One agent researches, one summarizes, one plans. Each writes its output back into the vault 4. Close the loop. Add one line to every agent's instructions: "read /brain before starting, write your result back when done." Now each task leaves the vault richer, and the next run reads that before it works. It compounds instead of resetting 5. You only steer. Review what the brain produces, point it at the next thing. The agents handle the reading, writing, and connecting The edge isn't better notes. It's a brain that feeds itself, so the work gets sharper every cycle instead of starting over Bookmark this

Yarchi

58,549 Aufrufe • vor 3 Monaten

CLAUDE CODE JUST SHIPPED THE FEATURE THAT SOLVES THE BIGGEST PROBLEM EVERY BUILDER HAS WITH AI AGENTS. The problem: Claude starts a task, gets distracted by a sub-problem, goes down a rabbit hole, and never finishes the original thing you asked for. The solution: /goal One command. You set the goal at the start of the session. Claude now has a north star it checks against every action it takes. Not just at the beginning. Throughout the entire session. Every time Claude is about to do something it asks: does this action move me toward the goal the user set or am I drifting? If it is drifting it corrects. If it completes a sub-task it returns to the primary goal. If it hits a blocker it reports back instead of spending 45 minutes solving the wrong problem. This sounds like a small feature. It is not. The reason most people do not trust Claude Code for long autonomous runs is not capability. It is reliability. A Claude Code session that reliably finishes what it started is worth 10 times more than one that is more capable but wanders. /goal is the feature that makes long autonomous sessions reliable. Set the goal. Let it run. Come back to a finished result. Not a result that got 70% done before Claude decided the sub-problem was more interesting. Done. The builders running overnight agent sessions are going to use this command on everything from today forward. Bookmark this. Follow CyrilXBT for every Claude Code feature the moment it ships.

CyrilXBT

19,586 Aufrufe • vor 4 Monaten

A DEVELOPER CONNECTED CLAUDE CODE TO OBSIDIAN SO HIS AI AGENT WOULD STOP FORGETTING THE PROJECT EVERY MORNING. Every coding session used to start the same way. Claude would understand the repo, fix the bug, explain the architecture, and then the moment the session ended, all of that context disappeared. Same codebase. Same decisions. Same architecture. Same mistakes repeated again. So he added a memory layer. Instead of treating Claude Code like a smart terminal, he connected it to a local Obsidian vault through MCP. Now Claude can read the repo, open the vault, create notes, link concepts, and write important decisions back into the system. When it studies the codebase, it does not just answer once and forget. It creates notes for the major services, maps how the architecture works, links auth to the database, connects APIs to storage, and records why certain migrations or design choices exist. Obsidian becomes the project graph. Now when he asks why something was built a certain way, Claude does not guess from the current prompt. It reads the decision notes. When he starts a new branch, Claude checks the active context file. When the work is done, it updates what changed, what is blocked, and what the next agent needs to know before touching the repo. That is the real loop: read context, write code, capture decisions, update memory. Most people are still using AI coding tools like disposable chat windows. Ask, patch, close, forget. This setup turns Claude Code into infrastructure. The repo gets a memory layer that survives every session, and multiple AI agents can work from the same project map without stepping on each other. The unlock is not better prompting. The unlock is giving the agent somewhere to remember what it already learned.

DegenCalls

20,124 Aufrufe • vor 2 Monaten