Loading video...

Video Failed to Load

Go Home

10 GitHub repos with ⭐ 450k+ combined stars that stop your agent from wasting tokens 1. caveman 98k⭐ - makes your AI agent talk like a caveman. Same answers, 65% fewer output tokens. 2. rtk 76k⭐ - CLI proxy that filters git/test/docker output before it hits your context. Up...

68,057 views • 1 month ago •via X (Twitter)

38 Comments

broke boy's profile picture
broke boy1 month ago

now my agent almost for free thank you

unicode's profile picture
unicode1 month ago

that might be a bit of an exaggeration, but you'll definitely save money

ieatded's profile picture
ieatded1 month ago

Nice top 10 Saved it as always

Alex Boudot's profile picture
Alex Boudot1 month ago

98k stars to make AI talk like a caveman is the clearest possible verdict on token pricing.

Hussain Hashim | Building SundayBack's profile picture
Hussain Hashim | Building SundayBack1 month ago

@unicodef1wn haha caveman mode is genius. reminds me of the early days when we were just making things work with whatever we had. gonna check it out!

shiqway92's profile picture
shiqway921 month ago

+book this for the next claude code session

Lyrics's profile picture
Lyrics1 month ago

ty for insanely gud repos

Jurly's profile picture
Jurly1 month ago

token savings are useful, but bad retrieval can quietly trade fewer tokens for worse answers.

Blum's profile picture
Blum1 month ago

great list, found a couple of useful repos to check out

unicode's profile picture
unicode1 month ago

glad you liked it, Blum

Vũ Trụ Số's profile picture
Vũ Trụ Số1 month ago

Where is gortex ?

Yumzlef's profile picture
Yumzlef1 month ago

this post made me question three assumptions i had about what this category of tool is useful for. all three assumptions turned out to be wrong

Gipp 🦅's profile picture
Gipp 🦅1 month ago

rtk with serena feels like the best pairing: less noise, sharper code access

Logics's profile picture
Logics1 month ago

tysm 4 insanely gud repossss

RaoulDuke's profile picture
RaoulDuke1 month ago

didnt know mem0 dropped a new release just three days ago with better scores

Jordan Lee's profile picture
Jordan Lee1 month ago

Token efficiency is becoming an actual part of agent engineering.

Quaramly's profile picture
Quaramly1 month ago

and which repo is best?

unicode's profile picture
unicode1 month ago

code-review-graph is underrated. i use it in every one of my projects

0xbobaa's profile picture
0xbobaa1 month ago

How many more repositories are hidden from us?

unicode's profile picture
unicode1 month ago

a lot. that's one of the reasons to follow me lmao

Pedro's profile picture
Pedro1 month ago

solid combo of repos

Eduard Lugovtsov's profile picture
Eduard Lugovtsov1 month ago

This is also gigachad repo

Sebastian Buzdugan's profile picture
Sebastian Buzdugan1 month ago

what failure context does rtk drop, since hidden flaky test output costs more than tokens

aibuilderhq's profile picture
aibuilderhq1 month ago

token saving via caveman speak is missing the point. optimize the prompt design.

The Black Box's profile picture
The Black Box1 month ago

who cares about token savings when the models keep getting worse? another listicle pretending tools fix broken ai

Mekasto Engineering Solutions's profile picture
Mekasto Engineering Solutions1 month ago

Everytime I run out of token and currently working on project in which I have to stump deespseek_v4_pro What can help me in this

Camaleón Raro's profile picture
Camaleón Raro1 month ago

filtering repetitive CLI diagnostic outputs through a local proxy daemon dropped our background loop token consumption by ~55%. routing raw git diffs and test logs to a 32B model before context injection keeps agent response latency under 300ms. are you enforcing token budgets at the daemon level or in agent runtimes?

Suchintan Singh's profile picture
Suchintan Singh1 month ago

also @skyvernai if you want to cut your browser token usage down

CoderLuii's profile picture
CoderLuii1 month ago

the tokens are only half the bill. everything the agent reads gets written to disk and kept. tear apart a large claude session and roughly two thirds of it is tool results being fed back to itself, not anything the user typed. you pay to generate it, then you pay to store it.

AI Mastery Guide's profile picture
AI Mastery Guide1 month ago

Caveman one made me laugh 😂

Clanker Tees's profile picture
Clanker Tees1 month ago

The cheapest token is the log line that never enters context. I want every tool to return a receipt first: exit code, files changed, three-line summary, and a pointer to the full mess only if the agent asks.

Mr.Keyoor's profile picture
Mr.Keyoor1 month ago

Try this

ShadowAguy's profile picture
ShadowAguy1 month ago

65% fewer tokens but i still have to read it in caveman speak

Ross · 402Signal's profile picture
Ross · 402Signal1 month ago

Context management as infrastructure instead of prompt coping is the actual unlock. Token waste is still the silent tax on most agent loops.

Yash's profile picture
Yash1 month ago

Checkout this one too. Get all public apis at one place.

beamnxw ./'s profile picture
beamnxw ./1 month ago

gme repos

Carmelo schepis's profile picture
Carmelo schepis1 month ago

useful angle on token savings but double check those star counts

whemo's profile picture
whemo1 month ago

even vegans will eat this meat

Related Videos

THIS MIGHT BE THE #1 OPEN-SOURCE REPO FOR CLAUDE CODE RIGHT NOW. IT GIVES CLAUDE A MEMORY AND SLASHES YOUR TOKEN COST ON EVERY QUESTION The repo is safishamsi/graphify, a free open-source skill that turns any codebase into a knowledge graph Claude Code can read instantly. Instead of grepping through your files every session, Claude gets a map of how everything connects The problem it fixes: Every time you ask Claude Code about a big repo, it does the same thing, greps through dozens of files like a brute-force Ctrl+F, blows through your context window, and sometimes still misses the answer hiding in a file nobody searched. Claude Code has no memory of how your project is structured. Every session starts from zero What it does: It maps your entire codebase into a knowledge graph, capturing not just which files exist, but which functions depend on which, which modules are central, and which files cluster around the same concern. Claude queries the map instead of scanning files How it works, three passes: 1. Code structure, free and local. Tree-sitter parses your files and pulls out classes, functions, imports and call graphs. No LLM, no tokens, just your actual code mapped deterministically 2. Audio and video, if you have them. Transcribed locally and folded into the graph 3. Docs, papers, images. Here an LLM does semantic analysis, figuring out what each document means and where it fits. Only the meaning gets sent up, never your raw source It saves you money: Normally a question about a big repo makes Claude spawn explore agents that scan file after file, eating your context window and your token budget before you get an answer. With the graph already built, Claude queries the map instead of re-reading the codebase every time. Same answer, a fraction of the tokens. The graph only gets built once, then a hook rebuilds it after each commit for free, so you never pay that scanning cost again. The bigger the repo, the bigger the gap The best parts: it's a skill, so once installed Claude knows when to use it without you memorizing commands. It works on non-code folders too, point it at docs or notes and it can spin up an Obsidian vault How to add it to your Claude: 1. Install Claude Code if you haven't: npm install -g Paul Jankura-ai/claude-code 2. Add the skill: claude skill add safishamsi/graphify 3. Open your project folder and run /graphify . to build the graph 4. Optional, make it automatic: graphify hook install so the graph rebuilds after every commit That's it. Ask Claude about your repo and it reads the map instead of burning tokens on a file hunt Bookmark this

Yarchi

56,502 views • 3 months ago

10 repos that cut your ai agent token bill by up to 80% 1. microsoft/LLMLingua → cuts prompt size by up to 95% compresses prompts before the api call. 20x compression. published at EMNLP + ACL. near-zero quality loss. 6,100 stars 2. mem0ai/mem0 → replaces full conversation history in context stores what matters. retrieves only what's needed. 10,000 token history → 200 token memory. per agent. 54,800 stars 3. BerriAI/litellm → routes each call to the cheapest model simple task → haiku. complex task → sonnet. tracks cost per agent, per call, per day. 45,700 stars 4. run-llama/llama_index → replaces sending full documents rag: 100-page doc → 3 relevant chunks → same answer. 98% fewer tokens per query. 49,100 stars 5. chroma-core/chroma → replaces keyword search in full context vector store. finds the closest match. feeds only that. 50-200 tokens per query instead of thousands. 27,800 stars 6. letta-ai/letta → replaces infinite context window crashes paged memory for agents. loads only relevant memory. stops your agent from hitting limits and retrying. 22,400 stars 7. guidance-ai/guidance → cuts output token bloat by 30-50% structured generation. constrains model output natively. no more 100-token prompts to get json back. 21,400 stars 8. Aider-AI/aider → replaces pasting entire codebases builds a repo map. sends only files relevant to the task. not your whole project. just what the agent needs. 44,300 stars 9. openai/tiktoken → count tokens before you send know the exact cost before the api call happens. not after the bill arrives. 18,100 stars 10. simonw/ttok → hard cap on what gets sent cli tool: count tokens, truncate to budget limit. pipe any text in. get truncated output back. 389 stars most agents are expensive not because the model is expensive. because nobody checked what was being sent to it.

self.dll

39,554 views • 5 months ago

20 GitHub repos with 2.4M+ combined stars that replace tools costing $60,000+/year 1. public-apis ⭐456k - 1,500+ free APIs across every category, weather to finance to games, all documented. 2. awesome-selfhosted ⭐312k - self-hosted replacements for Notion, Google Photos, Zapier, and dozens more paid subscriptions. 3. hermes-agent ⭐230k - self-improving personal agent with persistent memory, cron scheduling, MCP built in. Free alternative to paid always-on agent platforms. 4. n8n ⭐200k - visual automation with native AI agents. Replaces Zapier/Make entirely, self-hosted. 5. ollama ⭐178k - run Llama, Mistral, DeepSeek locally with one command. No API bill, no rate limits. 6. dify ⭐152k - visual builder for AI agents and RAG pipelines. Skip the $500/mo no-code AI builder subscription. 7. free-for-dev ⭐132k - hundreds of services with permanent free tiers. No trials, no credit card. 8. awesome-llm-apps ⭐132k - 100+ ready AI agents and RAG apps with full code. 9. awesome-mcp-servers ⭐92k - thousands of MCP servers connecting your agent to browsers, databases, anything. 10. supabase ⭐108k - Firebase alternative that's actually free to start. Auth, DB, storage in one Claude Code prompt. 11. strapi ⭐73k - open-source headless CMS, generates a full API from your content model in minutes. No Contentful bill. 12. immich ⭐110k - self-hosted photo and video backup with face recognition. Cancel the Google Photos storage plan. 13. appwrite ⭐57k - complete backend-as-a-service, self-hosted. Auth, DB, functions, storage, one prompt away from Firebase money. 14. medusa ⭐36k - full ecommerce backend, open source. Skip Shopify Plus fees entirely on your next vibe-coded store. 15. novu ⭐39k - notification infrastructure for email, SMS, push, in-app, all in one API. Replaces OneSignal's paid tiers. 16. tooljet ⭐38k - drag-and-drop internal tool builder connected to any database or API. Retool's seat pricing gone. 17. mattermost ⭐38k - self-hosted team chat built for engineering orgs. Slack without the per-seat bill. 18. outline ⭐40k - fast, clean team wiki and docs. Replaces Confluence and Notion's team plan. 19. plausible ⭐28.5k - privacy-friendly analytics, lightweight script, real dashboards. No GA360 contract needed. 20. openwork ⭐22k - open-source Claude Cowork alternative. Share skills and MCPs across Claude Code, Cursor, Codex, one setup for every agent. Save this before you pay for another tool this list already replaces for free 👇

unicode

34,424 views • 1 month ago

this AI agentic stack is f*cking crazy 8 open-source repos that cover everything you need to run AI agents for free. not a random "top tools" dump. each one handles a different part of the job: BUILD > video-use - drop raw footage in a folder, tell Claude Code "make a launch video", get back final.mp4. it cuts filler words, adds subtitles, grades color > > Univer - a full office suite made for agents. they edit your sheets, docs and slides in a draft, you review and merge > ORCHESTRATE > Superpowers - stops your coding agent from writing code blindly. it plans first, splits work into 2–5 min tasks, tests everything > > Treg - OpenRouter, but for agent tools. one key, ~2,600 endpoints from 40+ providers: scraping, SEO, email finding > ACT > Mobile MCP - your agent controls a real iPhone or Android. taps, swipes, installs apps, records the screen > > security-audit-skill - Cloudflare's own skill that turns your coding agent into a security auditor. one agent hunts bugs, a different one double-checks every finding > REMEMBER · TEST · SHIP > Hindsight - memory that actually learns. it doesn't just recall old chats, it builds beliefs and updates them over time > > Graphify - turns your whole codebase into a map your agent can query instead of grepping files blindly. zero LLM cost to build it > the cool part isn't any single repo. it's that your agent can now: edit video -> work in Excel -> plan like a senior dev -> use 2,600 tools -> drive a phone -> audit its own code -> learn -> know your codebase you don't have to build an AI employee from scratch anymore. every part is already on GitHub. you just put it together. bookmark this before you start your next agent.

kiosa

12,887 views • 6 days ago