Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Introducing Cognee v1.0: a major breakthrough in agentic intelligence. It is 145% better than Opus 4.8 and GPT 5.5 at long context memory retrieval. Cognee allows a 100 BILLION token context window 100,000x more than Claude. It's: - 6.9x cheaper than GPT 5.5 and Opus 4.8 - Cold starts...

848,750 Aufrufe • vor 3 Monaten •via X (Twitter)

71 Kommentare

Profilbild von Vasilije
Vasilijevor 3 Monaten

Book a demo or try today 👇

Profilbild von Vasilije
Vasilijevor 3 Monaten

Today, agents have poor memory recall, burning up to 90% of your tokens: - Agents get stuck in loops trying to retrieve context - Agents get forced into gargling more context than they can handle - Agents guess when limited context misses details, producing an inaccurate output

Profilbild von Vasilije
Vasilijevor 3 Monaten

Cognee connects your context to the right places, so your agent knows what matters right now, and why. When data is requested, Cognee activates its memory graph and updates how that information is connected, used, and understood over time.

Profilbild von Vasilije
Vasilijevor 3 Monaten

Better recall creates higher quality context, as Cognee dynamically activates and updates relevant memory to produce more accurate outputs with fewer errors. Cognee significantly outperforms base models and performs better than competitors’ memory frameworks.

Profilbild von Vasilije
Vasilijevor 3 Monaten

To celebrate our launch, we're giving away the secrets to boost Claude’s output by 400%. Reply to the first post and we'll send it to you.

Profilbild von greg
gregvor 3 Monaten

Yes or no: when I say “explain this in NBA terms” will it explain it in NBA terms or will I once again need to explain what i mean? YES OR NO

Profilbild von Vasilije
Vasilijevor 3 Monaten

YES. The repeating yourself meta is officially dead. Cognee saves your preferences to a memory graph so it actually remembers how you like things broken down across sessions. Agent amnesia is cooked. What's the first NBA breakdown you're making it run?

Profilbild von Robert Scoble
Robert Scoblevor 3 Monaten

and we were all just... restarting every session like that was normal. how easy is it to install Cognee?

Profilbild von Vasilije
Vasilijevor 3 Monaten

It's a 5-minute setup. Cognee acts as a plug-and-play memory layer for your existing AI setup. You just connect it to your current agents with a few lines of code, and they instantly stop forgetting context.

Profilbild von Michael Wondwossen
Michael Wondwossenvor 3 Monaten

@Scobleizer Great stuff! Memory/context is such a massive pain point people are suddenly waking up to en masse. Sounds almost too simple haha, but I’ve been seeing you grinding on this on LinkedIn for a while.

Profilbild von Vasilije
Vasilijevor 3 Monaten

@Scobleizer Gotta grind!

Profilbild von techbimbo
techbimbovor 3 Monaten

can agents share memory with this or do both get their own “brain”?

Profilbild von Vasilije
Vasilijevor 3 Monaten

Both. You can isolate their memory graphs for individual brains, or connect multiple agents to one shared graph for a hive mind setup. What kind of multi-agent architecture are you building?

Profilbild von techbimbo
techbimbovor 3 Monaten

honestly still getting into it, but thanks!

Profilbild von Bark
Barkvor 3 Monaten

so the machine remembers everything now. cool. who has access to what it remembers??

Profilbild von Vasilije
Vasilijevor 3 Monaten

You do! Because Cognee is fully open-source and self-hosted, your data stays wherever you deploy it, locally on your machine or inside your own private cloud.

Profilbild von WallStreetBets
WallStreetBetsvor 3 Monaten

I’m pulling up 👀

Profilbild von Machina
Machinavor 3 Monaten

we built machines that can pass the bar but can’t remember context from 10 minutes ago lol you cooked fr

Profilbild von Vasilije
Vasilijevor 3 Monaten

Real, the irony is wild. Smart enough to pass the bar but can't remember the last prompt. What kind of use case or workflow are you building that needs this long-term memory?

Profilbild von Boring_Business
Boring_Businessvor 3 Monaten

6.9x cheaper than GPT 5.5 and Opus 4.8. How?

Profilbild von Vasilije
Vasilijevor 3 Monaten

Context stuffing is cooked. Instead of dumping 100k tokens into the LLM every turn and burning cash, Cognee’s memory graph surgically retrieves only what’s needed for that exact prompt. Major token diet. What kind of daily token burn are you dealing with right now?

Profilbild von Luminara
Luminaravor 3 Monaten

so you're saying it has 145% more chances to remember how i was screaming at it "faster! shorter! punchier!"?

Profilbild von Vasilije
Vasilijevor 3 Monaten

Exactly. It hardcodes your formatting rules into the graph so you never have to scream at it again. What are you building?

Profilbild von Madhav 🦄
Madhav 🦄vor 3 Monaten

Reminds me of the LLM-as-compiler pattern Low latency and plug-and-play are smart Curious how graph evolution/consolidation holds up long-term (conflicts, forgetting, schema drift) and how inspectable it stays Excited to see memory treated as a real first-class layer

Profilbild von david 🔛⛓️
david 🔛⛓️vor 3 Monaten

Lmao every agent has been playing 50 first dates this whole time

Profilbild von Vasilije
Vasilijevor 3 Monaten

"Hi, nice to meet you for the 40th time today." 💀 Cognee completely puts an end to the madness. What's the first workflow you're locking in?

Profilbild von Madhav 🦄
Madhav 🦄vor 3 Monaten

Nice. Agent memory is still the weakest link, clever agents slowly degrade into loops and token waste. Graph + vector hybrid makes sense: compile raw observations into traversable entities/relations once, instead of stuffing everything into prompts.

Profilbild von Karan
Karanvor 3 Monaten

memory for agents startups finding out about cognee like

Profilbild von Vasilije
Vasilijevor 3 Monaten

Pack it up boys, Cognee just dropped

Profilbild von Just a Dude Who Invests
Just a Dude Who Investsvor 3 Monaten

we're funding AI companies at a trillion dollars and the breakthrough this week is "it can remember stuff." we are so early

Profilbild von Out of Context Human Race
Out of Context Human Racevor 3 Monaten

Developers taking the day off

Profilbild von Trevin Chow
Trevin Chowvor 3 Monaten

1. When adding @NousResearch Hermes support? 2. How does it compare to @garrytan gBrain and @supermemory ?

Profilbild von Vasilije
Vasilijevor 3 Monaten

@NousResearch @garrytan @supermemory Added, but @NousResearch doesn't integrate anymore memory providers closely, so we built our own plugin

Profilbild von Trevin Chow
Trevin Chowvor 3 Monaten

@NousResearch @garrytan @supermemory @Teknium is it true y’all aren’t supporting new memory providers?

Profilbild von Rachel Rapp
Rachel Rappvor 3 Monaten

Build me a better agentic memory system, make no mistakes (Nailed it -- congrats on the launch!! 🎉)

Profilbild von Vasilije
Vasilijevor 3 Monaten

Queen of Vegan Ramen strikes again!

Profilbild von Amir Valizadeh
Amir Valizadehvor 3 Monaten

these are extremely suspicious numbers...

Profilbild von Vasilije
Vasilijevor 3 Monaten

Healthy skepticism is always welcome 😄

Profilbild von Amir Valizadeh
Amir Valizadehvor 3 Monaten

well skepticism is warranted, especially when you’re comparing a memory framework to a large language model. They’re two different things, and posting numbers that say “our memory framework is a gajillion times better than this LLM” is completely misleading

Profilbild von shirish
shirishvor 3 Monaten

Does this mean Claude won’t waste 10 minutes compacting context every 10 prompts? I must be dreaming…

Profilbild von Vasilije
Vasilijevor 3 Monaten

should I pinch you or just try it out for yourself?

Profilbild von shirish
shirishvor 3 Monaten

wait..let me try it

Profilbild von Ciph
Ciphvor 3 Monaten

hey yh, here's a quick question: what’s the quickest way to try Cognee with a simple agent?

Profilbild von Vasilije
Vasilijevor 3 Monaten

The fastest route is hooking it directly into a dev environment like Cursor or Claude Code via their native MCP server. You just spin it up in your terminal, and your existing tools instantly read/write to the graph.

Profilbild von SAIF MR 🔺
SAIF MR 🔺vor 3 Monaten

Curious hw this compares to mem0, Can multiple agents share same memory graph ??

Profilbild von Vasilije
Vasilijevor 3 Monaten

Cognee is document-first Graph-RAG; it pipelines massive data into strict graphs. Multiple agents can share one graph using global datasets, keeping individual loops isolated via session_id.

Profilbild von CG
CGvor 3 Monaten

Why does this actually work/ genuinely unsettling. I’m scared.

Profilbild von Vasilije
Vasilijevor 3 Monaten

Because it builds a real graph that maps concepts like a brain instead of guessing keywords. Flawless recall is a trip. What workflow are you testing first?

Profilbild von Leo Grundström
Leo Grundströmvor 3 Monaten

every script i write, i spend the first ten minutes reminding the AI how i write. every single time. been doing this for two years now. Weird that cognee is only coming out now.

Profilbild von Vasilije
Vasilijevor 3 Monaten

Because the industry spent years selling bigger context windows as a cheap fix for "memory." Cognee changes the game by locking your coding style into a permanent graph node so you never have to repeat your guidelines. What language do you usually write your scripts in?

Profilbild von Vee
Veevor 3 Monaten

@barkmeta 👀

Profilbild von Vasilije
Vasilijevor 3 Monaten

@barkmeta 👀

Profilbild von Vaibhav Sisinty
Vaibhav Sisintyvor 3 Monaten

looks exciting. Congrats on the launch!

Profilbild von Vasilije
Vasilijevor 3 Monaten

Thanks! Appreciate you being here

Profilbild von PARSA
PARSAvor 3 Monaten

A memory layer that connects to agents you've already built is a smart approach

Profilbild von Vasilije
Vasilijevor 3 Monaten

100%. Decoupling memory from the core orchestration layer means you get to keep your current stack while completely fixing the agent amnesia problem.

Profilbild von 𝑺𝒉𝒂𝒌𝒆𝒔
𝑺𝒉𝒂𝒌𝒆𝒔vor 3 Monaten

cognee just turned AI agents from goldfish into elephants with 100B token memory at 1/7th the cost i’m definitely gonna book a demo today

Profilbild von kaize
kaizevor 3 Monaten

can two agents share one memory graph, or is it isolated per agent? cross-agent recall is exactly where most "memory" tools quietly stop

Profilbild von Vasilije
Vasilijevor 3 Monaten

Most tools fold here, but Cognee supports shared graphs natively. It uses global namespace routing so a whole swarm of agents can read/write to the same underlying ontology, while using unique session IDs to keep runtime context from leaking. Multi-agent synergy is real.

Profilbild von axe
axevor 3 Monaten

need this lol just asked my AI about something from 40 minutes ago and it was like

Profilbild von Vasilije
Vasilijevor 3 Monaten

Stopping that exact amnesia loop is why we built Cognee. No more session resets or wasted tokens. What's the main agent architecture you're running right now?

Profilbild von greb
grebvor 3 Monaten

100 billion? that’s crazy good Gonna keep an eye out on this one for sure

Profilbild von Vasilije
Vasilijevor 3 Monaten

Absolute game changer. Bypassing that context wall changes the entire landscape for long-term agent tasks.

Profilbild von Peanut
Peanutvor 3 Monaten

i’m ready to put it to use and blow some minds here 🤯

Profilbild von Vasilije
Vasilijevor 3 Monaten

let's goooo

Profilbild von Miko
Mikovor 3 Monaten

this will likely have a huge impact in the AI landscape

Profilbild von ashen
ashenvor 3 Monaten

how did it take this long to make memory a core layer of the AI stack lol

Profilbild von Ally
Allyvor 3 Monaten

Cognee v.10 will be huge

Profilbild von Vasilije
Vasilijevor 3 Monaten

100%. The jump from basic persistent memory to full autonomous cognitive networks will be wild. Ready to scale your setup?

Profilbild von Bukky (Builder Arc)
Bukky (Builder Arc)vor 3 Monaten

If Cognee has all these features including being cheap. How come people are just getting to know about this wonderful tool ? I think everyone should be scared right now, imo with the ban of Fable 5 lately. But this , I see upskilling.

Profilbild von Vasilije
Vasilijevor 3 Monaten

The recent Fable 5 ban is exactly why relying entirely on closed APIs is terrifying. When a frontier model can vanish overnight by government decree, open-source tools like Cognee are how you build resilient, sovereign tech. It’s the ultimate upskilling play.

Ähnliche Videos

New short course: LLMs as Operating Systems: Agent Memory, created with Letta, and taught by its founders Charles Packer and Sarah Wooders. An LLM's input context window has limited space. Using a longer input context also costs more and results in slower processing. So, managing what's stored in this context window is important. In the innovative paper MemGPT: Towards LLMs as Operating Systems, its authors (which include the instructors) proposed using an LLM agent to manage this context window. Their system uses a large persistent memory that stores everything that could be included in the input context, and an agent decides what is actually included. Take the example of building a chatbot that needs to remember what's been said earlier in a conversation (perhaps over many days of interaction with a user). As the conversation's length grows, the memory management agent will move information from the input context to a persistent searchable database; summarize information to keep relevant facts in the input context; and restore relevant conversation elements from further back in time. This allows a chatbot to keep what's currently most relevant in its input context memory to generate the next response. When I read the original MemGPT paper, I thought it was an innovative technique for handling memory for LLMs. The open-source Letta framework, which we'll use in this course, makes MemGPT easy to implement. It adds memory to your LLM agents and gives them transparent long-term memory. In detail, you’ll learn: - How to build an agent that can edit its own limited input context memory, using tools and multi-step reasoning - What is a memory hierarchy (an idea from computer operating systems, which use a cache to speed up memory access), and how these ideas apply to managing the LLM input context (where the input context window is a "cache" storing the most relevant information; and an agent decides what to move in and out of this to/from a larger persistent storage system) - How to implement multi-agent collaboration by letting different agents share blocks of memory This course will give you a sophisticated understanding of memory management for LLMs, which is important for chatbots having long conversations, and for complex agentic workflows. Please sign up here!

Andrew Ng

201,127 Aufrufe • vor 1 Jahr