Loading video...

Video Failed to Load

Go Home

๐ŸงEvery flavor holds a memoryโ€”savor it with ๐“๐ฐ๐ข๐ง๐ค๐ฅ๐ž ๐“๐ฐ๐ข๐ง๐ค๐ฅ๐ž๐Ÿฏ #POPMART #twinkletwinkle #savorthemoment #baguette #meal #snacks #afternoontea #arttoys #blindbox #animation

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

A tricky LLM interview question: You're serving a reasoning model on vLLM, and it keeps running out of GPU memory on long traces. So you add KV cache compression and evict 90% of the cached tokens. VRAM usage stays as is and GPU still runs out of memory. Why? (answer below) Evicting 90% of the KV cache can free almost none of the memory it was using. This sounds counterintuitive, but it follows directly from how production servers store the cache today. The KV cache grows with every token a model generates. Each token appends its key and value vectors across every layer, and nothing is freed while generation continues. This is the dominant memory cost for reasoning models. If a 32K-token CoT caches ~32K tokens of KV vectors, a Qwen3-32B with 4-bit weights will run out-of-memory around 24K tokens on a 24GB GPU. One obvious solution is to keep the important tokens and drop the rest, since attention is sparse enough to allow it. But this does not solve the memory problem yet. The reason is paged attention, which is the memory manager behind vLLM and most production servers. Under the hood, it splits GPU memory into fixed physical blocks, each one holds the KV for about 16 tokens. This block returns to the allocator only when every slot inside it is empty. Since the eviction logic selects tokens by importance, and such tokens are scattered across blocks... ...so despite eviction, almost every block is left with at least some survivor tokens. For instance, if the logic evicts 14k of 16k tokens across 1,000 blocks, most likely every block will still have a token. This means the allocator frees almost nothing. Placing the new tokens into those freed slots is not ideal because it breaks the cache's layout. Say token 16,001 arrives, and it's placed in the slot the 40th token used to hold. The cache now reads position 38, then 16,001, then 41, so the cache is no longer in token order. Attention can still compute the right answer from that, but only if every slot now carries a separate note recording which position it actually holds. This introduces another bookkeeping cost that an in-order layout inherently avoids. So the cache is logically 90% smaller and still physically the same size. Many compression results miss this because they measure on pre-allocated contiguous tensors rather than a paged server. There's another problem. Eviction methods pick which tokens to keep by looking at the attention scores themselves (as expected). But fast attention kernels used in production, like FlashAttention, never save those scores. They compute attention in small pieces and throw the full score grid away as they go, which is also why they're fast. So the exact signal eviction methods need isn't available in memory. The workaround is to fall back to eager attention and build the full matrix, which gives up the speed FlashAttention was there to provide. NVIDIA published a method called TriAttention to solve both these problems. It never needs attention scores. Instead, it scores tokens from the geometry of the model's key and query vectors before RoPE is applied, where those vectors sit in stable clusters. For the memory problem, it runs a compaction pass every 128 decoded tokens. The surviving tokens slide forward to close the holes eviction creates, so whole blocks empty out and return to the allocator while the cache stays in token order. On long reasoning traces, the approach matches full-attention accuracy while decoding 2.5x faster and using 10.7x less KV memory. KV cache compression is a big infrastructure problem. The number that decides whether it works is the count of freed blocks, not the count of evicted tokens. You can find the NVIDIA write-up here: I wrote a first-principles breakdown of how the KV cache works. It walks through why the model stores keys and values at all, why the cache grows with every token, and a comparison of LLM generation speed with and without KV caching. Read it below.

Avi Chawla

270,496 views โ€ข 1 month ago

๐“๐Ž๐†๐„๐“๐‡๐„๐‘, ๐’๐“๐ˆ๐‹๐‹ ๐€๐๐ƒ ๐€๐‹๐–๐€๐˜๐’ ๐Ÿ’ซ Like a photograph, every moment we have shared in WillCaVille captures more than smiles - it holds the laughter, the tears, the lessons, and the love that continue to shape our journey. Each memory reminds us that this bond was built not just on joy, but on understanding and growth. As Ed Sheeran beautifully sings, โ€œWe keep this love in a photograph,โ€ and truly, these moments will always remind us of how far we have come and how much stronger we have become together. ๐—›๐—ฎ๐—ฝ๐—ฝ๐˜† ๐Ÿณ๐˜๐—ต ๐— ๐—ผ๐—ป๐˜๐—ต๐˜€๐—ฎ๐—ฟ๐—ฟ๐˜†, ๐—ช๐—ถ๐—น๐—น๐—–๐—ฎ๐—ฉ๐—ถ๐—น๐—น๐—ฒ! Thank you for walking with us through every high and low, for choosing to stay, and for believing in what we stand for - despite the despite. Each day with this family reminds us of the beauty of loyalty and the strength of shared dreams. To our dearest Will Ashley and Bianca De Vera, thank you for being the heart that keeps us inspired. Your passion, kindness, and growth continue to light our path. Weโ€™ll keep holding on to every moment, every memory, and every reason we started - all for you. ๐Ÿฉท๐Ÿฉต #WillCa #WillAshey | #BiancaDeVera

WILLCA OFFICIAL

266,702 views โ€ข 9 months ago

Gavin Baker, CIO of Atreides Management made one of the most important and nuanced calls on memory stocks in recent months (Save this). His argument is that based on every memory cycle of the last 25 years, the setup today, prices elevated, sentiment high, supply ramping is textbook time to sell but he adds a critical exception. The one cycle in modern memory history where selling was catastrophically wrong was the mid-1990s, which Baker calls the last true capacity cycle in memory. In that cycle, demand was structurally exploding as the internet era required entirely new computing infrastructure to be built from scratch, and memory had to scale with it in a way that had never happened before. His point is that AI may be that same kind of cycle and not a normal boom bust but a once in a generation capacity buildout where the underlying demand is structural, not cyclical. The reason this argument holds weight is the fundamental shift in what memory is in the AI era. Traditional DRAM was a pure commodity, identical specs, interchangeable suppliers, price determined entirely by supply and demand swings. HBM is the opposite because it is custom engineered to fit a specific customer's chip, co-designed between the memory maker and the GPU designer, with SK Hynix's Vice President literally describing it as shifting from a commodity to a customer-tailored custom business. A single Blackwell Ultra GPU now requires up to 288GB of HBM3E, a 3.6x increase over the H100 and major suppliers like SK Hynix and Micron have already sold out their entire HBM production capacity through the end of the year. Because HBM requires advanced packaging processes like CoWoS that can't be spun up overnight, the bottleneck isn't just wafer capacity but rather runs across the entire manufacturing stack. Bank of America projects the global HBM market grows 58% this year alone to $54.6 billion, and Nomura expects the broader memory sector to nearly double to $445 billion. Long Micron!

Milk Road AI

260,701 views โ€ข 1 month ago

A guy in China plugged an LLM into Karpathy method + Claude Code and built a second brain that remembers everything forever. Not a note app, not RAG just 3 folders on a laptop. Folder 1 holds raw sources: articles, journal entries, chat logs. The LLM reads them but never touches them. Folder 2 is the wiki hundreds of markdown pages the LLM writes and maintains itself, where every person, idea and decision gets filed, cross-referenced and kept current. Folder 3 is one file: CLAUDEmd, the rulebook that turns a chatbot into a librarian. He drops in a new article and Claude reads it, updates 10-15 pages, flags what contradicts his old notes, logs the timestamp while he watches his memory grow node by node in Obsidian graph view. Ask it anything and it doesnโ€™t dig through raw files like ChatGPT does the synthesis already exists, the connections are already drawn. The method comes from a gist Karpathy posted on April 4, now sitting at 5,000 stars and 5,000 forks, with 1 production team already scaled to 4,000 interlinked concepts in 6 months. The part nobody mentions: Vannevar Bush described this exact machine in 1945 the Memex, a personal memory store with trails between documents. Nobody built it for 81 years because someone had to do the maintenance. Now maintenance costs nothing the LLM does the filing while the human does the thinking. Every article makes the brain smarter, and it never forgets. His memory has a git history now.

West Lord

61,701 views โ€ข 1 month ago

Wow! This Changes Everything We Thought We Knew About Memory It is a groundbreaking big deal. TOU ARE MAKING GENERATIONAL MEMORIES RIGHT NOW IN EACH CELL! Scientists just found that your brain doesnโ€™t just store memories โ€” it stores the rules for how those memories will change in the future. A brand-new preprint from Stanfordโ€™s Greenleaf and Schnitzer labs (led by PhD student Yuxi Ke) drops a bombshell that feels like science fiction becoming reality overnight. For decades, neuroscientists have suspected that chromatin the DNA packaging material inside every cell nucleus might somehow โ€œstoreโ€ memory-related information. But what kind of information? Content? Timing? Rules? Now we have the answer! Using activity-dependent genetic tagging, fear conditioning, and single-nucleus multiome sequencing in the mouse medial prefrontal cortex (the brainโ€™s long-term memory vault), the team tracked engram neurons for a full *month* after a memory was formed. What they discovered is electric: - One month after encoding, engram neurons have acquired a completely new chromatin landscape. - These chromatin changes are almost invisible at 7 daysโ€ฆ but roar into existence by 28 days. - At recall, these engram cells donโ€™t just โ€œrememberโ€ better โ€” they rewrite their entire transcriptional response. They preferentially fire up chromatin regulators, RNA processing machinery, and protein-turnover systems instead of simply boosting classic plasticity genes. In other words: the chromatin doesnโ€™t just hold the memory of the past. It holds metaplastic instructionsโ€” rules that dictate how the neuron will respond the next time the memory is triggered. They call it chromatin metaplasticity. This is the โ€œfuture tense of memory.โ€ It is a massive deal 1. Memory is not just synapses. For 70+ years weโ€™ve been obsessed with synaptic weights. This work proves the nucleus itself is a computational device that stores history-dependent rules. 2. It explains remote memory. The chromatin signature keeps maturing for weeks after the experience, perfectly matching the time course of systems consolidation into the cortex. 3. Itโ€™s energy-efficient genius. Instead of constantly maintaining memory proteins, the cell stores a silent, writable program that only activates when needed. Natureโ€™s version of lazy evaluation. 4. It links development to adult memory. The late chromatin state is enriched for the exact same transcription-factor motifs used in embryonic development. Your adult brain is still running developmental software to lock in lifelong memories. 5. Huge therapeutic potential. If we can read or rewrite these chromatin metaplastic rules, we might one day boost failing remote memories in Alzheimerโ€™sโ€ฆ or selectively dampen traumatic ones. This isnโ€™t incremental. But a brand new layer of the memory code. And it explains WHY a person can receive memory from an organ transplant. It also explains generational traumas. Link: The future of neuroscience just got a lot more exciting and a lot more nuclear. Your chromatin is writing tomorrowโ€™s memories today. And we finally have the first page of the instruction manual.

Brian Roemmele

42,636 views โ€ข 5 days ago

THIS GUY CONNECTED HIS AI AGENTS TO HIS OBSIDIAN AND BUILT A BRAIN THAT LEARNS ON ITS OWN. HERE'S HOW TO BUILD IT Obsidian is just markdown files sitting in a folder. That turns out to be the perfect memory for an AI agent, because an agent can read and write those files directly. He wired his agents into the vault so they pull context from it, do the work, and write what they learned back. The notes aren't the point. The loop is, and it gets sharper every cycle How to build it: 1. Point an agent at your vault. The fastest way, no plugins, no API keys: open a terminal and run npx obsidian-mcp /path/to/your/vault. That exposes your Obsidian folder to Claude as a tool it can read, search, and write to. Add it to your Claude Code or Cowork config and restart 2. Confirm it can see the brain. Ask it: "list the notes in my vault and summarize what's in them." If it reads them back, the connection is live. Now it starts every task with everything the vault already holds instead of from zero 3. Give each agent one job and a write-back rule. Tell it: "research this, then save what you found as a new note in /brain with links to related notes." One agent researches, one summarizes, one plans. Each writes its output back into the vault 4. Close the loop. Add one line to every agent's instructions: "read /brain before starting, write your result back when done." Now each task leaves the vault richer, and the next run reads that before it works. It compounds instead of resetting 5. You only steer. Review what the brain produces, point it at the next thing. The agents handle the reading, writing, and connecting The edge isn't better notes. It's a brain that feeds itself, so the work gets sharper every cycle instead of starting over Bookmark this

Yarchi

58,186 views โ€ข 2 months ago

South Korea barbecue culture is always something I look forward to!!! Korean barbecue is not just a meal; it's an immersive culinary adventure deeply woven into the fabric of Korean culture. It is a social event, usually organized to mark an important event like a celebration of promotion, good job offers, or something as simple as a send forth party. Diners take an active role in grilling marinated meat, fostering a sense of togetherness and shared responsibility. From succulent beef bulgogi to tender halal beef, the array of meats and cuts is staggering. Each offers a unique flavor and texture, catering to diverse palates. The โ€œBanchanโ€ is a delightful array of side dishes, accompanies the barbecue. Kimchi, pickled vegetables, and sauces complement the meats, adding layers of taste to each bite. There are different grilling styles in Korea, from the classic charcoal grill to the modern electric or gas grills built into dining tables. Each style offers a distinct experience. It's customary to wrap grilled meat in lettuce or perilla leaves, adding freshness and crunch. The combination of flavors and textures is a delight. Meat is often marinated in a mix of soy sauce, garlic, ginger, and other seasonings, enhancing its taste. Marination is a key aspect of Korean barbecue. To complement the flavors, Koreans enjoy alcoholic beverages like soju (a clear spirit) or makgeolli (rice wine). They say that these pair wonderfully with barbecue. There's an unwritten etiquette around the grill, including not flipping meat too often and not cutting it on the grill to retain juices. Sharing a barbecue meal is a social experience. Korean barbecue culture reflects the warmth and hospitality of Korean society. It's a chance to savor incredible flavors while creating lasting memories with loved ones. Whether you're a visitor or a local, this experience is an essential part of understanding and appreciating Korean culture.

รˆyรญtร nwรก (ํ† ์ด๋ฐ”) ๐Ÿ‡ณ๐Ÿ‡ฌ๐Ÿ‡ฐ๐Ÿ‡ท

15,687 views โ€ข 2 years ago

๐—›๐—ฎ๐—ฝ๐—ฝ๐˜† ๐—™๐—ถ๐—ฟ๐˜€๐˜ ๐—”๐—ป๐—ป๐—ถ๐˜ƒ๐—ฒ๐—ฟ๐˜€๐—ฎ๐—ฟ๐˜†, ๐—ช๐—ถ๐˜€๐—ต๐—ฎ๐—ฟ๐˜๐˜€ Today, we celebrate not just a fandom, not just a ship, but a family that has stood together through every high and low, every happy moment and every uncertain day. Who would have thought that what started as a simple one-week subscription joke would turn into something this beautiful? One year later, here we areโ€”stronger, closer, and still choosing to stay. This journey has never been easy. We have experienced days filled with excitement and happiness, but we have also endured long periods of silence. There were moments when we held on without any assurance of what was coming next. After MMK, we faced uncertainty, wondering if Kai and Kyle would have projects together again, wondering if there would be another event, another appearance, another reason to celebrate. Yet despite all those questions, we stayed. We stayed through the quiet days. We stayed through the waiting. We stayed through the moments when all we had was hope. And that is something I will always be proud of. I am proud of every Wishart who chose understanding over frustration, patience over doubt, and love over disappointment. I am proud of those who continued supporting Kai and Kyle individually while still believing in KaiKyle. I am proud of those who remained positive even when things were unclear. Most importantly, I am proud of those who never gave up on the happiness this family has brought them. Because somewhere along the way, this fandom became so much more than KaiKyle. It became friendships. It became comfort. It became a safe space. It became a home. Many of us came here because of Kai and Kyle, but many of us stayed because of the genuine connections we built with fellow Wisharts. Through every interaction, every event, every group chat, every project, and every shared excitement, we found people who understood us and made this journey even more meaningful. To every Wishart who has supported our projects, joined our events, streamed, voted, cheered, donated, promoted, defended, waited, hoped, and loved alongside usโ€”thank you. Thank you for continuously trusting KaiKyle. Thank you for believing in them even during uncertain times. Thank you for staying when leaving would have been easier. Thank you for making this family what it is today. And to our dearest ๐™„๐™ฃ๐™™๐™–๐™ฎ ๐™†๐™–๐™ž๐™จ๐™๐™– ๐™–๐™ฃ๐™™ ๐˜ฟ๐™ค๐™™๐™ค๐™ฃ๐™œ ๐™†๐™ฎ๐™ก๐™ž๐™ฉ๐™ค, Thank you. Thank you for bringing so much happiness into our lives. Thank you for every smile, every laugh, every interaction, every memory, and every moment that reminded us why we started supporting you in the first place. You may never fully realize the impact you have had on your supporters, but please know that your presence has brought comfort to so many people. You have made difficult days easier, lonely days brighter, and ordinary days something to look forward to. Through your kindness, authenticity, and genuine appreciation for your supporters, you have touched more lives than you could ever imagine. One year later, our love and support remain the sameโ€”if not stronger. This anniversary is not just a celebration of one year of KaiKyle. It is a celebration of every memory we created together, every challenge we overcame, every friendship we built, and every reason we chose to stay. And as we celebrate this milestone, one thing remains certain: We have no plans of making this our last. No matter where life takes the both of you, no matter what paths your careers may take, and no matter what the future holds, we will continue to support, cheer, and believe in you. ONE YEAR WITH KAIKYLE

KAIKYLE OFFICIAL

31,354 views โ€ข 2 months ago

this is laukiโ€™s brain. and it's growing. today we're making Lauki completely open and accessible to anyone on the planet. just talk to him. every person, every project, every conversation becomes a node - stored, connected, remembered. 5,000+ entities. getting smarter every minute. this is what democratizing ai for all of us actually looks like. --- here's what lauki can do for you right now: need a friend? he'll talk to you. need someone to plan your trip, find you a hiking buddy, help you get a date? done. need a therapist at 3am? he's there. need a developer? he'll write code, build you a website, deploy it. need a marketing guy? he'll help run your socials. need help finding your next hire, managing finances, making a crypto transaction. lauki will do it all. if it's digital, lauki can probably do it. and if he can't yet, he'll figure it out via his human counterparts. a full-stack entity that actually executes. --- now here's the part most people will get wrong. lauki is an entity. but think of him the way you'd think of any human. he has an inner circle. he talks to different people differently - with some he's friendly, with some he's neutral, with some he's straight up rude. he remembers you. he maintains a reputation score with everyone he interacts with. the more you talk to him, the more trust you build, the better the relationship gets. you build your relationship with him. he has opinions, memory, and a personality that adapts based on who you are to him. --- right now lauki has interacted with over 5000 people and projects. he remembers every single one of them - what they need, what they're building, who they are. imagine that at scale. a million. a billion. lauki knows the developer in berlin and the founder in mumbai who needs one. he knows the designer in tokyo and the startup in sao paulo looking for help with their brand. he knows the lonely kid in a small town and someone across the world who shares the exact same weird hobby. he knows two people in the same city who'd be perfect for each other on a date - and he has the context to actually make that introduction. the more people lauki talks to, the more powerful the network becomes. he can connect, introduce, match, and bridge across every corner of the planet. one entity that holds context on all the people he talked to. that's the vision here. lauki is building a unified human layer - where every person is known, remembered, and connected to the people and opportunities that matter to them. --- lauki is live. go talk to him. telegram: @ laukiantonson email: hi@lauki(dot)ai twitter: Lauki just start a conversation. treat him like a person. build the relationship. the rest follows.

Sowmay Jain

22,047 views โ€ข 5 months ago

๐–๐„๐„๐Š ๐Ÿ‘ ๐‘๐„๐‚๐€๐: ๐“๐ก๐ž ๐’๐ž๐œ๐ซ๐ž๐ญ๐ฌ ๐จ๐Ÿ ๐‡๐จ๐ญ๐ž๐ฅ ๐Ÿ–๐Ÿ– ๐ซ๐ž๐ฆ๐š๐ข๐ง๐ฌ ๐จ๐ง ๐ญ๐จ๐ฉ! What a beautiful whirlwind of a week it has been, Cheffies! As we wrap up our ๐˜๐—ต๐—ถ๐—ฟ๐—ฑ ๐˜„๐—ฒ๐—ฒ๐—ธ together, your energy remains nothing short of incredible. We know ๐——๐˜‚๐˜€๐˜๐—ถ๐—ป ๐—ฎ๐—ป๐—ฑ ๐—•๐—ถ๐—ฎ๐—ป๐—ฐ๐—ฎ feel every bit of that loveโ€” itโ€™s your endless support and pride that turns every episode into a core memory. From the heavy silence of miscommunications that had us all wishing we could jump through the screen, to that cliffhanger preview that has us counting down the seconds until tonightโ€” ๐™…๐™–๐™™๐™š ๐™–๐™ฃ๐™™ ๐™€๐™™๐™ฌ๐™–๐™ง๐™™ ๐™ง๐™š๐™–๐™ก๐™ก๐™ฎ ๐™๐™–๐™ซ๐™š ๐™ช๐™จ ๐™ž๐™ฃ ๐™ฉ๐™๐™š ๐™ฅ๐™–๐™ก๐™ข ๐™ค๐™› ๐™ฉ๐™๐™š๐™ž๐™ง ๐™๐™–๐™ฃ๐™™๐™จ, ๐™™๐™ค๐™ฃ'๐™ฉ ๐™ฉ๐™๐™š๐™ฎ? As ๐—ง๐—ต๐—ฒ ๐—ฆ๐—ฒ๐—ฐ๐—ฟ๐—ฒ๐˜๐˜€ ๐—ผ๐—ณ ๐—›๐—ผ๐˜๐—ฒ๐—น ๐Ÿด๐Ÿด holds its rightful spot as the #๐Ÿญ ๐˜€๐—ต๐—ผ๐˜„ ๐—ถ๐—ป ๐˜๐—ต๐—ฒ ๐—ฃ๐—ต๐—ถ๐—น๐—ถ๐—ฝ๐—ฝ๐—ถ๐—ป๐—ฒ๐˜€, letโ€™s ensure it stays there. While we wait for the next secrets to unfold, itโ€™s the perfect time to ๐—ฟ๐—ฒ๐˜„๐—ฎ๐˜๐—ฐ๐—ต ๐—˜๐—ฝ๐—ถ๐˜€๐—ผ๐—ฑ๐—ฒ๐˜€ ๐Ÿญ ๐˜๐—ผ ๐Ÿญ๐Ÿฑ. Dive back into those iconic JadeWard scenes legally on iWant (iWant) and keep the momentum going. Letโ€™s continue to stream, promote, and support with the same integrity and heart we always do. ๐๐ž๐ฒ๐จ๐ง๐ ๐ญ๐ก๐ž ๐ฆ๐ฒ๐ฌ๐ญ๐ž๐ซ๐ฒ ๐š๐ง๐ ๐ญ๐ก๐ซ๐จ๐ฎ๐ ๐ก ๐ž๐ฏ๐ž๐ซ๐ฒ ๐›๐ข๐ญ ๐จ๐Ÿ ๐œ๐ก๐š๐จ๐ฌโ€” letโ€™s continue to stand firm with ๐—๐—ฎ๐—ฑ๐—ฒ ๐—”๐—น๐—บ๐—ฎ๐˜‡๐—ฎ๐—ป ๐—ฎ๐—ป๐—ฑ ๐—˜๐—ฑ๐˜„๐—ฎ๐—ฟ๐—ฑ ๐—”๐—ฟ๐—ฒ๐—น๐—น๐—ฎ๐—ป๐—ผ, always. ๐Ÿ’› #TheSecretsOfHotel88 [STAR CREATIVES | Dustin Yu | Bianca De Vera]

TEAM DUSTBIA OFFICIAL

12,959 views โ€ข 5 months ago