Loading video...

Video Failed to Load

Go Home

This peanut-sized chinese model just dethroned Gemini at reading documents. It’s called glm-ocr. it’s a tiny 0.9b parameter vision-language model that is about to replace every expensive ocr api you use. → Handles text, tables, formulas, handwriting → Scored 94.62 on OmniDocBench V1.5 → 8 languages → vLLM, SGLang,...

37,577 views • 3 months ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

let me explain what Anthropic just did they built an AI model so good at finding security vulnerabilities that they have refused to release it meet Claude Mythos → it’s Anthropic’s newest frontier model and it’s not available to the public. not because it’s not ready. because it’s too dangerous → Mythos found tens of thousands of zero day vulnerabilities across every major operating system and web browser… many of them 1 to 2 decades old. for context… Opus 4.6 found about 500. Mythos found tens of thousands → it found vulnerabilities in the Linux kernel. a 27 year old vulnerability in OpenBSD. a 16 year old vulnerability in FFmpeg → it doesn’t just find bugs. it writes the exploits too. that’s the part that scared them → so instead of releasing it… Anthropic has created Project Glasswing. a cybersecurity initiative where they hand picked 40+ companies to use Mythos for defense only → the partner list reads like a who’s who of tech… Amazon, Apple, Microsoft, Google, Nvidia, Broadcom, Cisco, CrowdStrike, Palo Alto Networks, JPMorgan, the Linux Foundation → Anthropic is giving up to $100 million in usage credits to these partners and $4 million to open source security organizations → they’re briefing CISA and the Commerce Department on how to handle this → the benchmarks are truly insane… Mythos hit 77.8% on SWE-bench Pro where Opus 4.6 scored 53.4%. hit 93.9% on SWE-bench Verified where Opus 4.6 scored 80.8% → Anthropic’s head of frontier red team said this is “the first time a model is this good that we decided to approach release in a very different way” this is the first time an AI company has held back a model because it was too capable not too expensive. not too slow. too dangerous and instead of locking it in a vault they weaponized it for defense and gave it to the companies that run the internet that’s either the most responsible thing an AI company has ever done… or the scariest only time will tell

klöss

21,270 views • 3 months ago

Holy shit... Microsoft open sourced an inference framework that runs a 100B parameter LLM on a single CPU. It's called BitNet. And it does what was supposed to be impossible. No GPU. No cloud. No $10K hardware setup. Just your laptop running a 100-billion parameter model at human reading speed. Here's how it works: Every other LLM stores weights in 32-bit or 16-bit floats. BitNet uses 1.58 bits. Weights are ternary just -1, 0, or +1. That's it. No floats. No expensive matrix math. Pure integer operations your CPU was already built for. The result: - 100B model runs on a single CPU at 5-7 tokens/second - 2.37x to 6.17x faster than llama.cpp on x86 - 82% lower energy consumption on x86 CPUs - 1.37x to 5.07x speedup on ARM (your MacBook) - Memory drops by 16-32x vs full-precision models The wildest part: Accuracy barely moves. BitNet b1.58 2B4T their flagship model was trained on 4 trillion tokens and benchmarks competitively against full-precision models of the same size. The quantization isn't destroying quality. It's just removing the bloat. What this actually means: - Run AI completely offline. Your data never leaves your machine - Deploy LLMs on phones, IoT devices, edge hardware - No more cloud API bills for inference - AI in regions with no reliable internet The model supports ARM and x86. Works on your MacBook, your Linux box, your Windows machine. 27.4K GitHub stars. 2.2K forks. Built by Microsoft Research. 100% Open Source. MIT License.

Guri Singh

2,180,357 views • 4 months ago

Baby talk vs “parentese” - what’s the difference? Often in my replies I see parents explaining why - as a method of promoting language acquisition - they never use “baby talk.” I generally agree with this sentiment, but there are some important distinctions to be made here. If by “baby talk” you mean using cutesy nonsense words (like wa-wa for water or ba-ba for bottle) you’re on the right track. There’s no need for you to make up incorrect or overly simplified vocabulary on your child’s behalf. Use real words, even if your child isn’t quite ready to do so themself. It’s how they learn. But sometimes I see confusion between “baby talk” and what is known as “motherese” or “parentese” - which isn’t nonsensical, but simply slower and more varied in intonation. (Think of Ms. Rachel’s sing-songy voice.) And this isn’t something you need to shy away from at all. In fact, research suggests that parentese - with its prolonged vowel sounds and expressive facial expressions - can be a social hook that attracts children’s attention, encouraging them to attune not only to the language to which they are being exposed but how it is produced. Importantly, it’s complete and grammatically correct… just a little more performative than you might use elsewhere. This lovely video, shared to IG by tommypadula, is a nice example. Everything mom says is 100% correct… it’s simply exaggerated in ways that are clearly capturing her daughter’s rapt attention. Just look at the smiles and eye contact it’s attracting. This isn’t baby talk. It’s parentese. And it’s fantastic.

Dan Wuori

28,726 views • 11 months ago

Yes, Gen Z, we get it—you don’t care about data privacy. But that’s just a tiny part of the story. The real issue here is control. The Chinese Communist Party is using TikTok to build a blueprint of influence over you and our society. They’re not just tracking what you do; they’re deciding what you see, and therefore *how you think*. There is plenty of evidence this is already happening. Now imagine how much more powerful it would be as an information warfare weapon when the proverbial shit hits the fan in a potential conflict over Taiwan. Yeah, the Chinese will take control of all advanced semi-conductor chips in Taiwan (the ones that power your precious smartphone), and then use your addiction to TikTok to convince you it’s all a good idea. TikTok, as an entertainment platform, is an incredible product. It’s fun, it’s creative, and millions of Americans love using it. But we also need to recognize that it’s a Trojan horse. And right now, that horse is carrying Chinese Communist Party influence into our backyard and sending American’s data back to Beijing. That’s the problem we need to solve. The goal here is not about trying to stop American small businesses who use TikTok to promote themselves, the latest dances, or end “brain rot” trends—though maybe fixing that wouldn’t hurt either. It’s about making sure that enjoyment doesn’t come at the cost of our privacy, our freedom, or our national security. There’s a way to fix this. We don’t need to kill TikTok; we need to kick the Chinese Communist Party out of it. Ownership by someone who shares—or at least respects—our democratic values is the solution. People like Kevin O’Leary from Shark Tank and Dodgers owner Frank McCourt are already exploring a potential acquisition. But the fact that TikTok seems unwilling to even consider this solution tells us all we need to know about its true purpose.

Dan Crenshaw

139,497 views • 1 year ago