正在加载视频...

视频加载失败

THIS DEVELOPER RAN THE LARGEST AI MODEL IN THE WORLD ON 5 MAC STUDIOS - AND IT COST 100X LESS THAN WHAT OPENAI USES 27:47 he says it after hours of setup - Llama 3.1 405B running locally on five Mac Studios - a model that normally requires 42...

21,553 次观看 • 1 个月前 •via X (Twitter)

0 条评论

暂无评论

原始帖子的评论将显示在这里

相关视频

People made fun of Alex Finn for buying three Mac Studios to run AI at home. Then Fable got banned for a week, GLM 5.2 dropped, and those exact Mac Studios started reselling for 4x what he paid. He showed me how he built his home AI lab from scratch. Here's the playbook: 1) The hardware. three 512GB Mac Studios, an NVIDIA DGX Spark, a custom RTX 5090 build, and a few Mac Minis. ~$30k all in. 2) The buying framework... - Mac Studio: huge memory, runs GLM 5.2 (open weights, near Opus 4.8 on benchmarks), but slow. - DGX Spark ($4,800): the sweet spot for most people. - RTX 5090: smaller models at blazing speed (Qwen's 29B now hits Sonnet 4 level). 3) Tailscale networks every machine into one private network with root access to each other. Only one machine is plugged into a monitor. 4) A Nous Research Hermes agent is his IT guy. New model drops? It SSHs into the right box, loads 5 candidates, runs evals overnight, and reports back which task belongs on which machine. Alex has literally never loaded a model himself. 5) The whole point: achieving "ambient intelligence." Always-on jobs that would bankrupt you on per-token billing. A security sweep of his API endpoints every hour. Code optimization every 20 minutes. Database anomaly & churn detection. Hourly scraping of X, Reddit & Hacker News for business opportunities. 6) Running those workloads on frontier models would cost thousands a month. His actual cost: ~$60 more in electricity. 7) Btw he's not anti-frontier. He still maxes out his Claude plan. The way he sees it: frontier is for hard thinking, local is for the foot soldiers that never sleep. 8) "We own everything except for the intelligence. Why can't we own the intelligence?" 9) He thinks frontier-level intelligence runs on consumer hardware within 6 months.

Alex Lieberman

57,764 次观看 • 1 个月前

🚨PERPLEXITY JUST LAUNCHED SOMETHING THAT MAKES EVERY OTHER AI PRODUCT LOOK LIKE A TOY.. AND NOBODY IS TALKING ABOUT IT.. They built a Personal Computer.. Not an app.. Not a chatbot.. A full digital worker that runs 24/7 on a Mac mini even while you sleep.. You press both command keys.. And it wakes up.. Ready to work.. But here's where it gets insane.. This thing doesn't run on one AI model.. It runs on 19 of them.. At the same time.. It uses Claude Opus for complex reasoning.. Gemini 3.1 Pro for deep research with a 2 million token context window.. Nano Banana Pro for 4K images.. Grok for fast tasks.. It doesn't just pick one model and hope for the best.. It reads your task.. Breaks it into subtasks.. And routes each one to whichever model is best at that specific thing.. All running in parallel.. While ChatGPT is still thinking about your first question.. Perplexity has already split your project into 6 pieces and assigned each one to a different AI.. And here's the part that should worry OpenAI.. Perplexity hallucinates at 3.3%.. ChatGPT hallucinates at 12%.. Claude at 15%.. It's not even close.. Because Perplexity is built differently.. Every other AI tries to remember facts.. Perplexity searches for them first.. It's structurally forced to cite live sources before it's even allowed to generate a response.. OpenAI Operator launched with a 32.6% success rate on computer-use tasks.. People called it "the world's most anxious intern" because it pauses every 5 seconds to ask if it's doing the right thing.. Perplexity runs multi-hour and multi-day workflows independently.. Only interrupts you when it hits a decision that actually matters.. You can start a task from your iPhone on the train.. And it executes on your Mac mini at home.. The economics are wild too.. Internal studies show it saved teams an average of $1.6 million in labor costs.. Performing 3.25 years of work in four weeks.. And unlike every other AI company.. Perplexity dropped ads entirely.. They charge $200 a month because they said they're in the "accuracy business".. Not the advertising business.. They even launched a $42.5 million publisher program to pay media partners when their content gets cited.. While OpenAI is getting sued by every newspaper on earth.. Google and OpenAI want you locked into their ecosystem.. If a better model comes out tomorrow you're stuck.. Perplexity just updates its routing matrix.. You get the best model on earth automatically.. No switching.. No migrations.. No friction.. This isn't an AI assistant anymore.. This is the first real AI employee.. And it costs $200 a month.

Evan Luthra

1,097,438 次观看 • 4 个月前