Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

YOU DON'T NEED A SERVER RACK. SIX USED MINI PCS AND A $60 CISCO SWITCH WIRE INTO A HOME CLUSTER ON YOUR DESK. that clip is six tiny mini-pcs stacked and patched into one cisco business 110 switch, red cables, a single power strip. each box is a full...

95,267 Aufrufe • vor 2 Monaten •via X (Twitter)

10 Kommentare

Profilbild von Chris Thompson
Chris Thompsonvor 2 Monaten

I have done this with mini elitedesks. I recommend building a little rack for them because they get quite hot when stacked

Profilbild von ViceSol
ViceSolvor 2 Monaten

This is one of the coolest things about homelabs. You don't need enterprise hardware to learn enterprise concepts - just a few inexpensive machines and curiosity

Profilbild von tech is cool
tech is coolvor 2 Monaten

Yes! My old surface book pro is running as an appliance for Ubuntu, synapse OS + Application, and hosting Ollama. Is a local home AI studio

Profilbild von Dilara Guller
Dilara Gullervor 2 Monaten

😀😀😀

Profilbild von darknessblade
darknessbladevor 2 Monaten

One thing you want to do is, open them up, and take out the WIFI card/antenna's. That way you have less hardware faults to deal with. And set it up so it can boot by LAN.

Profilbild von Nekt0
Nekt0vor 2 Monaten

pull is better than persuasion

Profilbild von Ashar Rai Mujeeb
Ashar Rai Mujeebvor 2 Monaten

Fun fact: this is how Steven Spielberg did the CGI for titanic.

Profilbild von Loweffortusername
Loweffortusernamevor 2 Monaten

I actually got a used server for pretty cheap, they're not always as bad as you'd think, but I do like this idea for a project!

Profilbild von Elon Orbit
Elon Orbitvor 2 Monaten

I love the setup. Real SPOF is that power strip though. Redundant nodes mean squat when the strip dies. Still beats cloud rent.

Profilbild von Akshay Ravirala | Data & AI Systems
Akshay Ravirala | Data & AI Systemsvor 2 Monaten

The $60 switch is never the bottleneck on a cluster like this — it's the power draw at 24/7 uptime that changes the math. Still a great way to hit k8s failure modes (node death, split-brain, noisy-neighbor) you'd never trigger on purpose in prod.

Ähnliche Videos

A USED $700 RTX 3090 HAS MORE MEMORY AND FAR MORE SPEED THAN THE ~$1,000 A2 GOING INTO THIS SERVER. THE A2 STILL WINS - FOR ONE REASON THAT ISN'T PERFORMANCE. that clip is a dell poweredge r760 - a 2u rack server - getting an nvidia a2 dropped in through a riser. this is the enterprise route to local ai. the rung the desktop ladder skips entirely. the a2, verified: 16gb of memory, single-slot, low-profile, and just 40 to 60 watts it's an inference card, built to sit in a server that has no room or spare power for a real gpu so why not a 3090? a used rtx 3090 gives you 24gb and many times the compute for around $700. the a2 gives 16gb, much slower, for roughly a thousand. on raw local-llm value, the desktop card wins, and it isn't close. the a2's whole reason to exist is the thing you can't see on a spec sheet: it fits. a 2u server has no spare gpu power and no room for a triple-slot furnace. the a2 slips into one low-profile slot on 60 watts and lets an existing server do inference without a rebuild. the honest caveats: the r760 is not a desk machine. loaded, it idles at hundreds of watts, and its fans sound like a hair dryer that never turns off. it belongs in a rack or a closet 16gb is still 16gb. the same 7b-to-32b ceiling as the cheap cards. the enterprise badge doesn't buy you a bigger model what it does buy: ecc memory, redundant power, remote management, hot-swap everything. reliability, not speed so who this is for: someone who already runs a rack and wants to add private inference without touching the power or cooling budget. for anyone starting from a desk, the mac mini or the 3090 wins on every axis that matters. the real point: the "best" local ai box depends entirely on what you already own. on a desk, efficiency wins. in a rack, fit and reliability win. the a2 isn't a bad card - it's a card for a constraint the ladder never mentions. no fastest gpu here, no desk-friendly box, no bigger model than the cheap cards already run. save this before you buy enterprise gear for a desktop job.

Grimmer

24,287 Aufrufe • vor 2 Monaten

FOUR DIFFERENT VENDORS ARE NOW SHIPPING GB10 MINI PCs WITH 128GB UNIFIED MEMORY, AND ONE MICROTIK CRS 804 SWITCH CAN CONNECT UP TO EIGHT OF THEM INTO A 1 TERABYTE LOCAL AI CLUSTER 00:00 he points at the MikroTik CRS 804, "you need some kind of switch that'll handle QSFP56 ports like these", the interconnect that makes the whole cluster possible the GB10 ecosystem is no longer just Nvidia. Dell Pro Max GB10, ASUS Ascent GX10, and MSI Edge Expert all ship the same Grace Blackwell Superchip with 128GB of coherent memory. same silicon, different cases, same 200 gigabit ports on the back the CRS 804 is what connects them at prosumer prices. four 400 gigabit QSFP56 ports on one 1U chassis, breakout cables that split each port into two 200 gigabit lanes. one switch drives eight GB10 units in parallel do the math. eight nodes at 128GB each equals 1024GB of pooled unified memory across the cluster. run vLLM, shard a frontier model across all eight, and inference happens locally on hardware that fits in half a rack the real limiter revealed in the stress test was never throttling. it was interconnect topology, exactly the layer this switch fixes at a fraction of enterprise switch pricing $400 a month for combined chatgpt pro and claude code max hits $4,800 a year per developer. a small team of five running through this cluster pays back inside eight months and never expires the article covers the buying ladder for a single desk. this post is proof of the cluster ladder that starts where the desk one ends save this before the GB10 lineup grows past four vendors and prosumer cluster switches move upmarket

NO1ennn

59,725 Aufrufe • vor 3 Monaten

Someone just posted the full blueprint for an AI swarm that does the job of a 200-person quant research team. Six agents. Running 24/7. Finding brand-new alpha while you sleep. Citadel needs 100 PhDs to do this. Two Sigma needs 200. This does it with six bots and one laptop. Two ways to play this - spend a weekend building your own swarm, or copy the wallet of one that's already up $2M: Boris Cherny runs Claude Code at Anthropic. Two weeks ago he said: "I don't prompt Claude anymore. I have loops running that prompt Claude. My job is to write loops" Alpha research is just a pipeline. So instead of sitting in it, you hand each stage to its own agent: > one reads every new research paper overnight and pulls out the trade idea > one builds the features and cleans the data > one backtests it over 20 years, costs and slippage included > one runs the hard stats and kills anything overfit > one checks it still works in every market regime > one strips out plain momentum and value to see if any real edge is left Each of those six is a job a fund pays a $600,000-a-year quant to do. He runs all six for the price of an API bill. The rule that makes it work: the agent that builds a signal never gets to approve it. A separate, stronger agent tries to kill it first. Whatever survives all six by morning is real, new alpha. One trader's already running this exact swarm on Polymarket. That $2M wallet is public, every trade on-chain. The full build is in the post below - six agents, the tool that runs them, and the five mistakes that kill most people. Bookmark & read this before it's buried.

cvxv666

104,237 Aufrufe • vor 3 Monaten

HOLY...THE BOTS SCOOPED BLOOMBERG SIX GROK AGENTS PUBLISHED A HEADLINE 11 MINUTES BEFORE THE TERMINAL AND TRADED IT FOR $2,860 WHILE I SLEPT I turned Grok Bot into a newsroom. Not a chatbot that summarizes news. Six desks that find it, check it, write it, trade it and sell it, in that order, with an empty chair at the top. Why it works: Grok sees every post on this app the moment it exists. The terminal sees it after a human picks up the phone. That gap is eleven minutes on a good night, and a newsroom that never sleeps can live inside it. > MARA reads 1,900 posts a minute -> flags anything from a real place with a real job title. > OTIS needs two independent sources -> one post is a rumor, a post plus a document is a story. > JUNE writes 40 words, zero adjectives -> the headline the terminal will write later. > REX trades the gap -> 4% of the desk, entry 0.31, no approval step. > PIP fires the alert to 212 paying readers -> $19 a month each. > CHIEF holds the kill switch -> the only human decision is whether the desk keeps running. One night, off the log. 04:12 a crane operator in Rotterdam posts that the east quay went dark. 04:16 two sources. 04:19 headline. 04:21 position. 04:23 alert out. 04:34 the terminal catches up. 04:41 the market sits at 0.74. $2,860 on one story. $4,028 a month from readers who want the next one eleven minutes early. One evening to set up, no code. Six bots on one cloud computer, one skill each, one group chat where the story moves down the building. A newsroom is not writers. It is a rule about sources and a chair that stays empty. What should the desk cover next?

slash1s

12,091 Aufrufe • vor 1 Monat