Loading video...

Video Failed to Load

Go Home

A 27 YEAR OLD BANGALORE DEV BOUGHT A DECOMMISSIONED BANK SERVER RACK FOR $3,200 AT AUCTION AND NOW PULLS $24,000 A MONTH FINE TUNING LLMs FOR US SAAS COMPANIES raj is 27, bangalore, one bedroom flat with a ceiling fan, won the lot off a state bank IT refresh...

1,191,544 views • 3 months ago •via X (Twitter)

33 Comments

Samar's profile picture
Samar3 months ago

Utter BS !

Alex Rubenstein's profile picture
Alex Rubenstein3 months ago

I don't think that decommissioned bank server will be able to handle this many requests, considering the server doesn't usually have GPUs.

Calder's profile picture
Calder3 months ago

a bank server rack in a one bedroom flat means the kid is sleeping next to industrial cooling fans, that is the actual story

Mayur Panghaal's profile picture
Mayur Panghaal3 months ago

Claim 4: Fine-tuning LLMs for 11 US SaaS companies from a flat Possible. Many AI consultants operate solo. However, "11 US SaaS clients" sounds like a marketing number. Questions you'd want answered: What kind of fine-tuning?LoRA? QLoRA? Full fine-tuning? How large are datasets? How often are jobs run? Are these actual companies or one-off Upwork gigs? The claim is not impossible, but lacks evidence. Credibility: Medium.

Uves Arshad (montr.io)'s profile picture
Uves Arshad (montr.io)3 months ago

Time to surf bank auction sites 🌊

Aryan Singh's profile picture
Aryan Singh3 months ago

You mean this guy and not some random 27 yr old Raj. Get your facts correct please

Mayur Panghaal's profile picture
Mayur Panghaal3 months ago

A 70B model is large. Two RTX 3090s provide: 24 GB VRAM each 48 GB total VRAM A quantized 70B model can run across multiple GPUs and system RAM, but performance depends heavily on: quantization level inference engine context size throughput requirements For experimentation, yes. For serving 11 paying US SaaS clients simultaneously, that's much harder. The story omits critical details: What quantization? What tokens/sec? What uptime? What SLA?

Anish Bhanushali's profile picture
Anish Bhanushali3 months ago

Bescom will render this hardware useless with voltage fluctuations and powercuts before it can process a single token

Animesh Singh's profile picture
Animesh Singh3 months ago

Abhi ruko zara, sitaraman ji khud aayengi tax lene.

Drake's profile picture
Drake3 months ago

wait, can you actually buy a decommissioned bank rack at auction in india, is that public information?

Liquidden's profile picture
Liquidden3 months ago

anyone on the timeline actually running an old bank rack at home, what is the noise level?

Mayur Panghaal's profile picture
Mayur Panghaal3 months ago

@grok are these claims credible ? Can two custom towers with RTX 3090s under the desk, runs llama 3.3 70B smoothly ?

Ajay Prabhu's profile picture
Ajay Prabhu3 months ago

@MoShahx07 Let me sound like an uncle and say : seekho kuch 😬😬

Ai LLM Peasant's profile picture
Ai LLM Peasant3 months ago

When they want to sell the old crap but still want to increase the demand in order to sell at highest prices. Enter the fake stories and fuuusss hype !

Leet's profile picture
Leet3 months ago

idk why these retards compare llama 3.1 70B to claude models, your shitty local model is not even close to any claude model not even haiku

KarmaWarrior's profile picture
KarmaWarrior3 months ago

BS pro max 😅

Rakita 🧂's profile picture
Rakita 🧂3 months ago

It’s the crypto mining from your room hype all over again. This will definitely kill AI for sure

sanjay patel's profile picture
sanjay patel3 months ago

Can anyone explain in simple words what is this ?

Esha Gavaskar's profile picture
Esha Gavaskar3 months ago

Any proof that he makes that much.. or just storytelling

Hana Trần's profile picture
Hana Trần3 months ago

whoa, 24k a month just fine tuning llms? 😮

Ash_kay's profile picture
Ash_kay3 months ago

i dont that is true. electricity prices in India are quite high

Thanveer Karim's profile picture
Thanveer Karim3 months ago

Over engineered piece of garbage , nowadays a single dgx spark or a mac studio can do much more

Paarivaala's profile picture
Paarivaala3 months ago

I have run qwen 3.6 28b mixture of experts model 128k context on a laptop RTX 4060 8gb with turboquant and well optimised using llama.cpp with 33 t/s , it works well but everything goes to trash when it’s about agentic coding. You’ll again go back to cloud based frontier models.

rohithje.near's profile picture
rohithje.near3 months ago

@NammaBESCOM doing this without commercial Meter look at this scam

Dinesh Ram 🇮🇳's profile picture
Dinesh Ram 🇮🇳3 months ago

Even AI makes better engagement clips, than this fake bullshit

Md. Asaad Sayed 🪽's profile picture
Md. Asaad Sayed 🪽3 months ago

Homelab goals 🛐

Tsera's profile picture
Tsera3 months ago

Rate limits are dead

Kong Trading 🦍's profile picture
Kong Trading 🦍3 months ago

Decommissioned bank servers running LLMs is the best hustle

ArŰSH's profile picture
ArŰSH3 months ago

People are really today just dumb or what?

Chetan Kale's profile picture
Chetan Kale3 months ago

insane

Firas D's profile picture
Firas D3 months ago

Who is paying for fine tuned versions of llama 3.3 70b? Come on man

Kushagra Keshari's profile picture
Kushagra Keshari3 months ago

There is a voltage stabilizer but no power UPS or BESS? 😂

Ritik Verma's profile picture
Ritik Verma3 months ago

"The server rack is the easy part. The monthly invoices are the hard part."

Related Videos

A USED $700 RTX 3090 HAS MORE MEMORY AND FAR MORE SPEED THAN THE ~$1,000 A2 GOING INTO THIS SERVER. THE A2 STILL WINS - FOR ONE REASON THAT ISN'T PERFORMANCE. that clip is a dell poweredge r760 - a 2u rack server - getting an nvidia a2 dropped in through a riser. this is the enterprise route to local ai. the rung the desktop ladder skips entirely. the a2, verified: 16gb of memory, single-slot, low-profile, and just 40 to 60 watts it's an inference card, built to sit in a server that has no room or spare power for a real gpu so why not a 3090? a used rtx 3090 gives you 24gb and many times the compute for around $700. the a2 gives 16gb, much slower, for roughly a thousand. on raw local-llm value, the desktop card wins, and it isn't close. the a2's whole reason to exist is the thing you can't see on a spec sheet: it fits. a 2u server has no spare gpu power and no room for a triple-slot furnace. the a2 slips into one low-profile slot on 60 watts and lets an existing server do inference without a rebuild. the honest caveats: the r760 is not a desk machine. loaded, it idles at hundreds of watts, and its fans sound like a hair dryer that never turns off. it belongs in a rack or a closet 16gb is still 16gb. the same 7b-to-32b ceiling as the cheap cards. the enterprise badge doesn't buy you a bigger model what it does buy: ecc memory, redundant power, remote management, hot-swap everything. reliability, not speed so who this is for: someone who already runs a rack and wants to add private inference without touching the power or cooling budget. for anyone starting from a desk, the mac mini or the 3090 wins on every axis that matters. the real point: the "best" local ai box depends entirely on what you already own. on a desk, efficiency wins. in a rack, fit and reliability win. the a2 isn't a bad card - it's a card for a constraint the ladder never mentions. no fastest gpu here, no desk-friendly box, no bigger model than the cheap cards already run. save this before you buy enterprise gear for a desktop job.

Grimmer

24,287 views • 2 months ago