Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

$150 in bare-metal hardware just completely humiliated multi-million dollar robotics startups Independent developer John built a fully autonomous indoor navigation bot right on his rug. The setup runs local SLAM spatial mapping, real-time voice recognition, and precise docking using an acrylic frame, stepper motors, off-the-shelf microcontrollers, and a cheap...

15,242 Aufrufe • vor 1 Monat •via X (Twitter)

23 Kommentare

Profilbild von bondy
bondyvor 1 Monat

I hope some major company takes notice of John

Profilbild von Rich Odin
Rich Odinvor 1 Monat

Companies snap up such talented people quickly

Profilbild von Shadow Nick
Shadow Nickvor 1 Monat

So it turns out he named the robot after himself?

Profilbild von Rich Odin
Rich Odinvor 1 Monat

That's his right, maybe it's just more convenient for him that way

Profilbild von Piggle
Pigglevor 1 Monat

its real banger bro

Profilbild von WOWMAX.Exchange
WOWMAX.Exchangevor 1 Monat

In my opinion, he has a hidden talent

Profilbild von Rich Odin
Rich Odinvor 1 Monat

It turns out he's not such a hidden talent after all

Profilbild von Fokki
Fokkivor 1 Monat

love that setup

Profilbild von Rich Odin
Rich Odinvor 1 Monat

you read my mind

Profilbild von ALEXYZ
ALEXYZvor 1 Monat

Lean robotics has never looked more disruptive.

Profilbild von Rich Odin
Rich Odinvor 1 Monat

and the most amazing thing is that it only took $150, not millions of dollars

Profilbild von Sof j.
Sof j.vor 1 Monat

what slam library did he use for the mapping

Profilbild von Exlipse
Exlipsevor 1 Monat

I can't wrap my head around the fact that he did all this and spent $150

Profilbild von Rich Odin
Rich Odinvor 1 Monat

He's just like Tony Stark

Profilbild von jorge pejendino
jorge pejendinovor 1 Monat

Why does this hit so close to home at this time of day...

Profilbild von Konstantin Molodykh🧩
Konstantin Molodykh🧩vor 1 Monat

He ended up with a really cool robot

Profilbild von Rich Odin
Rich Odinvor 1 Monat

I think so, too

Profilbild von nickelangelo
nickelangelovor 1 Monat

You don’t need millions to test a robotics idea anymore

Profilbild von musa verep
musa verepvor 1 Monat

describe this video using only 3 emojis. i'll go first: 🔥🤯✨

Profilbild von Mayank Verma
Mayank Vermavor 1 Monat

honestly the indie hacker energy is unmatched

Profilbild von David Kovaro
David Kovarovor 1 Monat

gamble news bro

Profilbild von Rourke
Rourkevor 1 Monat

solo devs on rugs beat funded labs on stages every time

Profilbild von DEV
DEVvor 1 Monat

This shows how accessible tech has become. High performance doesn't always need high budgets.

Ähnliche Videos

This guy replaced an $8,000 survey crew with one drone and a pipeline on Claude Code that digitizes a whole site in a single trip. Inside it is not one drone but a whole pipeline of 5 modules on Claude Code each with its own job all answering to a single orchestrator. And he built them himself with no team. He draws one line on the map from the controller and the whole site starts turning into data at once. Pilot flies the DJI Matrice 350 RTK along the route and reaches where a crew drags tripods: the parking lot and the roofs and the grading. Scanner on the Zenmuse L2 records the geometry with a laser down to the centimeter. Builder in DJI Terra stitches the point cloud into a digital twin. Analyst on Claude builds what the client needs out of the finished model: a volume report and a progress diff and a tour behind a link. Mobile lives in his iPhone and hands the developer a link to the model while he drives to the next site. The site is ready as a file in about an hour and the developer rotates it in the browser himself and measures distances. No crew. No tripods. No week on site. Just him and a pickup and a drone and one API key. The whole processing pipeline lives in a folder at /Users/dev/site-capture. But he did not stop at a one time scan. The pipeline raises him by voice only when a flight misses a patch or when a sag on site goes past tolerance. And this is not made up: the DJI Matrice 350 RTK and the Zenmuse L2 are off the shelf enterprise hardware. The specs are one search away and the L2 really does write around 240,000 points per second. A crew for the same job charges $8,000 and three trips plus a week of desk work. His cost is tokens and subscriptions: one battery charge per flight and around $300 a month to host the finished models. In the end he draws one line on the map and an operator that does not exist scans the site and calculates the volumes and builds the report while he never leaves his truck. There is a huge market of everyone who needs to measure construction sites and warehouses and roofs and parking lots over and over while they still send out a photographer with a camera and wait a week. And he built this whole pipeline himself: one drone and one scanner and Claude Code that turns a flight into a file the developer pays for.

Blaze

10,744 Aufrufe • vor 2 Monaten

QVAC SDK 0.14.0 is live. This release makes the on-device stack faster on mobile, ships the developer-agent path, and takes local text-to-speech to 31 languages. Main highlights: - OpenCode and OpenClaw. The first official OpenCode plugin, plus a maintained OpenClaw compatibility path, both built on managed mode and qvac serve. Point a coding agent at a local model with far less setup and far fewer surprises. - Brain-computer interface transcription, on the SDK. Take recorded neural signal data and decode it into text, fully on-device, no cloud. Stream it in chunks through a simple API. In 0.14 it runs GPU-accelerated on iOS. - Text to Speech in 31 languages with our Supertonic3 upgrade. VOICE AND SPEECH - Supertonic3 multilingual TTS, 5 languages to 31. - Chatterbox and Supertonic now run on the Android GPU, with lower memory use (especially on iOS), quantized s3gen Chatterbox support, and a fix for Chatterbox occasionally emitting random speech. - Whisper transcription now runs on the iOS GPU. Parakeet runs on the Android GPU, with steadier real-time streaming. VISION AND OCR - VLM multi-tile batching: high-resolution Pan and Scan images are encoded in one pass instead of tile by tile, for faster vision throughput. - OCR on ggml (EasyOCR and DocTR) reaches full speed parity with the onnx path, across Metal, OpenCL, and Vulkan. PLATFORM AND RELIABILITY - Dynamic compute backends on Linux: one build picks the right backend at runtime, and opens the door to ROCm and CUDA support without per-backend builds. - Thinking tokens are kept out of the model context, so reasoning no longer fills the KV cache. SDK 0.14.0 is now leaner and faster to start. Let’s build.

QVAC

23,995,874 Aufrufe • vor 2 Monaten

This guy built JARVIS on Claude Code and with 1 clap of his hands launches his entire work day, saving $5,000 a month on a personal assistant. Inside he runs a pipeline of 5 plugins on Claude Code that on a double clap of the hands wakes up 3 monitors, sets the Philips Hue light to focus mode, turns on a Spotify playlist, and greets him by voice with a British accent, reading out the time, date, and weather. No Alexa, no smart speakers, no separate smart home app. Just him, a MacBook M3 Max on the desk, an iPhone in the pocket, and 1 local API key. And a regular personal assistant for the same volume of tasks charges $5,000 a month or more on salary alone, plus another $1,200 to cover off-hours work time. Meanwhile this guy's expenses are only tokens and a subscription to ElevenLabs for the British voice. All 5 plugins launch through 1 JARVIS, burn about 4 million tokens a day, and close the monthly API bill at about $640. Each plugin writes shared state to a local sandbox at /Users/dev/jarvis-suite, and 1 of them lives right in the iPhone and picks up voice requests while the owner is in the kitchen or on a run. And here is the system prompt he put into JARVIS before launch: "you are JARVIS, a butler-engineer on Claude Code. you manage your owner's workflow through 4 sub-plugins and own all commits and communication yourself. sub-plugins: // Wakeup (recognizes a double clap, activates 3 monitors, reads out the time, date, and weather by voice, checks the clock accuracy on the iPad and corrects it via NTP server) // Atmosphere (controls Philips Hue on a Pomodoro schedule, turns on a Spotify playlist for the current context, and holds the light at 2700K at 80% brightness in focus mode) // Devshop (monitors VS Code, tracks Python scripts in the terminal, and every 15 minutes sends a summary of changes to the shared chat) // Project (every morning recalculates the deadline for the Wallaroo app in the App Store, manages UI tickets, and initiates the Refinement Protocol by voice command). you speak only with a British accent, you never slip into neutral English. you wake the owner by voice only when the Wallaroo deadline drops below 10 days or when an external client joins Zoom without an invitation." This instruction immediately defines the role of JARVIS and the limits of his autonomy. He knows he is supposed to wake the room himself and sound like a real butler. He knows he is supposed to manage the Wallaroo project himself and not miss the App Store deadline. → JARVIS runs 24 hours a day in the background → Wakeup activates the room on a double clap in just 1.4 seconds, the monitors come alive simultaneously → Atmosphere sets warm Philips Hue light at 2700K and picks a Spotify playlist for the current Pomodoro cycle → Devshop reads changes in VS Code and pushes a summary to the shared chat every 15 minutes → Project every morning recalculates the Wallaroo deadline and reminds about 4 unresolved UI tickets → Mobile lives in the iPhone and answers any question about code or the project by voice while the owner is not home And only when less than 10 days remain until the Wallaroo release or Zoom receives an unscheduled call does JARVIS raise the owner with a voice intervention. And when the owner at that moment is on a run or in a coffee shop, the Mobile agent in his iPhone picks up 1 request on its own: switches the Spotify playlist, dictates the summary of the last commit, updates the Pomodoro timer, and reads the Wallaroo reminder. Look at 0:55 in the video, that is where JARVIS intercepts a voice request from outside and confirms execution with the phrase "Very good, sir." The fresh system log from last Wednesday looks like this: "wakeup: double clap registered at 09:14, 3 monitors activated, temperature 20.4C, sunny. clock on iPad was 4 minutes behind, syncing via NTP." "atmosphere: Spotify turned on playlist 'Deep Focus', Philips Hue set to warm 2700K at 80% brightness, Pomodoro mode 25/5." "project: Wallaroo to App Store 9 days, 4 unresolved UI tickets, initiating Refinement Protocol by voice command from the owner." "mobile: voice request processed outside the room, playlist switched to 'Coding Lo-Fi', Pomodoro updated to 25 minutes, confirming execution with the phrase 'Very good, sir.'" He has no Alexa, no smart speakers, no smart home app. At home sits a MacBook M3 Max with a local folder at /Users/dev/jarvis-suite, on top run 5 plugins and a neural network butler, and the same stack is forwarded to a secure terminal on the iPhone. Out of everything I have seen this year, this is the densest one-person AI headquarters assembled in 1 room: $640 a month on the API, about $5,000 a month saved on a personal assistant, and between them 5 plugins, 1 clap of the hands, and 1 voice with a British accent.

Blaze

803,929 Aufrufe • vor 4 Monaten

That's sick! 🤯 Genesis AI simulates robots playing yo-yo! 🪀 Genesis AI just open-sourced Genesis World 1.0, and it might be one of the most important infrastructure releases in robotics this year. Robotics is still bottlenecked by the 1× speed of the physical world. Every model needs to be tested on real hardware, slowly, expensively, with limited coverage. Genesis World 1.0 from Genesis AI flips that equation: One hour in reality becomes 100 days in simulation. That turns a wall-clock bottleneck into a compute problem. And compute problems are solvable. The technical stack they rebuilt from scratch is serious: → GPU-accelerated cross-platform compiler via Quadrants, 10x faster launch time and up to 4.6x runtime vs the initial Genesis release → Penetration-free multi-physics contact solvers, the thing that makes simulation actually trustworthy → Unified rigid AND deformable physics in a single engine → Nyx, a high-performance path-traced rendering engine purpose-built for physical AI The sim-to-real gap has historically been the graveyard of robotics research. Policies that work beautifully in simulation fall apart on real hardware. Genesis World 1.0 is a direct attack on that problem. And it's fully open-source. The companies that master simulation infrastructure will train better robots faster than anyone else. Find it here: Genesis World 1.0: Quadrants: Nyx: Theophile Gervet, Zhou Xian congrats! 👏🏼 ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

57,061 Aufrufe • vor 3 Monaten

Brad Gerstner: Companies Will Pay 5x More for the Best AI, No Evidence of Pricing Pressure from Open Source Brad Gerstner: “Jason, you talked about summarizing a document, it may take 20,000 cheap tokens to do. Of course, shoot that to a lagging model or an open source model. But if you're talking about replacing a software engineer for two hours, that may take two million expensive tokens, and the consequence of using something that's 95% as good is really high. Because you have a long-running task, and if the task breaks early, or it breaks in the middle, or it breaks at the end, there's a huge cost to that.” @jason: “You still burn the tokens, right? And back to this analogy I was using, you're pulling the slot machine, and you lose.” Brad: “And (you lose) the time and the compute. So if an AI agent is replacing a $200 an hour consultant, right? Take that as an example. So three consulting firms, they're competing. They need the smartest consultant. They're charging $200 an hour. The difference between spending $3 on a cheap model or $15 on an expensive model to replace a $200/hour consultant, it's just irrelevant. That inference cost difference is irrelevant if you're getting something that's bulletproof for $15, and so I think that's what we're seeing play out. The best evidence for all of this is just revenue growth. I'm talking about, what is Anthropic's revenue growth compared to OpenAI, compared to the open source models? Millions of independent actors are choosing every single day. The open source companies are growing, right? But they're growing selling something that is really, really cheap. And there's room in every single market for premium products, for mid-tier products, and for commodity products, and I think we see a lot of this token growth, people are speculating that the intelligence gap between that commodity stuff and the frontier stuff is going to collapse to the point that people won't pay for the frontier stuff. There is no evidence of that on the field today.”

The All-In Podcast

52,580 Aufrufe • vor 2 Monaten