Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Full-circle Test-driven Firmware Development with OpenClaw Ladyada: "I've only had OpenClaw installed on this Raspberry Pi 5 for a couple of days, but boy, have we burned through a lot of tokens and learned a lot. Including what I think is a really fun improvement in my development process:...

25,443 görüntüleme • 6 ay önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

OpenClaw vs. a 475-page datasheet: let the robot do the transcribing 🦞🤖 The u-blox SAM-M8Q has been sitting on my bench for months. This little GPS module has a built-in antenna, coin cell backup, speaks both NMEA and UBX binary protocol over UART or I2C. So why isn't it in the shop already? Well, it's mostly cause of the 475-page interfacing datasheet documenting every command, struct, and config register. Hundreds of message types. I got partway through by hand with some Claude Code Sonnet assistance, but ran out of time - plus it was still tedious when babysitting Sonnet. However, now we're living in an Opus + Codex era! So I pointed my Raspberry Pi OpenClaw at it. Here's the setup: Raspberry Pi 5 running OpenClaw, wired to a QT Py RP2040, which talks to the SAM-M8Q. Opus 4.6 reads the datasheet (converted to markdown first by Sonnet 4.6 with 1M context to minimize re-parsing that PDF every session) and builds the implementation plan. I review the plan to make sure it prioritizes the most common commands and reports, and flagged some unessential sections like automotive-assist or RTK-specific. Then Codex is assigned each message implementation task as a sub-agent and writes the actual C code for the Arduino library. Opus suggested using struct-based parsing rather than digging through each uint8_t array; we just memcpy the checksummed message raw bytes onto the matching struct and extract the typed bit fields. We've got four message types done so far. After each message is implemented, Codex also writes a test sketch that will exercise / pretty-print the results of each message, great for self-testing as well as regression testing later. Tonight I'm telling it to keep going while I sleep: code, parse, test against live satellite data, fix failures, commit and push on success, then move on to the next. To me this is a great usage of "agentic" firmware development: there's no creativity in transcribing 84 different structs from a 475-page datasheet. Once the LLMs are done, I can review the PRs as if it were an everyday contributor and even make revision suggestions.

adafruit industries

60,294 görüntüleme • 6 ay önce

I just compared Claude Code vs Codex vs Cursor CLI The task was to build a Next.js app with Tailwind 4 and shadcn components to collect customer feedback and showcase it with a widget. I gave all three the same prompt and let them go for 30 minutes to see what they came up with. Claude Code with Opus 4.1 Even though I told it to set up the app in the existing project folder, it tried to create a directory for it. After I interrupted and told it not to do that, it built a demo form and landing page with no errors. I had to ask it to make the demo interactive so users could submit a testimonial and preview it. The landing page looked like AI and was pretty basic, but it worked and it was done in a fraction of the time of the others. Total tokens used: 33k Codex with GPT-5 At the end of the 30 minutes I just could not get Codex to produce a working app. It got stuck in a loop of not being able to set up Tailwind 4 and despite many, MANY, attempts, I ended up with a "failed to compile" error. Total tokens used: 102k Cursor Agent with GPT-5 This was the slowest agent by far and a couple of times I actually thought it got stuck in a loop and was close to Ctrl+C'ing to cancel it. The TUI is really nice though, especially how it shows diffs and it did eventually build a working app (after one or two slight errors that needed fixing) The demo was interactive and it had a very minimal design that looked bare but also a lot less like an "AI generated" app than the Opus 4.1 design. It also wasn't too chatty and just did what it needed to do! Code quality was on a par with Opus 4.1, but it did use 5.5x as many tokens to get there. Still cheaper than Opus on a direct comparison but not when you factor in a Claude Code Max subscription. Total tokens: 188k I'll be able to do a proper comparison and record some videos when I'm back from holiday but for now, Opus is still the more capable model out of the box and Claude Code is the more complete CLI product. It will be interesting to see how Cursor evolve their CLI though with commands and subagents because I think with GPT-5 they have a real shot at providing competition for Claude Code if they can optimise output to get similar quality with less tokens. Jump to 0:40 in the video to see the two apps. Which do you think is which? ;)

Ian Nuttall

195,173 görüntüleme • 1 yıl önce

Congrats to EloShapes.com on their launch! On a similar note, is out of beta and available to the public. A lot of you have been asking me what my plan is now that EloShapes is also doing 3D scans... the answer is: nothing changed! I think that competition is good. I'm a mouse nerd and a LONG TIME EloShapes user, so the more options I get as a user, the better. I really think that my vision and EloShapes, though, is fundamentally different. I've had the domain for FMM for years, I have a much larger platform in mind compared to what I built so far. The 3D scans are the necessary step for my vision, not the end goal: they never were. I am also very proud of the fact that I've been able to scan at least 5 mice in full color EVERY DAY with my own pipeline, and I think I will easily have almost full coverage of the mice people would want on the website within months. Check the video for the mice I've added in just the last few days, all textured. Apart from some of the features I've mentioned before (like the virtual hand/grip), FMM is meant to be kind of a "MyAnimeList" for mice, so that you can share your profile, like this: and then later on leverage this curated list of mice you build for yourself to be shown even more mice you might like. FMM is also a tool for reviewers and anyone that wants to share their opinions on mice. Links like this: allow you to easily share your specific points about a mouse via the annotations, viewpoint sharing and measurements. And of course, as I've said since the beginning, I will use FMM to centralize all my hard testing about sensors, dpi, and such, so that you'll be able to find EVERYTHING about a mouse, including the raw results of my standardized sensor implementation testing. So, with this said, I will keep scanning as many mice as possible and adding as many features as possible until FMM becomes what I envisioned so long ago. Join my discord if you want to keep up! I hope you'll all give it a try! Ciao!

bardOZ (Giovanni Laerte Frongia)

25,920 görüntüleme • 4 ay önce

Someone ran Claude Code on a beach where any device overheats and that spot suddenly turned out to be the best home for the most powerful AI in the world. This is the reMarkable Paper Pro. A paper tablet for notes with no browser and no social media and not a single app. He sat down right on the sand in the open sun and brought up Claude Code on Opus 4.6 over the Claude API on the paper screen and opened his project ~/repos/webs while the waves broke a few steps away. For years every device had the same trouble outside. In direct sun the screen glares and washes out and heats up and instead of your work you see your own reflection. But e-ink does not blast its own light into your face. It reflects the sunlight like the page of a book. And here is what came out of it. The very thing that kills any normal screen outside turned into fuel for this one. The brighter the sun the sharper the picture because it has nothing to glare with and nothing to wash out. And then comes the thing no laptop on a beach will give you. Your eyes do not get tired. You can watch Opus think on max effort for an hour and it reads like a book in the sun and not a backlight you squint into. The picture only comes alive. In bright light it does not fade but turns sharper and higher in contrast than it ever was in a room. The charge lasts for days. E-ink barely touches the battery so there is no outlet anywhere on the sand and the tablet does not care. It weighs as much as a notebook. The whole setup folds into a beach bag like a pad with a pen on top. Everything on the screen is for real. Claude Code v2.1.110 and Opus 4.6 on the Claude API and the project ~/repos/webs open right on the e-ink in the middle of the sand. In my opinion this is the most unexpected home for an AI this year. Not an office with the blinds drawn and not a monitor cranked to full brightness but a quiet sheet of paper on the sand that open sun only makes better and on it the most powerful Claude writes code right on the page like a pen.

Blaze

89,297 görüntüleme • 2 ay önce

We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested what Opus 5 would do if I gave it the first paragraph of the Lord of the Rings, a 1M token budget (~$10) and asked for three js render of it. Opus went off for ~2 hours and wrote 5500 lines of code that (procedurally) rendered the story. It's kind of janky but fun. But it's a bit mindboggling that the LLM has to place and orchestrate various polygon assets in (x,y,z) coordinates and write code that animates it all, and that it even does anything at all. I also like this kind of examples because no one in their right mind would ever spend the time to write something this custom but LLMs have all the stamina and patience in the world, so it's an example where we go from "no one would ever do this" to "sure, why not, it's ~free". There might be a lot more. But I'm excited about creating hyper custom worlds that you can imagine dropping players into, e.g. here to participate in the LoTR story as a spectator NPC, or one of the characters, or etc. Something like an ephemeral GTA of X on demand. Last thought is that the domain of worlds/games exposes a weakness in LLMs: they can't easily audit their work because they aren't able to efficiently and natively perceive videos or play games within them. Here, Opus 5 had to very slowly and painstakingly take screenshots at different points, and it messed up a few times and created a bunch of jank. An example of raw capability (multimodal, gameplay) that I think is still quite lacking.

Andrej Karpathy

5,023,391 görüntüleme • 1 ay önce

Mark: It's been on my mind, and I just felt like I had to get this off my chest. And so I just felt like I had to turn this live on and to really just tell you guys how I felt regarding the T-shirt with the Confederate flag on it. I am aware of the pain that it brought to a lot of people, and especially to my black fans, to the black community. And I know that I should have. I am very responsible for it. But I really wanted to directly tell you guys that it really wasn't my intention, with me knowing the meaning of the flag, the historical significance of it. It really shouldn't have happened, and it really does not reflect who I am as a person. But seeing so many fans being hurt by it, it really hurt me as well. I had no intention knowing the meaning of the flag, and I should have. There are no excuses to that. But I just wanted to promise you guys directly that something like that will never happen again. With me and my whole team, we have made, we have taken steps to make sure that something like that doesn't happen. But I just, and I know that a lot of people have been waiting for me to kind of address this directly. And I've been sitting on this for a while, and I really wanted to get this off my chest. And so I thought today was the right time for me to tell you guys. It really breaks my heart to think that maybe the trust has been broken. But I promise you that I'll make it up to you guys, just with my actions and with my music. I'll make sure to make it up to you guys if you will allow me to have one more chance. I'll try my best in that. And so once again, I truly would like to apologize for the pain that that picture has brought to a lot of people. And I promise to educate myself more on that matter as well. When I first heard about it actually, I read about it. I was shocked, and we and our entire team made sure that it wasn't on any official content, but the picture got leaked, and we instantly uploaded an apology as well. But now I really felt like I had to personally take action for this. And so once again, I'd like to apologize for bringing so much pain to a lot of people. It really doesn't. It had no intention, no personal intention. So yeah. And thank you for taking the time to wait for me to really address this. I know a lot of fans have also been believing in me. But yeah, thank you for waiting till this moment.

먛¹⁹⁹⁹ #TheFirstfruit🫑

21,373 görüntüleme • 12 gün önce

This is probably the most complex workflow I’ve ever built, only with open-source tools. It took my 4 days. It takes four inputs: author, title, and style; and generates a full visual animated story in one click in ComfyUI . I worked on it for four days. There are still some bugs, but here’s the first preview. Here’s a quick breakdown: - The four inputs are sent to LLMs with precise instructions to generate: first, prompts for images and image modifications; second, prompts for animations; third, prompts for generating music. - All voices are generated from the text and timed precisely, as they determine the length of each animation segment. - The first image and video are generated to serve as the title, but also as the guide for all other images created for the video. - Titles and subtitles are also added automatically in Comfy. - I also developed a lot of custom nodes for minor frame calculations, mostly to match audio and video. - The full system is a large loop that, for each line of text, generates an image and then a video from that image. The loop was the hardest part to build in this workflow, so it can process either a 20-second video or a 2-minute video with the same input. - There are multiple combinations of LLMs that try to understand the text in the best way to provide the best prompts for images and video. - The final video is assembled entirely within ComfyUI. - The music is generated based on the LLM output and matches the exact timing of the full animation. - Done! For reference, this workflow uses a lot of models and only works on an RTX 6000 Pro with plenty of RAM. My goal is not to replace humans, as I’ll try to explain later, this workflow is highly controlled and can be adapted or reworked at any point by real artists! My aim was to create a tool that can animate text in one go, allowing the AI some freedom while keeping a strict flow. I don’t know yet how I’ll share this workflow with people, I still need to polish it properly, but maybe through Patreon. Anyway, I hope you enjoy my research, and let’s always keep pushing further! :)

Lovis Odin

58,841 görüntüleme • 11 ay önce

Ever since I wired Claude Code to WhatsApp 3 weeks ago, I built a stupidly large infra around it. I mean, opus built it. No clue how the code even looks. The entire thing was vibe coded using my phone. I wanted to see how far I could push it without touching the computer. Everything via WhatsApp. Build what I need on the fly. So the resulting infrastructure will already be battle tested for software development. The entire thing was streamlined with nearly no manual interventions, everything was communicated via WhatsApp using a single script establishing this connection. If the script is down, I need to get home to start it again to resume the development. Claude was upgrading it, debugging it, restarting it while maintaining constant uptime so it could keep communicating with me. I stressed Claude about it, telling it that it will be “in the dark” and other words that deliberately sound scary about losing communications if the script dies. I also refused git and refused cloning the code, I wanted to see Claude adapting to work on a *LIVING* system. The way this whole thing works: Claude has its own dedicated phone number that I am paying for. A real WhatsApp account for it is installed on a real iPhone that is sitting on my desk. All is registered under my name, this is legit setup with no hacks and tricks. I’ve set up a WhatsApp “Community” and multiple different groups under it. Both me and Claude are the admins, so Claude could edit it on my behalf. Each group is a project I am working on and has its own isolated context. The Group description is a system prompt that gets auto-appended to the larger system prompt explaining this setup in general. When I send a message it’s an instant interrupt to Claude Code’s process, just like in the terminal. Voice notes are seamlessly transcribed with a local Whisper model. Images are used with multimodal reading in an isolated parallel session. Multiple groups running in parallel so I can work on all projects at the same time. No cross-talking, everything has an isolated context and history. And because it’s local on my own machine: Everything is REAL. The browser is REAL. I am connected as myself on it to all services because I actually use it in real life. Claude has unlimited internet access, just like humans who use actual browsers. It utilizes custom-made browser tools that I made to control any browser session it wants. Depending on the situation, it can either connect to my existing session or create one for its own. (You can tell it ‘look at my browser for a sec’ then talk about the current page you are on and it just works, pretty cool) My custom browser tools are not perfect (not by a long shot) but I managed to make them work well to the point they are somewhat reliable. This gives Claude full access to my real creds and all the services I actually use. I’m productive AS HELL with this. It really feels like a personal assistant. I ask it to read my emails and msgs, check x .com for news, research arxiv papers, write code, run experiments for me, investigate and reverse engineer github repos, even use my credit card and order things. [I try not to do this one a lot lol so far no disasters]. All from my phone. Super convenient. This is not a product or an open source project (maybe soon of it will make sense). This is just an ugly script I hacked the entire thing is ~600 lines. (ok maybe i did look at the code, but i swear i didn’t edit!) You can also vibe code this from scratch pretty fast and it will probably even end up better. This is just a cool thing so I’m sharing. It is a real speed booster for many things I do on daily basis, mostly boring things. Forcing my routine into some new “agent platform” just didn’t feel right for me. WhatsApp is where I already communicate and look for messages, so I decided that my agents will live there too. AGI in my pocket 24/7.

Yam Peleg

420,060 görüntüleme • 8 ay önce

Pi was built when there were already agent harnesses around. Here’s why Mario Zechner(Mario Zechner), found them suboptimal and built Pi, a minimalist self-modifying agent: #1 - Mario initially was a believer in Claude Code: "I was a believer in Claude code because they were the first that packaged agentic search up in a really compelling package. And at the time that fit my workflow really well. Everything around the LLM was kind of nice and tidy and easy to understand. I was super happy. I was proselytising Claude code." #2 - Reverse engineering Claude Code highlighted the degradation that Mario felt as a user: "I personally like simple tools that are stable and that I can rely on. Even if they have non-deterministic parts, all the deterministic parts should be as stable as possible. That was just not the experience with Claude Code around summer 2025. They would take away your control of the context. They would inject stuff behind your back, which is bad. Then, your workflows stopped working because there's now a system reminder that you don't even see in the UI that would modify the behaviour of the model. They would also do this to the system prompt. I built a little service where I can track the progression or evolution of the system, prompt and tool definitions and, with every release, it was messing with stuff. That just messed with my workflows and I don't appreciate that." #3 - PI was built with an appreciation for simple and reliable tools: "If I commit to a development tool, I want it to be a stable, reliable thing like a hammer. I don't want my hammer to break a different spot every day. That's terrible. We need somebody who goes the full velocity kind of way. But I don't want to work with a tool like that."

The Pragmatic Engineer

62,825 görüntüleme • 3 ay önce

Day 8-9 of building an AMM. (repos in the next post) I added the protocol fee + the swap fees to LPs yesterday. It's the same as Uniswap v2, which is 0.25% (goes to LP) + 0.05% (goes to protocol) per swap. It was actually not that easy, as you want to spend the least amount of gas (CUs) possible, and you have to compound the fees into the pool. But I figured it out and in the end, it hopefully works. I don't transfer the fee out of the pool on every swap for the protocol, fees just go to the liquidity pool's LP token associated token account on deposits/withdraws, which I'll be able to withdraw later. (instruction still needs to be implemented) I still haven't written a test in Mollusk, and instead I started working on the frontend, for reason I won't tell you now... but just let me tell you that there may be good thing coming out of this whole AMM project besides just the AMM. So we'll put the tests on hold, for now. About the frontend... well yeah. This is the part I hate the most. I don't like frontend development. I'm not really good at it, and I can just get by fine, especially with Claude. If vibecoding was invented for something, it was frontend. God bless Claude for being a React beast. All I really did with the frontend today was initilize it from the solana template, and make the basic interface without hooking it up to the program. So everything you see is just mock data, I have attached a small little video on how it looks now. The recorder is laggy for some reason, so bear with me, but the actual site is fine, for the most part. (and also the video quality is really shitty on X, idk why) Just FYI, I'm gonna be taking a break from building this AMM for a couple days, and there will be less updates for a bit - I have exams soon for university and I have to study, and I also have to work on some other stuff. Maybe I'll stick to these updates to once a week? We'll see, I like doing it a lot so I might end up working anyways lmfaooooo. Your patience & attention is always appreciated, and thank you for following along!

8bitpenis.sol

60,037 görüntüleme • 8 ay önce

Mark's full statement regarding the recent issue : Before we actually get into the music though, there have been some things that I wanted to say to you guys directly. It spent on my mind and I just felt like I had to get this off my chest and so I just felt like I had to turn this live on on and to really just tell you guys how we felt regarding the T-shirt with the confederate flag on it, I am aware of the pain that it brought to a lot of people, especially to my black fans, to the Black community and I should have, I am very responsible for it, but I really wanted to directly tell you guys that it really wasn't my intention with me knowing the meaning of the flag, the historical significance of it, it really shouldn't have happened. And it really does not reflect who I am as a person. But seeing so many fans being hurt by it, it really hurt me as well. I had no intention knowing the meaning of the flag, and I should have, there are no excuses to that. But I just wanted to promise you guys directly that something like that will never happen again with me and my whole team we have made, we have taken steps to make sure that something like that doesn't happen, but I just, and I know that a lot of people have been waiting for me to kind of address this directly and I've been sitting on this for a while and I really wanted to get this off my chest and so I thought today was at the right time for me to tell you guys. It really breaks my heart to think that maybe the trust has been broken, but I promise you that I'll make it up to you guys, just with my actions and with my music. I'll make sure to make it up to you guys if you will allow me to have one more chance. I'll try my best to get and so once again I truly would like to apologize for the pain that that picture is brought to a lot of people and I promise to educate myself more on that matter as well. When I first heard about it, actually I read about it, I was shocked and we and our entire team made sure that it wasn't on any official content but the picture got leaked. We instantly uploaded an apology as well. But now I really felt like I had to personally take action for this. Once again I would like to apologize for bringing so much pain to a lot of people. It really doesn't. It's not, It had no intention, no personal intention. So yeah, and thank you for taking the time to wait for me to really address this. I know a lot of fans have also been believing in me.But yeah, thank you for waiting till this moment.

MARK LEE BASE

133,370 görüntüleme • 12 gün önce