
am.will
@LLMJunky • 30,466 subscribers
StarSwap // Lynk // GooeyPi // DevX Director of n number of agents DevRel @nutrientdocs https://t.co/48fTTYOkmg
Shorts
Videos

OH MY GOSH ASTRA. GAMING = SOLVED! This is absolutely INSANE! This was not built in Blender. Nor was it built in Godot, Unity, or Unreal Engine. This is just Astra flexing its muscles on the rest of the models to say: I AM THE BEST What can I say? Is RocketLeagueBench saturated? I honestly think so at this point. My jaw is on the floor. This looks so incredibly good. Just you wait to see what people are going to build with this model. WOW.
am.will789,146 просмотров • 1 день назад

WOW! I thought Opus 5 would just be a cheaper Fable. I was dead wrong. This model is absolutely BONKERS. I'm genuinely blown away. Opus 5 built the best Rocket League clone i've ever seen, and it did it while consuming only 27% of my 5x Max sub. Sound on. There's layers to this onion. I'm floored rn. At least for building games and 3D modeling, this has to be the best model in the world by a notable margin. Look at the car. The reflections. The details. Playable game & open source code in the comments.
am.will341,691 просмотров • 1 месяц назад

I tried this so you don't have to. I know this is going to absolutely shock you but no this does not match the performance of Mythos. A few early thoughts: 1. The limits are pretty bad. I used 100% of my 5-hour usage in less than 1 prompt. 2. I specifically gave it a threejs task because it is an area that SOTA models have made big strides in, that other models just are not great at. I asked it to build a replica of Rocket League. I'll put the prompt in the comments. The game was pretty bad and notably worse than GPT 5.5. Even after multiple fixes, it took 7-8 back and forth with Codex just to get it an almost playable condition. Prior to these fixes, the game was not playable. Maybe it's really strong in other disciplines. I'd love to test that but I hit my limit in 1 prompt lol. GPT 5.5 by contrast did a pretty good job and required no follow ups. Fable would have absolutely nailed this as well. But yeah, early impressions...not great. But I hope I'm wrong. More testing tomorrow.
am.will594,106 просмотров • 2 месяцев назад

🥧 Introducing GooeyPi - A GUI for the Pi Family of Agents. Supports: Pi, Oh-My-Pi (OMP), and Prime Agent Very excited to share this with you all. Being into local models, I wanted to have a separate harness for my local models so I don't have to constantly switch configs. And I didn't want to go back to the TUI (nerds). But, I didn't want to miss out on all the features that I know and love. Turns out, I didn't have to. I just needed to build it myself. So I did. Features include: > Fully featured GUI experience > Agentic Browser > Realtime Voice Agent > Voice transcription (local or API) > Computer Use > Automation Desk > Agent to Agent messaging > Ask Question Tool > Git Control > Terminal > Pets > and more I tried to capture all of the most popular features I could in this first pass, but we're only just getting started. Available in BETA on MacOS, Linux, and Windows. I need testers! I have only tested on Mac and Linux. If you want to help me with this project, please reach out. Available:
am.will145,452 просмотров • 23 дней назад

Look ma new Codex Updates! 0.119.0 and 0.120.0 are here. And with it, a HUGE number of quality of life updates and bug fixes! > Hooks now render in a dedicated live area above the composer. They only persist when they have output, so your terminal stays clean. If you're running PreToolUse or PostToolUse hooks, this is a huge readability win. > Hooks are now available again on Windows > CTRL+O copies the last agent output. Small but clutch when you're pulling a code block into another file or chat. > New statusline option: context usage as a graphical bar instead of a percentage. Easier to glance at mid-session when you're trying to gauge how much runway you have left. > Zellij support is here with no scrollback bugs. If you've been stuck on tmux just because Codex was broken in Zellij, you're free now (shout out Felipe Coury 🦀) > Memory extensions just landed. The consolidation agent can now discover plugin folders under memories_extensions/ and read their instructions.md to learn how to interpret new memory sources. Drop a folder in, give it guidance, and the agent picks it up automatically during summarization. No core code changes needed. This is the first real extension point for Codex's memory system, and it opens the door for third-party memory plugins. > Did you know, you can /rename a thread? But what's really cool about that is, after you rename it, you can resume it with the same name, no more UUIDs. codex resume mynewapp or directly from the TUI: /resume mynewapp > Multi agents v2 got an update to tool descriptions More reliable multi agent environments and inter agent communication > You can now enable TUI notifications whether Codex is in focus or not. Modify this in your config: [tui] notification_condition = "always" > MAJOR overhaul to Codex MCP functionality: 1. Codex Tool Search now works with custom MCP servers, so tools can be searched and deferred instead of all being exposed up front. 2. Custom MCP servers can now trigger elicitations, meaning they can stop and ask for user approval or input mid-flow. 3. MCP tool results now preserve richer metadata, which improves app/UI handoff behavior. 4. Codex can now read MCP resources directly, letting apps return resource URIs that the client can actually open. 5. File params for Codex Apps are smoother: local file paths can be uploaded and remapped automatically. 6. Plugin cache refresh and fallback sync behavior are more reliable, especially for custom and curated plugins. > Composer and chat behavior smoother overall, resize bugs remain though. > Realtime v2 got several significant improvements as well. > You're still reading? What a legend. 🫶 npm i -g @openai/codex to update
am.will742,451 просмотров • 4 месяцев назад

I'll bet you didnt know you could do this with Claude Code
am.will296,280 просмотров • 2 месяцев назад

THE WAY YOU LOOK AT AI GAME DEVELOPMENT WILL CHANGE AFTER YOU WATCH THIS VIDEO "The AI Slop is Coming for Games." That's right. Introducing... Slopet League Everything you see here was built by Opus 5 and GPT 5.6 Sol in _____🤫. This is not an html page, it is a functional game with a fair bit of polish working inside a real game engine. This has to be some kind of trick. NO. It downloaded a map. NO. Imported assets. Only the Octane meshes, I have other cars that were generated that look just as good. The balls, the map, the effects, the physics, animations, etc were all generated from scratch by Opus 5. GPT 5.6 Sol handled the car meshes. This is only the beginning. I made this in about a day and a half, and most of that time I was waiting on Opus to finish generation. Maybe 25 prompts. I will share my process, and many photos, proofs, evidence of this as we go. This is just a taste, I'm extremely busy this week and working on real life stuff so I'm not going to put any more time into this until next week, but believe me I'm gonna polish this into something beautiful. If you're interested, please share, bookmark, follow, and comment below. I want the world to see just how powerful the language models are TODAY. It's only going to get better from here. Buckle up folks. Oh, and What a save!
am.will131,731 просмотров • 1 месяц назад

Omarchy has the best screensavers I've ever seen, and there's so many of them
am.will34,552 просмотров • 10 дней назад

anything under deepseek v4 flash basically doesn't exist to me. with <200GB of vram, i'm getting ~800K context @ 200 tokens per second. It's not quite Cerebras fast, but damn. It is a joy to use, and its basically somewhere in the opus 4.7/4.8 range. It's incredibly good at browser control too. this model is a beast. w/ qwen 3.8 27B being multimodal though, it might just be worth switching to, especially if you want to run a lot of parallel agents will 300toks/s be possible? Pretty addicting let me tell you. Imagine a computer use model this fast. watch below, this is basically human speed.
am.will72,729 просмотров • 26 дней назад

This new Siri update is absolutely maddening. It's a tale as old as time, Siri being a complete joke. Google has historically just been far better at basically every stage of their lifecycle in terms of voice to text, and phone control. Google Assistant was completely dominant. So you would think considering that Google makes their own frontier models, and the fact they are one of the industry leaders in small models capable of running on Edge devices, that they would easily dominate the mobile space, right? I mean surely that would be true considering Apple has more or less been on the sidelines in the AI race, right? I seriously don't understand how this happened. On device Gemini, even with "Personal Intelligence" really has very little additional utility over the previous deterministic STT "Google Assistant" we've had for the last decade. What are we doing?
am.will206,441 просмотров • 2 месяцев назад

The Codex app server was such a brilliant stroke of foresight that really doesn't get enough love Not only are you allowed to use your chatgpt account with any harness, but you can build your own apps directly on top of theirs. They just make building on and with codex such a great experience To demonstrate this utility, I want to highlight the kitty litter app, made by SIGKITTEN. Instead of having to build the entire harness, and all the infrastructure, he's plugged into the app server for a unified experience between mobile and dev machine. When I create a session on my computer, it's automatically available on my phone. All of the chats you see in this video automatically populated when we connected to the app server. All my skills. My agents. My sessions. My folders. My prompts. They're all ready to use - automatically. Because they're exposed by the app server, along with many other endpoints. It's a great ux/dx that really deserves some love. It's almost like they want you to build on top of their products ;) Btw Litter is great 👍
am.will267,261 просмотров • 5 месяцев назад

Can confirm Bonsai 1.7B, works great locally on my Galaxy S26U with Lynk! Very, very fast even on CPU.
am.will95,951 просмотров • 1 месяц назад

Nice. Cursor just dropped their new "Glass" alpha, and they're leaning heavily into the simplified coding GUI trend that's been blowing up lately. First impressions are really positive. And just look at how insanely fast Composer 2 is. First impressions? Drop yours 👇
am.will217,130 просмотров • 5 месяцев назад

Introducing Lynk, a brand new way to interact with your favorite harnesses on the go. Compatible with OpenClaw, Hermes, Codex, and local edge models. Fully featured client allowing you to easily and quickly switch between your favorite agents. Lynk has been my absolute favorite way to kick off tasks on the go, completely replacing Telegram. Only available as a beta on Android at present, but iOS version is in development. Works via local network or Tailscale. Features: - Create and continue threads with your favorite harness - OpenClaw, Hermes, Codex, and local models - Android Phone control - Realtime voice agent - Speech-to-text transcription - Codex Pets + live notifications - Draw over screen quick access chat overlay - Open Source software Join the beta in the comments.
am.will111,376 просмотров • 3 месяцев назад

this video encapsulates parenthood in such a perfect way. it is truly the most rewarding human experience I've ever felt. it fills you with joy and breaks your heart all at the same time. and speaking of time. it absolutely flies. these early years are just so unbearably short. so much happy sad. and remember, they're only little once. yt/jaysharon
am.will16,919 просмотров • 15 дней назад

This is so cool. In the next Codex update, multi agents will get a massive flexibility upgrade. "Hey Codex, when you implement this plan, I want you to delegate all the lower complexity tasks to GPT 5.3 Spark subagents" Instead of needing to create 100 different custom agent roles for different situations, you can just prompt your agent to spawn whatever model or reasoning level you want. With only natural language. No config files. No pre-defined roles. Just tell the orchestrator what to use and it listens.
am.will128,125 просмотров • 5 месяцев назад