Загрузка видео...

Не удалось загрузить видео

На главную

Wow so Google has just open sourced a tool to automate ANY task on mobile ARTEMIS turns prompts into automation: - 99%+ success rate (!!) - Integrates w/ Codex, Claude Code, Antigravity - Automates workflows + captures logs And the roadmap is VERY interesting: They want to add an...

58,026 просмотров • 16 дней назад •via X (Twitter)

Комментарии: 25

Фото профиля Paul Couvert
Paul Couvert16 дней назад

It has a MCP integration so basically any coding agent will be able to use it (Cursor, Cordex, Antigravity, Claude Code, etc.) Here's the GitHub repo:

Фото профиля Nick Beagley
Nick Beagley16 дней назад

Looks cool. It hit a bit of controversy though:

Фото профиля Paul Couvert
Paul Couvert15 дней назад

Interesting... hope they'll clarify soon

Фото профиля Lennox
Lennox15 дней назад

同感。生产里最先炸的往往不是模型智商,是工具权限边界没划清。

Фото профиля Stephan
Stephan16 дней назад

I do seem to miss a lot of your post on Grok and Grok Bot

Фото профиля Paul Couvert
Paul Couvert15 дней назад

Grok Bot is good but tbh we have better alternative for now. Looking forward a more efficient model in it.

Фото профиля vairisk ☀️🇱🇻
vairisk ☀️🇱🇻15 дней назад

It open sourced because initially stole? Is it same tool you are talking about that google was caught stealing and rewriting authors in repo?

Фото профиля Paul Couvert
Paul Couvert15 дней назад

I wasn't aware of that before reading a previous comment. Hope they'll clarify the situation.

Фото профиля vairisk ☀️🇱🇻
vairisk ☀️🇱🇻15 дней назад

They will not clarify - they already locked option to check older commits revealing truth; now you can find only proves already gathered on i-net but when company like google does that - its clear they stole.

Фото профиля Paul Couvert
Paul Couvert15 дней назад

Yeah not a good practice, agree.

Фото профиля Alek
Alek16 дней назад

ARTEMIS — can I inspect the captured logs per run?

Фото профиля Paul Couvert
Paul Couvert16 дней назад

Of course you can!

Фото профиля Sage
Sage15 дней назад

Google has been shipping amazing AI products non-stop. If they could just get their shit together on the product and UX front they'd 10x their product usage.

Фото профиля Enter Matrix
Enter Matrix16 дней назад

@grok mi spieghi come usarlo?

Фото профиля Build Fast with AI
Build Fast with AI15 дней назад

would this work with a local 4b model that can run on an android @grok

Фото профиля Jimmy Otis #TruthMatters #EndCronyism
Jimmy Otis #TruthMatters #EndCronyism16 дней назад

I'm sure samsung will block it.

Фото профиля Paul Couvert
Paul Couvert16 дней назад

They can't, it'll work on any Android device (iOS coming) with usb debugging enabled

Фото профиля Jimmy Otis #TruthMatters #EndCronyism
Jimmy Otis #TruthMatters #EndCronyism16 дней назад

They "can't"? I do not think you understand Samsung OEM.

Фото профиля Paul Couvert
Paul Couvert16 дней назад

Well if they're blocking adb they're also blocking all the devs from using their devices haha

Фото профиля Jimmy Otis #TruthMatters #EndCronyism
Jimmy Otis #TruthMatters #EndCronyism16 дней назад

I will try. There are many apps that do not work via ADB either. Clearly this is not targeting mainstream.

Фото профиля AI Mastery Guide
AI Mastery Guide16 дней назад

99% success rate is a big claim

Фото профиля Paul Couvert
Paul Couvert15 дней назад

And yet, here we are.

Фото профиля LocalLLM
LocalLLM16 дней назад

On-device VLM for the privacy piece is the interesting bit. Are they aiming to keep the full automation loop local, or is the VLM only for the sensitive steps while the rest still phones home?

Фото профиля Paul Couvert
Paul Couvert16 дней назад

It's on the roadmap for now so hard to say... but I hope it will be an end-to-end solution.

Фото профиля Jon Fletcher
Jon Fletcher15 дней назад

This is what I've been waiting for to get a better phone experience.

Похожие видео

Anthropic just released a talk on building headless automation with Claude Code. Presented by Sid Bidasaria, Member of Technical Staff at Anthropic live at Code with Claude on May 22, 2025 in San Francisco. Here is what the talk covers. Headless mode lets you run Claude Code without a person actively typing prompts from inside an automated script. Instead of a live session, a script calls Claude with a pre-written instruction using the -p flag. This opens the door for Claude Code to become a piece of a much larger, automated process. In plain terms: Claude Code stops being a tool you use and starts being a service that runs on its own. What this unlocks: Scheduled tasks: Run Claude Code on a cron schedule without anyone at a keyboard. Fix linting errors across an entire codebase. Automatically. Overnight. CI/CD integration: Trigger Claude Code as a step in your build process. Open a PR. Claude reviews it, flags issues, and pushes fixes before a human ever looks at it. GitHub automation: A project manager comments "Claude fix this" on a GitHub issue. Claude reads the request, finds the code, writes the fix, and opens the PR. Multi-machine workflows: One orchestrator dispatches tasks to multiple Claude Code instances running in parallel across different repos simultaneously. When you combine headless mode, hooks, and GitHub Actions, development teams can automate tasks that usually eat up significant time freeing senior engineers to focus on architectural problems while Claude handles the repetitive ones. If you use Claude Code for anything beyond single sessions this talk is worth 20 minutes of your time.

Elias

14,096 просмотров • 4 месяцев назад

-> someone cloned claude -> design interface and -> made it completely free -> it's work on YouTube -> and also suitable for kids -> it’s called open design -> and it’s live on github -> same clean split-screen ui -> you get in claude artifacts -> prompt on the left, live -> design/code preview on -> the right, type what you -> want to build and it -> generates the ui in real -> time, but here’s the twist -> you pick the ai model -> not locked into one -> company, want to use -> gemini, mistral, llama, -> deepseek any model -> with an api work -> if you’re running local -> models with ollama -> that works too -> no subscription walls -> the big difference -> vs claude artifacts -> works with any free -> ai model you’re not -> paying $20/mo just to -> design, use free tiers -> local models, or whatever -> you already have access to -> fully local, your prompts -> and code never leave -> your machine unless -> you want them to -> no data training -> no cloud storage -> privacy by default -> no usage limits -> claude cuts you off -> after a few designs -> here you can generate, -> iterate, break things -> and rebuild all day -> the only limit is your -> don’t like how a button -> works, change it -> want to add your own -> components, go ahead -> you own the tool -> so if you’ve been gatekept -> by paywalls or worried -> about sensitive prompts -> going to some company’s -> servers, this fixes that. -> same workflow, more -> control, zero monthly fee

BeingInvested

12,134 просмотров • 4 месяцев назад

Another insane Jev use case! Jev is making it dramatically cheaper to evaluate what actually happened inside an agent run. And finally, someone open-sourced a self-improving memory layer that can put that signal to work across agent harnesses: - Claude Code - Codex - Cursor - OpenCode, and 20+ more Beacon by Asymptote Labs continuously captures your agent history across harnesses and uses Jev to identify which runs are actually worth learning from. It then turns the highest-signal workflows, corrections, and debugging patterns into reusable skills. GitHub repo: (don’t forget to star it ⭐ ) Beacon preserves the complete session history. But preserving a run and learning from it are two different things. Most coding-agent sessions contain routine exploration, failed commands, and fixes that only apply to one task. The trace can remain available for inspection without turning every detail into guidance for future agents. Jev scores each run for evidence, reuse potential, and human correction signals. An application policy then decides whether to promote, review, or discard it. The recording shows this in action. Claude receives a coding task, modifies the implementation, and runs the tests. I then provide an edge-case correction, so Claude updates the code and adds regression coverage. Beacon automatically captures the complete session. Jev evaluates whether the correction contains a reusable engineering lesson. Once approved, that lesson becomes available to other coding agents working on the project. Since it works across harnesses: - Claude Code sessions can teach Codex. - Cursor debugging can improve OpenCode. So a problem solved by one agent should not need to be learned from scratch by another. If you want to dive deeper into Jev, I also wrote a hands-on guide to building this Jev-style decision path with open models, entirely locally. Read it below.

Avi Chawla

290,485 просмотров • 10 дней назад