Loading video...

Video Failed to Load

Go Home

We just made Astra's computer use 2.5x faster with Stagehand. Astra works by executing code against the a11y tree, we built a translator that turns Astra's playwright commands into Stagehand. Astra chose to batch the entire logo into one command and one-shotted it.

24,181 views • 19 days ago •via X (Twitter)

20 Comments

Kyle Jeong's profile picture
Kyle Jeong19 days ago

See how it works:

idan levin's profile picture
idan levin19 days ago

@Stagehanddev Would it be interesting to explore adding code mode to webmcp? Its a bit early as not enough websites covers it but would be super efficient btw we created a webmcp benchmark and just added Astra computer use vs webmcp, really interesting results

Kyle Jeong's profile picture
Kyle Jeong19 days ago

@Stagehanddev stagehand does support webmcp already, we do recommend webmcp if it exists, but we're very far from that for a majority of the internet

sameel arif's profile picture
sameel arif19 days ago

@Stagehanddev holy

Kyle Jeong's profile picture
Kyle Jeong19 days ago

@Stagehanddev

Andrew Gazelka's profile picture
Andrew Gazelka19 days ago

@Stagehanddev why not write in stagehand directly? does astra not style transfer properly?

Kyle Jeong's profile picture
Kyle Jeong19 days ago

@Stagehanddev models atm are post-trained on playwright, we've experimented with skills, prompt optimization etc but realized we should just get out of the way and let the model write playwright (it's already good at this) and then port it to stagehand (so it's faster and more token efficient)

Andrew Gazelka's profile picture
Andrew Gazelka19 days ago

@Stagehanddev Did the model just need to think longer? Or was it doing obviously sub par stuff?

Kyle Jeong's profile picture
Kyle Jeong19 days ago

@Stagehanddev same thinking level, what's happening in the video is the model is just choosing to batch commands together in stagehand which is why it was able to one shot the painting the playwright version can't batch and ends up doing the task sequentially

Andrew Gazelka's profile picture
Andrew Gazelka19 days ago

@Stagehanddev I’m asking what the failure case is when it is given stagehand directly.

Kyle Jeong's profile picture
Kyle Jeong19 days ago

@Stagehanddev ah lol, yeah it was just generating malformed code and not able to thinking multiple steps in advance

arjun chintapalli 🌐👾's profile picture
arjun chintapalli 🌐👾19 days ago

@Stagehanddev Why even need stagehand instead of using astra directly?

Kyle Jeong's profile picture
Kyle Jeong19 days ago

@Stagehanddev this is using astra with the codex harness too, it just literally translates the playwright that the model generates into stagehand so that it's faster. no cost whatsover just pure performance improvements

Oleks's profile picture
Oleks19 days ago

@Stagehanddev 2.5x faster astra computer use through stagehand is the kind of receipt people screenshot

derek's profile picture
derek19 days ago

@Stagehanddev bruh

Shrey Pandya's profile picture
Shrey Pandya19 days ago

@Stagehanddev one-shotting is all you need

刘朝 Zhao Liu's profile picture
刘朝 Zhao Liu19 days ago

@Stagehanddev A11y-tree execution plus generated Playwright is a compelling split: the model handles intent and batching, while Stagehand provides deterministic plumbing. The production test is replay fidelity after DOM drift, retries, and partial side effects—not the one-shot logo demo.

Kisson's profile picture
Kisson19 days ago

@Stagehanddev Batching only holds while the a11y tree is stable - one virtualized list or a late-mounting portal and step 3 targets a node that already moved. Do you re-snapshot inside a batch, or commit the whole thing blind?

Preyforge's profile picture
Preyforge18 days ago

@Stagehanddev the speedup is fun, but the real artifact is the translator between intent and the accessible tree. i keep looking for the awkward handoff where a fast run still needs a human to take the mouse.

AI Mastery Guide's profile picture
AI Mastery Guide19 days ago

@Stagehanddev 2.5x speed is a great upgrade

Related Videos

gpt astra vs fable 5.1 at goldberg machine gpt 6 astra – openai, landed on OpenRouter less then hour ago, provider pinned to openai fable 5.1 – anthropic, shipped sep 1 we put the two models on one job: a rube goldberg machine in three.js that presses a button and detonates a bomb the setup: one self-contained html file, three.js from a cdn, everything else procedural – no textures, no models, no physics engine, every collision hand-written. the hard part sits in the brief: a domino may only fall once the previous one actually touches it, checked by real overlap every frame, never by a timer. same rule for the hammer hitting the button and the button firing the bomb. one continuous camera, its speed driven by whatever is moving. we recorded both scenes frame by frame – 1200 frames, 60 fps, exactly 20 seconds – and stepped both by hand to read the telemetry. - cost #1 astra – $1.84 #2 fable – $29.16 - time #1 astra – 9m 56s #2 fable – 1h 12m - tokens #1 astra – 45k #2 fable – 360k - lines of code astra – 881 fable – 744 observations: • we told it what we saw and nothing else – no diagnosis, no patch. we never edit a model's code. round two ran the whole chain to the blast. • both files are deterministic. two runs each, identical state to twelve decimals, and neither model reached for math.random. conclusion: 15.8x cheaper and 7.2x faster, and it still took a second round to get the ball into the bucket! follow thehype. for 24/7 ai news, analysis and breakdowns

thehype.

37,335 views • 23 days ago