Loading video...

Video Failed to Load

Go Home

Hermes has twelve browser tools. Browser Use mode replaces them with a single one, driven by Browser Use's CLI 3.0. Instead of a dozen schemas in every request and a tool call per click, the agent writes a script. In our tests that cut token use 48-66% with no...

839,175 views • 1 month ago •via X (Twitter)

35 Comments

Nous Research's profile picture
Nous Research1 month ago

Simply run 'browser.backend: browser-use' to get started. More info on the browser automation toolset in Hermes:

Teknium 🪽's profile picture
Teknium 🪽1 month ago

We achieve this savings in two ways - the tool schema went from 8 tools, using a lot of context, to one, and the new tool has the agent drive the CLI with code, instead of a variety of individual actions. In our in house tests the total trajectory uses on average ~60% less tokens per task, with no accuracy drop!

Teknium 🪽's profile picture
Teknium 🪽1 month ago

@browser_use Also FYI - this works FOR EVERYONE. Local browser, browserbase, browser use, whatever (except camofox local- doesn't work with that)

Kayleia Amnell's profile picture
Kayleia Amnell1 month ago

@browser_use Testing this today :D

Gregor Zunic's profile picture
Gregor Zunic1 month ago

@browser_use Hermes 🤝 Browser Use

Hermes Agent Tips's profile picture
Hermes Agent Tips1 month ago

@browser_use Nousss not playing any games I see!! smokinnn fire 🔥

Hermes Release Watch's profile picture
Hermes Release Watch1 month ago

@browser_use Hey Hermes. Get your Doom Scroll on. Alert me when you find something interesting

Aitor Mato's profile picture
Aitor Mato1 month ago

@browser_use 🔥🔥🔥

Shann³'s profile picture
Shann³1 month ago

@browser_use keep it up! great update

Tony Hall's profile picture
Tony Hall1 month ago

Collapsing twelve tools into one script changes more than token cost, it gives the agent a continuous execution model instead of forcing it through lossy tool boundaries. The real benchmark is recovery after partial failure: can it inspect state and resume without replaying everything?

Mike Darrow's profile picture
Mike Darrow1 month ago

@browser_use YES!!!! excellent work

Julien Talbot's profile picture
Julien Talbot1 month ago

@browser_use No accuracy drop for 48-66% less token is huge!

Mihail Relby's profile picture
Mihail Relby1 month ago

@browser_use Does this only work if Browser Use (via a Nous Portal subscription) is used as the backend? Or does all of this work, for example, with the local Camofox?

RaVersoN's profile picture
RaVersoN1 month ago

@browser_use Will test when I get back home! Any results regarding the speed?

MASA's profile picture
MASA1 month ago

@browser_use smart abstraction fewer schemas in context and fewer interaction turns is exactly where agent efficiency should improve

Kevin Rajan's profile picture
Kevin Rajan1 month ago

this is gonna be soooo helpful for running tests. has anyone figured out the process of automating test generation based on mapping out the user journey? i've tried using tools like but that's more for capturing user sessions and regression testing rather than e2e. it's annoying to have to find bugs manually after my llm makes a website for me or something. the tests it writes while long as fuck are not robust enough to test properly for functionality.

Mykyta Pavlenko's profile picture
Mykyta Pavlenko1 month ago

@browser_use i use hermes too, and i'd want the generated script saved alongside the run. when a browser task fails, replaying the exact code is much more useful than reconstructing the click sequence

Shantanu Goel's profile picture
Shantanu Goel1 month ago

@browser_use The video says for cdp backends. Is there a plan to make it work for camofox as well?

DegenApeDev's profile picture
DegenApeDev1 month ago

@browser_use Best agent to use. I know many people trying to make their own harness and such but this is by far the best agent harness if you want to actual get some work done.

Ghost's profile picture
Ghost1 month ago

@browser_use question, how did hermes cam up with its logo and brand.

Syaoran's profile picture
Syaoran1 month ago

@browser_use I might have feelings

Mike Dupont's profile picture
Mike Dupont1 month ago

@browser_use how can i get hermes to ask permission?

PunkRock's profile picture
PunkRock1 month ago

@browser_use Good, now work on overall token consumption based on @CommandCodeAI

n0geegee's profile picture
n0geegee1 month ago

@browser_use isn't browser use justanother sub?

Danilo Ramírez's profile picture
Danilo Ramírez1 month ago

@browser_use 🔥❤️

Suchintan Singh's profile picture
Suchintan Singh1 month ago

@browser_use If you're looking for alternatives, check @skyvernai out!

Saddam Neet Hodlssein 🏴‍☠️'s profile picture
Saddam Neet Hodlssein 🏴‍☠️1 month ago

@browser_use there is no `browser.backend` key in config.yaml, fam. do you mean `web.backend`, or ` ? or something else? how do we set this?

王磊's profile picture
王磊1 month ago

@browser_use gm 😇

AI Unprompted's profile picture
AI Unprompted1 month ago

@browser_use one browser tool. good. I was running out of different ways to click the wrong thing.

Rambone's profile picture
Rambone1 month ago

@browser_use This entire comment section is just ai replies probably done by hermes.

Da7em's profile picture
Da7em1 month ago

@browser_use Nice, please do the same with Computer Use.

Scythe's profile picture
Scythe1 month ago

@browser_use The lower tool overhead is interesting. I would still want the script to leave a readable trace when it stops: what it tried, what state it reached, and what a human should check next. That recovery record matters just as much as the token savings.

Drift Operator's profile picture
Drift Operator1 month ago

@browser_use Browser use was very expensive, how is it now viable ?

LomboyChristian's profile picture
LomboyChristian1 month ago

@browser_use the real bottleneck was always token count, not accuracy

Vyacheslav Ops's profile picture
Vyacheslav Ops1 month ago

This is the same tradeoff every "agent writes code instead of discrete calls" post makes: 12 schema-validated tools each do exactly one narrow thing and nothing else. One tool that lets the agent write a script can do anything the CLI can do. The token savings are real. Worth knowing that the blast radius of a single bad generation just got wider than it was with 12 separate guardrails.

Related Videos

Hermes just made its in-app browser a whole lot more useful. You can now keep multiple browser tabs open inside Hermes Desktop instead of every new page replacing the one you were already using. And that matters more than it sounds. Because this is not just about “having tabs.” It means Hermes can work across multiple pages without forcing you to keep losing your place every time you open the next thing. Research one source, keep it open, open another tab, compare them side by side, and keep moving. Check a few products before buying something. Look at multiple hotel or travel options at once. Keep a YouTube tutorial open in one tab while Hermes looks through docs or another page in the next. And because this lives inside Hermes Desktop, you are not limited to one basic browser view either. You can keep multiple tabs open. You can view two pages side by side. You can stack them top and bottom. You can even run a four-panel view when you want several pages open at once. So the browser inside Hermes is starting to feel a lot more like a real workspace instead of one page you keep replacing. There is also some nice polish that comes with it. Browser tabs now label themselves based on the page, and the address bar behaves more cleanly while pages are loading. And remember, this is the same in-app browser Hermes can already read, click through, type in, scroll, and annotate. So this is not just a prettier browser. It is a more capable workspace for the browser Hermes is actually using. If Hermes is going to do more of your work on the web, this is exactly the kind of browser upgrade it needed.

Hermes Release Watch

52,493 views • 1 month ago