Загрузка видео...

Не удалось загрузить видео

На главную

Using WebMCP tools, I can surface all sorts of functionality in my app. I can batch process, change themes and dark mode, check on higher level goals that I'm keeping track of. All of this functionality exists already! You can do it manually if you prefer. But you could...

20,831 просмотров • 6 дней назад •via X (Twitter)

Комментарии: 37

Фото профиля Samarth
Samarth6 дней назад

did the same for two apps that i now use frequently- BingeWatcher (movie and tv-show discovery with lineups and watchlists) and ROUGH//CUT (in browser video editor) :))

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

Nice!

Фото профиля Samarth
Samarth6 дней назад

thanks a lot!

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

A couple people asked me why I showed chat and not voice- the answer is simple: I was sick and I sound like a frog. 🐸 I'll record one when I am not making ribbit noises

Фото профиля Vishal Anton
Vishal Anton6 дней назад

This is cool.

Фото профиля Srinivasan K K
Srinivasan K K6 дней назад

yes, I did some experimentation earlier. What I understood is WebMCP for the frontend, MCP is for the backend but for AI agents.

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

Yes, you’re correct

Фото профиля Srinivasan K K
Srinivasan K K6 дней назад

👍

Фото профиля Abdul Wasey
Abdul Wasey6 дней назад

WebMCP makes the UI an agent API without rebuilding the backend.

Фото профиля Vinny
Vinny6 дней назад

There are some very creative use cases for it too:

Фото профиля Ritwik
Ritwik6 дней назад

WebMCP that still works as a normal UI is the right constraint. Agent-only surfaces rot the moment a human needs to take over.

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

Not necessarily- I think we’ll see more handoffs between experiences over time. There are certain batch actions I might have an agent do and then want to step in to augment individual values, for instance. However, just surfacing the existing functionality is just fine too!

Фото профиля Manu Schiller
Manu Schiller6 дней назад

Is WebMCP maybe a good fresh take on accessibility?

Фото профиля Nart Madi
Nart Madi6 дней назад

Hey Sarah, I'd love to show you what we're building for WebMCP @Senro_AI.

Фото профиля OneOrigine
OneOrigine6 дней назад

;🍎💯

Фото профиля Jarno
Jarno6 дней назад

Voice control for dev tools sounds great until you're debugging out loud in a coffee shop.

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

I’m not doing voice control here, I’m just typing. That’s sort of the point- it’s flexible to your situation ☺️

Фото профиля Vinicius Dallacqua
Vinicius Dallacqua6 дней назад

Do you have a link to this? I'd love to use it for benchmarking and metrics on an article I'm writting.

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

Yeah I’m getting some last features over the line and then I’ll open source it

Фото профиля Rohan Verma
Rohan Verma6 дней назад

What's the chat UI on the right?

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

I made it a part of the app. I embedded Gemini and surfaced a UI that can be voice or chat

Фото профиля rusa
rusa4 дней назад

the a11y angle is the sleeper. app actions as callable tools is basically a machine-readable menu, which is exactly what screen readers always wanted and rarely got reliably.

Фото профиля Sarah Drasner
Sarah Drasner4 дней назад

1000%

Фото профиля Sudhir Dudeja
Sudhir Dudeja6 дней назад

I made my app webmcp enable as well, Quiak question what are using apart from codex to test and it and does it require API key ?🤔

Фото профиля Max Villemure
Max Villemure6 дней назад

I really like the view, and the image on top. Great job

Фото профиля John Rood
John Rood6 дней назад

we run enough browser agents to feel this: a semantic tool surface beats teaching the model to click pixels every time. the accessibility path becomes an interface you can test, not a demo you hope works.

Фото профиля Ben Mo
Ben Mo6 дней назад

does one undo roll back the whole batch, or would it step through each task change?

Фото профиля Agrit Tiwari
Agrit Tiwari5 дней назад

This is good to do client side tool calling. Makes sense. Really like it so far. But is this Copilotkit for the sideChat? PS. tc.

Фото профиля Sarah Drasner
Sarah Drasner5 дней назад

Thanks! Just Gemini

Фото профиля Aivan Monceller
Aivan Monceller6 дней назад

Are you running an LLM on the browser? Or still calling a frontier in this case?

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

I've embedded Gemini and made it part of the app experience. You don't have to do it that way, though

Фото профиля Aivan Monceller
Aivan Monceller6 дней назад

What does "embedded Gemini" mean? Is the model actually running in the page and driving the existing features through WebMCP? What's the difference in practice between embedding Gemini in the app vs just calling the API?

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

Yeah it's running in the page and driving existing WebMCP tooling. I exposed some special tools to do things my app can't do as well like batch processing or checking against larger goals. No big difference here, and you can still bring your own/drive it another way. The app is for me, so I wanted it as part of my experience.

Фото профиля Saket Tawde
Saket Tawde6 дней назад

Woah! Imagine if the whole OS was like this?! Wait a second, wasn't there a ChromeOS?

Фото профиля LottieFiles
LottieFiles6 дней назад

The voice part is really interesting. Being able to just say what you want instead of clicking through menus would help a lot of people. Have you tested it with voice yet, or is that still the plan?

Фото профиля Sarah Drasner
Sarah Drasner6 дней назад

Yes! I almost did a voice one, and will probably still post it. I've been sick this week and my voice sounds like a frog 🐸 ribbit

Фото профиля Tony
Tony6 дней назад

The accessibility angle is what caught my attention. Being able to say what you want done could make a complicated app feel much more approachable.

Похожие видео