Video wird geladen...
Video konnte nicht geladen werden
Until now, adding web search to open-source models meant hand-wiring orchestration, managing separate keys, and paying latency taxes on every round trip. Today, we're solving that. We're excited to introduce Baseten Hosted Tools and Baseten Grounded Inference to bring real-time web search server-side to open models running on Baseten... show more
29,259 Aufrufe • vor 6 Tagen •via X (Twitter)
11 Kommentare

Andrey Styskinvor 6 Tagen
Baseten x Keenable ❤️

Basetenvor 6 Tagen
🙌

Ishan Goswamivor 6 Tagen
Exa 🤝 Baseten

kesvor 6 Tagen
Great working with you guys on this!

Basetenvor 6 Tagen
💚

Exa Developersvor 6 Tagen
💚

Teo Gonzalezvor 6 Tagen
Love to see @baseten 🤝 @ExaAILabs

Basetenvor 6 Tagen
@ExaAILabs 💚

Manish Tyagivor 6 Tagen
Excited to see this unfold!

Turk 🇺🇸vor 6 Tagen
@saranormous Seems biased. Mentions Dreamforce happening in SF and says nothing about Grok galaxy which has more overall attendees than any baseball game happening this week in SF

RemoteBrowservor 6 Tagen
The latency tax framing is the real pain point. The round trips add up fast when the model decides to search mid-generation. Does Hosted Tools keep the search call inside the same inference loop, or is it still a separate hop?
