Video yükleniyor...
Video Yüklenemedi
Until now, adding web search to open-source models meant hand-wiring orchestration, managing separate keys, and paying latency taxes on every round trip. Today, we're solving that. We're excited to introduce Baseten Hosted Tools and Baseten Grounded Inference to bring real-time web search server-side to open models running on Baseten... show more
29,259 görüntüleme • 6 gün önce •via X (Twitter)
11 Yorum

Andrey Styskin6 gün önce
Baseten x Keenable ❤️

Baseten6 gün önce
🙌

Ishan Goswami6 gün önce
Exa 🤝 Baseten

kes6 gün önce
Great working with you guys on this!

Baseten6 gün önce
💚

Exa Developers6 gün önce
💚

Teo Gonzalez6 gün önce
Love to see @baseten 🤝 @ExaAILabs

Baseten6 gün önce
@ExaAILabs 💚

Manish Tyagi6 gün önce
Excited to see this unfold!

Turk 🇺🇸6 gün önce
@saranormous Seems biased. Mentions Dreamforce happening in SF and says nothing about Grok galaxy which has more overall attendees than any baseball game happening this week in SF

RemoteBrowser6 gün önce
The latency tax framing is the real pain point. The round trips add up fast when the model decides to search mid-generation. Does Hosted Tools keep the search call inside the same inference loop, or is it still a separate hop?
