正在加载视频...
视频加载失败
Until now, adding web search to open-source models meant hand-wiring orchestration, managing separate keys, and paying latency taxes on every round trip. Today, we're solving that. We're excited to introduce Baseten Hosted Tools and Baseten Grounded Inference to bring real-time web search server-side to open models running on Baseten... show more
11 条评论

Andrey Styskin6 天前
Baseten x Keenable ❤️

Baseten6 天前
🙌

Ishan Goswami6 天前
Exa 🤝 Baseten

kes6 天前
Great working with you guys on this!

Baseten6 天前
💚

Exa Developers6 天前
💚

Teo Gonzalez6 天前
Love to see @baseten 🤝 @ExaAILabs

Baseten6 天前
@ExaAILabs 💚

Manish Tyagi6 天前
Excited to see this unfold!

Turk 🇺🇸6 天前
@saranormous Seems biased. Mentions Dreamforce happening in SF and says nothing about Grok galaxy which has more overall attendees than any baseball game happening this week in SF

RemoteBrowser6 天前
The latency tax framing is the real pain point. The round trips add up fast when the model decides to search mid-generation. Does Hosted Tools keep the search call inside the same inference loop, or is it still a separate hop?
