Загрузка видео...
Не удалось загрузить видео
Until now, adding web search to open-source models meant hand-wiring orchestration, managing separate keys, and paying latency taxes on every round trip. Today, we're solving that. We're excited to introduce Baseten Hosted Tools and Baseten Grounded Inference to bring real-time web search server-side to open models running on Baseten... show more
29,259 просмотров • 6 дней назад •via X (Twitter)
Комментарии: 11

Andrey Styskin6 дней назад
Baseten x Keenable ❤️

Baseten6 дней назад
🙌

Ishan Goswami6 дней назад
Exa 🤝 Baseten

kes6 дней назад
Great working with you guys on this!

Baseten6 дней назад
💚

Exa Developers6 дней назад
💚

Teo Gonzalez6 дней назад
Love to see @baseten 🤝 @ExaAILabs

Baseten6 дней назад
@ExaAILabs 💚

Manish Tyagi6 дней назад
Excited to see this unfold!

Turk 🇺🇸6 дней назад
@saranormous Seems biased. Mentions Dreamforce happening in SF and says nothing about Grok galaxy which has more overall attendees than any baseball game happening this week in SF

RemoteBrowser6 дней назад
The latency tax framing is the real pain point. The round trips add up fast when the model decides to search mid-generation. Does Hosted Tools keep the search call inside the same inference loop, or is it still a separate hop?
