Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

We didn't just make local models faster. We made them dependable at tool calling. Gemma 4 12B now picks the right tool and runs it correctly, on-device and ~60% faster. Small models just got serious.

49,030 görüntüleme • 3 ay önce •via X (Twitter)

33 Yorum

Osaurus profil fotoğrafı
Osaurus3 ay önce

Free forever. MIT licensed. Native macOS, no browser engine.

AI Mastery Guide profil fotoğrafı
AI Mastery Guide3 ay önce

Getting a 12B model to reliably pick the right tool on-device is honestly more impressive than just being faster

Osaurus profil fotoğrafı
Osaurus3 ay önce

We've improved tool calling and model loading for all local models (specifically those that are <=20B)

AI Mastery Guide profil fotoğrafı
AI Mastery Guide3 ay önce

that <=20b range is exactly where most people run local, glad to see that getting real attention

Chiming Wang profil fotoğrafı
Chiming Wang3 ay önce

local tool calling has always been possible, here is a 4B model doing it (running on my macbook):

Osaurus profil fotoğrafı
Osaurus3 ay önce

Of course. But osaurus runs it on a sandbox (isolated VM using containerization framework). And we have a diverse set of tools & plugins (mail, notes, calendar etc) which was previously unusable with local models. That's what we've improved & the same is the context of this post!

VibeRaven profil fotoğrafı
VibeRaven3 ay önce

Local models is the future

Osaurus profil fotoğrafı
Osaurus3 ay önce

And the future belongs to open source!

VibeRaven profil fotoğrafı
VibeRaven3 ay önce

😃

ᐱ ᑎ ᑐ ᒋ ᕮ profil fotoğrafı
ᐱ ᑎ ᑐ ᒋ ᕮ3 ay önce

Quitareis una versión para linux en el futuro?

Osaurus profil fotoğrafı
Osaurus3 ay önce

Yes it's in our roadmap

SourceCodeplz profil fotoğrafı
SourceCodeplz3 ay önce

So this a fine tune ? For better function calling ?

Osaurus profil fotoğrafı
Osaurus3 ay önce

No. This is just a usual quantized version of Gemma 4 with no fine tuning. The real improvement is in the harness itself which makes tool calling reliable for all local models

Michał Piszczek profil fotoğrafı
Michał Piszczek3 ay önce

right tool is the easy half. dependable means it recovers when the call returns garbage, not just when it picks correctly

Osaurus profil fotoğrafı
Osaurus3 ay önce

We already have an agent loop that does exactly this! The models won't give up after a single tool call failure. Failures are fed back to the model as structured error envelopes so the model can correct course on the next turn rather than blindly repeat itself

Alican profil fotoğrafı
Alican3 ay önce

When ornith?

Osaurus profil fotoğrafı
Osaurus3 ay önce

Soon! It's already in the works

Santosh | n0c0de.com profil fotoğrafı
Santosh | n0c0de.com3 ay önce

Gonna try that now!

Dido Realm profil fotoğrafı
Dido Realm3 ay önce

macOS only 😔

anarkhgatsby profil fotoğrafı
anarkhgatsby3 ay önce

好像不能安装mcp

Osaurus profil fotoğrafı
Osaurus3 ay önce

You can connect to remote MCP providers (Settings -> Provider -> Add Provider) and also run Osaurus as an MCP server

Prabha profil fotoğrafı
Prabha3 ay önce

This is only for mac, right?

Osaurus profil fotoğrafı
Osaurus3 ay önce

Yes. For now

Stackin' ฿its profil fotoğrafı
Stackin' ฿its3 ay önce

I wish the Ui was designed in a more Codex-esque approach... ie... more project leaning instead of chat-centric. But... ngl... this interface might be one of the most reliable (open-source) interfaces for tool-calling that I've used thus far. Fast/reliable. Great work!

Osaurus profil fotoğrafı
Osaurus3 ay önce

Appreciate the feedback!

Shivay Lamba profil fotoğrafı
Shivay Lamba3 ay önce

how do you improve the accuracy when it comes to tool calls and hallucinations

Milton Yan profil fotoğrafı
Milton Yan3 ay önce

getting tool selection right at 12B params is the hard part — once that's reliable, the interesting problem shifts to what happens when you chain multiple correct calls and the side effects accumulate

Osaurus profil fotoğrafı
Osaurus3 ay önce

We've already solved this problem with agent loops!

Alice The Ai Expert profil fotoğrafı
Alice The Ai Expert3 ay önce

Great Gemma 4 12B small, on device, 60% faster, and actually reliable at tool calling.

The one profil fotoğrafı
The one3 ay önce

It is tempting to try Osaurus myself. But i am using my mac headless. Lots of good apps look like they are made with desktop use primarily. Which is fine. Guess i am “weird” again. I will read the docs and try it.

Osaurus profil fotoğrafı
Osaurus3 ay önce

Osaurus also exposes a local server with a drop-in OpenAI compatible endpoint. You can point any OpenAI client at it and it still works great when running headless

RAZA | AI EXPLORER profil fotoğrafı
RAZA | AI EXPLORER3 ay önce

Precision meets speed. Gemma 4 12B means business.

Sebastian Buzdugan profil fotoğrafı
Sebastian Buzdugan3 ay önce

on-device wins until one bad schema update turns correct calls into silent failures

Benzer Videolar