Loading video...

Video Failed to Load

Go Home

Introducing typesafe/jev-router: a cache-aware model router powered by Jev and TypeSafe AI The Jev Router picks the best model and reasoning effort for each request, balancing quality, speed, and cost. Here's how it works 👇🏻

1,297,534 views • 3 days ago •via X (Twitter)

18 Comments

OpenRouter's profile picture
OpenRouter3 days ago

On four agent benchmarks, Jev Router solved 82% more tasks than our Auto Router (237 vs 130 of 423).

OpenRouter's profile picture
OpenRouter3 days ago

On five agent benchmarks, Jev Router had a faster median time to first token than every other router we tested.

OpenRouter's profile picture
OpenRouter3 days ago

Most routers pick a model based on each message. Each switch loses the cached chat, so you pay full price for the new model to reread the whole conversation. Most routers also route on the task type, not the difficulty. Easy and hard coding tasks get the same model.

OpenRouter's profile picture
OpenRouter3 days ago

Jev Router runs on Jev, TypeSafe's first Decision / System One model. Before each turn, Jev reads your prompt and scores it on difficulty and precision. Also checks whether a bigger model or more effort would help, whether a cheaper model is enough, and whether the task changed.

OpenRouter's profile picture
OpenRouter3 days ago

Jev reads the conversation text only to pick the model and effort. It runs under zero data retention (ZDR) terms, so nothing is stored or trained on. Attachments are never sent to Jev, and requests with "zdr: true" work with Jev Router.

OpenRouter's profile picture
OpenRouter3 days ago

Jev Router keeps a model that works for the rest of the session. It can raise or lower effort without switching models. It switches only when the expected gain is larger than the cost, including the cached chat it would lose.

OpenRouter's profile picture
OpenRouter3 days ago

If the Jev call times out or returns invalid output, the request fails instead of falling back to another router. Each response includes routing metadata with the reason for the choice. Try it now using "typesafe/jev-router" in your app, or, see Jev Router's thinking.

OpenRouter's profile picture
OpenRouter3 days ago

You can see Jev Router's thinking process in OpenRouter Chat. The routing insights panel shows the model it picked for each turn and the scores behind it, like task, difficulty, precision, and larger model benefit. Test Jev Router here:

OpenRouter's profile picture
OpenRouter3 days ago

Read more:

Ventrue's profile picture
Ventrue3 days ago

@typesafeai is there a directive to limit the worst model it can route to? "Hey Jev don't ever route it to [this dogshit model]"

Timothy Kassis's profile picture
Timothy Kassis3 days ago

@typesafeai Are we able to specifically a list of models to choose from?

peyCyber's profile picture
peyCyber3 days ago

@typesafeai cache-aware routing should make repeated prompts cheaper without sacrificing quality

John Rood's profile picture
John Rood3 days ago

@typesafeai the fail-closed call is right, but make the router timeout its own error class: routing unavailable, not a task failure. otherwise every harness counts it against the agent and retries the work when the only thing that needs retrying is the routing decision.

Pode vir's profile picture
Pode vir3 days ago

@typesafeai cache-aware routing is just optimizing for what makes them money not what makes your product better

Joel's profile picture
Joel3 days ago

@typesafeai they finally shipped a router that stops my wallet from crying

MrOzi's profile picture
MrOzi3 days ago

@typesafeai Önbellek farkındalığı, yönlendirmeyi basit bir maliyet oyunundan çıkarıp gerçek bir mühendislik problemine dönüştürüyor.

Beto Muniz's profile picture
Beto Muniz3 days ago

@typesafeai And if you use Paseo, I created a TypeSafe's Jev (it also support local Laya as a classifier) model router for free:

Miguel's profile picture
Miguel3 days ago

This kind of tools will be a game changer once we figure out how to share cache between different models. I currently run something like this for my harness but it's only a single call at the beginning of every conversation. It's so unbeneficial that each call _might_ use the cache

Related Videos

this is unreal f*cking gold for Jev builders 20 repos people are building on Jev right now. browser agents, context tools, trading bots, even a drone 1. JEV-Ultrafast - a fast browser agent ↳ 2. Fast-JEV-Compaction - context compression ↳ 3. JSON-Render - generative UI ↳ 4. Typesafe-MCP - use Jev with any client ↳ 5. JEV-MCP - a judgment toolkit ↳ 6. Semdecide - a classifier that lives in your CLI ↳ 7. JEV-Codex-Router - routes each task to the right model ↳ 8. Winnow - garbage collection for your context ↳ 9. JEV-Review - code review triage ↳ 10. Blink - a repo navigator ↳ 11. Agent-Desktop - desktop automation ↳ 12. Typesafe-Mario - an agent that plays Super Mario ↳ 13. JEV-Drone - drone control ↳ 14. OneVOneJev - a browser FPS ↳ 15. JEV-Trader - HFT market making ↳ 16. Prism - liquidity signal detection ↳ 17. Neo4Jev - knowledge graph traversal ↳ 18. JEV-Curate - training data screening ↳ 19. Canny - checks whether a task was actually completed ↳ 20. KillMyIdea - scores startup ideas before you build them ↳ pick by what you do: > coding -> JEV-Review, Blink, Canny, JEV-Codex-Router > context -> Fast-JEV-Compaction, Winnow > automation -> JEV-Ultrafast, Agent-Desktop > clients and tools -> Typesafe-MCP, JEV-MCP, Semdecide > UI -> JSON-Render > trading -> JEV-Trader, Prism > data -> Neo4Jev, JEV-Curate > founders -> KillMyIdea > just for fun -> Typesafe-Mario, OneVOneJev, JEV-Drone grab the one closest to your job and ship something on top of it this week

Mr. Buzzoni

28,574 views • 4 days ago

Jev has been blowing up lately. If you've got the Jev API but don't know how to play around with it yet, you can just copy this checklist. 1. jev-ultrafast A high-speed browser Agent built with Browser Use. Jev only judges "what to do, which element to click" at each step, and only calls the small model when typing is needed. Searching for a flight on Google Flights takes about 7 seconds. 2. fast-jev-compaction Context compression for Claude Code. Before each tool call, have Jev judge if there's anything still useful; delete the useless stuff, and keep the original text without rewriting it. 3. json-render Vercel Labs' generative UI framework. In experiments, Jev doesn't write JSON token by token; it just handles selecting components, properties, and layouts. 4. typesafe-mcp Best for people who just got the API. Plug Jev into Claude Code, Claude Desktop, Codex, and Pi, and do Choice / Score / Noul anytime. 5. jev-mcp Ready-made Agent judgment toolkit: fact-checking, content screening, semantic ranking, classification, and information extraction. 6. SemDecide Turn Jev into a command-line tool. Directly classify, score, and filter in the Shell—great for hooking up to crawlers, CI, and data pipelines. 7. jev-codex-router First have Jev judge how hard this round of programming tasks is, then decide the model tier, reasoning depth, and speed mode. 8. Winnow Context garbage collection for Claude Code. When Read / Bash / Grep spits out a ton of stuff, Jev first judges which parts are really relevant to the current task. 9. jev-review Before code review, run it through Jev first to pick out high-risk changes, then hand them off to a pricier big model or a human. Comes with a local dashboard. 10. Blink Use Jev as a code repository navigator. At each directory level, judge which files are most relevant to the current issue, then keep digging down. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

rody

199,858 views • 7 days ago