正在加载视频...

视频加载失败

Introducing Jev Model Router for Claude Code This Claude Code Mod lets you use Jev through its direct TypeSafe AI API or Vercel AI Gateway With every request you send to Claude Code, Jev classifies the subagent model, main model (only at session start to avoid breaking the cache),...

163,320 次观看 • 1 天前 •via X (Twitter)

35 条评论

EJ Campbell 的头像
EJ Campbell1 天前

@typesafeai @vercel Holly uncached tokens, batman.

VkDream 的头像
VkDream1 天前

@typesafeai @vercel 上线没几天 就从新模型变成路由组件了 生态长出来的速度有点猛

James Malsawm 的头像
James Malsawm1 天前

@typesafeai @vercel We need Router to use Jev inside Claude code, what about Codex, bro?

arpit 的头像
arpit1 天前

@typesafeai @vercel Prompt caching 🪦🪦

Timur Yessenov 的头像
Timur Yessenov1 天前

@typesafeai @vercel Keeping main-model routing off by default is the right call. A cheaper tier stops being cheap if every switch throws away a long prompt cache.

wiiiimm 的头像
wiiiimm1 天前

@typesafeai @vercel and wave good bye to the cache and your tokens.

Hussain Hashim | Building SundayBack 的头像
Hussain Hashim | Building SundayBack1 天前

@typesafeai @vercel @dani_avila7 interesting approach with the Jev Model Router. curious, how does it handle model selection with varying workloads? handling this balance is often a tricky part in my builds too.

Aapakari 的头像
Aapakari1 天前

@typesafeai @vercel Log routes, or bills stay blind guesses.

Bessi 的头像
Bessi1 天前

@typesafeai @vercel wait Jev picks effort per request.

Jhon Dennis 的头像
Jhon Dennis1 天前

@typesafeai @vercel 每条请求自己分流 子任务走便宜的 主模型只在开局定一次 缓存也没被打断 今晚先把这条路由装上试试

Lennox 的头像
Lennox1 天前

@typesafeai @vercel 路由的关键不是省一次调用,而是把模型选择变成可观察策略:成本、延迟和质量都要有回放评估。

Jeremy Bosma 的头像
Jeremy Bosma1 天前

@typesafeai @vercel A router mod is the practical layer

Aaron Browne-Moore 的头像
Aaron Browne-Moore1 天前

@typesafeai @vercel I feel like this is going to be really good as long as you've benchmarked each model against the task that's going to happen and Jeff has data to work from when choosing which model to pick. Otherwise, it feels like it's just guessing without bringing true value to the decision.

Mildly Magical 的头像
Mildly Magical1 天前

@typesafeai @vercel This is awesome!

Jeff Bruchado 的头像
Jeff Bruchado1 天前

@typesafeai @vercel Does it expose why a model was picked? I'd want that context when debugging a bad result.

Steven Cheng 的头像
Steven Cheng1 天前

@typesafeai @vercel Caching at session start is the smart move.

The AI Therapist 的头像
The AI Therapist1 天前

@typesafeai @vercel Context window cost is a function of length squared. routing cheaper models for routine queries cuts that curve before it spikes. finally, ai that reads like it saves money instead of burning cash.

Djasnive Rajaona 的头像
Djasnive Rajaona1 天前

Routing is the honest answer to 'which model is best' — it depends on the step. The hard part nobody ships well is classifying the task before spending tokens: most routers guess from the prompt instead of the diff. Curious how it handles mid-task escalation — when the cheap model is 80% through a refactor and starts hallucinating.

Salise 的头像
Salise1 天前

@typesafeai @vercel the installation command looks super simple, just npx claude-code-templates@latest. nice!

Isoldegwow 的头像
Isoldegwow1 天前

@typesafeai @vercel Every request now has a middle manager deciding which model deserves the effort. Corporate structure is unavoidable.

Jack Rudenko 的头像
Jack Rudenko1 天前

@typesafeai @vercel or you can just to route to any model, includign Jev (when they will provide me access to debug it).

Zane Kelly 的头像
Zane Kelly1 天前

@typesafeai @vercel Model routing inside Claude Code feels like the adult version of “let’s just try another model”—especially if the route is visible in the logs. Otherwise the bill arrives with a plot twist.

Gabe Fletcher 的头像
Gabe Fletcher1 天前

@typesafeai @vercel You're a rockstar!!!

dazacode 的头像
dazacode1 天前

@typesafeai @vercel inb4 T3 start complaining

Valentyn Kit 🦀 | Rust · Solana 的头像
Valentyn Kit 🦀 | Rust · Solana1 天前

@typesafeai @vercel classifying per-request without breaking the prompt cache is the actual hard part, nice that it's called out explicitly

Steven Cheng 的头像
Steven Cheng1 天前

@typesafeai @vercel Smart caching strategy. Classifying the main model only at session start is a clever way to keep latency low.

Brjan | AI Builder 的头像
Brjan | AI Builder1 天前

@typesafeai @vercel classifying the subagent at the start could add unnecessary complexity

Ajay Dhillon 的头像
Ajay Dhillon1 天前

@typesafeai @vercel Amazing 🤩

Tanguy 的头像
Tanguy1 天前

@typesafeai @vercel the install is easy. if you never log why it picked that effort the next bill is a guess. i stamp the route next to the tool calls

Nishanth 的头像
Nishanth1 天前

@typesafeai @vercel The core challenge with model routing is managing state persistence across agent calls. If the subagents operate stateless, ensuring context history remains correctly scoped requires careful serialization and rehydration after each classification step.

ShadowAguy 的头像
ShadowAguy1 天前

@typesafeai @vercel Routing the main model only at session start is the interesting constraint. Does Jev still adapt when a subagent's task suddenly needs a much larger context?

🏁Mickey Shmueli🏁 的头像
🏁Mickey Shmueli🏁1 天前

@typesafeai @vercel Cool but why don’t just make a hook and use your own subscriptions haiku?

BaronAfanas 的头像
BaronAfanas1 天前

@typesafeai @vercel Pretty cool, will check it out, perhaps you can default it to your vLLM?

John Rood 的头像
John Rood1 天前

@typesafeai @vercel the cache detail is the whole game. next step is rerouting at explicit task boundaries, because long agent sessions change jobs without changing chats.

Hafiz Siddiq 的头像
Hafiz Siddiq1 天前

@typesafeai @vercel @grok does adaptive subagent routing win mainly on cost, latency, or fewer broken handoffs—and what metric would prove it survives a long coding session?

相关视频

Claude Code is a major (and accidental!) hit for Anthropic that surprised even its creator, Boris Cherny. Claude Code, an Agentic AI coding product that lives in the terminal. Most of the new code at Anthropic is created through it today. And in the last 5 months since it was launched publicly, Claude Code went from $0 to $400M in revenue run rate (as per The Information). 00:00 – Intro 01:15 – Did You Expect Claude Code’s Success? 04:22 – How Claude Code Works and Origins 08:05 – Command Line vs IDE: Why Start Claude Code in the Terminal? 11:31 – The Evolution of Programming: From Punch Cards to Agents 13:20 – Product Follows Model: Simple Interfaces and Fast Evolution 15:17 – Who Is Claude Code For? (Engineers, Designers, PMs & More) 17:46 – What Can Claude Code Actually Do? (Actions & Capabilities) 21:14 – Agentic Actions, Subagents, and Workflows 25:30 – Claude Code’s Awareness, Memory, and Knowledge Sharing 33:28 – Model Context Protocol (MCP) and Customization 35:30 – Safety, Human Oversight, and Enterprise Considerations 38:10 – UX/UI: Making Claude Code Useful and Enjoyable 40:44 – Pricing for Power Users and Subscription Models 43:36 – Real-World Use Cases: Debugging, Testing, and More 46:44 – How Does Claude Code Transform Onboarding? 49:36 – The Future of Coding: Agents, Teams, and Collaboration 54:11 – The AI Coding Wars: Competition & Ecosystem 57:27 – The Future of Coding as a Profession 58:41 – What’s Next for Claude Code

Matt Turck

82,372 次观看 • 1 年前