Загрузка видео...

Не удалось загрузить видео

На главную

Introducing Jev Model Router for Claude Code This Claude Code Mod lets you use Jev through its direct TypeSafe AI API or Vercel AI Gateway With every request you send to Claude Code, Jev classifies the subagent model, main model (only at session start to avoid breaking the cache),...

163,320 просмотров • 1 день назад •via X (Twitter)

Комментарии: 35

Фото профиля EJ Campbell
EJ Campbell1 день назад

@typesafeai @vercel Holly uncached tokens, batman.

Фото профиля VkDream
VkDream1 день назад

@typesafeai @vercel 上线没几天 就从新模型变成路由组件了 生态长出来的速度有点猛

Фото профиля James Malsawm
James Malsawm1 день назад

@typesafeai @vercel We need Router to use Jev inside Claude code, what about Codex, bro?

Фото профиля arpit
arpit1 день назад

@typesafeai @vercel Prompt caching 🪦🪦

Фото профиля Timur Yessenov
Timur Yessenov1 день назад

@typesafeai @vercel Keeping main-model routing off by default is the right call. A cheaper tier stops being cheap if every switch throws away a long prompt cache.

Фото профиля wiiiimm
wiiiimm1 день назад

@typesafeai @vercel and wave good bye to the cache and your tokens.

Фото профиля Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBack1 день назад

@typesafeai @vercel @dani_avila7 interesting approach with the Jev Model Router. curious, how does it handle model selection with varying workloads? handling this balance is often a tricky part in my builds too.

Фото профиля Aapakari
Aapakari1 день назад

@typesafeai @vercel Log routes, or bills stay blind guesses.

Фото профиля Bessi
Bessi1 день назад

@typesafeai @vercel wait Jev picks effort per request.

Фото профиля Jhon Dennis
Jhon Dennis1 день назад

@typesafeai @vercel 每条请求自己分流 子任务走便宜的 主模型只在开局定一次 缓存也没被打断 今晚先把这条路由装上试试

Фото профиля Lennox
Lennox1 день назад

@typesafeai @vercel 路由的关键不是省一次调用,而是把模型选择变成可观察策略:成本、延迟和质量都要有回放评估。

Фото профиля Jeremy Bosma
Jeremy Bosma1 день назад

@typesafeai @vercel A router mod is the practical layer

Фото профиля Aaron Browne-Moore
Aaron Browne-Moore1 день назад

@typesafeai @vercel I feel like this is going to be really good as long as you've benchmarked each model against the task that's going to happen and Jeff has data to work from when choosing which model to pick. Otherwise, it feels like it's just guessing without bringing true value to the decision.

Фото профиля Mildly Magical
Mildly Magical1 день назад

@typesafeai @vercel This is awesome!

Фото профиля Jeff Bruchado
Jeff Bruchado1 день назад

@typesafeai @vercel Does it expose why a model was picked? I'd want that context when debugging a bad result.

Фото профиля Steven Cheng
Steven Cheng1 день назад

@typesafeai @vercel Caching at session start is the smart move.

Фото профиля The AI Therapist
The AI Therapist1 день назад

@typesafeai @vercel Context window cost is a function of length squared. routing cheaper models for routine queries cuts that curve before it spikes. finally, ai that reads like it saves money instead of burning cash.

Фото профиля Djasnive Rajaona
Djasnive Rajaona1 день назад

Routing is the honest answer to 'which model is best' — it depends on the step. The hard part nobody ships well is classifying the task before spending tokens: most routers guess from the prompt instead of the diff. Curious how it handles mid-task escalation — when the cheap model is 80% through a refactor and starts hallucinating.

Фото профиля Salise
Salise1 день назад

@typesafeai @vercel the installation command looks super simple, just npx claude-code-templates@latest. nice!

Фото профиля Isoldegwow
Isoldegwow1 день назад

@typesafeai @vercel Every request now has a middle manager deciding which model deserves the effort. Corporate structure is unavoidable.

Фото профиля Jack Rudenko
Jack Rudenko1 день назад

@typesafeai @vercel or you can just to route to any model, includign Jev (when they will provide me access to debug it).

Фото профиля Zane Kelly
Zane Kelly1 день назад

@typesafeai @vercel Model routing inside Claude Code feels like the adult version of “let’s just try another model”—especially if the route is visible in the logs. Otherwise the bill arrives with a plot twist.

Фото профиля Gabe Fletcher
Gabe Fletcher1 день назад

@typesafeai @vercel You're a rockstar!!!

Фото профиля dazacode
dazacode1 день назад

@typesafeai @vercel inb4 T3 start complaining

Фото профиля Valentyn Kit 🦀 | Rust · Solana
Valentyn Kit 🦀 | Rust · Solana1 день назад

@typesafeai @vercel classifying per-request without breaking the prompt cache is the actual hard part, nice that it's called out explicitly

Фото профиля Steven Cheng
Steven Cheng1 день назад

@typesafeai @vercel Smart caching strategy. Classifying the main model only at session start is a clever way to keep latency low.

Фото профиля Brjan | AI Builder
Brjan | AI Builder1 день назад

@typesafeai @vercel classifying the subagent at the start could add unnecessary complexity

Фото профиля Ajay Dhillon
Ajay Dhillon1 день назад

@typesafeai @vercel Amazing 🤩

Фото профиля Tanguy
Tanguy1 день назад

@typesafeai @vercel the install is easy. if you never log why it picked that effort the next bill is a guess. i stamp the route next to the tool calls

Фото профиля Nishanth
Nishanth1 день назад

@typesafeai @vercel The core challenge with model routing is managing state persistence across agent calls. If the subagents operate stateless, ensuring context history remains correctly scoped requires careful serialization and rehydration after each classification step.

Фото профиля ShadowAguy
ShadowAguy1 день назад

@typesafeai @vercel Routing the main model only at session start is the interesting constraint. Does Jev still adapt when a subagent's task suddenly needs a much larger context?

Фото профиля 🏁Mickey Shmueli🏁
🏁Mickey Shmueli🏁1 день назад

@typesafeai @vercel Cool but why don’t just make a hook and use your own subscriptions haiku?

Фото профиля BaronAfanas
BaronAfanas1 день назад

@typesafeai @vercel Pretty cool, will check it out, perhaps you can default it to your vLLM?

Фото профиля John Rood
John Rood1 день назад

@typesafeai @vercel the cache detail is the whole game. next step is rerouting at explicit task boundaries, because long agent sessions change jobs without changing chats.

Фото профиля Hafiz Siddiq
Hafiz Siddiq1 день назад

@typesafeai @vercel @grok does adaptive subagent routing win mainly on cost, latency, or fewer broken handoffs—and what metric would prove it survives a long coding session?

Похожие видео

Claude Code is a major (and accidental!) hit for Anthropic that surprised even its creator, Boris Cherny. Claude Code, an Agentic AI coding product that lives in the terminal. Most of the new code at Anthropic is created through it today. And in the last 5 months since it was launched publicly, Claude Code went from $0 to $400M in revenue run rate (as per The Information). 00:00 – Intro 01:15 – Did You Expect Claude Code’s Success? 04:22 – How Claude Code Works and Origins 08:05 – Command Line vs IDE: Why Start Claude Code in the Terminal? 11:31 – The Evolution of Programming: From Punch Cards to Agents 13:20 – Product Follows Model: Simple Interfaces and Fast Evolution 15:17 – Who Is Claude Code For? (Engineers, Designers, PMs & More) 17:46 – What Can Claude Code Actually Do? (Actions & Capabilities) 21:14 – Agentic Actions, Subagents, and Workflows 25:30 – Claude Code’s Awareness, Memory, and Knowledge Sharing 33:28 – Model Context Protocol (MCP) and Customization 35:30 – Safety, Human Oversight, and Enterprise Considerations 38:10 – UX/UI: Making Claude Code Useful and Enjoyable 40:44 – Pricing for Power Users and Subscription Models 43:36 – Real-World Use Cases: Debugging, Testing, and More 46:44 – How Does Claude Code Transform Onboarding? 49:36 – The Future of Coding: Agents, Teams, and Collaboration 54:11 – The AI Coding Wars: Competition & Ecosystem 57:27 – The Future of Coding as a Profession 58:41 – What’s Next for Claude Code

Matt Turck

82,372 просмотров • 1 год назад