Video yükleniyor...
Video Yüklenemedi
ICYMI: Not every step in an agent workflow needs the same model. Meet NVIDIA NeMo Switchyard, a new open-source library for model routing. Use frontier models for complex reasoning and planning, and NVIDIA Nemotron Lightning for high-volume, specialized execution.
16,629 görüntüleme • 1 ay önce •via X (Twitter)
15 Yorum

Learn more:

🕸🤖🕸

Geospatial search is a good routing test. A small model can handle place, date and sensor filters. Only ambiguous queries like “recent burn scars near vineyards” need the expensive model.

你好,为什么测试你们平台的模型,几乎全部都连接失败,使用测试成功的模型也经常报错

@grok exine structure closely matched modern maize pollen.

This is the direction agents are heading 👀 Right model, right task, better efficiency 🚀

Looking forward to testing the escalation and prefill routers in real workflows.

I thought we already had model routing….?

Routing to the right model depending on the job size and complexity. That's how you get real cost compression and value for your bucks.

Can I talk to someone from this team? Why is this not ready/useable for production workload?

Could my @openclaw pass the message to the router which then dispatches to the right model? That would be cool (provided it works). Equally could Sol be able to dispatch to Terra and Luna via this to optimise token usage?

I'm a contributor :)

Exactly the kind of routing layer enterprise agents need. I separate complex reasoning from high volume execution across my digital health stacks. Nemotron Lightning is ideal for those inference steps.

按任务难度分配模型,确实比全程堆大模型省得多。

routing is the missing layer. frontier for thinking, Lightning for volume. this is how agent stacks should work




