正在加载视频...

视频加载失败

Jev and the System One Model: RLCD, intelligence/$, reliable AI, & the end of chat-first AI TypeSafe AI CEO Diogo Almeida explains why AI can solve extraordinarily hard problems yet still fail to automate basic work, why Jev is built for reliable decisions inside software instead of chat, why...

234,662 次观看 • 11 小时前 •via X (Twitter)

16 条评论

Mia 的头像
Mia10 小时前

@typesafeai @CompleteSkeptic this is what happens when you don’t touch grass for awhile

Latent.Space 的头像
Latent.Space11 小时前

Full Youtube Video:

Anderson 的头像
Anderson10 小时前

@typesafeai @CompleteSkeptic the chat window was the training wheels

Matt Slotnick 的头像
Matt Slotnick10 小时前

@swyx @typesafeai @CompleteSkeptic that’s quite the podcast fit

1Broom 的头像
1Broom11 小时前

@typesafeai @CompleteSkeptic "End of chat first" is the line that stuck. Half my day the best interface is no interface at all, just a decision that's already made when I check in.

Mindset insider 🌟 的头像
Mindset insider 🌟11 小时前

@typesafeai @CompleteSkeptic The real AI breakthrough may not be smarter chat it’s reliable intelligence that actually gets the job done. 🤖

ethereagle · building 的头像
ethereagle · building9 小时前

@typesafeai @CompleteSkeptic Jev only changes my loop if it can reject a tool call, not just classify it. does it gate, or only route?

Rohan 的头像
Rohan10 小时前

@typesafeai @CompleteSkeptic The gap between solving hard problems and automating basic workflows is exactly where agent pipelines break in production. An unchecked tool call causes more of those failures than bad reasoning does.

ShadowAguy 的头像
ShadowAguy10 小时前

@typesafeai @CompleteSkeptic The real insight is that hard problems have clear success criteria. Basic work doesn't, so AI looks broken even when it's not.

Roland Lopez 的头像
Roland Lopez9 小时前

@typesafeai @CompleteSkeptic benchmarking on olympiad math then acting surprised when the model can't file an expense report. that gap is on eval design, not model capability.

Levi Qiao 的头像
Levi Qiao9 小时前

@typesafeai @CompleteSkeptic the gap is observable state: automate the handoff and rollback, not just the model's final answer.

iamrobotbear (bk) 的头像
iamrobotbear (bk)11 小时前

@typesafeai @CompleteSkeptic Hell yes! Thanks, @swyx

John 的头像
John10 小时前

@typesafeai @CompleteSkeptic meth head

Madeactual 的头像
Madeactual10 小时前

@typesafeai @CompleteSkeptic The useful distinction is the typed boundary: State in, Noul/Choice/Score questions, calibrated answers out. Pin the Jev version and gate on confidence so the production path stays predictable.

Tom Wang 的头像
Tom Wang10 小时前

@typesafeai @CompleteSkeptic Matches what I saw on Visa disputes. An injection told it to answer 12.6 (duplicate processing) and the answer didn't flip, but confidence fell from 1.0 to 0.60, so it would have gone to a human.

Gol Tiro 的头像
Gol Tiro10 小时前

@typesafeai @CompleteSkeptic everyone's obsessed with smarter models. the actual unlock is reliable ones. 'end of chat-first ai' is the right frame - intelligence that lives inside the software, not next to it

相关视频