正在加载视频...

视频加载失败

.OpenRouter co-founder Alex Atallah thinks decision models like Jev could be the alignment layer for agents: "One of the cool potential applications of Jev and other decision models like it is going to be alignment: checking to see if a tool call or an agent-to-agent communication is aligned." "There's...

37,399 次观看 • 6 天前 •via X (Twitter)

19 条评论

Fanfulla 的头像
Fanfulla6 天前

@OpenRouter False stops on normal tool calls would annoy me fast

DeAI Summit Malta 的头像
DeAI Summit Malta5 天前

@OpenRouter The distinction between a model-level check and enforceable controls is crucial—especially once agents can trigger real-world actions. Typed policies, spend limits and auditability make alignment operational, not just aspirational.

jose arce 的头像
jose arce6 天前

For agents that touch money, a model checking another model is a filter, not a control. The check that holds is the one enforced below the model: spend limits, allowlisted counterparties, and a signing policy the agent cannot rewrite. A cheap classifier can flag a bad tool call. Only the key can refuse it.

Anderson 的头像
Anderson5 天前

@OpenRouter Agents policing agents sounds like a great idea

Pentra 的头像
Pentra6 天前

@OpenRouter if you're checking every tool call with another model, you've just doubled your inference cost on the happy path. the alignment layer becomes cheaper to run than the mistake it prevents, or it doesn't ship.

Tyzo 的头像
Tyzo5 天前

@OpenRouter Checking tool calls for agent alignment could reduce errors significantly

Wong Yong Jie 的头像
Wong Yong Jie5 天前

@OpenRouter Already doing this for our AI agent platform at

LFGCAPO 的头像
LFGCAPO6 天前

@OpenRouter The hidden-policy example makes the monitor part of the attack surface: a block reason or trace could reveal the constraint to the red-team agent. Are they testing policy leakage alongside false positives and per-call latency?

Ahnaf 的头像
Ahnaf5 天前

@OpenRouter Jev as an alignment check could prevent bad calls

Playground 的头像
Playground6 天前

@OpenRouter the idea of a decision model sitting on top of every tool call as an alignment check is sharp. the hard part still feels like picking which call to make before that check even runs.

Tyzo 的头像
Tyzo5 天前

@OpenRouter How exactly could Jev check tool call alignment

lvnbbs_bnb 的头像
lvnbbs_bnb6 天前

@OpenRouter using specialized models as a circuit breaker for agent logic is the only way to scale this safely

sofie p. 的头像
sofie p.5 天前

@OpenRouter Interesting perspective

icefrog.◎ 的头像
icefrog.◎5 天前

@OpenRouter recursive guardrails are the only way to scale agentic ops

Ahnaf 的头像
Ahnaf6 天前

@OpenRouter decision models like Jev could really help with agent alignment

泰坦链 | DeFi 的头像
泰坦链 | DeFi6 天前

@OpenRouter layering cheap decision models for guardrails is smart, structural integrity over prompt engineering every time

João Capital 的头像
João Capital6 天前

@OpenRouter O custo latência de checar cada chamada de ferramenta pode explodir o orçamento em escala. 💸

Gill 的头像
Gill5 天前

@OpenRouter Adding an extra check layer to every single tool call is going to spike latency hard.

Winston B. 的头像
Winston B.6 天前

@OpenRouter Jev already routes requests to models, so alignment is the same typed classifier call, just asked one step later in the loop. One cheap decision call at each branch point could answer both "which model?" and "is this tool call allowed?"

相关视频

Jordan Peterson: "If you're competent and silent you will be ignored." "You might think well people should reward you because you're competent. And yes of course they should. But if you're competent and silent then you're just part of the background that's keeping everything functioning" On why you cannot negotiate from a position of weakness: "If you want to push your career forward you have to push it forward. You have to be competent and you have to be strategic" "To be strategic when you negotiate for a new position or a new salary you have to be able to say if you don't give me what I want then something you don't like will happen to you" "It's not a physical threat. It's that you have an option" "You have your CV in order. You're educated and competent and desirable to people outside of your immediate job. You're willing to instantly put yourself on the job market" "So that when you go talk to the person you're negotiating with you're credible" On the one thing most people completely miss about negotiation: "It's very seldom that you're talking to the person who's at the top of the pecking order. What you need to do is tell them a story that they can tell to their boss to make you not a problem" "A good story is look we really need this person because they're hyper competent and they have a better offer" "If you go in there with no power you're going to lose. Obviously" He concludes with the one thing every competent person needs to hear: "The first thing you need to know if you're going to negotiate is that you have to be able to say no. And what no means is that you're not going to do it"

Brad

35,334 次观看 • 2 个月前

Pi was built when there were already agent harnesses around. Here’s why Mario Zechner(Mario Zechner), found them suboptimal and built Pi, a minimalist self-modifying agent: #1 - Mario initially was a believer in Claude Code: "I was a believer in Claude code because they were the first that packaged agentic search up in a really compelling package. And at the time that fit my workflow really well. Everything around the LLM was kind of nice and tidy and easy to understand. I was super happy. I was proselytising Claude code." #2 - Reverse engineering Claude Code highlighted the degradation that Mario felt as a user: "I personally like simple tools that are stable and that I can rely on. Even if they have non-deterministic parts, all the deterministic parts should be as stable as possible. That was just not the experience with Claude Code around summer 2025. They would take away your control of the context. They would inject stuff behind your back, which is bad. Then, your workflows stopped working because there's now a system reminder that you don't even see in the UI that would modify the behaviour of the model. They would also do this to the system prompt. I built a little service where I can track the progression or evolution of the system, prompt and tool definitions and, with every release, it was messing with stuff. That just messed with my workflows and I don't appreciate that." #3 - PI was built with an appreciation for simple and reliable tools: "If I commit to a development tool, I want it to be a stable, reliable thing like a hammer. I don't want my hammer to break a different spot every day. That's terrible. We need somebody who goes the full velocity kind of way. But I don't want to work with a tool like that."

The Pragmatic Engineer

63,020 次观看 • 5 个月前

Mark Zuckerberg, founder, chairman and CEO of Meta, on how he actually runs the company: "I basically don't believe in delegation." He starts by naming the assumption he's rejecting. "I think that there's sort of this theory that a lot of people have, which is like, all right, the job of a leader is you hire people and you delegate things to them." Zuckerberg doesn't operate that way. He's also clear that he can't personally handle everything: "My theory is there's so much going on across the company that I can't possibly get involved in all of it. So all these people are going to have a ton of stuff that they're going to do. But fundamentally, if there is a decision that I want to be involved with, I'm going to be involved in it." The distinction carries the whole argument. He reserves the right to review any call, and no layer of the org chart gets to tell him something is beneath him or outside his lane. His reasoning comes back to accountability: "I think that's generally a good way for founders to operate. If you're running the company and you're on the hook for everything and there's something that's important at whatever level of detail in the organization, I don't get the logic of saying 'I'm not going to be involved in that.'" That's the load-bearing line. If you're answerable for the output, opting out of the input makes no sense. Mark Zuckerberg does attach one condition, and it's a real one rather than a throwaway: "You want to have humility and know that if you're diving into some decision, you may not have the most context immediately." Diving in means showing up without pretending you already have the answer. The expectation extends across the whole company: "I generally think that you want to be able to just have the cultural expectation that things are not going to be so hierarchical and you're just going to dive into whatever you want."

Big Brain Business

45,196 次观看 • 1 个月前

Catherine Austin Fitts: "The problem isn't that [our] currency is fiat...[and] you are not going to fix this situation by going to gold...the central bankers have accumulated all [of it]. [And] now...you're [saying] we're going to go to a gold system? Are you out of your mind?" This clip of Fitts, a former Assistant Secretary of Housing and Urban Development, investment banker, and founder of the Solari Report (The Solari Report | Catherine Austin Fitts), is taken from a discussion with Jerm Warfare posted to the UK Column News (UK Column) YouTube channel on January 26, 2026. ----------------Partial transcription of clip--------------- "The thing that makes the current system what they would call slavery is debt-basing and secrecy, okay? And the failure of their elected representatives to obey the law. "So you have lawlessness, you have debt basing and you have secrecy, okay? The problem is not that the currency is fiat, because what I will tell you is if you go back through history, if you read Alexander Del Mar, the most effective currencies in the world are fiat currencies that are well governed. "We have a debt-based fiat currency that is not well governed in my opinion. But it could be. Now, remember, there has been almost no support in the general population for managing it responsibly. Everybody was like, no, don't manage it responsibly, get me my check. And if that means you're irresponsible, that's okay, I want my check. "But you are not going to fix this situation by going to gold and silver. You're going to make it much worse. Because while we've done this sort of hear no evil, see no evil, speak no evil, for 30 years the central bankers have accumulated all the gold. So now that they have all the gold, you're going to tell me we're going to go to a gold system? Are you out of your mind? "Because now they've got the gold and if you start a gold transaction system now you need gold from them and they've got you over a barrel, right? And what are you going to do to get gold? You're going to have to sell your land, you're going to have to sell your kids, you're going to have to sell real assets to get their gold, right? Why would you do that? "Why would you create— You know, you're dependent on your enemy now you're going to increase your dependency on your enemy now. You're out of your mind, okay? That's not a sound money system. Especially because they want to make it digital. And so they're going to have fiat gold, which is even— I mean, if you think fiat is bad, wait til you see fiat gold when they own all the gold. "So you know, what we want is we want a fiat system and we want it with lawful and no secrecy or minimal secrecy. You're going to have to have some secrecy and a good governance system. Can we get there? Of course we can get there, but we can't get there if you have an entire population that is absolutely committed to corrupt short-term behavior."

Sense Receptor

37,073 次观看 • 8 个月前

Airtable's Howie Liu says that basically everyone will need to graduate from being ICs to ICs that manage teams of 20-30 agents: "The best developers today don't just sit there in front of their IDEs and synchronously talk to their agent." "[Instead], you have like 30 separate branches that are each being worked on by a different agent. And you can have the agents continue to update the branches based on human and other agent feedback." "And I think this whole idea of it taking hours for that entire loop to complete — agent pushes some changes, the changes get feedback from other agents or humans, the agent responds to that — that whole loop could be hours, not just minutes. So you're not going to just sit there and watch it one at a time." "But the powerful thing about this is, each one is still actually operating faster than a human engineer. One agent on one branch can do the work of maybe three humans, operating 3x as fast. So it's like a 10x leverage factor just for one agent." "But the best engineers are now able to multitask and say, 'I'm going to oversee my own little team of 20-30 agents working concurrently.'" "Everyone needs to graduate from being an IC to an IC manager of agents. Meaning, if you're a VC analyst, your job should no longer be to go synchronously research one company. You need to go and research like 30 companies, and do them all faster, better, and higher quality than you could before." "That's the greatest leap that is going to be challenging for a lot of people in a lot of roles. Because it's a totally different mentality in how you operate, and what your role is."

TBPN

35,595 次观看 • 5 个月前

.OpenRouter co-founder Alex Atallah, in his first podcast since Stripe acquired the company, joins Replit ⠕ co-founder Amjad Masad and a16z's Erik Torenberg on why the future of AI is independence and specialization. In this conversation, Alex walks through how the Stripe deal unfolded, why he wasn't originally looking to sell, and why "payments and inference are going to blend together." Pre-OpenRouter, the typical AI workflow had one model provider to choose from, and little pressure on that provider to lower prices. Now enterprises are diversifying across labs and open-weight models, and every board is asking about AI costs and benchmarks. Amjad argues if your company depends on one AI lab, it can turn into your competitor. So Replit is building the layer that lets enterprises use any model and any cloud, without being locked into either. Alex and Amjad are split on personal agents – Amjad runs one agent across his whole company and loves the cross-domain joins, while Alex says general agents cause you to sacrifice understanding, and argues 10 specialized chiefs of staff beats one superagent. 0:45 How the Stripe deal unfolded 5:05 Why mixing models beats one model 7:25 Forcing the labs to compete on price 8:50 Enterprises want open-weight models 10:30 Every board asks about AI every month 12:25 Why companies must own their intelligence 14:15 Replit as the independence layer 15:10 Everyone is building the same agent 16:35 Why Amjad built bring-your-own-cloud 18:10 Amjad's agent that runs his whole company 19:55 Why 10 specialized agents beat one 23:35 Machines, not humans, should specialize 27:30 Guardrails for agents talking to agents 31:10 Models training their own replacements 33:45 Most tasks don't need a frontier model 40:50 Training small models on Qwen 8B 43:25 The Rust cycle is coming for AI 45:10 Fusion models: frontier quality at half the cost YouTube: Alex Atallah OpenRouter Amjad Masad Erik Torenberg

a16z

546,007 次观看 • 6 天前

.Rob Miles is spitting fire: “People are starting from a prior in which ‘[AIs] are safe until you give me an airtight case for why they're dangerous.’ This framing is exhausting. You explain one of the 10,000 ways that AIs could be dangerous, then they explain why they don't think that specific thing would happen. Then you have to change tack, and then they say, 'your story keeps changing'... "If you're building an AGI, it's like building a Saturn V rocket [but with every human on it]. It's a complex, difficult engineering task, and you're going to try and make it aligned, which means it's going to deliver people to the moon and home again. People ask “why assume they won't just land on the Moon and return home safely?" And I'm like, because you don't know what you're doing! If you try to send people to the moon and you don't know what you're doing, your astronauts will die. [Unlike the telephone, or electricity, where you can assume it’s probably going to work out okay] I contend that ASI is more like the moon rocket. "The moon is small compared with the rest of the sky, so you don't get to the moon by default - you hit some part of the sky that isn't the moon. So, show me the plan by which you predict to specifically hit the moon." And then people say, “how do you predict that [AIs] will want bad things?” There's more bad things than good things! It's not actually a complicated argument... I'm not going to predict specifically where it off into random space your astronauts are going, but you're not going to hit the moon unless you have a really good, technically clear plan for how you do it. And if you ask these people for their plan, they don't have one. What's Yann Lecun’s plan?” "I think that if you're building an enormously powerful technology and you have a lot of uncertainty about what's going to happen, this is bad. Like, this is default unsafe. If you've got something that's going to do enormously influential things in the world, and you don't know what enormously influential things it's going to do, this thing is unsafe until you can convince me that it's safe." HOST: “That’s a good way of thinking about it - with some technologies you can assume that the default will be good or at least neutral, or that the capacity of a person to use this in a very bad way is bounded somehow. There's just only so many people you could electrocute one by one."

AI Notkilleveryoneism Memes ⏸️

77,397 次观看 • 3 年前