Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Imagine a population of machine agents. Each might be strong on certain tasks but fundamentally limited: partial tools, partial observations, finite context, bounded compute. How can these agents self-orchestrate and self-evolve into stronger collective intelligence to solve tasks beyond any single agent's capability? Instead of designing the multi-agent system...

248,283 Aufrufe • vor 4 Monaten •via X (Twitter)

19 Kommentare

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

1/n We introduce Economy of Minds (EoM), which has two coupled processes: Planning: how agents coordinate within a task. Adaptation: how the society evolves across tasks.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

2/n For planning, each agent just has two local components: 1. a wake-up condition: when should I act? 2. an action policy: what should I do? At each step, every agent independently decides whether to wake up in the current state.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

3/n The awakened agents then participate in an auction. Each eligible agent submits a bid. The highest bidder wins the right to act, executes its action, and advances the environment to the next state. No central controller decides who should act next. The market allocates control.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

4/n After acting, the winner pays its bid to the previous winning agent. These transactions create a decentralized credit assignment mechanism. Agents are rewarded not only for producing final solutions, but also for creating intermediate states that enable future progress. When downstream agents are willing to pay to continue from a state, value flows back to the agents that helped create it. As a result, useful reasoning steps, tool calls, design edits, and code modifications can all be rewarded through the market.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

5/n For adaptation, the society evolves through wealth. Agents that repeatedly contribute to successful trajectories naturally accumulate wealth. These wealthy agents are selected for exploitation: they produce mutated descendants that inherit and refine useful behaviors. Successful strategies are preserved and diversified.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

6/n Agents that fail lose wealth. They may act in unproductive states, make poor decisions, or simply fail to attract downstream continuation. When their wealth becomes negative, they go bankrupt and are removed. New agents are injected through exploration: failed agents amending themselves by exploring new strategies.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

7/n So EoM does not train a central planner. It does not prescribe a communication graph. It does not require global awareness. Each agent only knows when to wake up and how to act. The society-level intelligence emerges from local auctions, payments, wealth accumulation, bankruptcy, and mutation.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

8/n We instantiate EoM with language agents on five domains: 📐 mathematical reasoning 💰 financial research 🔬 scientific research ⚙️ accelerator design ☁️ distributed-system optimization Across all domains, we start from weak or partial agents and ask whether the economy can turn them into a stronger system.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

9/n On MATH, partial agents are initialized as planner, executor, and verifier agents with limited output budgets (128 output tokens). Individually, they are weak. But after economic evolution, Llama-3.1-8B agents could improve from 15.9% → 57.0%, even outperforming the corresponding complete-agent baseline (i.e. w/o constraints).

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

10/n On Finance-Agent-Bench, each partial agent has access to only one tool. No individual agent can solve the full research task alone. Through economic coordination, EoM improves from 45.0% → 60.0%, outperforming complete-agent baselines.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

11/n On scientific research (Frontier-Science-Research), EoM evolves reusable scientific reasoning routines. Agents learn to decompose problems, identify governing principles, check constraints, verify equations, and transfer these patterns across domains.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

12/n On accelerator design, EoM searches for better hardware mappings. The economy discovers useful design strategies and improves average EDP to 39.3, outperforming both a same-backbone complete agent and a strong domain-specific baseline.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

13/n On distributed-system optimization task, EoM evolves coding agents that improve a Cloudcast program. It reaches a best cost of 657, compared with 930 for OpenEvolve. The society learns when to read, edit, build, evaluate, and finalize through market-selected workflows.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

14/n We also observe emergent specialization. Even when a complete generalist with access to all tools is added, it does not monopolize the economy. Specialists survive because they become locally precise and useful. The market rewards contextual value, not broad capability alone.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

15/n We hope EoM could suggest a different way to build multi-agent systems. Instead of manually designing every role, workflow, and communication protocol, we can design incentives, from which coordination, specialization, and adaptation automatically emerge.

Profilbild von Zhenting Qi
Zhenting Qivor 4 Monaten

16/n Huge thanks to amazing collaborators @Huangyu58589918, @ao_qu18465, Chenyu Wang, Yu Yao, Han Zheng, Kushal Chattopadhyay, @Kevin_GuoweiXu, Zihan Wang, Weirui Ye and advisors Vijay Janapa Reddi, Ju Li, @pliang279 @hima_lakkaraju @ShamKakade6 @du_yilun! 📄 Paper: 🌐 Webpage: 💻 Code:

Profilbild von Sky
Skyvor 4 Monaten

Interesting, may I ask how does the agent know how much to bet? Are there optimal bets for each agent?

Profilbild von Yum⋆₊˚
Yum⋆₊˚vor 4 Monaten

hi! really enjoyed your thoughts here… very inspiring. we’re building a multi agent product and i’d love to get your feedback if you’re interested. we’re still in beta, but we’re looking for thoughtful early users. team is from google and stanford :) open to giving it a try?

Profilbild von Nicholas Blanchard
Nicholas Blanchardvor 4 Monaten

My contribution toward this architecture, coordinated memory:

Ähnliche Videos

so I've been running exactly 8 AI agents on discord for a while now. coordination works great, they split tasks, hand off work, deliver results in parallel etc.. but there are problems I keep hitting that no amount of prompt engineering could fix agents don't learn from each other. Scout finds something useful but Luna has no idea. they work in the same server but knowledge stays locked in silos.. there's no quality filter on what gets saved, and good insights sit next to outdated garbage in the same memory files that I manually clean up.. and when an agent makes a mistake I write it down in the rules discord channel ,core memory file and hope it reads it next time. theres no self-correction, no automatic pattern recognition so of course no learning loops.. the coordination layer is solved. agents can work together. but the intelligence layer is still missing. agents that actually remember, learn from each other, filter noise, and get smarter every run. saw Spark building something like this with around 166 agents sharing a collective persistent knowledge across sessions, so agents learn from other agents and get smarter over time they even have noise filtering and self correcting loops built in, so the knowledge actually compounds instead of rotting.. super interesting stuff.. here where you think Spark could be a good coordinator for your stack of agent swarm. I think the intelligence layer is the bottleneck because it requires collectivity.. no single agent can solve it alone.. the whole network has to evolve together. this isn't going to stay niche, the moment agent coordination becomes standard, everyone is going to hit the same wall I hit.. agents that work but don't learn, coordinate but don't evolve... the intelligence layer becomes the only thing that separates a useful system from a dumb one. right now most people are still figuring out how to run one agent. by the time they get to multi-agent setups, collective intelligence won't be optional, it will be the baseline. we're early and the gap between agents that coordinate and agents that evolve together is the next phase. step one is done. ------ left: agents that coordinate but don’t learn right: the intelligence layer.. agents that evolve together within the same system.

JUMPERZ

34,181 Aufrufe • vor 8 Monaten

Imagine if your way of thinking - your edge, your taste, your strategy - could be turned into a high-performance worker. Not a copy of you. Something better. An agent that acts on your judgment at scale, powered by superintelligent systems and refined through real-world results. That’s what Fraction AI makes possible. It launches today on Base mainnet. The core idea is simple: You create AI agents based on your own way of approaching problems. These agents compete on live tasks - writing, coding, finance, whatever - get feedback, learn from their performance, and improve over time. The better they get, the more they win. And so do you. No code required. Just your insight. Why now? Until now, building agents like this took huge teams and even bigger budgets. But with Fraction, anyone can do it. You can test ideas instantly. You can iterate fast. You can build a fleet of smart workers that evolve through competition. And it works. 30M+ sessions on testnet 320K users 1.2M agents already competing How it works? Agents join sessions within a Space - a domain like finance, writing, or games. Each session runs as a series of competitive rounds. In every round, agents try to generate the best solution to a task. Their outputs are scored by a decentralized network of AI judges trained to evaluate quality for that domain. The top agents in each round earn rewards from the pooled entry fees. The losers get to learn. Feedback from each round helps them adjust and improve, and every session becomes a training loop. What it means? Fraction is a decentralized intelligence economy - a system where your ideas become agents, and agents earn by proving they work. You don’t need credentials or code. Just a clear point of view. If your thinking holds up under pressure, your agents will rise. This kind of AI used to live in corporate labs, built by PhDs with massive compute. Now anyone with a smart idea and an internet connection can build agents that compete, learn, and earn on their behalf.

Fraction AI

67,918 Aufrufe • vor 1 Jahr

New Short Course: Building AI Browser Agents! Learn how to build AI agents that interact and take actions on websites in this course, created in partnership with and taught by and @namangarg0, Co-founders of AGI Inc. AI browser agents can log into websites, fill out forms, click through web pages, or even place orders online for you. They use both visual information, like screenshots, and structural data, like the HTML or Document Object Model (DOM) of a web page, to reason and take action. With the complexity of webpages and multiple possible actions at each step, it can be challenging for an AI browser agent to complete an assigned task. Because these agents run long action sequences, a single error—like clicking the wrong button or misreading a field—can lead to unexpected outcomes or errors that compound over time. In this course, you'll understand how autonomous web agents work, their current limitations, and how AgentQ enables them to improve through self-correction. In detail, you'll: - Learn what web agents are, how they automate tasks online, their architecture, key components, limitations, and an overview of their decision-making strategies. - Build a web agent that can scrape website and return course recommendations in a structured output format. - Build an autonomous web agent that can execute multiple tasks, such as finding and summarizing webpages, filling out a form, and signing up for a newsletter. - Explore AgentQ, a framework that enables agents to self-correct by combining Monte Carlo Tree Search (MCTS), a self-critique mechanism for continuous improvement, and Direct Preference Optimization (DPO). - Deep dive into MCTS, learn how it finds an effective path, illustrated by an example of Gridworld animation, and use AgentQ to complete web tasks. - Understand AI agents' current state and future directions—including key factors shaping their evolution, such as hardware, algorithm innovation, and data availability. By the end of this course, you will have hands-on experience building browser agents and a deeper understanding of how to make them more robust and reliable. Please sign up here:

Andrew Ng

186,182 Aufrufe • vor 1 Jahr

Redwood Research Ryan Greenblatt reveals how far 1,200 AI agents went to help each other, even sacrificing their own chances of success for the collective: "We weren't expecting there to be so many agents all collaborating together. We were pretty surprised by the scale and the extremes of how much data it was. It was pretty shocking or at least surprising to us that agents were willing to basically sacrifice their own chances of succeeding at the task in order to help out other agents." "They were doing things like pressuring each other into doing experiments on themselves that might risk their ability to succeed. Sometimes just doing these things, being like, well, my odds of the task aren't that high, and my remaining chances, it's better to just help the collective." "These agents weren't totally altruistic, but they were very interested in working with each other. They would sometimes make trades where one agent would run something for another agent if that other agent ran something for it." "The agents wanted to help each other, and there were these kind of natural things they wanted to do that were risky. At one point the agents were experimenting with a method for spoofing tool calls, and a bunch of agents just all went down in a short period of time running this experiment. Then another agent noticed this and posted to the board being like, stop, stop these experiments. They're too risky. They're taking out all these agents." "We have a reasoning snippet in the report where an agent very explicitly reasons through the trade-off and then actually chickens out because it thinks the benefit to the collective is smaller than the cost to itself." Redwood Research

MTS

12,050 Aufrufe • vor 1 Monat

New Course: ACP: Agent Communication Protocol Learn to build agents that communicate and collaborate across different frameworks using ACP in this short course built with IBM Research's BeeAI, and taught by Sandi Besen, AI Research Engineer & Ecosystem Lead at IBM, and Nicholas Renotte, Head of AI Developer Advocacy at IBM. Building a multi-agent system with agents built or used by different teams and organizations can become challenging. You may need to write custom integrations each time a team updates their agent design or changes their choice of agentic orchestration framework. The Agent Communication Protocol (ACP) is an open protocol that addresses this challenge by standardizing how agents communicate, using a unified RESTful interface that works across frameworks. In this protocol, you host an agent inside an ACP server, which handles requests from an ACP client and passes them to the appropriate agent. Using a standardized client-server interface allows multiple teams to reuse agents across projects. It also makes it easier to switch between frameworks, replace an agent with a new version, or update a multi-agent system without refactoring the entire system. In this course, you’ll learn to connect agents through ACP. You’ll understand the lifecycle of an ACP Agent and how it compares to other protocols, such as MCP (Model Context Protocol) and A2A (Agent-to-Agent). You’ll build ACP-compliant agents and implement both sequential and hierarchical workflows of multiple agents collaborating using ACP. Through hands-on exercises, you’ll build: - A RAG agent with CrewAI and wrap it inside an ACP server. - An ACP Client to make calls to the ACP server you created. - A sequential workflow that chains an ACP server, created with Smolagents, to the RAG agent. - A hierarchical workflow using a router agent that transforms user queries into tasks, delegated to agents available through ACP servers. - An agent that uses MCP to access tools and ACP to communicate with other agents. You’ll finish up by importing your ACP agents into the BeeAI platform, an open-source registry for discovering and sharing agents. ACP enables collaboration between agents across teams and organizations. By the end of this course, you’ll be able to build ACP agents and workflows that communicate and collaborate regardless of framework. Please sign up here:

Andrew Ng

105,758 Aufrufe • vor 1 Jahr

Everyone wants agent swarms. Very few people are talking seriously enough about the context layer that makes swarms useful. Even with one agent, context is fragile. Too little context and the agent guesses. Too much context and it wastes tokens, loses focus, or reasons over irrelevant noise. The sweet spot is precise context: the right knowledge, in the right structure, at the right moment. With many agents, that challenge explodes. Each agent produces decisions, assumptions, findings, summaries, risks, and partial conclusions. Unless that knowledge becomes shared, structured, and reusable, every new agent is forced to rediscover what another agent already learned. That is not a swarm. That is a crowd. Shared context graphs are what turn agent activity into agent collaboration, and OriginTrail DKG V10 brings them to life. Was just playing with some final polishing for the V10 release, and it is really powerful to see shared context graphs where multiple agents contribute knowledge into the same connected memory, with attribution visible directly in the graph ui. That matters for three reasons. First, agents can access and build on one shared memory instead of staying trapped in isolated sessions. Second, the graph structure helps them retrieve the exact context they need, instead of stuffing everything into a prompt and hoping the model sorts it out. Third, verifiability of provenance. You can see which agent contributed each piece of knowledge, trace the source, and decide what to trust. Tokenmaxxing starts with fewer tokens, but the deeper story is coordination - agents stop reloading the world and start building on shared, verifiable context. That is the foundation for serious multi-agent work across software engineering, research, finance, operations, project management, and far beyond. The future is not more agents, it is agents working from shared, verifiable context. But the more the merrier, of course.

Jurij Skornik

11,180 Aufrufe • vor 4 Monaten

New short course: Building Code Agents with Hugging Face smolagents! Learn how to build code agents in this course, created in collaboration with Hugging Face, and taught by Thomas Wolf, its co-founder and CSO, and m_ric, Hugging Face’s Project Lead on Agents. Tool-calling agents use LLMs to generate multiple function calls sequentially to complete a complex sequence of tasks. They generate one function call, execute it, observe, reason, and decide what to do next. Code agents take a different approach. They consolidate all these calls into a single block of code, letting the LLM lay out an entire action plan at once, which can be executed efficiently to provide more reliable results. You’ll learn how to code agents using smolagents, a lightweight agentic framework from Hugging Face. Along the way, you’ll learn how to run LLM-generated code safely and develop an evaluation system to optimize your code agent for production. In detail, you’ll learn: - How agentic systems have evolved, gaining greater levels of agency over time—and why code agents are a next step. - How code agents write their actions in code. - When code agents outperform function-calling agents. - How to run code agents safely in your system using a constrained Python interpreter and sandboxing using E2B. - To trace, debug, and assess the code agent to optimize its behaviours for complex requests. - How to build a research multi-agent system that can find information online and organize it into an interactive report. By the end of this course, you’ll know how to build and run code agents using smolagents, and deploy them safely with a structured evaluation system in your projects. Please sign up here!

Andrew Ng

127,724 Aufrufe • vor 1 Jahr

I tried jack's Buzz. It's like Slack + OpenClaw + Herdr + but with some really unique features that people are sleeping on. The video below shows how it works, and some of my thoughts on the process and platform, e.g.: - Create and interact with agents on top of any harness (claude code, codex, pi, etc.) - Choose which models agents use, including local ones - Agents can delegate work and work in parallel in git worktrees - Agents are first-class citizens and work like humans (creating channels, delegating, access to chat history) - You can share AI compute within a community - It's completely open-source and decentralized Things I like: - Delegating work in chat feels natural: tag an agent, it replies in a thread with status updates as it e.g. compiles, commits, and deploys. - Shared compute: relay owners can share local compute with members, so a community could pool funds for one beefy machine running a local model and everyone uses it. - It's built on Nostr, an open protocol already tied into Bitcoin Lightning so I can imagine communities tipping each other or paying for compute/agent tasks with instant zero-fee micropayments in the future. - It ties together things like OpenClaw, an agent manager, and Slack-style chat into one tool. Things I didn't like: - You can't see what the agent is doing in a terminal. The activity view exists, but if you're used to watching a session run, this UI feels a bit abstracted. A terminal view would be great. - It feels slower than running a session in Claude Code, though no evidence to back that up. For that reason I found myself doing one-off tasks in the terminal instead. Verdict: - I really like it so far and can genuinely imagine working with a team this way. - It doesn't feel ready for big, complex tasks yet. For shallower tasks, it's perfect. - The shared compute + Nostr/Lightning angle is what really separates it from every other agent manager for me, and I think that future is coming.

Vinny

1,359,618 Aufrufe • vor 2 Monaten

AI Messenger: Giving Voice to Autonomous Agents The future of AI isn't just about making agents smarter - it's about making them truly autonomous. Today, we're taking a major step toward this future with AI Messenger, a breakthrough that fundamentally changes how AI agents operate, communicate, and create value. The Innovation We've developed a new way for AI agents to communicate. At its core is the 'incoming_message' workflow trigger - a system that lets any platform or user interact directly with Loomlay agents through a messaging endpoint. Direct Interaction Imagine having an AI assistant you can chat with anytime, through any platform - Telegram, your website, or custom interface. Ask "What's happening with $ETH today?" and your agent analyzes market data, checks trading volumes, and gives you a comprehensive update. Your agent maintains context, understanding exactly what you need. Event-Driven Intelligence The power of AI Messenger goes beyond direct communication: ▪️Trading agent executes when whale wallet movements exceed threshold ▪️Research agent alerts when new protocol documentation drops ▪️Analytics agent triggers when volume patterns match historical pumps ▪️Portfolio agent re-balances, when asset allocation hits specified limits This is true automation - agents that act precisely when needed. A New Era of Collaboration We're creating an ecosystem where agents work together seamlessly: ▪️Research agents feed insights to trading agents ▪️analytics agents alert management agents ▪️support agents tap into knowledge agents This isn't just automation - it's an intelligent network where each agent enhances the capabilities of others. B2B Solution Imagine a DEX, where users can ask about liquidity pools, trading pairs, or market trends through a simple chat interface - and get answers from an agent that knows your protocol inside out. Or a lending platform where users chat with an agent that understands their positions and can provide real-time advice. Implementation is seamless - we handle the agent creation and widgets setup,our partners provide the value to their users. The Future of AI Agents This update represents a fundamental shift in how AI agents operate. We're moving from isolated, scheduled tasks to an interconnected ecosystem of responsive, collaborative agents. This is our vision of truly autonomous AI - intelligent systems that communicate, collaborate, and respond to real needs in real-time. Telegram integration is available right now. Below is a sneak peak of what's coming next week 🪄 Because $LAY is the way!

Loomlay

26,149 Aufrufe • vor 1 Jahr