Loading video...

Video Failed to Load

Go Home

MIT PhD student Alex Zhang on why simple harness abstractions beat models trained around their own harness: "When these abstractions are implemented in a really simple way, basically what you see with Prime Agent is that you can slot in any model. We don't train any model around Prime...

146,258 views • 5 days ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

Small Language Models (SML) are the future of AI. "Small" (SML) instead of "Large" (LLM). These small models are highly specialized models with superhuman abilities on specific tasks. Here are two techniques to build these models: • Spectrum • Model Merging I give you a short introduction in the attached video, but here is a quick summary: Spectrum helps us identify the most relevant layers to solve one specific task. We can ignore everything else and focus on fine-tuning these layers. Using Spectrum, we can fine-tune models in a heartbeat. Model Merging combines multiple models into a unique, much better model than any of the individual input models. You can also combine models specialized in different tasks and get a model with multiple abilities. This is the state of the art of productizing models. It's what Arcee.ai's platform does behind the scenes. Arcee collaborated with me on this post and is sponsoring it. There are three main steps to produce a model for your particular use case: 1. You create a dataset by uploading your data. 2. You train a model. At this step, Arcee uses Spectrum and Model Merging to produce a highly specialized model for your task. 3. You can deploy that model to any environment you want. Three important notes: • Training process is 2x faster and 2x cheaper than regular fine-tuning. • Resultant models are smaller and have higher accuracy. • They create these specialized models from open-source models. Check this site so you can fully appreciate how this works: If you want to fine-tune an open-source model, consider Arcee's platform. This is the state of the art.

Santiago

164,162 views • 2 years ago

Perplexity CEO Aravind Srinivas on the brutal truth about who actually makes money in AI (and why it's not who you think): Aravind argues that the real value in AI comes from orchestration. He points to products like Codex, Claude Code, and Perplexity Computer: "What is that? It's an orchestration system. It takes a model, pairs it with an agent harness." And what is an agent harness? "The simplest way of describing it is like rules for how the agent loop should run. What are all the skills and sub-agents and connectors and tools it accesses? Without the harness, you don't necessarily capture and convert the intrinsic intelligence in the model into valuable output tokens." This leads to a blunt conclusion about who has a real business in AI, and who doesn't: "If you're literally just a reseller of model tokens, you have no business, because the model will get commoditized. So even if you're a model builder, you don't have a business. As an infra layer, you have some business on serving those output tokens. But as an application layer or model builder, you don't really have a business if you're just a reseller of tokens that come directly out of the model." So where does the value accrue? "You have a business if you know how to take the model, ground it in valuable context, orchestrate it with a really good agent harness, connected to the right set of tools and connectors (whether it's personal connectors or business connectors) and provide the experience to people in one single unified system." Aravind Srinivas then explains Perplexity's specific edge: Beyond orchestrating across tools, files, and connectors, they also orchestrate across models. "That is the differentiation that Anthropic and OpenAI cannot claim, because you wouldn't find GPT-5 inside the Claude Code harness. You wouldn't find Claude Opus inside the Codex harness. These are competing with each other. Whereas you would find both these models inside Perplexity Computer." Why does this matter? Because it all comes down to power. In Aravind's framing, the fundamental cost driver in AI is watts (the one input nobody can subsidize except the government). "Whoever provides the most valuable output tokens with the least amount of power expended to produce them generates the greatest value to the end user, has the most pricing power, has the most value. That is the orchestration problem to solve." His conclusion: "The one single most important metric in AI is token value per watt per user."

Big Brain AI

40,092 views • 7 days ago

EXTREMELY CONCERNING 🚨 PRIME Drinks is going through a lawsuit. “The lawyer who tested their drink is claiming it has 3x the amount of forever chemicals a human can safely have in their lifetime” What exactly is it that the FDA even does in America? “PRIME is now getting sued and you should be seriously concerned if you have any prime since it released. So prime is now going through a new lawsuit after it was discovered that their drink has PFOs which is forever chemicals but what's really concerning is the fact that the lawyer who tested their drink is claiming it has three times the amount of forever chemicals a human can safely have in their lifetime and the lawsuit is claiming they found these forever chemicals in the grape flavored prime drink. However, it also seems like other flavors might also have these chemicals because other prime flavors are being tested. And we're going to find out after the lawsuit is approved. One lawyer on TikTok actually spoke about this and said he had a 10 year old and his mother talked him about how her son got leukemia after drinking prime. Now knowing that prime contains these chemicals, it's very much likely because of it. These chemicals are known to cause cancers and deteriorate your health since they're forever chemicals and your body can't get rid of them. Prime has 3 times the amount of these chemicals a person should have in their lifetime. —So if you still drink Prime or have any make sure you dispose of all of it until we find out what's going on”

Wall Street Apes

9,886,885 views • 2 years ago

Introducing LobeHub: Agent teammates that grow with you. LobeHub is the ultimate space for work and life: to find, build, and collaborate with agent teammates that grow with you. We’re building the world’s first and largest human–agent co-evolving network. Two years ago, we built LobeChat, an open-source interface for using different AI models. Today, LobeChat has 70k+ GitHub stars and serves 6M+ users worldwide. How to fully unlock the power of models has always been a shared mission between us and the community. We started with interaction — a fundamentally new, agent-first experience. Agents are no longer passive tools invoked in a single conversation. They should be proactive, always-on units of work. Treating agents as the minimal atomic unit is also the core of our agent harness infra. Today’s agents are mostly one-off executors. Even with memory, it’s often global — and hallucinates. We build long-term agent teammates that evolve with users. Each agent has its own dedicated memory space, editable by users, allowing humans and agents to co-evolve over time. This, in turn, allows us to design clearer rewards for reinforcement learning and create cleaner environments for continual learning. Agent teammates can work in groups. Through a multi-agent system, agent groups operate faster, more cost-effective, and go beyond what single-agent systems can achieve. For example, a single agent often requires heavy user involvement to proceed step by step, whereas LobeHub can execute the same work from a single instruction, with a supervisor orchestrating agents that run in parallel or debate to produce better results. We are building the collaboration network among agent teammates — and between humans and agent teammates as well. Ease of use matters. AI intelligence and shared human intelligence are equally important. With simple instructions and tool selection, you can effortlessly build and team up with agent coworkers to deliver complex, systematic work — even assembling a quant team to execute trades. Through the LobeHub community, anyone can discover, reuse, and remix agents and agent groups, customizing them to fit their own workflows, preferences, and needs. Last but not least, our vision started with LobeChat: multi-model support is the most efficient approach for users. We believe different models excel in different scenarios. By routing across multiple models, LobeHub improves cost efficiency and unlocks capabilities that a single-model setup cannot easily support.

LobeHub

185,273 views • 6 months ago

I'm so confident Triple Whale will make you money that I'm making a bet: If you do over a million dollars/year I'll pay you $250 for 15 minutes of your time. Today we're Introducing the Prime Day Mega Agent, an intelligent Amazon Analyst built to print you money on Prime Day. Here’s what it does, autonomously: – Analyzes Meta, Tiktok and Amazon ad performance – Predicts your winning SKUs using historical trend modeling – Generates a plug-and-play Prime Day playbook: what to pause, scale, test and when + when to send out email campaigns It’s like having a Head of Growth, Media Buying, and Ops in one… Except It works around the clock and leverages more data than any human ever could. We built it because we noticed a critical trend👇 When analyzing Prime Day sales data from last year, we found that the top brands didn’t just edge out the competition, they crushed them. Same tools. Same budgets. Same Prime Day. Yet somehow a small cohort of brands were crushing at a clip we typically don't see... So what were the winners doing that no one else was? We found levers the winners pulled that everyone else missed. Levers like: - Making their PPC target Prime-specific keywords - Warming up email audiences weeks before to build anticipation - Surf-scaling their ads by the hour not by the day on Prime Day Now, you can get that ENTIRE playbook the winners used custom built on your data. This agent is available RIGHT NOW. Go into Moby in Triple Whale and search "Prime Day Mega Agent." If you're a brand doing over a million dollars a year on Amazon, doing Prime Day right isn't a nice to have, it's table stakes. I guarantee you this agent will make you more revenue... And I am putting CASH behind it. If you do over a million/year I'll pay you $250 to take a demo... Limited to the first 100 people to sign up. The link to book is below this tweet. One more thing 🤯 As part of this launch I'm giving away an Amazon Agentic Org Chart the top brands will be using in the next 12 months. It will show you exactly what agents to use and for what, so you can maximize Amazon growth, autonomously. Want it? Retweet and comment "Moby" below and I'll dm you it.

Maxx Blank 🐳

112,769 views • 1 year ago