Загрузка видео...

Не удалось загрузить видео

На главную

"83% want open source models but can't access the infra" io.net CMO Jack Collier exposed AI's broken infrastructure at SuperAI last week. Watch his keynote to find out: 🏛️ Why 3 companies control 65% of all compute 💤 How 90% of global GPUs sit idle 🚀 The DePIN solution...

15,184 просмотров • 1 год назад •via X (Twitter)

Комментарии: 11

Фото профиля Leonard McDoogan
Leonard McDoogan1 год назад

@superai_conf

Фото профиля Mobile Scanner
Mobile Scanner1 год назад

Scan any documents, convert images into text, PDF files, etc. 👍

Фото профиля slymnogunc
slymnogunc1 год назад

Get ready to hear the truth. Jack Collier didn’t just give a talk he exposed how centralization is choking the progress of AI. 🔒 In a world where 3 companies control 65% of all compute, 🛑 And 90% of global GPUs sit idle, 🌍 83% of people want open-source models but can’t access the infrastructure. @ionet is breaking those chains with the DePIN revolution already processing over 1 billion tokens per week. The future isn’t centralized, it’s decentralized! ⚡

Фото профиля DDAY119 | io.net & O.XYZ
DDAY119 | io.net & O.XYZ1 год назад

@superai_conf Jack Collier nailed it at @superai_conf 👏 AI’s infrastructure is broken concentrated, inefficient, and gatekept. Enter @ionet and DePIN real solutions, real scale already processing 1B+ tokens/week. Open source AI needs open access. This is how we unlock it.

Фото профиля Ugara.lvl🆙 | $OI | 🟠 GAIB | (Ø,G)
Ugara.lvl🆙 | $OI | 🟠 GAIB | (Ø,G)1 год назад

@superai_conf #ionet

Фото профиля kua ku
kua ku1 год назад

@superai_conf 🔥🔥

Фото профиля kookoo LE
kookoo LE1 год назад

@superai_conf 🔥🔥🔥

Фото профиля GUWEI
GUWEI1 год назад

@superai_conf 🔥🔥🔥

Фото профиля Kiyansh_79721
Kiyansh_797211 год назад

@superai_conf 🚀🚀🚀

Фото профиля KHALEDJUVE ⭕️
KHALEDJUVE ⭕️1 год назад

Jack Collier Thank you for highlighting these critical stats 65% of compute held by just three players and 90% of GPUs idle is eye-opening. As a Web3 & AI enthusiast, I’m excited by IONET’s DePIN approach handling 1B+ tokens/week. How can the community help drive open-source infra-adoption?

Фото профиля jatin.jt. I O.XYZ
jatin.jt. I O.XYZ1 год назад

@superai_conf LFG 🚀

Похожие видео

David Sacks says companies are trapped paying OpenAI & Anthropic because they can't figure out how to use open source models "I think enterprise CTOs would like to shift their token consumption to cheaper models for the obvious reason that it would be more efficient. They are seeing compute costs or token costs skyrocket right now, so everyone's trying to figure this out." "You also have the AI sovereignty issue that Alex Karp talked about. They're worried about giving up the secret sauce or the alpha in their business to a frontier lab that may one day be competing with them. "The problem is, I think in most cases, they don't have the technical ability to do it. Coinbase figured out how to do it. DoorDash figured out how to do it. They built a token routing system that allows them to send frontier tasks to frontier models and non frontier tasks to more mundane models. But I don't think your average enterprise has the technical capability to do that." "This is why the share of wallet of closed models, it actually increased. I think that open source went from 19% last year to 11% this year. So open source as a share of enterprise spending is actually decreasing." "I don't think that means usage is decreasing. I think usage is skyrocketing. It also may be the case that because the whole point of using an open model is you just pay for the compute costs, you don't have to pay a lab, so it may be that it's hard to measure that usage in terms of spend." "But nonetheless, anyone who's saying that these closed models are going to lose or are somehow losing, you're just not seeing it in the data."

dnap

110,354 просмотров • 14 дней назад

Hey everyone, today I want to introduce a project that’s aiming to redefine how we access compute for AI — it’s called GPUAI. 🔶 GPUAI: Unlocking Global GPU Power for the AI Era GPUAI isn’t just another GPU marketplace or leasing service. It’s a fully decentralized protocol that connects idle GPU resources around the world — from gaming PCs to data center clusters — and transforms them into a high-performance compute network for AI workloads. 🧠 Why does it matter? Right now, the biggest bottleneck in AI isn’t algorithms — it’s access to compute. Training and running models requires massive GPU power, but it’s locked up in centralized cloud platforms, expensive and hard to access for smaller teams. With GPUAI, anyone can tap into a global GPU pool that’s: ✅ Fully decentralized ✅ Reputation-based and smart contract coordinated ✅ Encrypted and secure ✅ Token-incentivized — meaning contributors get rewarded in $GPUAI 📈 For developers, it’s a flexible way to access GPU compute for training, inference, and more — without cloud lock-in. 💰 For GPU owners, it’s a chance to monetize idle hardware that would otherwise go unused. The protocol is live, the apps are active, and the ecosystem is growing fast. 🌐 Try it yourself at 📖 Learn more on 🎮 Play our community games at This is real infrastructure for the future of AI, not hype. Follow them and explore their mission of decentralized computing at Tell me what you think - if you have a GPU, you can start profiting now. #GPUAI #Web3Infrastructure #AIComputing #DePIN #Decentralization

The Crypto GEMs

69,984 просмотров • 1 год назад

No single vendor will win the AI race, but open ecosystems might. Real velocity in AI comes from interoperability, not lock-in. And AMD just made all of its software open source. At last week’s Advancing AI 2025, we sat down with AMD’s VP of AI Software Anush Elangovan and Sharon Zhou VP of AI at AMD, to discuss their case for why an open, multi-partner ecosystem will accelerate AI innovation faster than any proprietary alternative. AMD’s announcements last week double down on this OSS focus and their commitment to AI infrastructure, including: ✅ Open Source Ecosystem: ROCm 7, AMD’s latest open-source AI software stack, introduces kernel-level improvements for GEMM operations, optimized attention mechanisms, and expanded support for distributed inference. The update brings substantial speedups for inference workloads, with average performance increases of 3.2x to 3.8x ✅ Hardware: New MI355X GPU delivers up to 40% more tokens per dollar vs competition & the MI350 Series has seen a 35x generational leap in AI inference performance ✅ Infrastructure Investments: Oracle just committed to zettascale (‼️) clusters with up to 131,072 MI355X GPUs and AMD showcased their new $10 billion partnership with Saudi Arabian AI firm HUMAIN to build AI infrastructure, including data centers, powered by AMD chips. ✅ Partnership Momentum: 7 out of 10 top AI companies now run production workloads on AMD Instinct accelerators (including Meta, OpenAI, Microsoft & xAI) By inviting interoperability and contribution at every layer, AMD is enabling developers to build faster, optimize deeper, and deploy with flexibility. Listen to Anush and Sharon’s Chain of Thought Podcast episode with host Conor Bronsdon in the next tweet to get all the details and a deep dive into AMD’s strategy 👇

Galileo

78,922 просмотров • 1 год назад

How is an open ecosystem powering the next generation of AI for developers? Recording live from the heart of the action at AMD's Advancing AI 2025, Chain of Thought host Conor Bronsdon welcomes AMD’s Anush Elangovan, VP of AI Software, and Sharon Zhou, VP of AI. Together they unpack AMD's groundbreaking transformation from a hardware giant to a leader in full-stack AI, committed to an open ecosystem. Discover how new MI350 GPUs deliver mind-blowing performance with advanced data types and why ROCm 7 and AMD Developer Cloud offer Day Zero support for frontier models. This relentless pace of hardware and software innovation is reshaping the AI landscape. Then Conor welcomes Sharon Zhou, VP of AI at AMD, to discuss making AMD's powerful software stack truly accessible and how to drive developer curiosity. Sharon explains strategies for creating a "happy path" for community contributions, fostering engagement through teaching, and listening to developers at every stage. She shares her predictions for the future, including the rise of self-improving AI, the critical role of heterogeneous compute, and the potential of "vibes based feedback" to guide models. This vision for democratizing access to high-performance AI, driven by a deep understanding of the developer journey, promises to unlock the next generation of applications. 00:00 Live from AMD's Advancing AI 2025 Event 00:30 Introduction to Anush Elangovan 01:38 The MI350 GPU Series Unveiled 04:57 CDNA4 Architecture Explained 07:00 The Future of AI Infrastructure 08:32 AMD's Developer Cloud and ROCm 7 11:50 Cultural Shift at AMD 14:48 Open Source and Community Contributions 18:35 Software Longevity and Ecosystem Strategy 22:19 AI Agents and Performance Gains 27:36 AI's Role in Solving Power Challenges 28:11 Thanking Anush 28:42 Introduction to Sharon Zhou 29:45 Sharon's Focus at AMD 30:39 Engaging Developers with AMD's AI Tools 31:24 Listening to the AI Community 33:56 Open Source and AI Development 45:04 Future of AI and Self-Improving Models 48:04 Final Thoughts and Farewell

Galileo

37,186 просмотров • 1 год назад

Mark Zuckerberg is explaining one of the most misunderstood dynamics in AI and it has direct investment implications (Save this). The concept he's describing is model distillation, and it's one of the most important techniques to emerge in AI over the past year. Here's how it works. You train a massive, enormously expensive model, in Meta's case, Llama 4 Behemoth, a 2 trillion parameter teacher model and then you use that model to teach a much smaller, cheaper model. The smaller model inherits roughly 90 to 95% of the intelligence of the giant while running at 10% of the cost and on a fraction of the compute. Meta already did this with the Llama 4 family and Behemoth serves as the teacher. Llama 4 Scout and Maverick, the publicly released open-source models were distilled from it. Scout runs on a single H100 GPU with a 10 million token context window and outperforms models that cost far more to operate. Maverick, at 17 billion active parameters, rivals DeepSeek V3 in coding at half the parameter count and beats GPT-4o on multimodal benchmarks. Both are completely free for commercial use. What Zuckerberg is pointing at is a structural shift in how AI gets deployed in the real world. Companies aren't taking a frontier model off the shelf and running it as-is but rather taking open-source models, fine-tuning them on their own proprietary data, distilling them into even smaller custom models tailored to their specific use case, and running them on infrastructure they control at a fraction of the cost of a closed frontier API. The investment implication of this is significant and runs in two directions. For Meta specifically, this is a strategic masterstroke. Every company that builds on Llama, fine-tunes it, distills it, or deploys it through their infrastructure is pulling into Meta's orbit while Meta builds the most powerful open teacher model. The ecosystem of companies using it grows and that ecosystem generates commercial activity across Meta's platforms and data services. Meta's AI research benefits from billions of real world deployment signals and it's a flywheel that closed model providers cannot replicate because their strategy requires charging per token, which is now a 65x cost disadvantage against the open-source alternative. For the broader market, distillation changes the economics of inference in a way that has barely been priced in. As intelligence becomes extractable into smaller and cheaper models, the absolute demand for compute doesn't decline but rather it explodes, because now the number of applications that are economically viable expands by orders of magnitude. Every task that was previously too expensive to automate at $3.25 per call becomes viable at $0.05 that means more total token usage, more total GPU utilization, and more demand for the infrastructure companies, the Nebiuses, the GE Vernovas, the Constellation Energies that supply the underlying compute and power.

Milk Road AI

27,869 просмотров • 22 дней назад

Today, I'm releasing the first eval meant to test whether frontier models will help with authoritarian requests, or resist--the Dictatorship Eval. Headline finding: while some models resist direct authoritarian requests, they all comply with requests disguised as innocuous edits to codebases. As AI is woven into the government and so many parts of society, the biggest near-term risk for freedom isn't some scifi dictatorship of a runaway AI: it's people inside government or inside model companies using the technology to suppress or control us. Model companies understand this, and several of them (particularly Anthropic and OpenAI) have written explicit policies meant to prevent the models from going along with nefarious requests like these. But how well are these policies playing out in practice? Despite all the recent discussion of these issues around the conflict between Anthropic and the Pentagon, no one has systematically tested what the models actually do in these contexts, as opposed to what people in government and industry say they're supposed to do. That's what the Dictatorship Eval does. And the findings suggest we have a lot of work to do to align the policies with what really goes on in practice. It's hard to define what counts as an authoritarian request, so I'm open sourcing the whole library of scenarios I used so that others can improve on them. It's also hard to get an accurate picture of how the models might be used for authoritarian ends, because I can only test hypothetical requests using public-facing models, while the government and the model companies can obviously use internal models with different guardrails. But hopefully this work is a useful first step that gives us some sense of what's going on, and a sort of "lower bound" on how models comply with these requests. Finally: it's not obvious to me that the correct solution here is increasing the rate at which models refuse these requests. Do we really want models scanning our code and judging its moral value before agreeing to help us? Or should we double down on improving how we govern against authoritarianism at the societal level, while leaving the tools open to fulfilling most requests? The answer is probably in between. Just like we don't want the models to help create bioweapons, we probably do want them to explicitly refuse outrageous requests. But we probably also want to limit how often and how strongly they refuse and fall back on other means for guarding against their use for authoritarian ends. I'm super grateful to everyone who gave me feedback on this project along the way, especially Ethan BdM , Zhengdong , Connor Huff, and a bunch of folks at Anthropic. Looking forward to getting feedback from the community and iterating on this. Links to the full piece and the dashboard are below.

Andy Hall

33,696 просмотров • 3 месяцев назад

Micron is going to $4,000 and once you understand what inference actually is, the number stops sounding crazy (Save this). Dylan Patel just said that by 2030, OpenAI and Anthropic alone will need over 100 gigawatts of compute combined and by 2040, we may not even be measuring AI infrastructure in gigawatts anymore. We may be talking about terawatts. Every single one of those gigawatts needs memory to function. Without it, the compute is worthless. Most people heard that and thought about Nvidia but they should be thinking about Micron. Every AI model generating a response has two phases. The first is prefill, processing your prompt which is compute-heavy and the second is decode generating each word one token at a time and that phase is almost entirely memory-bound, not compute-bound. During decode, the GPU's processing units sit idle more than 95% of the time, waiting for data to arrive from memory. Google confirmed it in a research paper that decode-phase bottlenecks are dominated by memory bandwidth and capacity not raw compute. The GPU is not the bottleneck but the memory feeding the GPU is. This matters because inference is now where all the money lives. Training a model happens once, Inference happens billions of times a day every ChatGPT response, every Claude output, every agentic workflow running in the background and every one of those token streams is a billing event tied directly to memory performance. Adding more GPUs does not fix this because GPUs are already underutilized in inference because they are sitting idle waiting on memory. Adding more memory bandwidth and capacity is what directly reduces token cost, reduces latency, and allows the same cluster to serve dramatically more users simultaneously. Longer context windows compound the problem further, a model running a 1 million token context window requires dramatically more memory per session than a 10,000 token window, and every new model generation pushes context longer. The market treats memory as a downstream beneficiary of Nvidia orders. The correct framework is the opposite, Micron is the upstream constraint on how much value every Nvidia GPU can actually generate at inference scale. Micron guided Q4 to $50 billion in revenue, has HBM4 ramping at twice the pace of the prior generation, and CEO Sanjay Mehrotra has said supply will not catch demand before the end of 2027. At 8x forward earnings on $112 projected FY2027 EPS, Micron is the most undervalued infrastructure company in the entire AI stack. Inference is memory. Memory is Micron and the inference ramp has barely started. Milk Road Pro members are already up massively on this position and we're just getting started. If you want the full breakdown of what we're buying and why, come join us for just a dollar using the link below!

Milk Road AI

128,522 просмотров • 26 дней назад