正在加载视频...

视频加载失败

CES Report #2: Asygn claims it has the best computer vision chips for glasses. Uses very little power. I include this because these chips are going to run the vision workloads in glasses of the future. They will demonstrate how glasses can recognize everything in the world in a...

13,524 次观看 • 8 个月前 •via X (Twitter)

0 条评论

暂无评论

原始帖子的评论将显示在这里

相关视频

Americans are being priced out of playing videos games and the reason is Data Centers The Nintendo Switch 2 is $500 A high-end Xbox is $800 The PlayStation 5 Pro is $900 It’s causing major shifts: - PlayStation is ending physical disks - NVIDIA GeForce is implementing a gaming cap for a maximum amount of hours played per month - XBox is divesting in 5 gaming studios Data Centers are buying and using so much RAM memory that one stick that used to cost $200 now costs over $1,000 “Huge increases across the board” 3 companies make almost all memory chips “these chips are really only made by about 3 companies in the entire world, and those 3 companies have limited output. So they have to decide, do they want to build high-end memory chips for AI data centers or consumer-grade memory chips for all of our devices? As you can probably guess, They didn't choose us. OpenAI reportedly committed to buying 40% of all the memory chips in the entire world just for their Stargate multi-state data center project. For memory chip companies, this was a huge payday, but it meant absolute scarcity for consumer memory chips, which of course means massive price increases for you and I” Gaming companies aren’t going to invest in next gen consoles that will only sell a fraction of the amount of previous consoles because of price. It wouldn’t make sense because game developers are going to focus on the consoles with all the users, the older consoles So gaming console manufacturers are moving towards new revenue streams like controlling how you can buy a game, more subscription costs and more

Wall Street Apes

80,910 次观看 • 1 个月前

Neil thinks 90% of AI workloads will run in the background versus real time. In that future, latency matters less and cost matters more, and he's configured his company around a unique way of serving tokens at the cheapest possible price. "There's no bad chips. There's only bad pricing, and I will make any chip work at the right price. Let's talk about AMD, great chips overall. People don't understand how to program them very well. That's music to my ears. I'm happy for them to sleep on this chip and for me to buy as much as I can." "One of the ways I describe what we do is, we will buy any chip anywhere in the world for any duration of time. That is a level of flexibility and liquidity that no one else has right now. You'd have basically zero buyers for a data center that is ninety-five percent uptime. I'm that first buyer. First we scavenge chips, and then we scavenge power for those chips. The idea is, in both cases, I do not want to be bidding against Anthropic or OpenAI for compute capacity. I'm not gonna win against them, and I don't want to. I want to be more creative and use the supply that they don't find legible today. Over time I amass enough aggregate supply...and build my aggregate factory that is unbeatable in economics." "My whole goal is to so dramatically expand the supply of power across the United States that I have a home for a lot of chips that otherwise would not have earned their place in a data center." "We've really pushed AI to be an interactive chatbot tool, and everyone has chosen latency optimization because the shape of usage was chatbot oriented. That's the most profound change we're going to see in the next year, we're gonna move away from chatbots to more proactive or background agents, and in that world, it makes a lot more sense to build a stack around throughput." "I love this market because it's unbounded. There's no human in the loop, so you can consume as many tokens as you like in the background versus human attention span. So long term, we're going to end this year at maybe fifty-fifty background and real time workloads, but I see this going to ninety ten in favor of background." "The best latency is no latency at all. When you wake up in the morning, the work's already been done overnight, you didn't even have to ask for it. What we'd like is the agent to operate on more human time scales. You don't manage your colleagues every five minutes. You come back and check in maybe once a week." "My job is to make the tokens as cheap as humanly possible, and I will do it through every layer in the stack available to me."

Patrick OShaughnessy

82,491 次观看 • 1 个月前