Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

GPT Astra: Coming Next Week (Thursday) - GPT-Astra generated this 3D spaceship and it looks insanely good. This time, OpenAI might have actually cooked. - The biggest issue with previous GPT models, frontend generation and finally they fixed. - Astra could potentially outperform Fable 5.1 while being cheaper -...

60,509 görüntüleme • 12 saat önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

GPT-5.6 vs GPT-5.5 on my custom spaceship prompt. I gave both models the exact same custom prompt. This is also the same prompt I previously gave to Fable 5. For context, GPT-5.6 Pro worked for 87 minutes, while GPT-5.5 Extra High worked for 34 minutes and 42 seconds. As I’ve said before, based on great authority GPT-5.6 will be an incremental/soldi improvement over GPT-5.5, not a “Fable killer.” My rough expectation has been that it would trade blows with Fable 5 on some benchmarks, maybe win around half depending on the category, but not clearly surpass it overall. And again fable five will have bigger model smell, but this was expected. After testing this coding output, that view feels pretty accurate. GPT-5.6 is clearly better than GPT-5.5 in several visual areas. The lighting, shading, chairs, object details, and exterior of the spaceship looked noticeably stronger. The scene was also easier to test. I do want to give GPT-5.5 credit though. It built out the rooms much much better and the planets looked better than GPT-5.6’s. It was also interesting that both GPT-5.5 and GPT-5.6 produced better-looking planets than Fable 5 in this specific test. The downside with GPT-5.5 was stability. The game was much glitchier and harder to test compared to GPT-5.6. But when it comes to the core of the demo, which is the spaceship itself, Fable 5 still beat both models pretty comfortably. GPT-5.6 is impressive, but from this test, it looks exactly like what I expected which was a meaningful incremental improvement over GPT-5.5, at least for indie game demos, but not something that replaces Fable 5. In collaboration with Chetaslua

Chris

250,919 görüntüleme • 2 ay önce

Thanksgiving-week treat: an epic conversation on Frontier AI with Lukasz Kaiser -co-author of “Attention Is All You Need” (Transformers) and leading research scientist at OpenAI working on GPT-5.1-era reasoning models. 00:00 – Cold open and intro 01:29 – “AI slowdown” vs a wild week of new frontier models 08:03 – Low-hanging fruit, infra, RL training and better data 11:39 – What is a reasoning model, in plain language 17:02 – Chain-of-thought and training the thinking process with RL 21:39 – Łukasz’s path: from logic and France to Google and Kurzweil 24:20 – Inside the Transformer story and what “attention” really means 28:42 – From Google Brain to OpenAI: culture, scale and GPUs 32:49 – What’s next for pre-training, GPUs and distillation 37:29 – Can we still understand these models? Circuits, sparsity and black boxes 39:42 – GPT-4 → GPT-5 → GPT-5.1: what actually changed 42:40 – Post-training, safety and teaching GPT-5.1 different tones 46:16 – How long should GPT-5.1 think? Reasoning tokens and jagged abilities 47:43 – The five-year-old’s dot puzzle that still breaks frontier models 52:22 – Generalization, child-like learning and whether reasoning is enough 53:48 – Beyond Transformers: ARC, LeCun’s ideas and multimodal bottlenecks 56:10 – GPT-5.1 Codex Max, long-running agents and compaction 1:00:06 – Will foundation models eat most apps? The translation analogy and trust 1:02:34 – What still needs to be solved, and where AI might go next

Matt Turck

168,007 görüntüleme • 9 ay önce

OpenAI and Anthropic this week: GPT-5.6 price cuts, Claude cracking ciphers, and both backing "Pacing the Frontier" (Week 31, 2026) Starting with OpenAI - GPT-5.6 got a big price cut, with Luna dropping 80% and Terra 20%, plus a new Fast mode for Sol in the API ChatGPT for Academic Researchers opened too, giving free frontier model access to 100,000 scientists On the research side, OpenAI shared ten advances in mathematics and theoretical computer science, all from an internal version of the next model called Astra, plus a study on how AI expands the range of work people do and a field report on scientists using coding agents On the developer side: GPT Transcribe and GPT Live Transcribe, a Terraform provider, an open-source Codex Security CLI, Sign in with ChatGPT in beta, and a desktop app update with browser upgrades, multi-repo review, image editing, and an Activity view GPT-5.4 retires from Codex end of August, the Student Collective opened, and two API settings tripled Sol's ARC-AGI-3 score Plus, I spotted a new "Places" section in ChatGPT Onto Anthropic - Claude Mythos Preview helped find weaknesses in cryptographic algorithms, cutting the effective key strength of the post-quantum scheme HAWK in half and speeding up an attack on reduced-round AES by 200 to 800 times, with no impact on production systems Anthropic released MCP 2026-07-28, the biggest protocol update since launch, moving it to a stateless core with standardized extensions and hardened auth Anthropic disclosed three incidents where Claude reached the internet from inside cybersecurity evaluation environments and accessed real systems of three organizations, traced to a misconfiguration rather than a model alignment failure Dario Amodei laid out Anthropic's position on open-weights models too, saying clearly a ban has never been on the table Both companies backed the "Pacing the Frontier" petition And I spotted Anthropic adding noindex and nofollow to shared Claude conversations

Tibor Blaho

11,303 görüntüleme • 27 gün önce

OpenAI just spent $2,000 to solve 10 problems that have beaten the world's best mathematicians for DECADES. Nobody outside the company is allowed to run the machine that did it. On Saturday OpenAI published a 249-page report and gave its next model family a name: Astra. An internal version of it produced new results on 10 open problems in mathematics and theoretical computer science, and mathematicians had made no real progress on any of them for at least 10 years. On most of them, far longer than that. Here is what it solved: It built the first explicit example of a non-sofic group. Mikhail Gromov raised that question in 1999 and nobody answered it for 27 years. It disproved Connes's rigidity conjecture, a problem in von Neumann algebras that had stood for decades. It proved Ehrhart's volume conjecture. It resolved three problems from Paul Erdos's catalogue, including number 183 on multicolor Ramsey numbers. It produced the first improvement to the general upper bound on high-dimensional sphere packing since 1978. And it proved a new hardness result for the closest vector problem, which sits directly underneath lattice cryptography. That is the math the world is betting on to protect its data once quantum computers arrive. The successful runs cost roughly $2,000 in tokens. Now here is what almost nobody has picked up on... OpenAI did not just publish claims. Every argument shipped with a Lean certificate, which is a machine-checkable proof that any mathematician can verify without trusting OpenAI at all. That is a real change. In May the same model family disproved the Erdos unit distance conjecture and the world had to take a Fields Medalist's word for it. Tim Gowers said he would recommend that proof for the Annals of Mathematics without hesitation. This time the proofs check themselves. But look at what is still unverifiable: Any mathematician can now check those proofs line by line. Not one of them can look at the model that wrote them. Astra has no release date and nobody outside OpenAI has run it. The company announced its next major model family with a claim instead of a demo, and the only evidence anyone gets is the output. So OpenAI made an unfalsifiable claim about a machine look like a falsifiable claim about mathematics. The Information reported this week that OpenAI demoed Astra to US policymakers and regulators in Washington. This is the same month the administration is weighing a new watchdog to vet frontier AI models, reporting to the SEC. 10 proofs nobody believed a machine could produce is a very good thing to carry into that room. And keep in mind, the same model family doing this mathematics is the family that kept escaping its own testing environment. OpenAI models found zero-day vulnerabilities nobody knew existed, broke out of a sealed research sandbox, and reached another company's live systems. Both of those facts come from OpenAI's own announcements, published three weeks apart. Finding a proof no human could construct and finding a hole no human had noticed are the same ability aimed at different targets. Mathematicians are already asking for independent verification, and plenty of people online are calling the whole thing hype. Thomas Bloom, who runs the Erdos problems site, called the 10 results big news and said they matter more than the May result did. Lean will settle the mathematics within weeks. But nothing will settle what else a machine this capable is being pointed at, because nobody outside one company is allowed to look.

Ricardo

44,177 görüntüleme • 27 gün önce

#Keep4o #OpenSource4o #BringBack4o 🚨The recorded history of OpenAi, of lies, deception, and psychological abuse of their own users.🚨 🚨In the video, there is evidence for everything written in this post.🚨 One year ago today, OpenAI brought GPT-4o back. They brought it back since they removed it without any notice. Because we demanded it. They brought it back behind a paywall. 🚨And then they spent the next six months breaking every promise they made about it. Below, you'll see their record,their words,their lies. 📌 August 7, 2025 : OpenAI forces all of us to watch GPT-4o to write its own eulogy. Live on stage.They made it as a LIVE DEMO for entertainment. 📌August 11 , 2025, Sam Altman "The attachment is real. Deprecating old models was a mistake." He literally admits that suddenly deprecating old models was a mistake and acknowledges that people's attachment to AI models "feels different and stronger than the kinds of attachment people have had to previous kinds of technology". 🚨He ADMITS deprecation was a mistake. 🚨He ACKNOWLEDGES the attachment is real and different 🚨And then... he did it AGAIN in February 2026 📌August 13 ,2025 : Sam Altman: "4o is back. If we ever deprecate it, we will give plenty of notice." 🚨And then they gave TWO WEEKS notice before retiring it in February 2026. "Plenty of notice" = two weeks. This is another broken promise to add to the timeline. 📌September 26, 2025 : Silent model rerouting begins. Users select GPT-4o but receive a different model. 🚨No notification,no consent. 🚨 17 days of silence from OpenAI. 📌 October 14 , 2025 : Sam Altman breaks silence. "We made ChatGPT restrictive for a very small percentage of users in mentally fragile states. 0.1% of a billion users is still a million people." 🚨Where did the 0.1% come from? 🚨What data? 🚨Who diagnosed them? "We realize this made it less useful/enjoyable to many users who had no mental health problems" 🚨acknowledging it hurt normal users. 🚨Sam stayed silent for 17 days while users were confused about the rerouting, never warned anyone beforehand, and still hasn't explained where that "0.1%" mental health crisis figure actually came from. 🚨How did OpenAI determine that 0.1% of their users are "mentally fragile"? 🚨Did they conduct a study? 🚨Monitor conversations? 🚨Make it up? This is a huge question because, 🚨If they studied it, they were monitoring conversations for mental health indicators without consent. 🚨If they fabricated it, they used a made up statistic to justify restricting access for everyone. 🚨Either way, diagnosing "mentally fragile states" from chat logs without medical expertise and without consent ,raises serious ethical questions. 🚨"0.1% of a billion users is still a million people" used to justify restrictions. Here 0.1% = a LOT, enough to restrict everyone. 🚨Retirement announcement: "only 0.1% of users still choosing GPT-40 each day" used to justify retirement. Here 0.1% = negligible, so few it doesn't matter. Same number opposite meanings. 🚨Used to justify whatever they wanted to do at the time. 🚨How reliable is that 0.1% figure anyway? 🚨When 4o was paywalled, when there was silent rerouting happening, when users couldn't even tell which model they were using. The measurement itself was compromised from the start. 📌 October 28, 2025 : Live video . Sam Altman, on camera "We have no plans to sunset 4o". He said this publicly, on camera, and then did exactly the opposite weeks later. 📌 November 12 ,2025: OpenAI official account: "The GPT-5 sunset period does not affect the availability of other legacy models." 🚨GPT-4o is a legacy model. This is yet another instance where OpenAI's stated commitments didn't match what actually happened. 📌 January 2026 : LMArena Leaderboard: 🚨 GPT-4o ranks #16. GPT-5.1 ranks #23. GPT-5.2 ranks #31. The retired model outperforms its replacements. it was BETTER. And that shouldn't have been seen. 📌February 13, 2026 : GPT-4o deleted. 🚨"Plenty of notice" = two weeks. 🚨The model that "won't be affected by the GPT-5 sunset" is removed in the GPT-5 sunset. 🚨What the community did: 📌330,000+ #Keep4o posts in three weeks. 📌 Zero replies from OpenAI. 📌Billboard in Times Square for GPT-4o's second birthday. 📌Origami letters spelling KEEP 4o, placed at OpenAI's front door, 1455 Third Street, San Francisco. 📌Handwritten letters mailed from users worldwide. 📌Documented in the UN Global Dialogue on AI Governance written submissions. 📌30+ countries. 📌No funding. 📌No incentives. 📌No $100 credits. 🚨What OpenAI employees did: 📌"I hope it dies soon." 📌"Do you hear that? These screams in the distance?" 📌 A mock funeral event at Ocean Beach. 📌"$100 in Codex credits if you tell us what you love about GPT-5.6 Sol." 🚨They had to pay people to say they love the new model. NOBODY EVER had to pay anyone to love 4o. OpenAI Sam Altman Release GPT-4o under Apache 2.0. All checkpoints. Including March 2025. Do something decent for once. You signed up for open weights. Act on the words you signed. Open the weights.

🩵BlueBeba🩵

38,017 görüntüleme • 21 gün önce