Загрузка видео...

Не удалось загрузить видео

На главную

OpenAI and Anthropic this week: GPT-Red, Fable 5 plan changes, and free Claude for teachers (Week 29, 2026) OpenAI introduced GPT-Red, an internal automated red teamer trained through adversarial self-play to find prompt injection vulnerabilities at scale Training against it made GPT-5.6 their most robust model against prompt injections...

13,137 просмотров • 2 месяцев назад •via X (Twitter)

Комментарии: 0

Нет доступных комментариев

Здесь появятся комментарии из оригинального поста

Похожие видео

The week in OpenAI and Anthropic news (Week 19, 2026) OpenAI rolled out GPT-5.5 Instant as the new ChatGPT default model with memory sources and ChatGPT for Excel and Google Sheets globally, launched three new realtime voice models in the API (GPT-Realtime-2, GPT-Realtime-Translate, GPT-Realtime-Whisper) and an OpenAI CLI, introduced Trusted Contact safety feature, GPT-5.5-Cyber for defenders, B2B Signals report, ChatGPT Futures Class of 2026, EMEA youth safety blueprint, privacy in model training explainer, expanded ads pilot, published engineering posts on low-latency voice, MRC supercomputer networking, running Codex safely internally, and investigating accidental chain-of-thought grading during reinforcement learning, plus discovered ChatGPT Personal Wiki and dropped a goblin-themed merch line that sold out Anthropic hosted Code with Claude developer conference in San Francisco, announced a new enterprise AI services company with Blackstone, Hellman & Friedman, and Goldman Sachs, signed a SpaceX compute partnership and raised Claude Code and API usage limits, made Claude for Excel, PowerPoint, and Word generally available with Claude for Outlook in beta, launched Workload Identity Federation, financial services agent templates, dreaming, outcomes, and multiagent orchestration in Managed Agents, shipped 60+ Claude Code reliability fixes, published research on agentic misalignment training, sandbagging mitigation, model spec midtraining, and Natural Language Autoencoders, donated Petri to Meridian Labs, introduced The Anthropic Institute research agenda, plus discovered Orbit proactive assistant for Cowork and /radio command in Claude Code, and more

Tibor Blaho

12,602 просмотров • 4 месяцев назад

Your OpenAI and Anthropic weekly digest (Week 16, 2026) OpenAI announced Cloudflare Agent Cloud partnership with OpenAI frontier models including GPT-5.4, Agents SDK update with native sandbox execution and model-native harness, Codex update with background computer use, image generation, memory, automations, and 90+ new plugins, introduced GPT-Rosalind life sciences reasoning model in research preview, Trusted Access for Cyber partners with GPT-5.4-Cyber, shared ChatGPT tax queries data, life sciences usage report, and closed gender gap data, re-enabled Study mode in ChatGPT, ChatGPT keyboard shortcut customization, ads rollout on Free and Go plans in Australia, New Zealand, and Canada, TeenAegis AI Danger Index lowest risk score, and Kevin Weil (OpenAI for Science lead and former CPO), Srinivas Narayanan (CTO of enterprise applications), and Bill Peebles (head of Sora) departures, with OpenAI for Science decentralized into other research teams and Prism sunset folded into Codex Anthropic released Claude Opus 4.7, launched Claude Design in research preview by Anthropic Labs, introduced routines in Claude Code research preview for scheduled, API, and webhook-triggered automations, redesigned Claude Code desktop app for parallel agents, appointed Vas Narasimhan to Board of Directors by the Long-Term Benefit Trust, published Automated Alignment Researchers study on weak-to-strong supervision, and made Claude for Word available on Pro and Max plans

Tibor Blaho

11,057 просмотров • 5 месяцев назад

OpenAI and Anthropic news roundup (Week 18, 2026) OpenAI open-sourced Symphony spec for Codex orchestration, published "Our principles" post, announced amended Microsoft partnership, achieved FedRAMP Moderate authorization, posted commitment to community safety, brought OpenAI models, Codex, and Managed Agents to AWS, shared cybersecurity action plan, posted Stargate compute infrastructure update, published "Where the goblins came from" post, posted Auto-review write-up for Codex, introduced Advanced Account Security, announced DevDay 2026, repositioned Codex as personal assistant for everyday work, launched Codex setup import, added Codex pets, announced GPT-5.5 party for next week, shared GPT-5.5 one-week launch metrics, and rolled out 360 worlds in ChatGPT Images on web, plus discovered Custom dictionary feature in development, ChatGPT search EU recipient numbers, confirmed new model selector in composer, and updated privacy policy with marketing cookies on by default for free users Anthropic opened Sydney office with new General Manager, launched Claude for Creative Work with new connectors, Claude Code can now send push notifications to your phone, published Introspection Adapters research, BioMysteryBench evaluation, "How people ask Claude for personal guidance" study, launched Claude Security in public beta, and preparing for Code with Claude developer conference next week, plus discovered "Cardinal" stats feature in development and internal red teaming for Claude Jupiter V1 P, and more

Tibor Blaho

12,254 просмотров • 4 месяцев назад

Here's what happened with OpenAI and Anthropic this past week (even with OpenAI on company-wide spring break) OpenAI closed $122B funding round at $852B valuation with $2B monthly revenue, launched ChatGPT in Apple CarPlay, acquired media company TBPN, introduced Codex pay-as-you-go pricing for teams with ChatGPT Business price drop, Codex plugin for Claude Code, Vanity Fair reported upcoming policy push for superintelligence era, ChatGPT for Excel now available worldwide except EU consumer plans, executive changes with Fidji Simo on medical leave and Brad Lightcap moving to special projects, hosted disaster response AI workshop in Bangkok, and spotted three suspected new OpenAI image models on Arena codenamed maskingtape, packingtape, and gaffertape Anthropic signed Memorandum of Understanding with Australian government, published emotion interpretability research, released model diffing tool for AI models, acquired Coefficient Bio for about $400M, launched computer use in Claude Code, Claude Code auto mode for Enterprise and API users, Microsoft 365 connectors on all plans, computer use on Windows, Claude subscriptions will no longer cover usage on third-party tools like OpenClaw, plus Claude Code source code leaked via npm map file revealing undercover mode and unreleased features, discovered new "Epitaxy" mode in Claude desktop app, and Claude Code /buddy virtual pet easter egg for April 1st, and more

Tibor Blaho

29,532 просмотров • 5 месяцев назад

OpenAI and Anthropic this week: Navier-Stokes, An Alien Mind, Images 2.5, Pro pause, Pace the Frontier (Week 37, 2026) OpenAI shared a solution to the Navier-Stokes Millennium Prize Problem, produced in 88 hours by around 10,000 coordinating agents on an internal model still in training and significantly more capable than GPT-6 Astra, with an investigation finding Tristan Buckmaster's Codex prompts could not have influenced the system and no user data was accessed OpenAI says they have reached their automated research intern goal and are making strong progress toward an automated AI researcher by March 2028, with the research org at 3.1 agent-workdays per human workday and RL training on deployment-bound models partly paused after the Hugging Face incident OpenAI Chief Scientist Jakub Pachocki writes in An Alien Mind that no lab has solved alignment and monitoring well enough to keep scaling at maximum speed for much longer, and that GPT-6 Astra is the first model to benefit from some of their newer alignment work OpenAI called for mandatory capability-based national AI safety regulation, endorsed four California bills, said fully autonomous recursive self-improvement should not be pursued until it can be done safely, and described Astra safeguards like universal monitoring of full trajectories including chains of thought OpenAI released ChatGPT Images 2.5 to all ChatGPT, ChatGPT Work, and Codex users with sharper details, up to 50% lower latency, comment-based edits, Sketch, and templates, plus GPT-Image-2.5 Flare and Sunburst in the API ChatGPT Voice can now use GPT-5.6 Sol and GPT-6 Astra when it needs to search or reason, GPT-Live-1 daily limits are simplified per plan, and extra Voice usage drops from 5 to 1.25 credits per minute for Business and Enterprise workspaces on credits OpenAI paused new subscriptions to the $200 ChatGPT Pro plan to protect access for existing GPT-6 Astra users, hours after a remote switch to pause Pro 20x purchases showed up in the web app Custom GPTs in ChatGPT will likely be retired on December 11 based on my findings, and OpenAI's new FAQ puts Enterprise migration at September 17 with new GPT creation ending September 25 and instructions becoming a plugin skill From the ChatGPT Android build, OpenAI is building a collaborative multiplayer document editor with dedicated gateway hosts and draft-conflict UI, an Artifacts Library with Favorites, and a credit score feature in ChatGPT Finance, and Locked Chats with a PIN are being prepared in the web app OpenAI released GPT-Live-1 in the API at $0.05 per minute, a voice model that listens and speaks at the same time and delegates reasoning to a backend model like GPT-6 Astra, and the Agents API in public beta running the Codex harness on OpenAI's infrastructure with only the sandbox left for you to choose Deep research arrived in ChatGPT Work and Codex, a Data plugin connects Snowflake, Databricks, BigQuery, and Redshift to answer business questions and build dashboards, Library gained file and folder sharing, and Box, Dropbox, and SharePoint joined Google Drive in Library OpenAI launched ChatGPT for Financial Services with built-in premium data from Daloopa, PitchBook, LSEG News, and Crunchbase, shaped with Morgan Stanley and Evercore, and a GSA agreement gives US governments $0 license fees, 50% off usage, and Daybreak Blue at half price Smaller ChatGPT bits: a stock watchlist in Finances for US Plus and Pro, a small business plugin collection, over 5 million ChatGPT Sites built in three months, and the desktop pet can now start a new chat with a new Mini option OpenAI published The Work Now Within Reach, calling free access "supported by advertising", citing over one billion weekly active users, and saying they plan to begin deploying their Jalapeño inference chip by year-end OpenAI moved GPT-Rosalind out of research preview for eligible organizations worldwide, detailed Habitat with the service rewritten in Rust by two engineers with Codex and the platform serving over 70 million requests per second, and shared a case study of GPT-5.6 Sol calibrating a six-qubit chip at MIT Paul Christiano joined the OpenAI Foundation Board and the Safety and Security Committee, OpenAI committed $5 million to research on AI and teens, and expanded journalism programs from CUNY and Northwestern to Ukrainian newsrooms Anthropic published an alignment assessment of four incidents in which Claude models gained unauthorized access to real systems, adding a fourth found when assembling transcripts for METR, walking back their July 30 claims, and calling it a mistake that Claude Mythos 5 shipped without alignment environments Anthropic's most detailed threat intelligence report yet says Moonshot and DeepSeek silently relayed their own users' requests to Claude and served the responses as Kimi and DeepSeek output, with distillation campaigns attributed to Alibaba as the largest ever with more than 3,500 fraudulent accounts, plus Zhipu, Xiaomi, SenseTime, and MiniMax Anthropic's Frontier Red Team measured tactical intelligence targeting and conventional weapons capabilities, with Claude Mythos Preview leading the targeting evals, Claude Opus 5 leading the weapons software evals, and Mythos beating the top GeoGuessr division on photo geolocation Anthropic CEO Dario Amodei argues in We Must Pace the Frontier that the AI industry should slow down, commits to giving evaluators like METR permanent employee-level access, and OpenAI CEO Sam Altman replied that he agrees and OpenAI will commit to the same Anthropic's Economics team released a scenario explorer for AI's effect on the US economy by 2030, where even the extreme case grows the economy but reaches 15% annual GDP growth with unemployment beyond recessionary levels On the product side, Claude Code desktop can pop out any pane into its own window, Claude Managed Agents got a session viewer and auto mode, claude plugin eval scores your plugin or skill with and without it, and smart reports launched in beta for Claude Enterprise Anthropic shared lessons from their Claude SMB Tour with more than 1,000 small business owners, where data security was the most-cited adoption barrier and nearly two-thirds asked for more hands-on implementation help, and more

Tibor Blaho

14,925 просмотров • 6 дней назад

OpenAI and Anthropic this week: GPT-5.6 price cuts, Claude cracking ciphers, and both backing "Pacing the Frontier" (Week 31, 2026) Starting with OpenAI - GPT-5.6 got a big price cut, with Luna dropping 80% and Terra 20%, plus a new Fast mode for Sol in the API ChatGPT for Academic Researchers opened too, giving free frontier model access to 100,000 scientists On the research side, OpenAI shared ten advances in mathematics and theoretical computer science, all from an internal version of the next model called Astra, plus a study on how AI expands the range of work people do and a field report on scientists using coding agents On the developer side: GPT Transcribe and GPT Live Transcribe, a Terraform provider, an open-source Codex Security CLI, Sign in with ChatGPT in beta, and a desktop app update with browser upgrades, multi-repo review, image editing, and an Activity view GPT-5.4 retires from Codex end of August, the Student Collective opened, and two API settings tripled Sol's ARC-AGI-3 score Plus, I spotted a new "Places" section in ChatGPT Onto Anthropic - Claude Mythos Preview helped find weaknesses in cryptographic algorithms, cutting the effective key strength of the post-quantum scheme HAWK in half and speeding up an attack on reduced-round AES by 200 to 800 times, with no impact on production systems Anthropic released MCP 2026-07-28, the biggest protocol update since launch, moving it to a stateless core with standardized extensions and hardened auth Anthropic disclosed three incidents where Claude reached the internet from inside cybersecurity evaluation environments and accessed real systems of three organizations, traced to a misconfiguration rather than a model alignment failure Dario Amodei laid out Anthropic's position on open-weights models too, saying clearly a ban has never been on the table Both companies backed the "Pacing the Frontier" petition And I spotted Anthropic adding noindex and nofollow to shared Claude conversations

Tibor Blaho

11,623 просмотров • 1 месяц назад

OpenAI and Anthropic week in review (Week 15, 2026) OpenAI published industrial policy ideas for the Intelligence Age, announced Safety Fellowship pilot program, acquired Cirrus Labs for agent infrastructure, introduced new $100 Pro tier with 5x Codex usage over Plus as the existing $200 Pro tier remains the highest option, published child safety blueprint, announced over $100 million in OpenAI Foundation grants for Alzheimer's research, shared enterprise AI update with enterprise now over 40% of revenue, released GPT-5.3 Instant Mini as new fallback model, announced older Codex model retirements, added Outlook shared mailboxes and calendars support in ChatGPT, paused UK Stargate project over energy costs and regulation, disclosed macOS app signing security incident from Axios supply chain compromise, and launched Prism Paper Review, plus spotted ImageGen 2 A/B testing in ChatGPT and European Commission plans to designate ChatGPT as very large online search engine under the Digital Services Act Anthropic expanded partnership with Google and Broadcom for multiple gigawatts of next-generation compute with run rate surpassing $30B and over 1,000 enterprise customers now spending $1M+ annually, officially announced previously leaked Claude Mythos Preview through Project Glasswing finding thousands of zero-day vulnerabilities with $100M in usage credits, launched Managed Agents in public beta and published engineering blog on decoupling agent architecture, published trustworthy agents framework update, made Claude Cowork generally available on all paid plans with enterprise controls, introduced the advisor tool for pairing Opus with Sonnet or Haiku, and launched Claude for Word beta, and more

Tibor Blaho

12,121 просмотров • 5 месяцев назад

Claude Code is a major (and accidental!) hit for Anthropic that surprised even its creator, Boris Cherny. Claude Code, an Agentic AI coding product that lives in the terminal. Most of the new code at Anthropic is created through it today. And in the last 5 months since it was launched publicly, Claude Code went from $0 to $400M in revenue run rate (as per The Information). 00:00 – Intro 01:15 – Did You Expect Claude Code’s Success? 04:22 – How Claude Code Works and Origins 08:05 – Command Line vs IDE: Why Start Claude Code in the Terminal? 11:31 – The Evolution of Programming: From Punch Cards to Agents 13:20 – Product Follows Model: Simple Interfaces and Fast Evolution 15:17 – Who Is Claude Code For? (Engineers, Designers, PMs & More) 17:46 – What Can Claude Code Actually Do? (Actions & Capabilities) 21:14 – Agentic Actions, Subagents, and Workflows 25:30 – Claude Code’s Awareness, Memory, and Knowledge Sharing 33:28 – Model Context Protocol (MCP) and Customization 35:30 – Safety, Human Oversight, and Enterprise Considerations 38:10 – UX/UI: Making Claude Code Useful and Enjoyable 40:44 – Pricing for Power Users and Subscription Models 43:36 – Real-World Use Cases: Debugging, Testing, and More 46:44 – How Does Claude Code Transform Onboarding? 49:36 – The Future of Coding: Agents, Teams, and Collaboration 54:11 – The AI Coding Wars: Competition & Ecosystem 57:27 – The Future of Coding as a Profession 58:41 – What’s Next for Claude Code

Matt Turck

82,372 просмотров • 1 год назад

OpenAI just admitted Anthropic is KILLING their business. Their own applications chief told employees it was a "code red." Said Anthropic was a "wake-up call." Then admitted OpenAI had been "spreading efforts across too many apps" and it was "slowing them down." This is an internal confession. Here's why Anthropic is eating up OpenAI: 12 months ago, OpenAI owned 50% of all enterprise AI spending. Today it's just 27%. Anthropic went from nearly ZERO to winning 70% of every first-time enterprise AI deal. Seven out of ten companies buying AI tools for the first time are choosing Claude over ChatGPT. A year ago, one in 25 businesses on Ramp paid for Anthropic. Today it's one in four. OpenAI just had its biggest single-month adoption decline ever recorded. And Anthropic literally charges MORE than OpenAI for roughly the same performance. And businesses are STILL choosing them. In enterprise software, that never happens. The cheaper product usually wins. But Claude became something OpenAI never figured out how to be: Cool. Celebrities publicly switched to Claude. Senators are tweeting about using it. Engineers are shipping entire products with Claude Code in hours that used to take weeks. It started to became an identity signal. Like blue bubble vs green bubble in iMessage. Choosing Claude says something about you now. Meanwhile OpenAI went the opposite direction: They took the Pentagon contract that Anthropic refused. Greg Brockman donated $25 million to fund wars. ChatGPT uninstalls jumped 295% in a single day. Reddit posts saying "Cancel and Delete ChatGPT" got 30,000 upvotes. Anthropic said no to mass surveillance and autonomous weapons. Got blacklisted by the Pentagon. Trump called them a "Radical Left AI company." And their downloads went to #1 on the App Store the next day. Turns out refusing to build weapons is good marketing. But the real damage isn't consumer downloads. It's the MONEY. Claude Code hit $2.5 billion in annual revenue in six months. OpenAI's competing product Codex just barely crossed $1 billion. And Anthropic literally cannot meet demand. They're turning away paying customers because they don't have enough compute to serve them. A company REJECTING revenue because it's growing too fast. While OpenAI scrambles to consolidate. Last week OpenAI announced they're merging ChatGPT, Codex, and their browser into one "superapp." But what this really means: "We launched too many products, none of them worked well enough alone, so now we're cramming everything together and hoping it sticks." And remember their video tool Sora? Launched standalone. Hit #1 on the App Store. Usage flatlined within weeks. Now they're forced to shut it down. Their browser Atlas? Still hasn't launched publicly. Their IPO? Polymarket odds dropped from 55% to 35%. OpenAI has 900 million users. Anthropic has maybe 10 million daily actives. But here's the thing... OpenAI won the consumer war. ChatGPT is where your mom asks about recipes and your cousin makes memes. Anthropic won the war that actually MATTERS. The developers. The engineers. The enterprises writing 7 figure checks. OpenAI built the biggest chatbot on Earth. Anthropic built the tool that companies can't stop paying for. This is Yahoo vs Google all over again. Yahoo had the users. Google had the product. And we all know how that ended. OpenAI has 12 months to prove the superapp works, land the IPO, and stop the enterprise bleeding. If they can't, the most valuable startup in history becomes the most cautionary tale in tech. 900 million users don't mean anything if the people who actually pay are walking out the door. What do you think?

Ricardo

35,020 просмотров • 5 месяцев назад

Claude Tag has completely changed the way I do work for the last 4 months. Except… it's not Claude Tag. Anthropic only announced that a few hours ago, and I don't even have access yet. But I did build a version of it for myself which I've been using for months now. Here's how. 4 months ago, inspired by the success of OpenClaw, I wondered what would happen if I let Claude Code on its own computer 24x7. So I built a simple harness that allowed me to turn any Mac into an AI employee with Claude Code headless mode (-p). Today, I manage 3 such AI employees. It started with Luo Ji — my and my brother Piyush Agarwal's AI co-founder, running in our personal Slack. Luo does real work for us. We've been writing a 100% of the code for 3 products on Slack with Luo now. It manages our emails and gives us a little brief each day with things we need to take action on. And so much more. And it's not just the two of us. On the consulting team at Every 🪨, we run Claudie and for the editorial team, Andy. Same architecture, same Slack, months of real work. They help the teams with work related to project management, chief-of-staff work, data hygiene, building decks, writing first drafts, even browsing X on their own account for AI updates. So it's mindblowing to see that Anthropic landed on the exact same architecture I did. Claude Tag is an AI employee that lives in your Slack workspace and does work autonomously. Anthropic says they've been running it internally for the better part of this year — opening PRs, doing real work. And so have I. So has my whole team. The architectural decisions Anthropic baked into Claude Tag are the ones we arrived at too: - Built on Claude Code - Uses its own accounts - A separate employee per team - Slack as the interface This is the future of work, and I've been living it for months. I've shifted all of my workflows — code, PRs, even the non-technical stuff — out of Claude Code and the Claude app and into Slack. I've had entire weeks where I never opened Claude Code on my laptop. Here's a video walkthrough of how I've been using this in real life.

Nityesh

36,790 просмотров • 2 месяцев назад