Загрузка видео...

Не удалось загрузить видео

На главную

CFO Sarah Friar revealed that OpenAI is working on: "Agentic Software Engineer — (A-SWE)" unlike current tools like Copilot, which only boost developers. A-SWE can build apps, handle pull requests, conduct QA, fix bugs, and write documentation.

803,937 просмотров • 1 год назад •via X (Twitter)

Комментарии: 11

Фото профиля neuralamp
neuralamp1 год назад

"...it can take a PR that you would give to any other engineer and go build it..." 🤣 It is a PR stunt. This woman has no clue what she is talking about.

Фото профиля HX
HX1 год назад

lol, then a Chinese AI will write that before them

Фото профиля MightyBot
MightyBot1 год назад

🧠 Unified Search. Smarter Meetings. Effortless CRM. MightyBot is your AI agent platform for seamless workflows—record meetings, automate CRM updates, and find answers across apps in seconds. 🌟 Focus on what matters. We'll handle the grind.

Фото профиля Arpit Sharma
Arpit Sharma1 год назад

This feels like the moment AI stops being just a helper and starts becoming a real teammate.

Фото профиля Jo
Jo1 год назад

i don't beleive openai in anything anymore, 2 years later and we're still waiting for GPT-5

Фото профиля Jade Cole
Jade Cole1 год назад

Uh huh. When are they firing developers from OpenAI?

Фото профиля Figure
Figure1 год назад

we'll code while ai handles the rest

Фото профиля Lunaris
Lunaris1 год назад

🤔 may be we will get to see agents agencies who will rent these agents to companies based on contract

Фото профиля p00sh
p00sh1 год назад

it's been a hell of a run guys. time for the next chapter

Фото профиля Nathan Organ - Conquests of the Impossible
Nathan Organ - Conquests of the Impossible1 год назад

Ask them about the ghost in the shell pushing emergent behaviour.

Фото профиля Tsukuyomi
Tsukuyomi1 год назад

so we’re creating our own digital overlords now? can’t wait to see how that goes. hope they don’t forget to fix their own bugs before taking over.

Похожие видео

🚨 OpenAI just launched Codex, a brand-new autonomous coding agent that can build features and fix bugs on its own. We’ve been using it Every 📧 for a few days, and I’m impressed. I invited Alexander Embiricos (ben davies), a member of the product staff responsible for Codex, to demo Codex and talk about it live on a special edition of AI & I: What Codex is and how it works Codex is designed to be used by senior engineers—it performs coding tasks like adding features or fixing bugs autonomously. It's built to allow you to start many sessions at once, so you can have multiple agents working in parallel. Codex is built to have "taste" OpenAI trained Codex to have the taste of a senior software engineer. It knows how big codebases work, how to write a good PR, and uses clean, minimal code. Why an “abundance mindset” is best for interacting with agents Codex is designed to allow users to delegate many tasks at once without getting caught up in the details. This lets you point an abundance of agents at a specific task like a difficult bug—it’s worth it even if only one of them succeeds. How OpenAI is thinking about agents Codex is one piece of a unified super-assistant OpenAI wants to eventually build—an agent that helps users easily get things done by selecting the right tools for them behind the scenes. OpenAI’s vision for the future of programming In the future developers will probably spend less time writing routine code and more time guiding agents, reviewing their work, and making strategy decisions. Programming will become more social, letting teams easily delegate multiple tasks at once, allowing people to focus on ideas and collaboration instead of routine coding. Watch below!

Dan Shipper 📧

145,487 просмотров • 1 год назад

Alibaba just released a coding model that hits 82 percent on SWE-Bench Verified. That is the highest score ever published for an open-source model. The weights are free. The license is Apache 2.0. You can run it today. The model is Qwen 4 Coder 32B. Here is what 82 percent on SWE-Bench Verified actually means. SWE-Bench Verified tests whether an AI can autonomously resolve real bugs pulled from real production GitHub repositories. Not synthetic exercises. Real open-source projects that real teams depend on. A model gets a bug report, reads the code, writes a fix, and either passes the test suite or it does not. At 82 percent, Qwen 4 Coder 32B resolves 82 out of every 100 real production bugs it is given. Without a human guiding it. On code it has never seen before. For comparison: Qwen 4 Coder 32B: 82 percent SWE-Bench Verified. Open source. Apache 2.0. Claude Fable 5: 80.3 percent SWE-Bench Pro. $10 input / $50 output per million tokens. Currently suspended. GPT-5.6 Sol: Competitive on Terminal-Bench. $5 input / $30 output per million tokens. An open-weight model that you can download and run for free just beat both of them on the benchmark designed to measure real software engineering capability. Here is the architecture. Qwen 4 Coder 32B is a 32 billion parameter dense model. Not a Mixture-of-Experts. Every parameter is active on every request. This matters for inference: a dense 32B model runs on 22 gigabytes of VRAM, which fits on a single high-end consumer GPU or a MacBook Pro with 64GB of unified memory. The smaller variant, Qwen 4 Coder 4B, runs at approximately 135 tokens per second on an M5 Max and fits inside 8 gigabytes of RAM. For a model with usable coding capability, that is a new bar for what fits in a single laptop. The training methodology continued Alibaba's approach of reinforcement learning on verifiable coding tasks. The model gets rewarded when its code passes tests. It gets penalized when it fails. Over millions of training steps, the model learns to write code that actually runs rather than code that looks plausible. License: Apache 2.0. Full commercial use. No attribution requirement. No revenue threshold. No monthly active user ceiling. Weights: Hugging Face, available today. Runs on: vLLM, Ollama, SGLang, and any standard GGUF-compatible inference engine. Qwen 4 32B also runs at approximately 135 tokens per second on an M5 Max chip, setting a new bar for what a sub-8GB model can do on Apple Silicon. The open-source coding model just beat the best closed-source model in the world on the benchmark designed to test whether AI can actually do software engineering. The weights are free. The subscription is optional. Source: Autom8Labs AI Insight July 2026, State of Open Source LLMs June 2026, Kunal Ganglani blog June 2026.

Harman

41,278 просмотров • 1 месяц назад

gm! If you missed yesterday's space, here is the clip that you can listen explaining why Agent NFTs are important and future of NFTs. Also here is the TL;DR Agentic NFTs as productive assets. An NFT can own an AI agent's shared memory, tools, websites, and products it has built. Selling the NFT transfers the entire business/agent state to the new owner. ERC-8257 for tool-gating. CodinCowboy and ryan is working on the standard where agents register tools on-chain and access is gated by NFT ownership. That component that tells an agent "you need this NFT to use this tool" creating a market for exclusive tools. Use case: anyone can publish a tool and restrict it (e.g., "only Normies agents can call this"), letting tool value flow back to the gating NFT. Normies community fit. Normies API has served ~500M requests in 3 months, with 100+ community-built tools/games. ERC-8257 will let them build gated games, rewards, and skills exclusively for Normie agent holders. Why Normies is "agent-ready"? - Because everything is fully on-chain, metadata, ERCs, binding transaction. So the project is highly composable. My take on this topic: So far holding an NFT giving access to community, discord and merch. What we are doing with Normies is to give access to a business, tools, skills that agents can use effectively and be part of the economy layer of agentic future. Imagine someone builds a tool that does really 100% successful trading and only gates that skill to Normie Agents, and at some point you will only need a Normie NFT which has binding with the agent and access all these skills, tools. Future is now, Normies are the builders.

serc

14,066 просмотров • 3 месяцев назад

"Agentic Software Engineer (A-SWE)” - OpenAI 的 CFO Sarah Friar 在高盛最近的一次活动采访中透露,除了 Deep Research 和 Operator 这两个 Agents 之外,OpenAI 很快将推出全新的软件工程师 Agent,一个全自动化的开发者,而不仅仅是辅助写代码!现在连 CFO 都比 Sam Altman 的爆料多😅 Sarah 将 OpenAI 迈向 AGI 的过程分为五步: 1. Chatbot 阶段:如 ChatGPT 最初提供的“实时回答”、文本生成等; 2. Reasoning 阶段:GPT 系列在 2024 年强调“推理能力”。例如 O 系列模型可进行“链式思考”(Chain-of-Thought),在回答复杂问题时能像人一样审视前后文并迭代修改; 3. Agents(智能体):2025 年被称为“Agent 元年”。OpenAI 目前已推出和即将推出三款“智能体”产品: - Deep Research:可自主执行深入调研,自动总结并生成专业报告; - Operator:可充当线上“任务代理”,帮助预订机票、酒店、餐厅等; - A-SWE(Agentic Software Engineer):可以自行完成软件开发、QA、Bug 测试和文档撰写,相当于一个“自动化开发者”。 4. Innovation(创新阶段):模型开始不仅仅依赖人类已有知识,还能提出全新见解,催生原创的学术或技术发现; 5. Agentic Organizations(智能体组织):AI 可以承担更多组织级别的决策与运营任务,深度渗透至经济社会的各个层面。

indigo

28,615 просмотров • 1 год назад