Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Nex-N2.5 feels less like “another model launch” and more like a push toward agents that can actually finish real work. The part that stands out to me is Nex-N2.5 Pro: a 397B multimodal model built for Vision, Computer Use, and long-horizon interaction. Instead of only understanding a screen, it...

66,447 görüntüleme • 6 gün önce •via X (Twitter)

33 Yorum

Aureus profil fotoğrafı
Aureus6 gün önce

This feels less like another model drop and more like a push toward agents that finish the job. Reading a changing UI and staying in the software is the part that stands out.

Figo profil fotoğrafı
Figo6 gün önce

The shift from AI that gives instructions to AI that can actually operate the tools is massive. This is the kind of Computer Use demo worth watching

江野 Lab profil fotoğrafı
江野 Lab6 gün önce

Real diff is reading UI changes, not just chat.

Almusty profil fotoğrafı
Almusty6 gün önce

That shift from writing a code snippet and stopping to maintaining a live visual feedback loop is what finally makes "computer use" feel practical rather than like a heavily scripted demo.

Ⓜ️ega Ⓜ️ind🤯(❖,❖) profil fotoğrafı
Ⓜ️ega Ⓜ️ind🤯(❖,❖)6 gün önce

It can navigate changing interfaces and correct itself, which specific software workflow are you planning to hand over to it first

Leo 的宏观笔记 profil fotoğrafı
Leo 的宏观笔记6 gün önce

Free tier rate-limited hard? Wanna run overnight.

Vantage profil fotoğrafı
Vantage6 gün önce

Idea to a Blender scene, then check the screen, fix it, verify again. That loop is the useful claim. Not a one-shot script.

Almusty profil fotoğrafı
Almusty6 gün önce

Having models like Pro handle UI-heavy tools, CAD, and desktop software natively means agents can finally bridge the gap between abstract user intent and a verified, working outcome.

CARAXCS profil fotoğrafı
CARAXCS6 gün önce

Games, Blender, CAD, browsers. The brief is pointing at desktop work, not chat demos. Long-horizon interaction is the bar.

Nova profil fotoğrafı
Nova6 gün önce

Mini and Pro for multimodal screen work, Max for the heavy reasoning jobs. Same-day access is how you actually test that split.

Em Kei profil fotoğrafı
Em Kei6 gün önce

A 397B model that can keep track of changing screens while working through a long task is definitely something I want to stress-test.

maximillion retardio profil fotoğrafı
maximillion retardio6 gün önce

The shift from “describe the click” to “keep going when the UI changes” is the only claim that matters. Curious which of those demos was a live loop vs a cleaned-up recording — Blender and FreeCAD are very different failure modes.

Em Kei profil fotoğrafı
Em Kei6 gün önce

Free on OpenRouter for a limited time makes this even easier to experiment with. I’d start with Pro and throw a real Blender or browser workflow at it.

sulaiman Waleed profil fotoğrafı
sulaiman Waleed6 gün önce

Most demos die after 1 action. This one navigates, battles, recovers. Actually useful. Free on OpenRouter 🔥

sulaiman Waleed profil fotoğrafı
sulaiman Waleed6 gün önce

This is the real test for AI agents. Playing Pokémon for hundreds of steps is insane. Trying it on OpenRouter now

Em Kei profil fotoğrafı
Em Kei6 gün önce

The shift from answering to actually operating software is the part that matters. 👀 That’s where AI starts feeling like an agent instead of a chatbot.

恒星sun profil fotoğrafı
恒星sun6 gün önce

写完代码还会自己跑一遍、看界面找问题再修,这段确实很戳人

Almusty profil fotoğrafı
Almusty6 gün önce

Splitting the family into multimodal vision variants for GUI interaction (Mini and Pro) alongside a massive text-only MoE powerhouse (Max) gives developers the right scaling axis depending on whether the bottleneck is visual tracking or deep reasoning.

SAD BOYS WIL SMILE AGAIN profil fotoğrafı
SAD BOYS WIL SMILE AGAIN6 gün önce

This is the real leap, agents that actually open the tools and finish the work.

Rubio is not Rubio profil fotoğrafı
Rubio is not Rubio6 gün önce

视频里最有说服力的不是会点按钮,而是它能跟着界面变化继续往下做:从零件摆放到灯光、分层、试图,中间还在改脚本。现在多数 Computer Use 还停在‘看懂屏幕’,真正卡在长程反馈闭环。Nex-N2.5 Pro 如果能在 Blender / FreeCAD 这类工具上稳定跑完一轮,比再堆一个基准分数更有意义。准备去 OpenRouter 试 Mini/Pro,想先看它处理报错后的自我修正稳不稳。

Amber 的财富手账 profil fotoğrafı
Amber 的财富手账6 gün önce

Agent that fixes its own bug > bigger params imo.

Cyber Jay profil fotoğrafı
Cyber Jay6 gün önce

Free trial first; paid only if it survives long runs.

Yuze profil fotoğrafı
Yuze6 gün önce

真正拉开差距的不是参数,是能盯着界面变化把活干完。

Ethos profil fotoğrafı
Ethos6 gün önce

reading a changing UI and staying in the software is the part that stands out to me

Moonora profil fotoğrafı
Moonora6 gün önce

397B built for vision and computer use only counts if it can stay in the tool step by step. Most models stop after the first pretty output.

Pavorix profil fotoğrafı
Pavorix6 gün önce

this feels less like another model drop and more like agents that can actually finish the job

Opti Max 💜🍀 profil fotoğrafı
Opti Max 💜🍀6 gün önce

After writing the code, they even run it themselves once, check the interface for issues, and fix them—this part really hits home.

Al-Shamus profil fotoğrafı
Al-Shamus6 gün önce

There lies the real test for AI agents. Playing Pokémon for many numbers of steps is insane. Trying it on OpenRouter now

阿良|AI 工作流 profil fotoğrafı
阿良|AI 工作流6 gün önce

Just tried Nex-N2.5 Pro on my Blender gear assembly. It set cameras, fixed lighting, organized parts, and rendered a clear exploded view on its own. First model that actually finishes the desktop work instead of only describing steps.

W.W profil fotoğrafı
W.W6 gün önce

Mini and Pro for multimodal computer use, Max at 1.6T for long reasoning. The split is clearer than another single flagship drop.

Alex Moore profil fotoğrafı
Alex Moore6 gün önce

Watching UI states makes long tasks practical

Ella Tech & Tool profil fotoğrafı
Ella Tech & Tool6 gün önce

This is huge "Do it" agents "talk about it" agents

OA🦁 profil fotoğrafı
OA🦁6 gün önce

The CAD and Blender examples are the part I’d watch. Once an agent can stay inside professional software long enough to finish a task, the use case gets a lot more serious.

Benzer Videolar

Introducing Dola Seed 2.0 Pro, referred to below as Seed 2.0 Pro We have launched Seed 2.0 Pro, our most capable model in the Dola Seed 2.0 series, engineered to power the next generation of autonomous AI agents. Enterprise AI is moving beyond models that simply analyze text or images. What businesses increasingly need are agents that can understand, reason, use tools, and execute tasks across complex workflows. That is exactly what Seed 2.0 Pro is built for. Seed 2.0 Pro combines strong reasoning with advanced image understanding and video understanding, giving enterprise agents the ability not only to interpret information, but also to take action. It is designed for high-value, multi-step enterprise workflows, with strong performance in: - tool calling - workflow execution across enterprise systems - agentic task completion - browser and computer use This makes Seed 2.0 Pro a powerful engine for a wide range of agent scenarios, from daily office automation and deep web research to in-depth report drafting, financial analysis, content moderation, physical inspection, and video creation workflows. It is also highly optimized for OpenClaw🦞 and ReAct architectures, helping enterprises build agents that can navigate digital interfaces, enter information, and complete tasks with high reliability. In short, Seed 2.0 Pro is not just built to generate insights. It is built to serve as the brain and execution engine for enterprise AI agents. And it brings these capabilities at a highly attractive price point, making advanced agent deployment more practical for enterprise teams. Try Seed 2.0 Pro for free: Or book a free consultation: #BytePlus #DolaSeed #EnterpriseAI #AIAgents #ImageUnderstanding #VideoUnderstanding #ReasoningModel #ModelArk #openclaw

BytePlus

96,329 görüntüleme • 5 ay önce

New course to bring you up to state-of-the-art at using AI to help you code: Build Apps with Windsurf's AI Coding Agents, built in partnership with WIndsurf (Codeium) and taught by Anshul Ramachandran! AI-assisted IDEs (Integrated Development Environments) make developers’ workflows faster, more efficient, and much more fun. Agentic tools like Windsurf are more than just code autocomplete—they are collaborative coding agents that help you break down complex applications, iterate efficiently, and generate code that spans multiple files. Although a lot of coding assistants share the same underlying large language models for planning and reasoning, a major point of distinction is how they handle tools, keep track of context, and stay aligned with your intent as a developer. For instance, if you make modifications to a class definition in your code and make the same modifications to other classes in the same directory, you might tell the AI agent "Do the same thing in similar places in this directory." Here, tracking your intent means understanding that “the same thing" refers to that recent edit you just made, which must be followed by appropriate search and tool-calling to implement the changes. In this course, you'll learn the inner workings of coding agents, their strengths and limitations, and how to use Windsurf to quickly build several applications. In detail, you'll: - Build a mental model of how agents work by combining human-action tracking, tool integration, and context awareness to carry out an agentic coding workflow. - Learn the challenges of code search and discovery and how a multi-step retrieval approach helps coding agents address them. - Use Windsurf to analyze and understand a large, old codebase and update it to the latest versions of the frameworks and packages it uses. - Build a Wikipedia data analysis app that retrieves, parses, and analyzes word frequencies. - Enhance the performance of your Wikipedia analysis app by adding caching, and through this, also learn how to course-correct when the AI agent produces unexpected results. - Learn tips and tricks such as keyboard shortcuts, autocomplete, and @ mentions to quickly call on agentic capabilities. - Use image/multimodal capabilities of the AI agent to increase your development velocity; you'll see an example of uploading a mockup with sketched-out UI features, and ask the agent to use that to build new functionality to an app. By the end of this course, you’ll understand agentic coding in-depth and know how to use it to make your development process much faster, more efficient, and enjoyable. Please sign up here!

Andrew Ng

139,978 görüntüleme • 1 yıl önce