正在加载视频...

视频加载失败

Cursor, Windsurf … all cool. But @AugmentCode just dropped `Augment Agent` 200K context. Persistent memory. Deep tool integrations. ... and it just hit #1 on SWE-bench Verified (65.4% on real tasks). It’s kind of a big deal. Let me show you 🧵 ↓

63,421 次观看 • 1 年前 •via X (Twitter)

20 条评论

Charly Wargnier 的头像
Charly Wargnier1 年前

If you haven’t heard of SWE-bench, it’s the gold standard for testing AI agents on real @github issues. Agents must read codebases, reason through tasks, fix bugs unsupervised. Augment’s Agent hit 65.4%, *the* best open-source score to date! Results:

Charly Wargnier 的头像
Charly Wargnier1 年前

I use @cursor_ai a lot, so I honestly wasn’t expecting much from @augmentcode’s Agent. But wow - it plans, edits across files, runs tests, opens PRs - all from a single prompt, inside your IDE. Tried it on a scene data structure w/ confidence scores - nailed it on the first go!

Charly Wargnier 的头像
Charly Wargnier1 年前

.@augmentcode’s Agent also has next-level visual debugging: 1️⃣ Drag in a screenshot 2️⃣ Agent spots the UI issue (CSS, layout, logic) 3️⃣ Suggests a fix 4️⃣ Runs only the relevant tests Pretty ace!

Charly Wargnier 的头像
Charly Wargnier1 年前

🧩 It has persistent memory: - it can learn your coding style - remembers previous refactors - adapts to your infra and conventions It gets smarter the more you use it, no need to start from scratch each time.

Charly Wargnier 的头像
Charly Wargnier1 年前

🧰 Integrated Dev Workflows You can use @augmentcode’s Agent to go from ticket → code → PR without switching tools! e.g.: – @GitHub: branch, commit, PR – @Linear: issue detection + resolution – @Notion, @JIRA, @Confluence: context → implementation Pretty rad.

Charly Wargnier 的头像
Charly Wargnier1 年前

MCP is 🔥 and becoming a must-have in AI tooling. @augmentcode’s Agent supports it natively. – @supabase, @figma, @vercel, @Cloudflare, @getsentry – plug in APIs, SQL, CLI tools via settings.json You can even add your own tools to the agent’s context. →

Charly Wargnier 的头像
Charly Wargnier1 年前

Right now, you can try @augmentcode’s Augment Agent with unlimited agent calls. The folks at AugmentCode told me they’ll switch to usage-based pricing later - but for now, there’s absolutely no cap! Try it here →

Charly Wargnier 的头像
Charly Wargnier1 年前

That’s a wrap! If you found this thread useful, a quick RT goes a long way, especially for small teams like @augmentcode taking on the giants! 🦾 Follow me → @datachaz for takes on AI, LLMs, and agents. No fluff I promise.

Jeremy Wickersheimer 的头像
Jeremy Wickersheimer1 年前

@augmentcode The one thing I thought was missing, they added it. Very cool

Charly Wargnier 的头像
Charly Wargnier1 年前

@augmentcode Glad to hear, Jeremy! 🤗

Rama Jha (🏗️📝🎓) 的头像
Rama Jha (🏗️📝🎓)1 年前

@augmentcode 200K context AND persistent memory?! 👀 That's next level!

Charly Wargnier 的头像
Charly Wargnier1 年前

@augmentcode Man... I'm switching to it from Cursor! 🔥

xavier lois 的头像
xavier lois1 年前

@augmentcode Let me know when one of these tool finally manages to implement a proper software architecture ... and is able to stick to it!

Charly Wargnier 的头像
Charly Wargnier1 年前

@augmentcode what do you mean by "and is able to stick to it"?

BitFuturist 的头像
BitFuturist1 年前

@augmentcode AugmentCode is the next level 👍🏾

Mike H (eboss/acc) 的头像
Mike H (eboss/acc)1 年前

@augmentcode Love @augmentcode - def a tool that deserves more love

GamePrompts 的头像
GamePrompts1 年前

@augmentcode ohh this this looks nice, good little qol with the drag and capture thing

VibeCode 的头像
VibeCode1 年前

@augmentcode interesting

Mr Mastrianni (🚀,🔴) 的头像
Mr Mastrianni (🚀,🔴)1 年前

@augmentcode I have a hard time logging in

Alexander Myasoedov 的头像
Alexander Myasoedov1 年前

INTRODUCING: Agentic Security - LLM Security Scanner! 🔍 🔑 Features: Scans for prompt injections, jailbreaking & more. Provides detailed reports & options to customize attack rules. 🔗access the GitHub Link ↓

相关视频

🚀New Amazon Q Developer agent for software development is available to customers: This agent is based on a new agent architecture that has exciting results coming from the SWE-bench scores (on the full and verified benchmarks) representing AI models’ ability to resolve real-world coding problems. Interesting aspect of Q Agent is that with these newest updates, Q drove nearly 50% more successful coding tasks completed. What makes Q Dev Agent remarkable? The agent architecture is not just about using the best LLMs (which we do), but also giving the agent the ability to constantly explore multiple paths to find the best way to resolve a particular problem (and back tracking when it has reached dead end like a developer would do). Needless to say, we are just getting started on the developer agent and we are constantly pushing to advance our AI capabilities while maintaining quality, security, privacy, and reliability to keep Amazon Q Developer an innovative and trusted option available to our customers using agents for software development. We highlighted the results of our first SWE-bench submission of Amazon Q Developer back in June blog post; with these updates, our new agent resolves 51% more coding tasks than its previous iteration on the SWE-bench verified dataset, and 43% more on the full dataset. That’s the difference a few months make, and I can’t wait to share what our teams will deliver at re:Invent this December. Here's a quick demo showcasing our new Agent in action:

Swami Sivasubramanian

28,946 次观看 • 1 年前