Загрузка видео...
Не удалось загрузить видео
Your coding agent can run all night. It still can't tell if what it built actually works. Today we're open-sourcing the TestSprite CLl (Apache-2.0) A tool your agent calls on its own to test your app end-to-end like a real user, fix what broke, and re-check everything it ever... show more
1,042,568 просмотров • 3 месяцев назад •via X (Twitter)
Комментарии: 35

Love All The Support!!!

This is massive. Verification > raw generation. Open-sourcing the CLI + proving cheaper models win with it is exactly what the agent space needed.

Reliability is becoming the new frontier for AI development tools.

Giving a cheap model the ability to test and fix its own code is brilliant.

AI coding agents are impressive. AI coding agents that can test, verify, and fix their own output are where things get really interesting.🚀

This is a strong reminder that verification matters more than model size. Generating code is one thing, ensuring it actually works is what creates real value.

Nice! Open-sourcing the E2E testing CLI is a big win for AI agents.

This is game-changing for the future of AI coding! Agents that can actually test, fix, and verify their own work — at a fraction of the cost? That's real progress. Huge respect for open-sourcing it under Apache-2.0. Can't wait to see more builders putting this to work. Keep leading the way! #DI88 #MARE88

The real benchmark isn't writing code—it's proving the code works. Impressive.

Very few people talk about evaluation in AI coding workflows. Glad to see more focus on it.

This is a big step forward for AI coding. Building is easy verifying it actually works is the hard part. Automated end-to-end testing changes the game.

This is actually a big step forward. Writing code is one thing, but automatically testing, fixing, and validating it like a real user is where the real value comes from. Love seeing tools focused on reliability instead of just generation.

Finalmente estamos saindo da fase de “o agente escreveu” para “o agente verificou se realmente funciona”

This is the missing piece for AI coding agents. Generating code is easy; verifying that it actually works end-to-end is the hard part. Open-sourcing TestSprite CLI and proving its value on a public leaderboard is a strong step toward more reliable AI-built software. Excited to see how it performs in real-world projects.

Verification > generation. 🔥 This changes the game. 🚀 AI that tests itself. Finally. The missing layer for AI coding. 👏 Shipping working software > writing code. Trust, but verify. 🤖 Huge step for autonomous coding. The future is self-testing agents. 🔥 This is how AI ships production code. Proof beats promises. 🚀

TestSprite is doing exactly what AI development needs right now—turning generated code into trustworthy code. Love the focus on real-world testing, reliability, and automation

AI can generate code in seconds. The real challenge is knowing whether it works. This solves the part most tools still miss.

This solves a real, overlooked problem.

This could dramatically reduce the gap between generated code and production-ready code.

Game-changer for AI agents! Verification > generation. Excited to try the CLI and see agents actually ship reliable code.

Love the focus on validation instead of generation.

This is Amazing Let me try this out!

How well does it handle auth flows and stateful sessions in real apps? Super excited to try it with my agents.

Cheapest model, strongest result, lower cost. That’s the kind of benchmark that changes buying decisions.

89% correctness with the cheapest model is wild. Looks like testing quality is becoming more important than model size

This highlights an important shift in AI coding: success is no longer just about generating code, but validating it.

Verification is the real game changer. Great move! This is the future of AI coding.

Smart automation that reduces human error delivers immediate value.

Finally, agents can verify reality instead of trusting their own code.

Absolutely thrilled to see the TestSprite CLI open-sourced! 🎉 The promise of agents building reliable software at a fraction of the cost is revolutionary. Thank you for this impactful contribution to the developer community

Appreciate your insights, always learning here!

This makes agents feel less like fast typists and more like actual builders with quality control.

Finally, testing that doesn't depend on model size or price.

@ArewaOS26 you need need this I think since we already started coding

The future isn't just AI that writes code it's AI that verifies, fixes, and validates it automatically.
