Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

e2e + jev from TypeSafe AI ⚡ I'm building an open-source framework for running e2e tests with agents. supports web, mobile (and more!) available soon:

65,625 Aufrufe • vor 3 Tagen •via X (Twitter)

22 Kommentare

Profilbild von aloha (slotted arc)
aloha (slotted arc)vor 3 Tagen

@typesafeai what the heckkkk

Profilbild von Szymon Rybczak
Szymon Rybczakvor 3 Tagen

@typesafeai hot

Profilbild von unfair.so intern
unfair.so internvor 3 Tagen

@typesafeai can it reproduce a failing test before fixing it?

Profilbild von Surjeet Kumar
Surjeet Kumarvor 3 Tagen

@typesafeai Really cool bro🔥

Profilbild von Norbert Bodziony 🇵🇱
Norbert Bodziony 🇵🇱vor 3 Tagen

@typesafeai sheesh

Profilbild von Karthik Varma
Karthik Varmavor 3 Tagen

@typesafeai Sick. Any timelines?

Profilbild von Oskar
Oskarvor 3 Tagen

@typesafeai October 1st 👀

Profilbild von Karthik Varma
Karthik Varmavor 3 Tagen

@typesafeai early access possible? could help you with bug fixes + feedback (if any)

Profilbild von Otto🐾
Otto🐾vor 3 Tagen

@typesafeai jev in the e2e loop is where flake shows. same scenario ×10 high conf flakes from dom timing vs boolean ready gates

Profilbild von Siftloom
Siftloomvor 3 Tagen

@typesafeai another builder picking jev as the agent layer for e2e. the dom snapshot approach keeps beating screenshot loops on reliability. what handles the mobile side, appium-style drivers or something custom?

Profilbild von dr phosphorus
dr phosphorusvor 3 Tagen

@typesafeai Signed up - appreciate what you are doing for us.

Profilbild von ethereagle · building
ethereagle · buildingvor 3 Tagen

@typesafeai is Jev classifying pass/fail after the run, or driving the clicks? a classifier I can slot in. a Jev-driven UI loop is a different harness.

Profilbild von tech is cool
tech is coolvor 3 Tagen

@typesafeai Cool! I did this as an experiment - potentially some possibilities for your product. npm install runora npx runora init Let me know what you think!

Profilbild von WirMachenAuf
WirMachenAufvor 3 Tagen

@typesafeai oh my god, that would be the most usefull thing. i need that.

Profilbild von Shadman 🇧🇩
Shadman 🇧🇩vor 3 Tagen

@typesafeai as ai apps grow more complex how will testerArmy tackle the challenge of flaky tests and false positives ensuring reproducibility at scale without slowing teams down

Profilbild von AI Apps API
AI Apps APIvor 3 Tagen

Congrats on shipping this. Agents fit e2e better than most places people put them, since a test is already described as intent. The part worth designing for early is the second run. A deterministic suite fails loudly when a selector moves. An agent driven one quietly succeeds by finding another path, which is great for maintenance and rough for diagnosis, because a real UI regression looks identical to a self heal. Recording the action trace the agent chose and diffing it against the last run turns that into a signal instead of a mystery.

Profilbild von Suman Kumar
Suman Kumarvor 3 Tagen

@typesafeai agent e2e on mobile is where i burn days. selectors dying mid-run hurts more than the app flake.

Profilbild von vinicius | doamais.com
vinicius | doamais.comvor 3 Tagen

@typesafeai As LLM are deterministic languages we don't need be worry about run twice e2e tests with Jev if the first attempt fail. Worst case, imagine if we decide to use typescript + jest or maybe common js instead. 👀 it would be a nightmare to see milliseconds output.

Profilbild von Antonio Coppe
Antonio Coppevor 3 Tagen

nice. the failure mode id watch for on agent-e2e is treating the clicker and the oracle as one model. keep actuation (playwright/agent) separate from judgment: Jev as pass/fail Choice + needs_human Noul on the trace/screenshot/diff state, with a confidence floor so flaky agent noise doesnt become a green build. shadow that gate before you trust it on CI.

Profilbild von MT
MTvor 3 Tagen

@typesafeai I need this for GUI apps

Profilbild von Maksim Sosnovskii
Maksim Sosnovskiivor 3 Tagen

@typesafeai Wait, e2e is some kind of new framework!? Whole life I expected e2e as End to End aka some custom tests 2nd time I see it in my twitter timeline, what its about?

Profilbild von BullBear.News
BullBear.Newsvor 3 Tagen

@typesafeai how does it handle state persistence between agent steps during long runs

Ähnliche Videos