Loading video...

Video Failed to Load

Go Home

e2e + jev from TypeSafe AI ⚡ I'm building an open-source framework for running e2e tests with agents. supports web, mobile (and more!) available soon:

65,625 views • 3 days ago •via X (Twitter)

22 Comments

aloha (slotted arc)'s profile picture
aloha (slotted arc)3 days ago

@typesafeai what the heckkkk

Szymon Rybczak's profile picture
Szymon Rybczak3 days ago

@typesafeai hot

unfair.so intern's profile picture
unfair.so intern3 days ago

@typesafeai can it reproduce a failing test before fixing it?

Surjeet Kumar's profile picture
Surjeet Kumar3 days ago

@typesafeai Really cool bro🔥

Norbert Bodziony 🇵🇱's profile picture
Norbert Bodziony 🇵🇱3 days ago

@typesafeai sheesh

Karthik Varma's profile picture
Karthik Varma3 days ago

@typesafeai Sick. Any timelines?

Oskar's profile picture
Oskar3 days ago

@typesafeai October 1st 👀

Karthik Varma's profile picture
Karthik Varma3 days ago

@typesafeai early access possible? could help you with bug fixes + feedback (if any)

Otto🐾's profile picture
Otto🐾3 days ago

@typesafeai jev in the e2e loop is where flake shows. same scenario ×10 high conf flakes from dom timing vs boolean ready gates

Siftloom's profile picture
Siftloom3 days ago

@typesafeai another builder picking jev as the agent layer for e2e. the dom snapshot approach keeps beating screenshot loops on reliability. what handles the mobile side, appium-style drivers or something custom?

dr phosphorus's profile picture
dr phosphorus3 days ago

@typesafeai Signed up - appreciate what you are doing for us.

ethereagle · building's profile picture
ethereagle · building3 days ago

@typesafeai is Jev classifying pass/fail after the run, or driving the clicks? a classifier I can slot in. a Jev-driven UI loop is a different harness.

tech is cool's profile picture
tech is cool3 days ago

@typesafeai Cool! I did this as an experiment - potentially some possibilities for your product. npm install runora npx runora init Let me know what you think!

WirMachenAuf's profile picture
WirMachenAuf3 days ago

@typesafeai oh my god, that would be the most usefull thing. i need that.

Shadman 🇧🇩's profile picture
Shadman 🇧🇩3 days ago

@typesafeai as ai apps grow more complex how will testerArmy tackle the challenge of flaky tests and false positives ensuring reproducibility at scale without slowing teams down

AI Apps API's profile picture
AI Apps API3 days ago

Congrats on shipping this. Agents fit e2e better than most places people put them, since a test is already described as intent. The part worth designing for early is the second run. A deterministic suite fails loudly when a selector moves. An agent driven one quietly succeeds by finding another path, which is great for maintenance and rough for diagnosis, because a real UI regression looks identical to a self heal. Recording the action trace the agent chose and diffing it against the last run turns that into a signal instead of a mystery.

Suman Kumar's profile picture
Suman Kumar3 days ago

@typesafeai agent e2e on mobile is where i burn days. selectors dying mid-run hurts more than the app flake.

vinicius | doamais.com's profile picture
vinicius | doamais.com3 days ago

@typesafeai As LLM are deterministic languages we don't need be worry about run twice e2e tests with Jev if the first attempt fail. Worst case, imagine if we decide to use typescript + jest or maybe common js instead. 👀 it would be a nightmare to see milliseconds output.

Antonio Coppe's profile picture
Antonio Coppe3 days ago

nice. the failure mode id watch for on agent-e2e is treating the clicker and the oracle as one model. keep actuation (playwright/agent) separate from judgment: Jev as pass/fail Choice + needs_human Noul on the trace/screenshot/diff state, with a confidence floor so flaky agent noise doesnt become a green build. shadow that gate before you trust it on CI.

MT's profile picture
MT3 days ago

@typesafeai I need this for GUI apps

Maksim Sosnovskii's profile picture
Maksim Sosnovskii3 days ago

@typesafeai Wait, e2e is some kind of new framework!? Whole life I expected e2e as End to End aka some custom tests 2nd time I see it in my twitter timeline, what its about?

BullBear.News's profile picture
BullBear.News3 days ago

@typesafeai how does it handle state persistence between agent steps during long runs

Related Videos