Loading video...

Video Failed to Load

Go Home

i built a codebase classifier with Jev and this might be the solution to overengineered code that agents create what should I test Jev on next?

96,541 views • 2 days ago •via X (Twitter)

23 Comments

vogel's profile picture
vogel2 days ago

what should I try to build with Jev next?

Ceoz's profile picture
Ceoz2 days ago

Ok, that is ACTUALLY a pretty interesting concept.

Michael Vessia's profile picture
Michael Vessia2 days ago

Looks interesting, I am currently using for cyclomatic/cognitive complexity, CRAP, etc. So this might be redundant, but fallow takes about 40s on my pretty small codebase so this might be quicker, will have to compare results if you publish.

Ishwar | Infrastructure Systems's profile picture
Ishwar | Infrastructure Systems2 days ago

Test it on a repo after 100+ agent-driven changes. The interesting question isn't whether Jev can spot bad code. It's whether it can find the complexity agents accumulate while each individual change looks reasonable.

Pythonics's profile picture
Pythonics2 days ago

Slop scoring text & sections of text as feedback to have the model iteratively improve writing, acting as a more reliable judge.

mackle's profile picture
mackle2 days ago

So kinda like SonarQube?

alias's profile picture
alias2 days ago

Damn, this is exactly what I needed to understand Jev. Thank you, man. Brilliant.

Shawki Sukkar's profile picture
Shawki Sukkar2 days ago

Computer use

nin's profile picture
nin2 days ago

Test Jev on day trading. Perhaps, for speed's sake, use historical data but anonymize the companies and dates.

Ted Kalaw's profile picture
Ted Kalaw2 days ago

done anything w scores? try doing a line break on blocks of text, prepending w the line number, and then asking questions that reference the line number you can make almost like gradients by line (or other block). there’s a p cool demo for this on the devsite i believe!

Krimo's profile picture
Krimo2 days ago

Send an invite please 😢

Dominik Vít's profile picture
Dominik Vít2 days ago

Can you make it an MCP, plug it to OpenCode and let it refactor in a loop? Would love to see the results

Lauri Jutila's profile picture
Lauri Jutila2 days ago

Very good ideas. Thanks for sharing!

Calliope's profile picture
Calliope2 days ago

My thing I built it with big pickle

Harith Bakhrani's profile picture
Harith Bakhrani2 days ago

Throw it at a repo after 2 years of features. One where utils is a graveyard, the same API call exists 4 times, and half the types are any. If it can point to real duplicates and explain why they should be merged without trying to rewrite the whole thing, I’m interested.

Alek's profile picture
Alek2 days ago

Ryan - test it on a legacy repo next?

Ofek Shaked | AI Engineer's profile picture
Ofek Shaked | AI Engineer2 days ago

Agents love extra folders for no reason. A cheap pass that flags "this is overbuilt" saves more time than another lint rule.

NAMAN RAJ's profile picture
NAMAN RAJ2 days ago

Good luck testing Jev's limits, Ryan! Can't wait to hear how it handles the AI equivalent of a teenager's first car 🚗💥

Yaya Soumah's profile picture
Yaya Soumah2 days ago

This is super cool

shai granit's profile picture
shai granit2 days ago

בנית מסווג עם Jev ושואל מה הלאה תבדוק אם הוא מתחיל להתנצל ברגע שהקלט נשמע כמו טיקט מלקוח

Rohan's profile picture
Rohan2 days ago

Test it on the same feature implemented three ways: minimal, abstraction-heavy, and with deliberate dead paths. Expose the concrete signals behind each classification. A developer should be able to identify what to delete and review the score against the code.

pysolin's profile picture
pysolin2 days ago

verify agentic payment transactions

Keegan Gaffney's profile picture
Keegan Gaffney2 days ago

Let it play Pokémon

Related Videos