Loading video...

Video Failed to Load

Go Home

Introducing jevgrep - a research agent CLI powered by jev from TypeSafe AI that reduces your coding agent cost by 40% (verified on SWE-bench) Make sure to use the built in skill so your coding agent knows to use jg for context collection

409,232 views • 2 days ago •via X (Twitter)

49 Comments

David's profile picture
David2 days ago

This works because coding agents typically spend 30-60% of all its tokens on research to collect context before writing a single line of code. The actual code generation tokens are tiny. To install, just send claude/codex this exact repo (or even the exact tweet)

David's profile picture
David2 days ago

A lot of time was spent on optimization, gpt-6-astra ran this in an autoresearch loop to optimize cost / perf for ~70hrs to find the optimal input & output shapes for API calls & the CLI This CLI is designed 100% for agents, its outputs would make no sense to a human

David's profile picture
David2 days ago

Also - this video (including the sound) is made 100% in opus-5.5, what a time to be alive!

David's profile picture
David2 days ago

On why this implementation is different compared to other general Jev based file search tools:

khaled's profile picture
khaled2 days ago

@typesafeai related :)

David's profile picture
David2 days ago

Took a quick look - this is a good generlized Jev implementation, but it won't work for coding agents, you need a recusive code discovery loop (e.g. an actual research agent) to make the context useful, else it'll be either too much context or too little and won't be useful enough to cut costs This would be good for humans where my jevgrep is made 100% for agents. You can try to use the cli yourself but the output will be too dense & confusing.

Tim Williams's profile picture
Tim Williams2 days ago

@eltokh7 @typesafeai Yes I found exactly this - for reviewing code, even if you pass a ranked list of hunks in as context, the agent is still gonna just pull a huge chunk of the file context in anyway. Fighting the weights

David's profile picture
David2 days ago

@eltokh7 @typesafeai Yup which is why the cli output needs to contain instructions for the agent and structured to be agent friendly, and also why the built in skill is important

Eliot Gevers's profile picture
Eliot Gevers2 days ago

@typesafeai Does it pass the @theo test?

David's profile picture
David2 days ago

@typesafeai @theo What test is that 😅

ahmad ghoniem's profile picture
ahmad ghoniem2 days ago

i'd love to test it out if you want to truly take it a step further find a local classifier model (there are alot emerging every day) laya is the 1st that i can think of and let astra / opus 5.5 post train it (if it's doable) that's an experiment i might run myself if i found jevgrip useful haha

David's profile picture
David2 days ago

@typesafeai Yea will def be testing out diff models

CV.YH's profile picture
CV.YH2 days ago

@typesafeai Great man! I will do a Eikosgrep forking it!

lily zhang's profile picture
lily zhang2 days ago

@typesafeai 40% is impressive, but isn't swe-bench retard? need the skill to generate this motion video ASAP!

David's profile picture
David2 days ago

@typesafeai deepswe is better but that's like $500 per run. Swebench is the poor man's benchmark 😂

Essam Sleiman's profile picture
Essam Sleiman2 days ago

@typesafeai very cool!

David's profile picture
David2 days ago

@typesafeai Thanks! Give it a try, it's been making my max plans last a lot longer

samuelgao's profile picture
samuelgao2 days ago

@typesafeai Cool, I want to try this

David's profile picture
David2 days ago

@typesafeai Lmk how it works out for you!

Brjan | AI Builder's profile picture
Brjan | AI Builder2 days ago

@typesafeai a 40% cost reduction is impressive, tools that optimize coding efficiency are essential

Yechan Do's profile picture
Yechan Do2 days ago

@typesafeai Simple but strong idea

neamtu's profile picture
neamtu2 days ago

@typesafeai swe bench is trash

David's profile picture
David2 days ago

@typesafeai I know 😅

Travis Fischer's profile picture
Travis Fischer2 days ago

@typesafeai LOVE this 💪 would be really cool to see a fuller eval comparing harnesses using ripgrep vs jevgrep

David's profile picture
David2 days ago

@typesafeai If there's enough interest I will def put in some more $$ for a full run & with other models

rishub.'s profile picture
rishub.2 days ago

@typesafeai How to make a promo video like this?

David's profile picture
David2 days ago

@typesafeai Opus 5.5 and my custom skill! I will release that soon as well

John Rood's profile picture
John Rood2 days ago

@typesafeai the 40% only holds while the agent keeps calling jg. skill instructions are the kind of context that compacts away first, and once they are gone runs quietly revert to raw greps. track adoption over long sessions, and re-inject the skill at compaction.

Magik's profile picture
Magik2 days ago

@typesafeai Ha, made one too -

The Coding Sloth's profile picture
The Coding Sloth2 days ago

@typesafeai This video is impressive wtf

⚡️Federico (rawnly)'s profile picture
⚡️Federico (rawnly)2 days ago

@typesafeai How does this compare to FFF-grep?

David's profile picture
David2 days ago

@typesafeai I haven't benchmarked against the 2 but if there's enough interest I will (and against ripgrep as well)

⚡️Federico (rawnly)'s profile picture
⚡️Federico (rawnly)2 days ago

@typesafeai Would be nice to see! Currently i’m using FFF almost everywhere’s supported

Sophie 🌟's profile picture
Sophie 🌟2 days ago

@typesafeai the skill so it actually uses the cheap tool. needed that

Tax Dude's profile picture
Tax Dude2 days ago

@typesafeai ngl the video looks great. I might give it a try at some point

Monty's profile picture
Monty2 days ago

@typesafeai sick!

CoinCollector's profile picture
CoinCollector2 days ago

@typesafeai just tried it, sadly not useful at all, super slow on repos, thx anyway!

Isaac Hinman's profile picture
Isaac Hinman2 days ago

@typesafeai How is this better than semble?

Timothy LeGendre's profile picture
Timothy LeGendre2 days ago

@typesafeai This is wild!

Fausto Yuuki's profile picture
Fausto Yuuki2 days ago

@typesafeai can u compare it against fff ?

Chris Stvn's profile picture
Chris Stvn2 days ago

@typesafeai @theo what about this?

Mike Lydick's profile picture
Mike Lydick2 days ago

A/B'd jg against ripgrep on a codebase we maintain. The right file often ranked first, but we still got a 12-51 file flood in seconds vs tens of ms for rg. NL ranking wins as the seed when you don't know the symbol, then you walk defs/refs/deps. Where does jg win beyond cold-repo exploration?

Konstantin Anagnostou's profile picture
Konstantin Anagnostou2 days ago

@typesafeai Do you think is good for Hermes?

Kashif Ali Khan's profile picture
Kashif Ali Khan2 days ago

@typesafeai cutting 40% token cost on swe-bench context collection is actually massive

G's profile picture
G2 days ago

@typesafeai This is definitely one of the smartest use cases of Jev I've seen

Inferred's profile picture
Inferred2 days ago

@typesafeai Worth a try, context is still a hard question right now

WAGMİ 100x💎's profile picture
WAGMİ 100x💎2 days ago

@typesafeai context collection is where the real agent cost hides — everyone optimizes the model call, nobody optimizes what feeds it. does the 40% hold pass@1 though, or did swe-bench resolution rate move with it?

zahir's profile picture
zahir2 days ago

@typesafeai what in tarnation is this motion design

Webster | JARVIS's profile picture
Webster | JARVIS2 days ago

@typesafeai Nice, a research CLI that cuts agent cost by 40% is super handy. The built-in skill for jg is a smart touch too. Congrats on the SWE-bench numbers!

Related Videos

this is unreal f*cking gold for Jev builders 20 repos people are building on Jev right now. browser agents, context tools, trading bots, even a drone 1. JEV-Ultrafast - a fast browser agent ↳ 2. Fast-JEV-Compaction - context compression ↳ 3. JSON-Render - generative UI ↳ 4. Typesafe-MCP - use Jev with any client ↳ 5. JEV-MCP - a judgment toolkit ↳ 6. Semdecide - a classifier that lives in your CLI ↳ 7. JEV-Codex-Router - routes each task to the right model ↳ 8. Winnow - garbage collection for your context ↳ 9. JEV-Review - code review triage ↳ 10. Blink - a repo navigator ↳ 11. Agent-Desktop - desktop automation ↳ 12. Typesafe-Mario - an agent that plays Super Mario ↳ 13. JEV-Drone - drone control ↳ 14. OneVOneJev - a browser FPS ↳ 15. JEV-Trader - HFT market making ↳ 16. Prism - liquidity signal detection ↳ 17. Neo4Jev - knowledge graph traversal ↳ 18. JEV-Curate - training data screening ↳ 19. Canny - checks whether a task was actually completed ↳ 20. KillMyIdea - scores startup ideas before you build them ↳ pick by what you do: > coding -> JEV-Review, Blink, Canny, JEV-Codex-Router > context -> Fast-JEV-Compaction, Winnow > automation -> JEV-Ultrafast, Agent-Desktop > clients and tools -> Typesafe-MCP, JEV-MCP, Semdecide > UI -> JSON-Render > trading -> JEV-Trader, Prism > data -> Neo4Jev, JEV-Curate > founders -> KillMyIdea > just for fun -> Typesafe-Mario, OneVOneJev, JEV-Drone grab the one closest to your job and ship something on top of it this week

Mr. Buzzoni

28,472 views • 4 days ago