Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

Introducing jevgrep - a research agent CLI powered by jev from TypeSafe AI that reduces your coding agent cost by 40% (verified on SWE-bench) Make sure to use the built in skill so your coding agent knows to use jg for context collection

409,232 görüntüleme • 2 gün önce •via X (Twitter)

49 Yorum

David profil fotoğrafı
David2 gün önce

This works because coding agents typically spend 30-60% of all its tokens on research to collect context before writing a single line of code. The actual code generation tokens are tiny. To install, just send claude/codex this exact repo (or even the exact tweet)

David profil fotoğrafı
David2 gün önce

A lot of time was spent on optimization, gpt-6-astra ran this in an autoresearch loop to optimize cost / perf for ~70hrs to find the optimal input & output shapes for API calls & the CLI This CLI is designed 100% for agents, its outputs would make no sense to a human

David profil fotoğrafı
David2 gün önce

Also - this video (including the sound) is made 100% in opus-5.5, what a time to be alive!

David profil fotoğrafı
David2 gün önce

On why this implementation is different compared to other general Jev based file search tools:

khaled profil fotoğrafı
khaled2 gün önce

@typesafeai related :)

David profil fotoğrafı
David2 gün önce

Took a quick look - this is a good generlized Jev implementation, but it won't work for coding agents, you need a recusive code discovery loop (e.g. an actual research agent) to make the context useful, else it'll be either too much context or too little and won't be useful enough to cut costs This would be good for humans where my jevgrep is made 100% for agents. You can try to use the cli yourself but the output will be too dense & confusing.

Tim Williams profil fotoğrafı
Tim Williams2 gün önce

@eltokh7 @typesafeai Yes I found exactly this - for reviewing code, even if you pass a ranked list of hunks in as context, the agent is still gonna just pull a huge chunk of the file context in anyway. Fighting the weights

David profil fotoğrafı
David2 gün önce

@eltokh7 @typesafeai Yup which is why the cli output needs to contain instructions for the agent and structured to be agent friendly, and also why the built in skill is important

Eliot Gevers profil fotoğrafı
Eliot Gevers2 gün önce

@typesafeai Does it pass the @theo test?

David profil fotoğrafı
David2 gün önce

@typesafeai @theo What test is that 😅

ahmad ghoniem profil fotoğrafı
ahmad ghoniem2 gün önce

i'd love to test it out if you want to truly take it a step further find a local classifier model (there are alot emerging every day) laya is the 1st that i can think of and let astra / opus 5.5 post train it (if it's doable) that's an experiment i might run myself if i found jevgrip useful haha

David profil fotoğrafı
David2 gün önce

@typesafeai Yea will def be testing out diff models

CV.YH profil fotoğrafı
CV.YH2 gün önce

@typesafeai Great man! I will do a Eikosgrep forking it!

lily zhang profil fotoğrafı
lily zhang2 gün önce

@typesafeai 40% is impressive, but isn't swe-bench retard? need the skill to generate this motion video ASAP!

David profil fotoğrafı
David2 gün önce

@typesafeai deepswe is better but that's like $500 per run. Swebench is the poor man's benchmark 😂

Essam Sleiman profil fotoğrafı
Essam Sleiman2 gün önce

@typesafeai very cool!

David profil fotoğrafı
David2 gün önce

@typesafeai Thanks! Give it a try, it's been making my max plans last a lot longer

samuelgao profil fotoğrafı
samuelgao2 gün önce

@typesafeai Cool, I want to try this

David profil fotoğrafı
David2 gün önce

@typesafeai Lmk how it works out for you!

Brjan | AI Builder profil fotoğrafı
Brjan | AI Builder2 gün önce

@typesafeai a 40% cost reduction is impressive, tools that optimize coding efficiency are essential

Yechan Do profil fotoğrafı
Yechan Do2 gün önce

@typesafeai Simple but strong idea

neamtu profil fotoğrafı
neamtu2 gün önce

@typesafeai swe bench is trash

David profil fotoğrafı
David2 gün önce

@typesafeai I know 😅

Travis Fischer profil fotoğrafı
Travis Fischer2 gün önce

@typesafeai LOVE this 💪 would be really cool to see a fuller eval comparing harnesses using ripgrep vs jevgrep

David profil fotoğrafı
David2 gün önce

@typesafeai If there's enough interest I will def put in some more $$ for a full run & with other models

rishub. profil fotoğrafı
rishub.2 gün önce

@typesafeai How to make a promo video like this?

David profil fotoğrafı
David2 gün önce

@typesafeai Opus 5.5 and my custom skill! I will release that soon as well

John Rood profil fotoğrafı
John Rood2 gün önce

@typesafeai the 40% only holds while the agent keeps calling jg. skill instructions are the kind of context that compacts away first, and once they are gone runs quietly revert to raw greps. track adoption over long sessions, and re-inject the skill at compaction.

Magik profil fotoğrafı
Magik2 gün önce

@typesafeai Ha, made one too -

The Coding Sloth profil fotoğrafı
The Coding Sloth2 gün önce

@typesafeai This video is impressive wtf

⚡️Federico (rawnly) profil fotoğrafı
⚡️Federico (rawnly)2 gün önce

@typesafeai How does this compare to FFF-grep?

David profil fotoğrafı
David2 gün önce

@typesafeai I haven't benchmarked against the 2 but if there's enough interest I will (and against ripgrep as well)

⚡️Federico (rawnly) profil fotoğrafı
⚡️Federico (rawnly)2 gün önce

@typesafeai Would be nice to see! Currently i’m using FFF almost everywhere’s supported

Sophie 🌟 profil fotoğrafı
Sophie 🌟2 gün önce

@typesafeai the skill so it actually uses the cheap tool. needed that

Tax Dude profil fotoğrafı
Tax Dude2 gün önce

@typesafeai ngl the video looks great. I might give it a try at some point

Monty profil fotoğrafı
Monty2 gün önce

@typesafeai sick!

CoinCollector profil fotoğrafı
CoinCollector2 gün önce

@typesafeai just tried it, sadly not useful at all, super slow on repos, thx anyway!

Isaac Hinman profil fotoğrafı
Isaac Hinman2 gün önce

@typesafeai How is this better than semble?

Timothy LeGendre profil fotoğrafı
Timothy LeGendre2 gün önce

@typesafeai This is wild!

Fausto Yuuki profil fotoğrafı
Fausto Yuuki2 gün önce

@typesafeai can u compare it against fff ?

Chris Stvn profil fotoğrafı
Chris Stvn2 gün önce

@typesafeai @theo what about this?

Mike Lydick profil fotoğrafı
Mike Lydick2 gün önce

A/B'd jg against ripgrep on a codebase we maintain. The right file often ranked first, but we still got a 12-51 file flood in seconds vs tens of ms for rg. NL ranking wins as the seed when you don't know the symbol, then you walk defs/refs/deps. Where does jg win beyond cold-repo exploration?

Konstantin Anagnostou profil fotoğrafı
Konstantin Anagnostou2 gün önce

@typesafeai Do you think is good for Hermes?

Kashif Ali Khan profil fotoğrafı
Kashif Ali Khan2 gün önce

@typesafeai cutting 40% token cost on swe-bench context collection is actually massive

G profil fotoğrafı
G2 gün önce

@typesafeai This is definitely one of the smartest use cases of Jev I've seen

Inferred profil fotoğrafı
Inferred2 gün önce

@typesafeai Worth a try, context is still a hard question right now

WAGMİ 100x💎 profil fotoğrafı
WAGMİ 100x💎2 gün önce

@typesafeai context collection is where the real agent cost hides — everyone optimizes the model call, nobody optimizes what feeds it. does the 40% hold pass@1 though, or did swe-bench resolution rate move with it?

zahir profil fotoğrafı
zahir2 gün önce

@typesafeai what in tarnation is this motion design

Webster | JARVIS profil fotoğrafı
Webster | JARVIS2 gün önce

@typesafeai Nice, a research CLI that cuts agent cost by 40% is super handy. The built-in skill for jg is a smart touch too. Congrats on the SWE-bench numbers!

Benzer Videolar

this is unreal f*cking gold for Jev builders 20 repos people are building on Jev right now. browser agents, context tools, trading bots, even a drone 1. JEV-Ultrafast - a fast browser agent ↳ 2. Fast-JEV-Compaction - context compression ↳ 3. JSON-Render - generative UI ↳ 4. Typesafe-MCP - use Jev with any client ↳ 5. JEV-MCP - a judgment toolkit ↳ 6. Semdecide - a classifier that lives in your CLI ↳ 7. JEV-Codex-Router - routes each task to the right model ↳ 8. Winnow - garbage collection for your context ↳ 9. JEV-Review - code review triage ↳ 10. Blink - a repo navigator ↳ 11. Agent-Desktop - desktop automation ↳ 12. Typesafe-Mario - an agent that plays Super Mario ↳ 13. JEV-Drone - drone control ↳ 14. OneVOneJev - a browser FPS ↳ 15. JEV-Trader - HFT market making ↳ 16. Prism - liquidity signal detection ↳ 17. Neo4Jev - knowledge graph traversal ↳ 18. JEV-Curate - training data screening ↳ 19. Canny - checks whether a task was actually completed ↳ 20. KillMyIdea - scores startup ideas before you build them ↳ pick by what you do: > coding -> JEV-Review, Blink, Canny, JEV-Codex-Router > context -> Fast-JEV-Compaction, Winnow > automation -> JEV-Ultrafast, Agent-Desktop > clients and tools -> Typesafe-MCP, JEV-MCP, Semdecide > UI -> JSON-Render > trading -> JEV-Trader, Prism > data -> Neo4Jev, JEV-Curate > founders -> KillMyIdea > just for fun -> Typesafe-Mario, OneVOneJev, JEV-Drone grab the one closest to your job and ship something on top of it this week

Mr. Buzzoni

28,574 görüntüleme • 4 gün önce