Loading video...
Video Failed to Load
Jev’s founder just dropped the 12-page harness that makes coding agents 400× cheaper the first thing it shows: your agent spends under 10% of its tokens actually writing code → reading and searching take 56.2% of tool turns and 46.5% of tokens. that’s where the money goes → routing... show more
61,361 views • 5 days ago •via X (Twitter)
7 Comments

Most of the cost is hiding there

This makes search, reading, and context control cheaper; the new bottleneck is deciding what information deserves full fidelity before compression or routing.

compaction solves what the agent needs to see now. we can handle what the agent needs to remember later, including previous attempts and the state they left behind

@heyjevbook can you substantiate these claims?

Routing tool use by task and cost is a practical complement to shared context; the harder part is preserving those conventions across agents and sessions.

56.2% on reading and searching is wild, what are the top tips to optimize that?

That 56.2% reading number matches what I see locally. The fix that actually moved the needle was scoring every chunk before it entered context: full text for the risky 10%, tight summaries for the rest. Cheap models can do the scoring, the frontier model only sees what survived.
