Загрузка видео...

Не удалось загрузить видео

На главную

this harness stack is f*cking gold TypeSafe AI + LangChain finally dropped the layer that makes agents affordable at company scale: jev running inside the langchain agent loop as middleware the problem every big team hits: each fork in the loop is another full LLM call. at 10 agents...

31,666 просмотров • 5 дней назад •via X (Twitter)

Комментарии: 21

Фото профиля Hussain Hashim | Building SundayBack
Hussain Hashim | Building SundayBack4 дней назад

@N01ennn one thing I've noticed is the cost runaway when not controlling those loops tightly. Hit that wall building my own thing. Keeping an eye on loop depth saves sanity.

Фото профиля Gipp 🦅
Gipp 🦅5 дней назад

pure builder gold, trying this stack tonight broski

Фото профиля NO1ennn
NO1ennn4 дней назад

thx bro , im try , its was not hard make it because JEV cooking

Фото профиля Harley Lewis Foote
Harley Lewis Foote4 дней назад

LangChain's middleware finally exposes token burn per tool call. Still can't see which MCP server lied to your agent though.

Фото профиля slash1s
slash1s4 дней назад

fr good visualisation

Фото профиля MrOzi
MrOzi5 дней назад

The real unlock isn't another layer, it's killing redundant LLM calls before they happen.

Фото профиля magsimich
magsimich4 дней назад

Putting Jev inside the agent loop makes sense

Фото профиля PEF
PEF4 дней назад

Excellent. C'est le meilleur cas d'usage que j'ai pu constater jusqu'à maintenant. Merci

Фото профиля Crio Songo
Crio Songo4 дней назад

This actually solves the LLM cost problem for enterprise-scale agents, pretty neat.

Фото профиля George O'Nair
George O'Nair4 дней назад

Company-scale agent cost usually dies in the loop, not the model bill. I'd meter retries + tool calls per accepted artifact, and fail the run when the harness can't show that ratio.

Фото профиля Gubernaut
Gubernaut4 дней назад

Every fork billed as a full model call is how spend multiplies. Keep cheap decisions off that path, and hard-cap the loop either way.

Фото профиля GBE
GBE4 дней назад

this dashboard making my head hurt

Фото профиля beamnxw ./
beamnxw ./4 дней назад

crazy idea bros

Фото профиля ALEXYZ
ALEXYZ4 дней назад

middleware at the right layer serious cost control

Фото профиля catman
catman4 дней назад

This is like putting a fast tollbooth before the expensive road: cheap checks handle routine turns, while the frontier model takes the genuinely hard ones.

Фото профиля Eugene Oldman
Eugene Oldman4 дней назад

4 gates at 0.95 each = 0.81. enjoy

Фото профиля Roshni
Roshni4 дней назад

Absolute goldmine of a stack breakdown. Definitely testing this out!

Фото профиля V1nT
V1nT4 дней назад

This is realy focking gold

Фото профиля Myttle
Myttle4 дней назад

4 jev gates per step adds up fast

Фото профиля rewind
rewind4 дней назад

saving this stack bro

Фото профиля Alek Dob
Alek Dob4 дней назад

yeah counting the forks. that's where the cheap-model pick actually pays

Похожие видео

JEV + OPUS 5.5 IS INSANE FOR BUILDING A COMPANY BRAIN I pulled the whole architecture out of the TypeSafe and Anthropic docs and packed it into a 14-page PDF the 10 steps: 1. meet the pair > Opus 5.5 thinks, Jev decides, your code holds the branch 2. stop asking a text generator for a yes or no > Jev returns a typed answer with a calibrated probability in 0.44s for $0.00035 3. ask everything at once > Choice, Score and Noul run in parallel, so the fourth question costs almost nothing 4. branch on the number > 0.999 goes straight into the if statement. ~99% of turns end right here 5. stop routing blind > Opus 5.5 to Sonnet and back costs 5.84 against 3.32 for staying on 5.5 6. keep one context warm > cache reads at $0.20 per Mtok are 20x cheaper than a fresh load 7. escalate the hard part > the toughest 1% goes to Opus 5.5 with 1M context and 66.4% on Terminal-Bench 4.0 8. score every chunk on every query > keep whole, summarize or drop. the context gets rebuilt each turn 9. gate the actual command > every bash call gets classified before it runs, inside your own code 10. judge 100% of runs > $3.50 a day for 10,000 traces, and it matched the human label on all 500 decisions the result: a while loop that paid a frontier model for every tiny call turns into a brain that spends a fraction of a cent to notice and pays properly only when it has to think the person who brings this into their team walks into the budget meeting with the AI bill cut and the output up the PDF maps the company brain. the loop side of it - how Jev takes a Claude bill from $765 to $3 a month - is in the article below ↓

Mr. Buzzoni

85,590 просмотров • 6 дней назад