Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Stop using your most expensive model for every decision. In our Minesweeper experiment (open-source code below), JEV crushed every other model we tested on speed and cost. JEV: 8s / ~$0.003 Fable 5.1: 100s / ~$2.41 Astra: 128s / ~$8.75 Grok 4.6: 158s / ~$0.55 Where would you put...

171,498 Aufrufe • vor 19 Tagen •via X (Twitter)

14 Kommentare

Profilbild von Buddy Works
Buddy Worksvor 19 Tagen

The repo is open source, so you can run your own benchmark and check the results for yourself. Inside, you'll find „Run in sandbox”, so you can play with it in a minute, no digging through config. Run it and share your results 👇

Profilbild von Bartosz Wróbel
Bartosz Wróbelvor 19 Tagen

Fun little test. In my experience, Jev is fast but less accurate than SOTA LLMs on the tasks I tried. It also requires framing problems as structured decisions with predefined outputs. It looks promising for fast data labeling or powering dynamic dashboards.

Profilbild von Raphael Sztwiorok
Raphael Sztwiorokvor 19 Tagen

I’d use it to flag PRs that need an extra security review before we hit merge.

Profilbild von Raphael Sztwiorok
Raphael Sztwiorokvor 19 Tagen

In short, how does it work: the model sees the board + a frontier of unknown cells with "mines still needed" per number. Rule: flag only proven mines, reveal only proven-safe cells, guess lowest-probability only when stuck. Jev agent: probabilities, flag ≥0.85, reveal ≤0.07 More info in the repo:

Profilbild von Lukasz Czulak
Lukasz Czulakvor 19 Tagen

I'd use JEV for ticket classification like urgency, topic, or team. It's fast, consistent, and much cheaper than wasting high-end model tokens on thousands of small daily decisions

Profilbild von Buddy Works
Buddy Worksvor 18 Tagen

That’s a solid use case, especially at scale.

Profilbild von Darek Sztwiorok
Darek Sztwiorokvor 19 Tagen

Almost 3000x cheaper and 16x faster than Astra? That’s not a tweak, that’s a total game changer 🚀

Profilbild von Buddy Works
Buddy Worksvor 18 Tagen

Yeah, the numbers are pretty wild.

Profilbild von BartShoot
BartShootvor 19 Tagen

Wonder if it's fast enough to play Tetris?

Profilbild von Paweł Kapała
Paweł Kapałavor 19 Tagen

Feels like a great fit for a gating layer, deciding which tasks actually need deeper reasoning from more expensive models, or routing each task to the model best suited for it.

Profilbild von Buddy Works
Buddy Worksvor 19 Tagen

That’s a great use case!

Profilbild von Brjan | AI Builder
Brjan | AI Buildervor 18 Tagen

how did you measure the speed and cost for each model's performance?

Profilbild von Krzysztof Stryczek
Krzysztof Stryczekvor 18 Tagen

Regression triage. Sorting 100 failed tests into flake / test fix / infra / real bug is a decision, not an essay. JEV does the first pass, the expensive model only sees what's left.

Profilbild von Buddy Works
Buddy Worksvor 18 Tagen

Yep, makes a lot of sense.

Ähnliche Videos

this is f*cking gold 20 GitHub repos with 500K+ combined stars that will level up your JEV workflow AGENTS > jev-ultrafast: Jev picks every click and DOM target, a small LLM only types > hermes-jev-skills: routing, memory, compaction and skill picks in one pack > typesafe-computer-use: OCR reads your Mac screen, Jev picks the next click MEMORY > fast-jev-compaction: scores every tool call keep, truncate or drop instead of summarizing > jevmem: project memory for Claude Code, Cursor and Codex, updated every turn > jev-second-brain: your Obsidian vault, Jev judges which notes duplicate, revise or contradict SAFETY > jev-guard: a gate before every tool call, Jev scores the risk, you set allow, ask or deny TOOLS > skills: the official TypeSafe skill for Claude Code and Codex > system-one-adapter-python: dry-run your questions on an ordinary LLM before you burn a Jev key > jev-mcp: claim checks, screening and ranking as MCP tools > typesafe-mcp: plug Jev into any MCP client > json-render: Vercel's generative UI, where Jev picks the components OPEN MODELS > SemIf-OpenJev: semantic ifs from frozen open models kev: Jev-like models on Qwen that run on your MacBook > laya-mlx: the Laya decision engine on Apple silicon > clm: an open System One model with Choice, Noul and Score > jevlike: train your own Jev-like model TRADING > jev-trader: one buy or sell decision per Monad block START HERE > awesome-jev: the biggest map of everything built on Jev > awesome-jev-by-typesafe: use cases, patterns and starter code bookmark it before your next build

NO1ennn

18,445 Aufrufe • vor 12 Tagen

this is unreal f*cking gold for Jev builders 20 repos people are building on Jev right now. browser agents, context tools, trading bots, even a drone 1. JEV-Ultrafast - a browser agent built for speed ↳ 2. Fast-JEV-Compaction - squeezes your context down ↳ 3. JSON-Render - UI generated on the fly ↳ 4. Typesafe-MCP - plugs Jev into any client ↳ 5. JEV-MCP - a toolkit for judgment calls ↳ 6. Semdecide - a classifier right in your terminal ↳ 7. JEV-Codex-Router - sends every task to the model that fits it ↳ 8. Winnow - clears the junk out of your context ↳ 9. JEV-Review - sorts code reviews by what needs eyes first ↳ 10. Blink - finds your way around any repo ↳ 11. Agent-Desktop - runs your desktop for you ↳ 12. Typesafe-Mario - an agent playing Super Mario ↳ 13. JEV-Drone - flies a drone ↳ 14. OneVOneJev - a shooter in your browser ↳ 15. JEV-Trader - high-frequency market making ↳ 16. Prism - spots liquidity signals ↳ 17. Neo4Jev - walks a knowledge graph ↳ 18. JEV-Curate - screens training data ↳ 19. Canny - confirms a task is really finished ↳ 20. KillMyIdea - scores a startup idea before you sink time into it ↳ start where your work is: > coding -> JEV-Review, Blink, Canny, JEV-Codex-Router > context -> Fast-JEV-Compaction, Winnow > automation -> JEV-Ultrafast, Agent-Desktop > clients and tools -> Typesafe-MCP, JEV-MCP, Semdecide > UI -> JSON-Render > trading -> JEV-Trader, Prism > data -> Neo4Jev, JEV-Curate > founders -> KillMyIdea > for fun -> Typesafe-Mario, OneVOneJev, JEV-Drone pick one, build on it this week, and you'll be the person on your team who actually knows Jev engineering when it gets asked for

Mr. Buzzoni

11,667 Aufrufe • vor 10 Tagen

Jev model has been causing a sensation on GitHub ever since its release, and I am going to tell you about the best repositories that nobody is talking about How does this work: browser-use/browser-use⁠ - 32,500+ stars The core web automation library that houses the Jev integration ecosystem (including ⁠jev-ultrafast⁠). It enables AI decision models to directly control web browser actions, click DOM elements, and automate complex workflows without heavy LLM latency langchain-ai/langchain⁠ - 98,500+ stars Integrates Jev as an ultra-fast model router, guardrail evaluator, and safety firewall. Jev evaluates task risk, intent, and complexity before passing execution to heavy LLMs LiteLLM/litellm⁠ - 59,200+ stars A lightweight proxy and routing library for LLMs. It features native Jev decision-model routing to classify incoming prompts and instantly direct requests to the cheapest or fastest model endpoint dify-ai/dify⁠ - 58,000+ stars An open-source LLM application development platform that incorporates Jev-style decision nodes for instant conditional branching and typed classification in multi-agent workflows vllm-project/vllm⁠ - 36,000+ stars A high-throughput LLM serving engine widely used for self-hosting custom decision-model checkpoints (such as Qwen3.5-based Jev fine-tunes) with ultra-low latency logits evaluation QwenLM/Qwen⁠ - 18,500+ stars The open base model family (Qwen3.5/3.8) used to train open-source Jev-like decision models (like Kev and RSI-Jev) powering typed probabilistic choices directly from custom model weights 🔗 You will definitely find what you need. I recommend adding this to your bookmarks

Bober_smart

17,773 Aufrufe • vor 4 Tagen

Jev has been exploding in popularity recently. If you already have access to the Jev API but aren't sure how to start experimenting with it, just copy this checklist: 1. jev-ultrafast Browser Use's fastest agent. Jev decides the next action and which element to click, and a language model is only called when text has to be typed. 2. typesafe-mario Jev plays Super Mario Bros. from structured emulator state, choosing every action from features pulled out of the game. 3. jev-plays-pokemon Reads Pokémon Red's game state as text, answers typed questions each turn, and lets plain code turn the answers into moves. 4. jev-drone A camera-only autonomous drone in MuJoCo with a Jev judgment model sitting in the control loop at 2.5 Hz. 5. robo-harness A real SO-101 robot arm workbench where Jev picks bounded joint steps from typed candidate actions under a spend budget. 6. fast-jev-compaction Claude Code plugin that replaces the compaction summary with Jev decisions, scoring every tool call for whether it is still needed. 7. jev-claude Routes Claude Code's own judgment calls through Jev: typed choices with probabilities at plan approval, on questions, and before risky commands. 8. is-malicious Supply-chain check before you run anything: Jev Noul checks over source and build files, returning the implicated files and lines. 9. sqlite-jev Jev inside SQL. Noul, Choice and Score judgments exposed as SQLite functions, with confidence on every row. 10. jevinci Paints images by having Jev predict every pixel's colour in parallel, with confidence deciding how wide each stroke is drawn. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

Hanako

29,894 Aufrufe • vor 15 Tagen

this is pure f*cking treasure Engineer at TypeSafeAI just mapped 389 public Jev builds in one list: skills, MCP servers, SDKs, agents, benchmarks and guides ROUTING > jev-router sends every Claude Code task to the cheapest model that can do it > Switchboard picks the model and reasoning effort per task, then keeps it stable so the prompt cache survives > jev-oncall triages alerts: 418 ms p50, $0.04 per 1,000 alerts GUARDRAILS > jev-axi gates shell commands for Claude Code and Codex, 44/44 on its labeled set > hermes-jev-approvals: 8.7x faster approvals, 4.4x fewer prompts to the user > Sniff Test lints AI slop out of your writing at 182 ms a paragraph AGENTS > Jev Ultrafast: Jev picks every click, a small LLM only types > fast-jev-compaction scores your context instead of summarizing it > jev-browser-use reports 5 to 10x faster browser runs inside Codex SEARCH & RAG > jev-retrieval placed 2nd of 90 models on a reranking leaderboard > jevsearch: 83% Hit@1 against 41% for keyword search alone SDKs & MCP > clients for Swift, Go, Rust, Kotlin, Ruby, Elixir, .NET, Laravel and Spring > MCP servers for Claude Code, Cursor and Codex > Jev inside SQLite, DuckDB and Postgres, straight from SQL EVALS > pytest-jev checks LLM replies in 5.3 s where Claude took 27.1 s > a medical hallucination check at 92.9% accuracy, 204 ms, $0.03 per 1,000 OPEN MODELS > Laya answers in one ~35 ms forward pass > kev trains and runs on a MacBook GAMES, ROBOTS, TRADING > Jev plays Mario, Pokemon and chess > it drives a robot arm and a drone > it trades on Monad with 81 ms decisions 389 builds. one decision layer. go steal the ones you need

NO1ennn

25,798 Aufrufe • vor 12 Tagen

Jev has been blowing up lately. If you've got the Jev API but don't know how to play around with it yet, you can just copy this checklist. 1. jev-ultrafast A high-speed browser Agent built with Browser Use. Jev only judges "what to do, which element to click" at each step, and only calls the small model when typing is needed. Searching for a flight on Google Flights takes about 7 seconds. 2. fast-jev-compaction Context compression for Claude Code. Before each tool call, have Jev judge if there's anything still useful; delete the useless stuff, and keep the original text without rewriting it. 3. json-render Vercel Labs' generative UI framework. In experiments, Jev doesn't write JSON token by token; it just handles selecting components, properties, and layouts. 4. typesafe-mcp Best for people who just got the API. Plug Jev into Claude Code, Claude Desktop, Codex, and Pi, and do Choice / Score / Noul anytime. 5. jev-mcp Ready-made Agent judgment toolkit: fact-checking, content screening, semantic ranking, classification, and information extraction. 6. SemDecide Turn Jev into a command-line tool. Directly classify, score, and filter in the Shell—great for hooking up to crawlers, CI, and data pipelines. 7. jev-codex-router First have Jev judge how hard this round of programming tasks is, then decide the model tier, reasoning depth, and speed mode. 8. Winnow Context garbage collection for Claude Code. When Read / Bash / Grep spits out a ton of stuff, Jev first judges which parts are really relevant to the current task. 9. jev-review Before code review, run it through Jev first to pick out high-risk changes, then hand them off to a pricier big model or a human. Comes with a local dashboard. 10. Blink Use Jev as a code repository navigator. At each directory level, judge which files are most relevant to the current issue, then keep digging down. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

rody

202,492 Aufrufe • vor 19 Tagen