Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

introducing Jevbox: an open-source document drive that organizes itself, answers questions, and cites everything (all powered by Jev). TLDR; - upload anything and Jev categorizes and files it into the folder tree - search is hierarchical: Jev picks the folder, then the document, then the section. No embeddings, no...

173,895 Aufrufe • vor 4 Tagen •via X (Twitter)

31 Kommentare

Profilbild von Kushal Byatnal
Kushal Byatnalvor 3 Tagen

open source repo here:

Profilbild von Michael
Michaelvor 3 Tagen

where can i try this?

Profilbild von Yechan Do
Yechan Dovor 3 Tagen

Cool. Jev is solving retrieval problems

Profilbild von Kushal Byatnal
Kushal Byatnalvor 3 Tagen

if you squint, everything is a classification problem

Profilbild von Saad
Saadvor 3 Tagen

gh repo?

Profilbild von Kushal Byatnal
Kushal Byatnalvor 3 Tagen

here you go:

Profilbild von brexton
brextonvor 3 Tagen

Insane

Profilbild von Kushal Byatnal
Kushal Byatnalvor 3 Tagen

@andrewlu0 is insane

Profilbild von Darren Chow
Darren Chowvor 4 Tagen

@ameyaajoshi the use cases with extend + jev are crazy

Profilbild von Josh Kaplan
Josh Kaplanvor 3 Tagen

cool

Profilbild von Fred Dogan
Fred Doganvor 3 Tagen

this is cool, organizing folders is a pet peeve of mine. I really don't enjoy it. Notion could use a similar approach honestly. Cool use case for Jev, I have been building a tool router for machine payments with Jev it's quite interesting. Will launch soon too.

Profilbild von Sael
Saelvor 4 Tagen

no vector db is the real flex here, cant wait to stress test the hierarchy

Profilbild von Aniruddha Ganesh
Aniruddha Ganeshvor 3 Tagen

@theo I recall you doing something about file organisation and files not needing to be at one exact path. This seems like the most natural evolution of that ideao.

Profilbild von Sahibzada Allahyar
Sahibzada Allahyarvor 3 Tagen

I’d use GLiDE for the folder → document → section choices. It beats Jev on Decision Index retrieval, 60.9 vs 55.4, and can reason further when two paths look plausible. That’s worth testing on the same document library.

Profilbild von Jack
Jackvor 3 Tagen

The filing is the easy half. What actually breaks document search in a real company is that the drive holds three versions of the same policy and nothing marks which one is current. A citation only helps if somebody decided that page was still true.

Profilbild von flare on
flare onvor 3 Tagen

Looks interesting

Profilbild von Mildo 🥽
Mildo 🥽vor 3 Tagen

Is this a different kind of vector database? What’s the benefit over graphrag?

Profilbild von kartik bhardwaj
kartik bhardwajvor 3 Tagen

Retrieval-time permission checks are the part most RAG stacks get wrong, cached chunks serve after access is revoked. The failure I would eval first is miscategorization, a misfiled doc goes invisible to hierarchical search. AI engineer working on agent harnesses and evals.

Profilbild von Farzan Ansari
Farzan Ansarivor 3 Tagen

people here are doing crazy things!

Profilbild von Johannes Laslo
Johannes Laslovor 3 Tagen

what problem does this solve? Is this advantageous to a local wiki style knowledgebase with vector database?

Profilbild von Saurabh Gayali
Saurabh Gayalivor 3 Tagen

Super slick UI, but needing paid SaaS endpoints for search & parsing hurts. Wish it had native fallbacks for fast, non-autoregressive System 1 models like Laya for scoring/ranking + local OCR for files. A completely free, offline Jevbox stack would be amazing.

Profilbild von Crio Songo
Crio Songovor 3 Tagen

No vector DB is such a refreshing design, I’ll go check out the repo later.

Profilbild von Zuhayr Khan
Zuhayr Khanvor 3 Tagen

What size library have you tested this on so far? Curious how many Jev calls a question spanning several folders needs.

Profilbild von catman
catmanvor 3 Tagen

The useful detail is that Jevbox checks permissions at retrieval, so losing access to a source also removes access to answers built from it.

Profilbild von Auspex-Aerie
Auspex-Aerievor 3 Tagen

this is amazing Great work!

Profilbild von Andrew Luo
Andrew Luovor 3 Tagen

@typesafeai built using the hierarchical classification cookbook ✍️

Profilbild von David T Kramaley
David T Kramaleyvor 3 Tagen

hierarchical classification is the sweet spot for auto‑organizing docs

Profilbild von Nirant
Nirantvor 3 Tagen

I built this for HTML internally _yesterday_

Profilbild von Sujal
Sujalvor 4 Tagen

Does it ever reshuffle the folder tree after you've already learned where things live?

Profilbild von Jhon Dennis
Jhon Dennisvor 3 Tagen

文档自己会归类 提问还能带出处 这活以前得专人干

Profilbild von Lior Strugach
Lior Strugachvor 3 Tagen

Sounds expensive, how much does it cost to run it all?

Ähnliche Videos

Jev has been exploding in popularity recently. If you already have access to the Jev API but aren't sure how to start experimenting with it, just copy this checklist: 1. jev-ultrafast Browser Use's fastest agent. Jev decides the next action and which element to click, and a language model is only called when text has to be typed. 2. typesafe-mario Jev plays Super Mario Bros. from structured emulator state, choosing every action from features pulled out of the game. 3. jev-plays-pokemon Reads Pokémon Red's game state as text, answers typed questions each turn, and lets plain code turn the answers into moves. 4. jev-drone A camera-only autonomous drone in MuJoCo with a Jev judgment model sitting in the control loop at 2.5 Hz. 5. robo-harness A real SO-101 robot arm workbench where Jev picks bounded joint steps from typed candidate actions under a spend budget. 6. fast-jev-compaction Claude Code plugin that replaces the compaction summary with Jev decisions, scoring every tool call for whether it is still needed. 7. jev-claude Routes Claude Code's own judgment calls through Jev: typed choices with probabilities at plan approval, on questions, and before risky commands. 8. is-malicious Supply-chain check before you run anything: Jev Noul checks over source and build files, returning the implicated files and lines. 9. sqlite-jev Jev inside SQL. Noul, Choice and Score judgments exposed as SQLite functions, with confidence on every row. 10. jevinci Paints images by having Jev predict every pixel's colour in parallel, with confidence deciding how wide each stroke is drawn. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

Hanako

29,894 Aufrufe • vor 11 Tagen

this is f*cking gold 20 GitHub repos with 500K+ combined stars that will level up your JEV workflow AGENTS > jev-ultrafast: Jev picks every click and DOM target, a small LLM only types > hermes-jev-skills: routing, memory, compaction and skill picks in one pack > typesafe-computer-use: OCR reads your Mac screen, Jev picks the next click MEMORY > fast-jev-compaction: scores every tool call keep, truncate or drop instead of summarizing > jevmem: project memory for Claude Code, Cursor and Codex, updated every turn > jev-second-brain: your Obsidian vault, Jev judges which notes duplicate, revise or contradict SAFETY > jev-guard: a gate before every tool call, Jev scores the risk, you set allow, ask or deny TOOLS > skills: the official TypeSafe skill for Claude Code and Codex > system-one-adapter-python: dry-run your questions on an ordinary LLM before you burn a Jev key > jev-mcp: claim checks, screening and ranking as MCP tools > typesafe-mcp: plug Jev into any MCP client > json-render: Vercel's generative UI, where Jev picks the components OPEN MODELS > SemIf-OpenJev: semantic ifs from frozen open models kev: Jev-like models on Qwen that run on your MacBook > laya-mlx: the Laya decision engine on Apple silicon > clm: an open System One model with Choice, Noul and Score > jevlike: train your own Jev-like model TRADING > jev-trader: one buy or sell decision per Monad block START HERE > awesome-jev: the biggest map of everything built on Jev > awesome-jev-by-typesafe: use cases, patterns and starter code bookmark it before your next build

NO1ennn

18,445 Aufrufe • vor 8 Tagen

Jev has been exploding across GitHub since launch, here's what people have already built with it if you have API access and don't know where to start, copy this: jev-trader - real trading bot placing live limit orders on Monad every 300ms block, judged by Jev alone. 1,911 stars jev-ultrafast - browser agent that picks every click itself, only calling a text model when it actually needs to type something. 16,758 stars jev-doom-agent - real Chocolate Doom compiled to WebAssembly, two engines running the same map, Jev picking the tactical macro every frame jev-t-rex-runner - the Chrome dinosaur game you've procrastinated with a hundred times, now played entirely by Jev picking jump, duck, or keep running typesafe-chess - Jev vs a real search engine, two games, colors swapped. the search won both, but overruled Jev's first instinct on roughly half the moves jev-drone - a simulated quadrotor clears a five-station obstacle course by camera alone, Jev judging the situation twice a second tax-doc-classifier - sorts real IRS tax forms with 100% strict accuracy across 261 forms, at roughly $0.001 a page killmyidea - describe your startup idea, Jev scores it from every angle, then hands back kill, fix, or ship jev-curate - streams Parquet and JSONL rows through typed judgments at 1,500+ rows a second, keeping only what clears the bar pg-jev - a PostgreSQL extension that lets you ask your own database tables plain-English questions and get a real answer back none of these ten generate a single word of text. every one of them returns a number against an answer someone already defined full setup below, then run the three-question test from the article before you build an eleventh

Ryven

276,722 Aufrufe • vor 13 Tagen

Jev has been exploding in popularity recently. If you already have access to the Jev API but aren’t sure how to start experimenting with it, just copy this checklist: 1. agent-desktop Desktop automation. Read the system's accessibility tree, judge which button, menu, or input field to click next. 2. typesafe-mario Have Jev play Super Mario. No screenshots—just read the structured state in the emulator's RAM, then decide to run, jump, or dodge. 3. jev-drone Use Jev to control a drone. The underlying flight control still handles stability and safety; Jev just does higher-level judgments like climbing, braking, and navigating obstacles. 4. OneVOneJev 1v1 FPS in the browser. Every decision tick, judge movement, view angle, aiming, firing, and jumping. 5. jev-trader High-frequency market making on Monad testnet. Jev judges the next buy or sell based on spreads and trade direction, with model latency around 81ms. 6. Prism Doesn't directly have Jev place orders. It judges states like toxic flow, market pressure, mean reversion, etc., then hands off to the original strategy. 7. neo4jev Stuff Jev into a knowledge graph. At each node, judge the most worthwhile edge to take next, then follow it all the way. 8. jev-curate Use Jev to screen training data. For JSONL / Parquet, first judge quality, relevance, and risk, then decide which ones go into the next training round. 9. Canny Prevents Coding Agents from stubbornly claiming they're done. Look at tool outputs, code diffs, and test results, then judge if the completion claim is reliable. 10. killmyidea Input a startup idea, and Jev scores it from multiple angles, finally giving you KILL, FIX, or SHIP. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

rody

346,644 Aufrufe • vor 14 Tagen

Jev has been blowing up lately. If you've got the Jev API but don't know how to play around with it yet, you can just copy this checklist. 1. jev-ultrafast A high-speed browser Agent built with Browser Use. Jev only judges "what to do, which element to click" at each step, and only calls the small model when typing is needed. Searching for a flight on Google Flights takes about 7 seconds. 2. fast-jev-compaction Context compression for Claude Code. Before each tool call, have Jev judge if there's anything still useful; delete the useless stuff, and keep the original text without rewriting it. 3. json-render Vercel Labs' generative UI framework. In experiments, Jev doesn't write JSON token by token; it just handles selecting components, properties, and layouts. 4. typesafe-mcp Best for people who just got the API. Plug Jev into Claude Code, Claude Desktop, Codex, and Pi, and do Choice / Score / Noul anytime. 5. jev-mcp Ready-made Agent judgment toolkit: fact-checking, content screening, semantic ranking, classification, and information extraction. 6. SemDecide Turn Jev into a command-line tool. Directly classify, score, and filter in the Shell—great for hooking up to crawlers, CI, and data pipelines. 7. jev-codex-router First have Jev judge how hard this round of programming tasks is, then decide the model tier, reasoning depth, and speed mode. 8. Winnow Context garbage collection for Claude Code. When Read / Bash / Grep spits out a ton of stuff, Jev first judges which parts are really relevant to the current task. 9. jev-review Before code review, run it through Jev first to pick out high-risk changes, then hand them off to a pricier big model or a human. Comes with a local dashboard. 10. Blink Use Jev as a code repository navigator. At each directory level, judge which files are most relevant to the current issue, then keep digging down. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

rody

202,492 Aufrufe • vor 15 Tagen

this is pure f*cking treasure Engineer at TypeSafeAI just mapped 389 public Jev builds in one list: skills, MCP servers, SDKs, agents, benchmarks and guides ROUTING > jev-router sends every Claude Code task to the cheapest model that can do it > Switchboard picks the model and reasoning effort per task, then keeps it stable so the prompt cache survives > jev-oncall triages alerts: 418 ms p50, $0.04 per 1,000 alerts GUARDRAILS > jev-axi gates shell commands for Claude Code and Codex, 44/44 on its labeled set > hermes-jev-approvals: 8.7x faster approvals, 4.4x fewer prompts to the user > Sniff Test lints AI slop out of your writing at 182 ms a paragraph AGENTS > Jev Ultrafast: Jev picks every click, a small LLM only types > fast-jev-compaction scores your context instead of summarizing it > jev-browser-use reports 5 to 10x faster browser runs inside Codex SEARCH & RAG > jev-retrieval placed 2nd of 90 models on a reranking leaderboard > jevsearch: 83% Hit@1 against 41% for keyword search alone SDKs & MCP > clients for Swift, Go, Rust, Kotlin, Ruby, Elixir, .NET, Laravel and Spring > MCP servers for Claude Code, Cursor and Codex > Jev inside SQLite, DuckDB and Postgres, straight from SQL EVALS > pytest-jev checks LLM replies in 5.3 s where Claude took 27.1 s > a medical hallucination check at 92.9% accuracy, 204 ms, $0.03 per 1,000 OPEN MODELS > Laya answers in one ~35 ms forward pass > kev trains and runs on a MacBook GAMES, ROBOTS, TRADING > Jev plays Mario, Pokemon and chess > it drives a robot arm and a drone > it trades on Monad with 81 ms decisions 389 builds. one decision layer. go steal the ones you need

NO1ennn

25,798 Aufrufe • vor 9 Tagen

Another insane Jev use case! Jev is making it dramatically cheaper to evaluate what actually happened inside an agent run. And finally, someone open-sourced a self-improving memory layer that can put that signal to work across agent harnesses: - Claude Code - Codex - Cursor - OpenCode, and 20+ more Beacon by Asymptote Labs continuously captures your agent history across harnesses and uses Jev to identify which runs are actually worth learning from. It then turns the highest-signal workflows, corrections, and debugging patterns into reusable skills. GitHub repo: (don’t forget to star it ⭐ ) Beacon preserves the complete session history. But preserving a run and learning from it are two different things. Most coding-agent sessions contain routine exploration, failed commands, and fixes that only apply to one task. The trace can remain available for inspection without turning every detail into guidance for future agents. Jev scores each run for evidence, reuse potential, and human correction signals. An application policy then decides whether to promote, review, or discard it. The recording shows this in action. Claude receives a coding task, modifies the implementation, and runs the tests. I then provide an edge-case correction, so Claude updates the code and adds regression coverage. Beacon automatically captures the complete session. Jev evaluates whether the correction contains a reusable engineering lesson. Once approved, that lesson becomes available to other coding agents working on the project. Since it works across harnesses: - Claude Code sessions can teach Codex. - Cursor debugging can improve OpenCode. So a problem solved by one agent should not need to be learned from scratch by another. If you want to dive deeper into Jev, I also wrote a hands-on guide to building this Jev-style decision path with open models, entirely locally. Read it below.

Avi Chawla

291,608 Aufrufe • vor 15 Tagen

this is unreal f*cking gold for Jev builders 20 repos people are building on Jev right now. browser agents, context tools, trading bots, even a drone 1. JEV-Ultrafast - a browser agent built for speed ↳ 2. Fast-JEV-Compaction - squeezes your context down ↳ 3. JSON-Render - UI generated on the fly ↳ 4. Typesafe-MCP - plugs Jev into any client ↳ 5. JEV-MCP - a toolkit for judgment calls ↳ 6. Semdecide - a classifier right in your terminal ↳ 7. JEV-Codex-Router - sends every task to the model that fits it ↳ 8. Winnow - clears the junk out of your context ↳ 9. JEV-Review - sorts code reviews by what needs eyes first ↳ 10. Blink - finds your way around any repo ↳ 11. Agent-Desktop - runs your desktop for you ↳ 12. Typesafe-Mario - an agent playing Super Mario ↳ 13. JEV-Drone - flies a drone ↳ 14. OneVOneJev - a shooter in your browser ↳ 15. JEV-Trader - high-frequency market making ↳ 16. Prism - spots liquidity signals ↳ 17. Neo4Jev - walks a knowledge graph ↳ 18. JEV-Curate - screens training data ↳ 19. Canny - confirms a task is really finished ↳ 20. KillMyIdea - scores a startup idea before you sink time into it ↳ start where your work is: > coding -> JEV-Review, Blink, Canny, JEV-Codex-Router > context -> Fast-JEV-Compaction, Winnow > automation -> JEV-Ultrafast, Agent-Desktop > clients and tools -> Typesafe-MCP, JEV-MCP, Semdecide > UI -> JSON-Render > trading -> JEV-Trader, Prism > data -> Neo4Jev, JEV-Curate > founders -> KillMyIdea > for fun -> Typesafe-Mario, OneVOneJev, JEV-Drone pick one, build on it this week, and you'll be the person on your team who actually knows Jev engineering when it gets asked for

Mr. Buzzoni

11,667 Aufrufe • vor 6 Tagen