正在加载视频...

视频加载失败

introducing Jevbox: an open-source document drive that organizes itself, answers questions, and cites everything (all powered by Jev). TLDR; - upload anything and Jev categorizes and files it into the folder tree - search is hierarchical: Jev picks the folder, then the document, then the section. No embeddings, no...

173,895 次观看 • 4 天前 •via X (Twitter)

31 条评论

Kushal Byatnal 的头像
Kushal Byatnal3 天前

open source repo here:

Michael 的头像
Michael3 天前

where can i try this?

Yechan Do 的头像
Yechan Do3 天前

Cool. Jev is solving retrieval problems

Kushal Byatnal 的头像
Kushal Byatnal3 天前

if you squint, everything is a classification problem

Saad 的头像
Saad3 天前

gh repo?

Kushal Byatnal 的头像
Kushal Byatnal3 天前

here you go:

brexton 的头像
brexton3 天前

Insane

Kushal Byatnal 的头像
Kushal Byatnal3 天前

@andrewlu0 is insane

Darren Chow 的头像
Darren Chow4 天前

@ameyaajoshi the use cases with extend + jev are crazy

Josh Kaplan 的头像
Josh Kaplan3 天前

cool

Fred Dogan 的头像
Fred Dogan3 天前

this is cool, organizing folders is a pet peeve of mine. I really don't enjoy it. Notion could use a similar approach honestly. Cool use case for Jev, I have been building a tool router for machine payments with Jev it's quite interesting. Will launch soon too.

Sael 的头像
Sael4 天前

no vector db is the real flex here, cant wait to stress test the hierarchy

Aniruddha Ganesh 的头像
Aniruddha Ganesh3 天前

@theo I recall you doing something about file organisation and files not needing to be at one exact path. This seems like the most natural evolution of that ideao.

Sahibzada Allahyar 的头像
Sahibzada Allahyar3 天前

I’d use GLiDE for the folder → document → section choices. It beats Jev on Decision Index retrieval, 60.9 vs 55.4, and can reason further when two paths look plausible. That’s worth testing on the same document library.

Jack 的头像
Jack3 天前

The filing is the easy half. What actually breaks document search in a real company is that the drive holds three versions of the same policy and nothing marks which one is current. A citation only helps if somebody decided that page was still true.

flare on 的头像
flare on3 天前

Looks interesting

Mildo 🥽 的头像
Mildo 🥽3 天前

Is this a different kind of vector database? What’s the benefit over graphrag?

kartik bhardwaj 的头像
kartik bhardwaj3 天前

Retrieval-time permission checks are the part most RAG stacks get wrong, cached chunks serve after access is revoked. The failure I would eval first is miscategorization, a misfiled doc goes invisible to hierarchical search. AI engineer working on agent harnesses and evals.

Farzan Ansari 的头像
Farzan Ansari3 天前

people here are doing crazy things!

Johannes Laslo 的头像
Johannes Laslo3 天前

what problem does this solve? Is this advantageous to a local wiki style knowledgebase with vector database?

Saurabh Gayali 的头像
Saurabh Gayali3 天前

Super slick UI, but needing paid SaaS endpoints for search & parsing hurts. Wish it had native fallbacks for fast, non-autoregressive System 1 models like Laya for scoring/ranking + local OCR for files. A completely free, offline Jevbox stack would be amazing.

Crio Songo 的头像
Crio Songo3 天前

No vector DB is such a refreshing design, I’ll go check out the repo later.

Zuhayr Khan 的头像
Zuhayr Khan3 天前

What size library have you tested this on so far? Curious how many Jev calls a question spanning several folders needs.

catman 的头像
catman3 天前

The useful detail is that Jevbox checks permissions at retrieval, so losing access to a source also removes access to answers built from it.

Auspex-Aerie 的头像
Auspex-Aerie3 天前

this is amazing Great work!

Andrew Luo 的头像
Andrew Luo3 天前

@typesafeai built using the hierarchical classification cookbook ✍️

David T Kramaley 的头像
David T Kramaley3 天前

hierarchical classification is the sweet spot for auto‑organizing docs

Nirant 的头像
Nirant3 天前

I built this for HTML internally _yesterday_

Sujal 的头像
Sujal4 天前

Does it ever reshuffle the folder tree after you've already learned where things live?

Jhon Dennis 的头像
Jhon Dennis3 天前

文档自己会归类 提问还能带出处 这活以前得专人干

Lior Strugach 的头像
Lior Strugach3 天前

Sounds expensive, how much does it cost to run it all?

相关视频

Jev has been exploding in popularity recently. If you already have access to the Jev API but aren't sure how to start experimenting with it, just copy this checklist: 1. jev-ultrafast Browser Use's fastest agent. Jev decides the next action and which element to click, and a language model is only called when text has to be typed. 2. typesafe-mario Jev plays Super Mario Bros. from structured emulator state, choosing every action from features pulled out of the game. 3. jev-plays-pokemon Reads Pokémon Red's game state as text, answers typed questions each turn, and lets plain code turn the answers into moves. 4. jev-drone A camera-only autonomous drone in MuJoCo with a Jev judgment model sitting in the control loop at 2.5 Hz. 5. robo-harness A real SO-101 robot arm workbench where Jev picks bounded joint steps from typed candidate actions under a spend budget. 6. fast-jev-compaction Claude Code plugin that replaces the compaction summary with Jev decisions, scoring every tool call for whether it is still needed. 7. jev-claude Routes Claude Code's own judgment calls through Jev: typed choices with probabilities at plan approval, on questions, and before risky commands. 8. is-malicious Supply-chain check before you run anything: Jev Noul checks over source and build files, returning the implicated files and lines. 9. sqlite-jev Jev inside SQL. Noul, Choice and Score judgments exposed as SQLite functions, with confidence on every row. 10. jevinci Paints images by having Jev predict every pixel's colour in parallel, with confidence deciding how wide each stroke is drawn. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

Hanako

29,894 次观看 • 11 天前

this is f*cking gold 20 GitHub repos with 500K+ combined stars that will level up your JEV workflow AGENTS > jev-ultrafast: Jev picks every click and DOM target, a small LLM only types > hermes-jev-skills: routing, memory, compaction and skill picks in one pack > typesafe-computer-use: OCR reads your Mac screen, Jev picks the next click MEMORY > fast-jev-compaction: scores every tool call keep, truncate or drop instead of summarizing > jevmem: project memory for Claude Code, Cursor and Codex, updated every turn > jev-second-brain: your Obsidian vault, Jev judges which notes duplicate, revise or contradict SAFETY > jev-guard: a gate before every tool call, Jev scores the risk, you set allow, ask or deny TOOLS > skills: the official TypeSafe skill for Claude Code and Codex > system-one-adapter-python: dry-run your questions on an ordinary LLM before you burn a Jev key > jev-mcp: claim checks, screening and ranking as MCP tools > typesafe-mcp: plug Jev into any MCP client > json-render: Vercel's generative UI, where Jev picks the components OPEN MODELS > SemIf-OpenJev: semantic ifs from frozen open models kev: Jev-like models on Qwen that run on your MacBook > laya-mlx: the Laya decision engine on Apple silicon > clm: an open System One model with Choice, Noul and Score > jevlike: train your own Jev-like model TRADING > jev-trader: one buy or sell decision per Monad block START HERE > awesome-jev: the biggest map of everything built on Jev > awesome-jev-by-typesafe: use cases, patterns and starter code bookmark it before your next build

NO1ennn

18,445 次观看 • 8 天前

Jev has been exploding across GitHub since launch, here's what people have already built with it if you have API access and don't know where to start, copy this: jev-trader - real trading bot placing live limit orders on Monad every 300ms block, judged by Jev alone. 1,911 stars jev-ultrafast - browser agent that picks every click itself, only calling a text model when it actually needs to type something. 16,758 stars jev-doom-agent - real Chocolate Doom compiled to WebAssembly, two engines running the same map, Jev picking the tactical macro every frame jev-t-rex-runner - the Chrome dinosaur game you've procrastinated with a hundred times, now played entirely by Jev picking jump, duck, or keep running typesafe-chess - Jev vs a real search engine, two games, colors swapped. the search won both, but overruled Jev's first instinct on roughly half the moves jev-drone - a simulated quadrotor clears a five-station obstacle course by camera alone, Jev judging the situation twice a second tax-doc-classifier - sorts real IRS tax forms with 100% strict accuracy across 261 forms, at roughly $0.001 a page killmyidea - describe your startup idea, Jev scores it from every angle, then hands back kill, fix, or ship jev-curate - streams Parquet and JSONL rows through typed judgments at 1,500+ rows a second, keeping only what clears the bar pg-jev - a PostgreSQL extension that lets you ask your own database tables plain-English questions and get a real answer back none of these ten generate a single word of text. every one of them returns a number against an answer someone already defined full setup below, then run the three-question test from the article before you build an eleventh

Ryven

276,722 次观看 • 13 天前

Jev has been exploding in popularity recently. If you already have access to the Jev API but aren’t sure how to start experimenting with it, just copy this checklist: 1. agent-desktop Desktop automation. Read the system's accessibility tree, judge which button, menu, or input field to click next. 2. typesafe-mario Have Jev play Super Mario. No screenshots—just read the structured state in the emulator's RAM, then decide to run, jump, or dodge. 3. jev-drone Use Jev to control a drone. The underlying flight control still handles stability and safety; Jev just does higher-level judgments like climbing, braking, and navigating obstacles. 4. OneVOneJev 1v1 FPS in the browser. Every decision tick, judge movement, view angle, aiming, firing, and jumping. 5. jev-trader High-frequency market making on Monad testnet. Jev judges the next buy or sell based on spreads and trade direction, with model latency around 81ms. 6. Prism Doesn't directly have Jev place orders. It judges states like toxic flow, market pressure, mean reversion, etc., then hands off to the original strategy. 7. neo4jev Stuff Jev into a knowledge graph. At each node, judge the most worthwhile edge to take next, then follow it all the way. 8. jev-curate Use Jev to screen training data. For JSONL / Parquet, first judge quality, relevance, and risk, then decide which ones go into the next training round. 9. Canny Prevents Coding Agents from stubbornly claiming they're done. Look at tool outputs, code diffs, and test results, then judge if the completion claim is reliable. 10. killmyidea Input a startup idea, and Jev scores it from multiple angles, finally giving you KILL, FIX, or SHIP. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

rody

346,644 次观看 • 14 天前

Jev has been blowing up lately. If you've got the Jev API but don't know how to play around with it yet, you can just copy this checklist. 1. jev-ultrafast A high-speed browser Agent built with Browser Use. Jev only judges "what to do, which element to click" at each step, and only calls the small model when typing is needed. Searching for a flight on Google Flights takes about 7 seconds. 2. fast-jev-compaction Context compression for Claude Code. Before each tool call, have Jev judge if there's anything still useful; delete the useless stuff, and keep the original text without rewriting it. 3. json-render Vercel Labs' generative UI framework. In experiments, Jev doesn't write JSON token by token; it just handles selecting components, properties, and layouts. 4. typesafe-mcp Best for people who just got the API. Plug Jev into Claude Code, Claude Desktop, Codex, and Pi, and do Choice / Score / Noul anytime. 5. jev-mcp Ready-made Agent judgment toolkit: fact-checking, content screening, semantic ranking, classification, and information extraction. 6. SemDecide Turn Jev into a command-line tool. Directly classify, score, and filter in the Shell—great for hooking up to crawlers, CI, and data pipelines. 7. jev-codex-router First have Jev judge how hard this round of programming tasks is, then decide the model tier, reasoning depth, and speed mode. 8. Winnow Context garbage collection for Claude Code. When Read / Bash / Grep spits out a ton of stuff, Jev first judges which parts are really relevant to the current task. 9. jev-review Before code review, run it through Jev first to pick out high-risk changes, then hand them off to a pricier big model or a human. Comes with a local dashboard. 10. Blink Use Jev as a code repository navigator. At each directory level, judge which files are most relevant to the current issue, then keep digging down. Copy these complete Jev blueprints - then read full Jev setup below ↓ ↓

rody

202,492 次观看 • 15 天前

this is pure f*cking treasure Engineer at TypeSafeAI just mapped 389 public Jev builds in one list: skills, MCP servers, SDKs, agents, benchmarks and guides ROUTING > jev-router sends every Claude Code task to the cheapest model that can do it > Switchboard picks the model and reasoning effort per task, then keeps it stable so the prompt cache survives > jev-oncall triages alerts: 418 ms p50, $0.04 per 1,000 alerts GUARDRAILS > jev-axi gates shell commands for Claude Code and Codex, 44/44 on its labeled set > hermes-jev-approvals: 8.7x faster approvals, 4.4x fewer prompts to the user > Sniff Test lints AI slop out of your writing at 182 ms a paragraph AGENTS > Jev Ultrafast: Jev picks every click, a small LLM only types > fast-jev-compaction scores your context instead of summarizing it > jev-browser-use reports 5 to 10x faster browser runs inside Codex SEARCH & RAG > jev-retrieval placed 2nd of 90 models on a reranking leaderboard > jevsearch: 83% Hit@1 against 41% for keyword search alone SDKs & MCP > clients for Swift, Go, Rust, Kotlin, Ruby, Elixir, .NET, Laravel and Spring > MCP servers for Claude Code, Cursor and Codex > Jev inside SQLite, DuckDB and Postgres, straight from SQL EVALS > pytest-jev checks LLM replies in 5.3 s where Claude took 27.1 s > a medical hallucination check at 92.9% accuracy, 204 ms, $0.03 per 1,000 OPEN MODELS > Laya answers in one ~35 ms forward pass > kev trains and runs on a MacBook GAMES, ROBOTS, TRADING > Jev plays Mario, Pokemon and chess > it drives a robot arm and a drone > it trades on Monad with 81 ms decisions 389 builds. one decision layer. go steal the ones you need

NO1ennn

25,798 次观看 • 9 天前

Another insane Jev use case! Jev is making it dramatically cheaper to evaluate what actually happened inside an agent run. And finally, someone open-sourced a self-improving memory layer that can put that signal to work across agent harnesses: - Claude Code - Codex - Cursor - OpenCode, and 20+ more Beacon by Asymptote Labs continuously captures your agent history across harnesses and uses Jev to identify which runs are actually worth learning from. It then turns the highest-signal workflows, corrections, and debugging patterns into reusable skills. GitHub repo: (don’t forget to star it ⭐ ) Beacon preserves the complete session history. But preserving a run and learning from it are two different things. Most coding-agent sessions contain routine exploration, failed commands, and fixes that only apply to one task. The trace can remain available for inspection without turning every detail into guidance for future agents. Jev scores each run for evidence, reuse potential, and human correction signals. An application policy then decides whether to promote, review, or discard it. The recording shows this in action. Claude receives a coding task, modifies the implementation, and runs the tests. I then provide an edge-case correction, so Claude updates the code and adds regression coverage. Beacon automatically captures the complete session. Jev evaluates whether the correction contains a reusable engineering lesson. Once approved, that lesson becomes available to other coding agents working on the project. Since it works across harnesses: - Claude Code sessions can teach Codex. - Cursor debugging can improve OpenCode. So a problem solved by one agent should not need to be learned from scratch by another. If you want to dive deeper into Jev, I also wrote a hands-on guide to building this Jev-style decision path with open models, entirely locally. Read it below.

Avi Chawla

291,608 次观看 • 15 天前

this is unreal f*cking gold for Jev builders 20 repos people are building on Jev right now. browser agents, context tools, trading bots, even a drone 1. JEV-Ultrafast - a browser agent built for speed ↳ 2. Fast-JEV-Compaction - squeezes your context down ↳ 3. JSON-Render - UI generated on the fly ↳ 4. Typesafe-MCP - plugs Jev into any client ↳ 5. JEV-MCP - a toolkit for judgment calls ↳ 6. Semdecide - a classifier right in your terminal ↳ 7. JEV-Codex-Router - sends every task to the model that fits it ↳ 8. Winnow - clears the junk out of your context ↳ 9. JEV-Review - sorts code reviews by what needs eyes first ↳ 10. Blink - finds your way around any repo ↳ 11. Agent-Desktop - runs your desktop for you ↳ 12. Typesafe-Mario - an agent playing Super Mario ↳ 13. JEV-Drone - flies a drone ↳ 14. OneVOneJev - a shooter in your browser ↳ 15. JEV-Trader - high-frequency market making ↳ 16. Prism - spots liquidity signals ↳ 17. Neo4Jev - walks a knowledge graph ↳ 18. JEV-Curate - screens training data ↳ 19. Canny - confirms a task is really finished ↳ 20. KillMyIdea - scores a startup idea before you sink time into it ↳ start where your work is: > coding -> JEV-Review, Blink, Canny, JEV-Codex-Router > context -> Fast-JEV-Compaction, Winnow > automation -> JEV-Ultrafast, Agent-Desktop > clients and tools -> Typesafe-MCP, JEV-MCP, Semdecide > UI -> JSON-Render > trading -> JEV-Trader, Prism > data -> Neo4Jev, JEV-Curate > founders -> KillMyIdea > for fun -> Typesafe-Mario, OneVOneJev, JEV-Drone pick one, build on it this week, and you'll be the person on your team who actually knows Jev engineering when it gets asked for

Mr. Buzzoni

11,667 次观看 • 6 天前