Jev report · radar

Jev radar

Curated twice a day from three sources: TypeSafe AI's #show-and-tell Discord channel (read in full from 2026-07-22), GitHub search, and X posts found through web search. Deliberately wide; use the search and filters.

Jev primitives: Choice (pick one), Noul (yes/no), Score (0 to 1). No temperature. No open-value extraction.

1682 entries1131 distinct ideas across 1486 tools and demoslast update 2026-10-11refreshed twice daily
Score professional usefulness, 1 to 5
  • 5Production-proven with numbers; adoptable today
  • 4Solid tool numbers thin
  • 3Useful but niche or unmeasured
  • 2Prototype
  • 1Demo only
Start here · updated 2026-10-08

Picked from every entry below by reading each candidate's repository: license, tests, last push, what the code does, and what it sends off the machine. Re-judged weekly against the week's new entries.

Worth a trial 8

  1. Sniff Test pre-commit, CI, Claude Code skill

    Prose linter. Regex rules run locally; judgment rules ask Jev. Custom rules load from .snifftest.yaml.

    Sends Nothing for regex rules or --dry-run. Judgment rules ask before sending text.Checked MIT, tests, on npm. 182 ms and $0.013 per 100 paragraphs (author).

  2. jev-belay Claude Code Stop hook

    Claude Code Stop hook. Blocks a "done" when files changed and no test, build or lint passed afterwards; one Jev call, only on those stops.

    Sends On flagged stops only: the turn's prompt and closing message, secret-redacted.Checked MIT, offline test suite. AUROC 0.976 on 100 model-labelled stops (author).Note CI runs on Linux only, and the test-run detector reads Bash tool calls, not PowerShell. Has a shadow mode.

  3. jev-align any Jev judge with a labelled set

    Rewrites a Jev question's wording against your labels, asking you to label only the uncertain rows each round.

    Sends The evaluated rows to TypeSafe; reflection prompts to a second model.Checked Apache-2.0, tests, 286 stars.

  4. pytest-jev pytest suites that test LLM output

    pytest fixture that asserts what model output says (holds, lacks, choice, score), all claims in one Jev call.

    Sends The text under test and the context passed with it.Checked MIT, tests. 110x cheaper than Sonnet 5 as the judge (author).

  5. Sando Claude Code PostToolUse hook

    Trims large tool output in Claude Code: redacts secrets, caps per content type, keeps error lines and test totals, stores the full output on disk.

    Sends Nothing. No model calls; telemetry is opt-in.Checked MIT, about 60 test files. 14.5 to 18.9% fewer tokens on tool results (author).

  6. jevwire Claude Code PostToolUse on WebFetch, WebSearch and MCP results

    Screens fetched pages and MCP results for injected instructions before the agent reads them. Its other hooks can be switched off.

    Sends Each screened result, up to about 128,000 characters.Checked MIT, tests. 210 ms p50 warm (author).

  7. jev-lint CI or pre-commit

    Asks Jev one plain-English question about each ast-grep match (a comment, a catch block, a test) against a fitted cutoff.

    Sends The matched source code. --dry-run lists the questions and the price and sends nothing.Checked MIT, tests, on npm, 78 stars.

  8. jev-engineering Claude Code plugin hook before each command; GitHub Action on pull requests

    Two guards: jev-gate approves clearly safe agent commands, blocks clearly dangerous ones and defers the rest to the normal prompt; review-router tells an AI PR reviewer when to read the whole pull request.

    Sends jev-gate: each command, not the conversation. review-router: the pull request diff. Both through OpenRouter.Checked MIT. 0 of 12 dangerous commands ran against 8 of 12 under Claude Code auto mode, median 0.28 s; 66% of 3,622 commands auto-approved with no rm, git push, git reset or sudo approved; 13 of 13 CVE fixes on 561 commits sent to full review (author).Note Ran 9 of 12 safe commands without a prompt. The README states it catches mistakes, not a determined attacker. Starts in watch-only mode.

Worth knowing 8

Useful for the stack

New 2026-10-11

  • EdgeDecisionLocal cross-encoder picks Home Assistant actions, benchmarked against default agent.classificationopsdiscord★ 143
  • MOMO CODESelf-evolving coding agent adds a Jev decision layer for routing.coding-agentroutinggithub★ 1.1k3
  • vegamlFrozen LM plus physics engine gives calibrated, conformal typed decisions locally.classificationevalgithub★ 1514
  • skimmerCLI filters pages and files with Jev, cutting token usage sharply.ragcligithub★ 144
  • better-typescriptTypeScript linter asks Jev Noul questions for semantic, type-aware rules.classificationcoding-agentgithub★ 83
  • rulesmithDSPy searches cheap rule plus judge decision graphs, with GEPA tuning.classificationevalgithub★ 54
  • ai-agents-masterclass-modsToggleable Claude Code plugin adds Jev classify/filter/judge mode on demand.classificationcoding-agentgithub★ 42
  • Reflex (monotykamary)Routes requests across your coding agent sessions using Jev, not chat.routingcoding-agentgithub★ 43
  • MochiMac file organizer offers a direct TypeSafe Jev decision adapter option.opsclassificationgithub★ 32
  • Jev in LangSmith EvalsLangSmith evals use Jev to score difficulty and correctness per trace.evalclassificationx4
  • Unsloth Decision Model TrainerFree notebook fine-tunes your own small decision model like Jev locally.sdkclassificationx4
  • jev-deterministic-benchmark1000-question deterministic benchmark comparing Jev accuracy against other decision models.evalclassificationdiscord★ 23
  • FrenzyScores referral signups as human or bot using Jev classification.classificationsecuritydiscord2
  • Jev Showcase (cobusgreyling)Explains Jev's fan-out, pricing and confidence semantics beyond basic docs.educationdemogithub★ 1362
  • flockBlackboard multi-agent framework now routes on calibrated decision models.routingcoding-agentgithub★ 1213
  • llmsimRust load-test simulator for LLM APIs, including TypeSafe System One.opssdkgithub★ 63
  • ClaudeRouterClaude Code mod routing requests to haiku/sonnet/opus, optionally via Jev.coding-agentroutingcligithub★ 23
  • skill-router-modClaude Code mod where Jev ranks which skills fit each prompt.coding-agentroutinggithub★ 23
  • ejTiny on-device typed-decision model, calibrated probabilities, no token generation.classificationsdkgithub★ 23
  • Jev Skill TypeaheadPredicts which Claude Code skill will fire as you type.coding-agentroutingx3

New 2026-10-10

  • decisionkitCLI, MCP server and Claude Code plugin wrapping Jev and OpenAI Decisions.climcproutinggithub★ 44
  • llamacpp-jevServes Jev's structured decision API locally in front of llama-server.sdkclassificationgithub★ 83
  • basal-rsRust inference engine implementing System One API for Basal/Bielik models.sdkclassificationgithub★ 73
  • tink-routeRoutes tasks to specialist agent skills using Jev confidence scoring.routingcligithub★ 53
  • jev-docs-benchmarksBenchmarks Jev against Clef and Laya on accuracy, calibration and cost.evalgithub★ 43
  • TypeSafeJevPlaygroundUnofficial .NET port of TypeSafe's Jev SDK with MCP server guide.sdkgithub★ 92
  • feed-gardener-project-packLocal content aggregator filtering feed items by Jev relevance score.dataclassificationgithub★ 272
  • Arize Decision Model BenchmarkBenchmarks eight decision models on accuracy, latency, cost, prompt sensitivity.evalclassificationroutingdiscord4
  • discord-policy-checkerChrome extension flags Discord messages against policy using Jev.guardrailsbrowserclassificationdiscord★ 02
  • Sausage Battlefield case studyGame strategist benchmark: Jev 3-23x faster, 3.5-225x cheaper than 9 LLMs.evalgamediscord4
  • ZabioRoutes LLM requests via Jev classification, claims 30-40% cheaper than OpenRouter auto.routingdiscord3
  • HapslandReviews coding agent's decisions in realtime against your AGENTS.md rules.coding-agentevaldiscord3
  • OpenRouter Benchmark HarnessCLI harness for reproducible LLM benchmarks with Jev router cost tiers.evalroutingcligithub★ 224
  • Paper Trellis Citation VerifierChecks if cited papers support the sentence citing them; Jev scores it.ragresearchgithub3
  • decide (obie)Ruby decision layer turning typed yes/no/choice answers into policy verdicts.sdkclassificationgithub★ 83
  • jev-browser-wingmanJev-driven browser automation; completed 34/34 tasks vs Playwright's 32/34, 30% cheaper.browsercoding-agentgithub★ 54
  • jev-agent-controlOpenCode plugin routes Plan/Build agents and handles handoffs using Jev.coding-agentroutinggithub★ 33
  • neatrootsLocal-first agent trace viewer; Jev answers yes/no checks on each run.opsevalgithub★ 23

New 2026-10-09

  • grounded_ai CascadeEvaluatorCascades Jev before an LLM judge, cheaper and faster evalsevalclassificationgithub★ 714
  • WarrantModel-agnostic decision ledger logging agent actions with evidence and costopsevaldiscord★ 03
  • JevPromptShieldJev scores text at tool-call time to block prompt injectionguardrailssecuritygithub★ 2023
  • quicke2ePlain-English browser e2e tests, decision model picks clicks, exports Playwrightcoding-agentbrowserevalgithub★ 924
  • jevotronCLI flags anomalous fields in tabular, text and SQL data via Jevdataclassificationgithub★ 124
  • JEVDBDuckDB extension running semantic SQL filters and joins through Jevdataraggithub★ 94
  • optim-jevClaude Code plugin compacting context, routing models and skills via Jevcoding-agentroutinggithub★ 64
  • Jev Engineering CookbookOfficial 60-recipe notebook curriculum for Jev's Choice, Noul, Score patternseducationsdkgithub★ 84
  • pi-automode-classifierClassifies shell commands as auto-run, confirm or block before executionguardrailscoding-agentgithub★ 63
  • coderelay (GeWu-academy)Routes coding tasks across Claude Code, Codex, Pi CLIs with Jevcoding-agentroutinggithub★ 53
  • jev-router (dominicrico)Routes each Claude Code task to the right model and effortcoding-agentroutinggithub★ 33
  • gtm-engineer-starterScores prospect lists against your ICP rules with Jev evidence packetsopsclassificationgithub★ 53
  • omarchy-laserJudges your screen against your stated task to block distractionopsclassificationgithub★ 73
  • pagegradeGrades webpage sections on ten rubric criteria using Jevclassificationwritinggithub★ 73
  • carve-jevOpen-source Jev-like model family, one-token typed decisions via Qwen3.8classificationsdkgithub★ 103
  • jev.nvimNeovim semantic grep: scores every function against a plain-language questionsearchcoding-agentgithub★ 83
  • jev-maxBrowser agent piercing shadow DOM and iframes, with MCP supportbrowsermcpgithub★ 33
  • awesome-jev-use-cases (walidboulanouar)Curated list of 74 Jev demos and 150+ repos with costsresearchdiscord★ 4182
  • postgres-searchAlgolia-style Postgres search with optional Jev step to drop bad resultssearchragdiscord★ 23
  • CaretMac autocomplete; Jev decides abstain, inline edit or proposed actioncoding-agentroutinggithub★ 1573
  • Bud Decision StudioDesktop app to run open decision models locally, like LM Studiosdkopsgithub★ 1083
  • think-slow-act-fastPattern: LLM plans slowly, fast decision model acts in millisecondsroutingresearchgithub★ 53
  • JevLightJev picks traffic-signal phases from live intersection state each cycleresearchclassificationgithub★ 92

New 2026-10-08

  • Complee EU AI Act Compliance CheckerClassifies EU AI Act risk from text, asks one clarifying question.classificationdiscord3
  • OpenComputerUseComputer-use agent tool that can delegate action decisions to Jev or Clef.browsermcpdiscord★ 143
  • jev-router (kalowery)Routes coding-agent requests to cheaper models using Jev prompt classification.routingcoding-agentdiscord★ 04
  • jev-playground (deepdave98)GTM workflow benchmarks: lead triage, deal risk, dedup, reactivation, via Jev.classificationmcpgithub★ 293
  • TypeWrightCompiles declarative decision specs into typed Jev question programs via DSPy.classificationsdkgithub★ 73
  • dsh-tokenslashUses Jev to prune unneeded tool schemas from coding-agent system prompts.opscoding-agentgithub★ 53
  • pi-reply-guardScores agent replies against skill rules with Jev, prompts edits if low.guardrailscoding-agentgithub★ 44
  • SoulsBenchBenchmarks decision models playing Dark Souls combat on win rate and cost.evalgamegithub★ 33
  • laya-security-guardrailLocal security guardrail suite comparing Laya, Jev and LLMs on latency.guardrailssecuritygithub★ 124
  • could-jevClaude skill that judges whether Jev fits your use case.coding-agentclassificationgithub★ 33
  • graph-memory-system-one-demoDistills Claude Code agent traces into a Jev-run System One skill.ragcoding-agentgithub★ 14
  • jevgrep (dzhng)Research agent CLI using Jev for context collection, cuts coding costs 40%.coding-agentsearchx★ 2.6k4

New 2026-10-07

  • libsemopC++17 library for typed Jev decisions, flows, and reranking.sdkclassificationdiscord★ 02
  • brave-jev-mcpFilters Brave search snippets via packed Jev Choice questions, cuts redundant tokens.ragmcpsearchdiscord★ 03
  • Surely - AI TriageTriages monday.com board items with Jev; 74% automatic, $0.04 per 1000.classificationopsdiscord4
  • SOLO DecisionReorders rows and fields for prefix caching, up to 11.7x decision throughput.dataclassificationgithub★ 3603
  • Claude Model Router (alexei-led)Picks Claude Code model tier per turn using Jev or Clef.coding-agentroutinggithub★ 163
    1 similar
    • ClaudeRouterClaude Code mod routing requests to haiku/sonnet/opus, optionally via Jev.coding-agentroutingcligithub★ 23
  • OpenJevSwiftNative Swift OpenJev server runs Jev-compatible decisions on Apple silicon.sdkgithub★ 102
  • Evaluating and Benchmarking Jev (AML-Lab)Benchmarks Jev zero-shot on 37 datasets against two open-weight LLMs.evalclassificationgithub★ 64
  • jev-effortSets Claude Code reasoning effort per prompt from a Jev score.coding-agentroutinggithub★ 33
  • Decision 2.0 (vllm-sr)Open decision models, 0.6B to 27B, Jev-style classifiers on Hugging Face.classificationsdkx3
  • Lyzr Jev Router/Controller BenchmarkJev as RAG stop-or-search controller beats specialized models, 0.90 vs 0.69 AUROC.ragroutingevaldiscord5
  • aiduMEIMemory engine's optional Jev classifier cuts retrieval latency 78%, raises accuracy.ragdatagithub★ 204
  • ollajevDrop-in local server runs Hugging Face decision models behind Jev's wire API.sdkopsgithub★ 134
  • agent-thread-toolsJev decides when to hand off long Claude Code/Codex sessions, cuts tokens 30-40%.coding-agentcligithub★ 94
  • Taste BudsScores film/book/game taste with Jev; AUC 0.76 beats raw audience score.classificationdatamediadiscord3
  • jevlineExpands one confirmed-malicious indicator into a full incident timeline via Jev relatedness scoring.securityclassificationgithub★ 373
  • TopicJevZero-shot entailment check refines topic-model clusters without ground truth.ragclassificationgithub★ 93
  • llmClassificRR toolkit spanning bag-of-words to Jev classification, with calibration tools.classificationraggithub★ 103
  • dcisionSchema-defined typed-decision layer returning confidence before an agent runs.classificationroutinggithub★ 53
  • jev-job-helperJev screens job listings on BOSS, cut review time 59% in a 300-job test.classificationopsgithub★ 23
  • Irina: Knowing When to Re-RepresentCompares sparse state, event history, and derived mental-state representations for Jev agents.researchragdiscord2
  • Analisa USARoutes plain-English questions to maps and fact-checks federal contract data via Jev.dataroutingdiscord2
  • poordjaevin (Devin ACP fork)Local-first decision layer fork adds a keyless Devin ACP backend.classificationcligithub★ 62
  • TypeSafeSharp.NET client for TypeSafe's System One API with DI and OpenTelemetry support.sdkgithub★ 52

New 2026-10-06

  • repo-stateSemantic retrieval layer shrinks repo context for Jev reasoningragcoding-agentdatadiscord3
    1 similar
    • jevgrep (dzhng)Research agent CLI using Jev for context collection, cuts coding costs 40%.coding-agentsearchx★ 2.6k4
  • jev-inbox-triageJev decision model sorts store inbox emails into lanesclassificationopsdiscord★ 04
  • astrbot_plugin_jevJev decision model rechecks whether bot should chime into group chatclassificationopsgithub★ 33
  • jev-autoJev classifier judges each Pi agent tool call, allow block or askguardrailscoding-agentgithub★ 44
  • jurlcurl variant using Jev and Clef to pick relevant page contentsearchbrowsergithub★ 23
  • tea-fabrikaState machine dev loop uses Jev for narrow yes/no ticket decisionsclassificationcoding-agentgithub★ 22
  • kfchow-llm-value-routerJev routes each chat turn to cheapest model good enoughroutingopsgithub★ 23
  • MultiTool Office NextWindows office workstation uses Jev to assist OCR field judgmentopsdatagithub★ 202
  • awesome-jev (yibie)Curated list of agent harnesses, skill routers and Jev workflowscoding-agentdemox★ 2.3k2
  • Databricks ai_decide()Runs typed decision models natively over governed Databricks datadataclassificationx3
  • Jev vs Claude: A Measured Guide (Maven course)Paid course teaching Jev versus Claude tradeoffs end to end.educationdiscord2
  • nextjs-math-course (AI feedback classifier)Uses Jev to classify LMS lesson feedback by category and sentiment.classificationeducationgithub★ 113
  • jiwoOpen decision models compatible with Jev's API, benchmarked on Decision Index.classificationevalgithub★ 864
  • antigravity-feishu-botFeishu bot uses Jev as a safety gate and intent router.guardrailsroutinggithub★ 53
  • AgrDecision model alternative to Jev, scores 58.15 on Decision Index benchmark.classificationevalgithub★ 73
  • jev-270mBuilds a Jev-style decision interface on Gemma 270M, measures calibration gains.classificationresearchgithub★ 23
  • Jev in Vercel AI SDK for PythonAdds Jev decision model support to the AI SDK for Python.sdkx3

New 2026-10-05

  • dailypaper-skillsFilters and ranks daily research papers using Jev for cheap scoring.researchclassificationdatagithub★ 1.3k4
  • genai (maruel)Go AI package exposing typed decision calls across Jev, Clef, local models.sdkclassificationgithub★ 323
  • langchain-skill-routerPer-turn skill routing for LangChain with Jev judge, cuts prompt tokens 77%.routingclassificationgithub★ 64
  • PistonDecompilerGhidra binary analysis pipeline using Jev for candidate triage, unmeasured.securityclassificationgithub★ 52
  • ThoughtZero (pyschic-vla)Jev judge scores math tree-search steps; code complete, not yet run at scale.researchclassificationgithub★ 31
  • Laya WorkbenchOne-click local workbench for testing Laya and Jev-compatible decision APIs.sdkopsgithub★ 43
  • coding-router-jevRoutes Codex and Claude Code turns through Jev for model and effort selection.coding-agentroutinggithub★ 23
  • claude-deckClaude Code HUD showing live usage, subagents, and Jev routing decisions.coding-agentopsgithub★ 113
  • quiz-reviewerChecks quiz answer keys against course materials using local Jev-style decision models.educationraggithub★ 24
  • OpenRouter Model Router BenchmarksCompares 7 model routers, including Jev Router, on quality, speed, cost.routingevalx4
  • cai (ad-si)Multi-provider AI CLI with a Jev yes/no classification command built in.cliclassificationgithub★ 2034
  • jev-graph-searchJev reranks Obsidian/Logseq notes, raised top-5 source recall from 63.9% to 81.8%.ragsearchgithub★ 124
  • jev-mcp-router (ini8labs)Jev answers yes/no per MCP tool so the LLM only sees a few, cuts routing cost 32x.mcproutinggithub★ 34
  • jev-snips-benchmarkMeasures Jev intent and slot-filling accuracy on SNIPS across prompting conditions.evalclassificationgithub★ 13
  • hotel-reviews (svpino)Scrapes reviews with Apify then classifies them against custom criteria with Jev.classificationdatax★ 23
  • Amazon Strands DeciderOpen-weight 2B decision model, runs locally, ~153ms median latency on an M3 Mac.routingclassificationx3
  • Ollama Nimble (/v1/systemone)Ollama adds a local decision-model API and Nimble model for triage and routing.classificationclix3

New 2026-10-04

  • tenjin-agentRoutes Claude Code tool calls through Jev and pays for tools via x402/USDC.coding-agentroutingsdkdiscord★ 153
  • jev4pgAdds natural-language-to-SQL and Jev semantic operators to PostgreSQL queries.dataragclassificationgithub★ 2384
  • pi-model-routerRoutes Pi coding-agent turns across model tiers with optional Jev advice.routingcoding-agentcligithub★ 63
  • WaterSheepOpen-source calibrated yes/no, choice, score and multi-label model, a Jev alternative.classificationevalgithub★ 123
  • Jev Benchmark (ywchiu)Benchmarks Jev against open alternatives like Clef and Cygnet on agent routing.evalroutinggithub★ 63
  • JEV Research MCPFilters web search results and content blocks with Jev before feeding agent context.mcpragsearchgithub★ 34
    1 similar
    • brave-jev-mcpFilters Brave search snippets via packed Jev Choice questions, cuts redundant tokens.ragmcpsearchdiscord★ 03
  • DoubtBenchScores decision models' probabilities against human annotator disagreement, not just accuracy.evalclassificationgithub★ 24
  • agent-routerPicks Cursor, Claude Code, Codex or OpenCode plus model via Jev.routingcoding-agentclidiscord★ 1133
  • jev-scriptcCompiles Jev CLI to native binary, 2.43x faster than Node.clisdkdiscord★ 03
  • JCT LabsVisual builder, playground and regression suites for Jev questions.evalclidiscord3
  • jevsd-pgSelf-developing Postgres database answers natural language queries via Jev operators.dataclassificationgithub★ 2383
  • workshop-jevLangGraph sales agent compares Jev vs GPT/Claude on cost and latency.evalroutinggithub★ 123
  • stfuClaude Code hook scores reply length with Jev, caps verbosity.coding-agentguardrailsgithub★ 84
  • sysone-benchByte-identical benchmark compares Laya, Jev and Qwen decision models directly.evalgithub★ 64
  • macos-computer-use-kitAccessibility-tree computer use for agents with optional Jev safety guards.coding-agentmcpgithub★ 54
  • thinkthenShell CLI returns true, false, label or number from Jev.cliclassificationgithub★ 73
  • RaidCEL policy engine gates agent actions, escalates ambiguous cases to Jev.guardrailscoding-agentgithub★ 44
  • JeffortScores each Claude Code prompt with Jev to set effort level.coding-agentroutinggithub★ 23
  • ConjevtureTypeScript library combines Jev probabilities with explicit Boolean logic circuits.classificationdatagithub★ 22
  • cleffaNative Metal engine serves Clef models in Jev's API format.sdkcligithub★ 53
  • Clef / Clef-flashCloudflare's open decision models are fully Jev-API compatible on Workers AI.sdkroutingx4

New 2026-10-03

  • jev-permission-gateJev pre-screens Claude Code auto-mode tool calls, 2x faster, zero unsafe allows in tests.guardrailscoding-agentgithub★ 294
  • fastcampus-jevKorean tutorial: Jev for tool selection, injection/PII guardrails, and risky-action gates.educationguardrailscoding-agentgithub★ 373
  • c9r-dev/plugins (browser-check)Claude Code plugin runs plain-English QA checklists in a browser, judged by Jev.coding-agentbrowserevalgithub★ 74
  • Hermes SwitchyardJev picks skills, reasoning effort, models and UI actions for a Hermes agent session.routingcoding-agentgithub★ 73
  • cypridRuns local open-weight decision models behind one Jev-style API on a Mac.sdkclassificationgithub★ 33
  • jev-job-searchJev fills and verifies job application forms; 100 applications in under 10 minutes for about $1.browsercoding-agentgithub★ 54
  • jev-attack-surface-analysisRanks codebase files and lines by vulnerability risk using Jev, about a cent per repo.securityclassificationgithub★ 24
  • JevImpactScores a vulnerability report's real security impact with Jev, returns a 0-1 impact score.securityclassificationgithub★ 43
  • HumanCalBenchmarks whether a decision model's probabilities match human annotator disagreement, 7500 questions.evalclassificationgithub★ 23
    1 similar
    • DoubtBenchScores decision models' probabilities against human annotator disagreement, not just accuracy.evalclassificationgithub★ 24
  • Pydantic AI + JevPydantic AI agents can run yes/no and pick-one decisions on Jev, shown in Logfire.sdkroutingx3
  • restaurant-review-jev-classifierClassifies food safety violations from Google reviews using Jev.classificationdatadiscord★ 02
  • RakshaQuantPaper-trading platform measures whether Laya plus Jev beats baseline, net of cost.financeevaldatagithub★ 604
  • JevboxPermission-aware document library with Jev-powered search and source-grounded chat.ragsearchdatagithub★ 7053
  • nanny-agency-workshopWorkshop notebook uses Jev as confidence-gated judge in an eval pipeline.evaleducationguardrailsgithub★ 63
  • spending-effort-with-jevClaude Code plugin uses Jev to recommend per-prompt effort level.coding-agentclassificationcligithub★ 74
  • awesome-ai-rabbit-holesCatalog of AI agent tools with a section on Jev-style decision models.researchgithub★ 72
  • jev-pr-qualityJev-assisted pull request quality reviews with a live multi-repo dashboard.coding-agentevalgithub★ 74
  • pi-jfilesPi extension answers typed per-file questions via Jev, returns no source text.classificationcoding-agentcligithub★ 53
  • tev1-50-use-cases50 local use cases for Together AI's tev1 decision model via Ollama.classificationdemoroutinggithub★ 63

New 2026-10-02

  • TypeSafe Official Claude Code SkillOfficial skill teaches Claude Code to build Jev integrations correctly.coding-agentsdkx4
  • VibefilterFilters admin tables by a plain-English statement using Jev's calibrated probabilities.classificationdatadiscord4
  • grind.fyi AutopilotTracks new software job postings and autofills applications, using Jev to label job levels.opsclassificationdiscord2
  • JevMadeSearchable catalogue of 2,300+ Jev projects, guides and experiments with source credit.searchdatadiscord2
  • pi-web-search (Pi Agent)Web and code search tool for Pi agents using Jev for ranking, filtering and prompt-injection detection.mcpsecuritysearchdiscord★ 204
  • Competitor Price Monitor (Apify + Jev)Apify actor monitoring competitor Shopify prices, first actor to ship a Jev feature.financedatadiscord2
  • Awesome Decision ModelsCurated list of System One decision model APIs, open weights, runtimes, SDKs and benchmarks.researchdatagithub★ 6363
  • Daily News (AI intel board)Multi-source AI news dashboard with optional Jev classification for relevance and prompt injection.dataclassificationgithub★ 1112
  • GoEventBusGo event bus that optionally uses Jev to pick the event type before deterministic routing.routingopsgithub★ 733
  • ZengLian AI Tutor (EduAgent)Open-source RAG AI tutor for K-9 curricula with a Jev decision layer and failover.rageducationgithub★ 623
  • agent-stackMulti-agent dev team wiring Claude Code, Codex and Jev for planning, review and merge gates.coding-agentroutinggithub★ 313
  • Alfred (computer-use)Voice-controlled Mac agent where Jev picks the typed action after local speech recognition.cliopsgithub★ 253
  • gutsyLocal 0.8B GGUF decision model with Jev-style API, 0.021 calibration error on held-out data.sdkclassificationgithub★ 844
  • jev-doc-searchSearches long documents page by page with Jev's Choice calls, no vector DB needed.ragsearchgithub★ 1384
  • Jev ArchitectAgent skill that helps decide where Jev decision loops fit and how to test them first.coding-agentevalgithub★ 73
  • jesOpen-source guardrails using decision models to block prompt injection, jailbreaks and PII leaks at every agent step.guardrailssecuritygithub★ 144
  • Meta Awesome JevMerges and deduplicates 106 awesome-jev lists into 18,330 cross-checked links.researchdatagithub★ 93
  • jev-demos (mani-aiml)Demo folders including Jev vs Claude Haiku as a prompt-injection review gate against real attacks.guardrailsevalgithub★ 73
  • jev-agent-toolsMulti-provider Jev transport layer with fail-closed validation and automatic retries.sdkgithub★ 63
  • Jev LinkedIn Saved ClassifierReads your LinkedIn saved posts and classifies each one with Jev into a filterable board.classificationdatagithub★ 93
  • adhereLinter for non-deterministic business rules checked by Jev against markdown-defined policies.classificationcoding-agentgithub★ 83
  • jev-search-mcp (mukiwu)Jev-powered web search as an MCP server, Claude Code plugin and CLI with zero dependencies.mcpsearchcoding-agentgithub★ 74
  • jev-compaction-plusCompacts Claude Code sessions in 0.5s vs 35s, keeping needed context word for word.coding-agentopsgithub★ 263
  • Jev Skills (n23eos)Fourteen Jev-powered advisory skills for picking skills, models and next steps in coding agents.coding-agentroutinggithub★ 53
  • awesome-trustworthy-jevCurated papers and projects on Jev trustworthiness and cybersecurity applications.securityresearchgithub★ 53
  • JevBench (metamorphic coherence)Tests whether decision models' probabilities stay coherent across 50 metamorphic laws, no gold labels needed.evalclassificationgithub★ 113
  • Hola, PerúRoutes described situations to the matching gob.pe government procedure using Jev's Choice.classificationroutinggithub2
  • jev-ultrafast AI UI TestsSelector-free UI tests that pick operations and elements via a local System 1 model.evalbrowsergithub★ 13

New 2026-10-01

  • Jev Guard (gulbaki)Scores text against OWASP LLM Top 10 risks via Jev, with live demo.guardrailsclassificationgithub★ 143
  • reflex-routerProxies Claude Code, uses Jev to judge task difficulty and route to cheaper model.routingcoding-agentgithub★ 53
  • QevFine-tunes Qwen into an open decision model handling Choice, Noul, Score.classificationsdkgithub★ 263
  • JevAltOpen 4B decision model, calibrated answers in English, Turkish, German, runs on CPU.classificationgithub★ 53
  • Jev Chat GateLets an agent join group chats unsolicited when Jev judges it useful.routingclassificationgithub★ 43
    1 similar
    • astrbot_plugin_jevJev decision model rechecks whether bot should chime into group chatclassificationopsgithub★ 33
  • Jev local CasaNorteLoads an Excel, gets Jev decisions with probabilities, exports results locally.dataclassificationgithub★ 22
  • claude-refereeClaude Code plugin checks agent's done claims with Jev and logs results.coding-agentevalgithub★ 64
  • fileroute (fr)Jev scores file relevance to a question, live and in parallel, no indexing.searchcoding-agentgithub★ 54
  • iMessage Spam BlockerNative macOS app uses Jev to classify and block spam iMessages.classificationsecuritygithub★ 22
  • doppProxies Jev-shaped calls, logs them, trains a small owned model as fallback.routingsdkgithub★ 23
  • CloakBrowser AgentJev decides each browser step in 0.3s, stealth browser executes it.browsermcpgithub★ 33
  • s1-tuiTerminal UI for testing Jev and local Laya typed decisions side by side.cliclassificationgithub★ 23
  • VevOpen vision-capable decision model, Jev-API compatible, runs on your own GPU.classificationmediagithub★ 173
  • Medplum Provider + JevJev flags conflicts between a clinical note and hospital discharge summary.healthclassificationgithub★ 42
  • KBC TectonicInfers banking client life-stage profiles from transactions, with optional Jev scoring.financeclassificationgithub★ 22
  • D1ASmall Gemma-based decision model built on Kev, calibrated typed answers.classificationevalgithub★ 23
  • Hunch (browser extension)Browser extension highlights paragraphs answering your search, not just keyword matches.browsersearchdiscord3
  • SedumRuns AI browser e2e tests on every PR, matched via Jev.coding-agentbrowserevaldiscord★ 114
  • IntentSQLConverts natural language to SQL via chained Jev decisions over schema state.dataclassificationdiscord★ 33
  • yardsortDesktop app runs parallel coding agents, flags off-task, secret or weakened-test changes.coding-agentguardrailsdiscord★ 104
    1 similar
    • tomarigi-desktopShows Claude Code/Codex sessions as birds, flags abusive messagescoding-agentdemogithub★ 172
  • ResurfaceJev screens resumes against job questions with quoted evidence, 31x cheaper.classificationragdiscord★ 04
  • langgraph-jevLangGraph node wrapping Jev decisions with routing and confidence thresholds for review.sdkroutingdiscord★ 34
  • POK-AgentWindows computer-use agent uses Jev for actions, caching successes as motor programs.coding-agentdemodiscord★ 14
  • jev-vs-layaBenchmarks Jev and Laya decision models playing 2048 against Expectimax and humans.evalgamediscord3
  • hermes-tool-slimmerTrims tool schemas sent per turn; Jev mode cuts tokens 71%.coding-agentsdkgithub★ 394
  • vybridRust coding assistant routes turns among models using optional Jev routing.coding-agentroutinggithub★ 122
  • dspy-system-one-agent-patternsSix examples show a coding agent using Jev for safety checks.coding-agenteducationgithub★ 383
  • verdict (khimaros)Self-hosted System One API server serving any llama-server as a Jev replacement.sdkclassificationgithub★ 84
  • sensoredPII redaction library with optional Jev semantic confirmation to cut false positives.securityguardrailsgithub★ 103
  • jev-idsJev-based intrusion detector matches NSL-KDD accuracy 22x cheaper than an LLM.securityclassificationgithub★ 184
  • jev-playwright (cooper667)Plain-English browser QA checklist judged by Jev on Cloudflare Workers, 35 seconds.coding-agentbrowserevalgithub★ 74
  • Bud-Decision-EngineDesktop app runs open decision models locally, like LM Studio for Jev.sdkcligithub★ 1084
    1 similar
    • Bud Decision StudioDesktop app to run open decision models locally, like LM Studiosdkopsgithub★ 1083
  • agents-mailSelf-hosted agent mailbox with optional Jev spam, phishing and injection labeling.guardrailssecuritymcpgithub★ 54
  • playwright-jev (AdriaanVE)Adds Jev-based failure triage, retries, locator healing and test ordering to Playwright.coding-agentbrowserevalgithub★ 24
  • TrancheJev clusters duplicate PRs into reviewable batches for a huge GitHub backlog.opsgithub★ 32
  • anofox-decideDuckDB extension evaluates natural-language predicates using Jev or local decision models.dataclassificationgithub★ 23
  • JevTraceMCP server retrieves only relevant TS/JS code for coding agents using Jev.coding-agentmcpgithub★ 34
  • skill-scannerScans Agent Skills offline for prompt injection and exfiltration before install.securityguardrailscligithub★ 24
  • Jev Haiku/Opus Game Theory BenchmarkBenchmarks Jev Haiku and Opus models playing game theory scenarios.evalresearchdiscord2

New 2026-09-30

  • sf-permit-helperDeterministic permit flowchart calls Jev only when it needs more input.classificationopsdiscord★ 02
  • Best FitJev-judged leaderboard picks the cheapest model for your exact job.routingclassificationmcpdiscord4
  • In Other WordsChrome extension uses Jev to rewrite verbose LinkedIn posts in plain English.writingclassificationgithub★ 293
  • jev-contextJev decides what to keep when Claude Code or Codex compacts context.coding-agentsdkgithub★ 83
  • ReflexJev screens every tool call and turn for a Pi coding agent.coding-agentbrowsersdkgithub★ 53
  • Jev Web AnalyzerJev scores a SaaS homepage's clarity and differentiation like a first-time visitor.writingclassificationdemogithub2
  • Meta-ArchitectQuality-gate workflow layer uses Jev for autonomous routing across coding agents.coding-agentopsgithub★ 52
  • jev-medical-benchBenchmarks Jev's typed decisions against chat LLMs on clinical judgment tasks.evalhealthclassificationgithub★ 113
  • what-the-jevReproducible experiments probing Jev on math, vision, ethics and injection detection.evalsecurityclassificationgithub★ 73
  • jev-humanizerJev flags AI-sounding phrasing patterns so an agent rewrites only those.writingclassificationcoding-agentgithub★ 23
  • WebJevSpecialized decision model beats Jev at picking browser actions, ties on benchmarks.browserevalgithub★ 33
  • Deep RecallJev reranks Markdown note passages for a Claude Code memory plugin.ragsearchcoding-agentgithub★ 124
  • cargo-crapUses Jev to judge which duplicate Rust functions are worth merging.coding-agentclassificationdiscord3
  • Home Assistant TypeSafe Conversation AgentHome Assistant voice agent picks device actions via typed Jev questions.classificationopsdiscord★ 73
  • jev-proof-selectorBenchmarks Jev choosing Lean proof lemmas against heuristics and other models.evalresearchdiscord★ 02
  • Sez Where?Points to relevant document passages for a question without summarizing.ragsearchdiscord3
  • system-one-foundation-modelsSwift bridge runs Jev-style decision models on-device in 15-150ms.sdkclassificationgithub★ 993
  • JevstillerDistills a repeated Jev classification call into a local model live.routingclassificationgithub★ 663
  • claude-subagent-routerJev classifier picks Sonnet vs Opus per Claude Code sub-task, cuts tokens.coding-agentroutinggithub★ 104
  • JevaroBatches Jev calls and streams answers as Arrow for pandas or DuckDB.datasdkgithub★ 164
  • WorkflowEvalsOfficial code to reproduce TypeSafe's evals.typesafe.ai workflow benchmark results.evalsdkgithub★ 164
  • riffRuff-style prose linter flags writing issues using calibrated Jev judgments.writingcligithub★ 73
  • typed-decisions (kotoba-lang)Reproduces Jev's typed-decision shape on ModernBERT, DeBERTa, LLaDA-MoE with benchmarks.researchclassificationgithub★ 103
  • usejevRuns Laya locally on Bun with ONNX and a TypeSafe-compatible API.sdkclassificationgithub★ 133
  • Startlux-DecisionOpen typed decision models, 0.8B to 27B, drop-in Jev API compatible.classificationsdkgithub★ 3673
  • jev-in-the-wildTracks real Jev use cases and criticism across GitHub, Reddit, and blogs.researchdemogithub★ 73
  • jevrR client for sending typed questions to Jev's System One API.sdkgithub★ 52
  • opencode-jev-router (robertn702)Jev picks reasoning effort per OpenCode step; matched quality with fewer tokens.coding-agentroutinggithub★ 104
  • DocketGo TUI files documents after Jev scores category, sensitivity, and urgency.classificationcligithub★ 63
  • DecisSelf-hosted server speaks Jev's API for open models like Laya and kev.sdkopsgithub★ 94
  • go-system-oneNative Go decision runtime hits 84.85% on JevBench's public subset.evalclassificationgithub★ 93
  • warped-sys1-labSvelteKit workbench and CLI for authoring and running Jev experiments.clievalgithub★ 33
  • Jev ModeClaude Code skills that rate and image-prep Jev projects, plus a lens.coding-agentevalgithub★ 63
  • jevforestBagged trees vote on which feature to buy next under a budget.researchclassificationgithub★ 22
  • BiliBili Jev FilterFilters Bilibili recommendations by metadata using TypeSafe Jev in Quantumult X.classificationopsgithub★ 22
  • Hermes AgentJev triages scraped job postings before larger models fill applications.classificationopsgithub★ 03
  • MnemonFast Jev judgments filter memory search hits for agents under 4k tokens.ragclassificationgithub★ 43
  • semantic-find-userscriptBrowser userscript does semantic Ctrl+F on page text via Jev directly.browsersearchgithub★ 22

New 2026-09-29

  • jev-engineeringCoding-agent guard tools: review router and command lock, both benchmarked.coding-agentguardrailsgithub★ 94
  • n8n-nodes-jevn8n community node that runs Jev classify/route/score steps in workflows.classificationroutinggithub★ 62
  • oh-my-plumbUses Jev to catch coding-agent edits that break unenforceable AGENTS.md rules.guardrailscoding-agentgithub★ 54
  • jev-opusJev reflex re-picks Claude Opus effort level every turn without breaking prompt cache.coding-agentroutinggithub★ 133
    2 similar
    • JeffortScores each Claude Code prompt with Jev to set effort level.coding-agentroutinggithub★ 23
    • jev-effortSets Claude Code reasoning effort per prompt from a Jev score.coding-agentroutinggithub★ 33
  • meetvconMeet transcription extension classifies intent and sentiment per line with Jev in ~300ms.mediaclassificationgithub★ 52
  • model-switcherRoutes Claude Code prompts to cheap or heavy models, with optional Jev pre-check.routingcoding-agentgithub★ 53
  • greenwash-ossGitHub app that asks Jev per-hunk whether a PR fakes a passing test.coding-agentguardrailsgithub★ 64
  • DutygateFlags legal/compliance obligations in support messages before a bot replies, via Jev.guardrailsclassificationgithub★ 43
  • jevtok-tsOffline Jev token counter and request accounting for Node, no API calls.sdkopsgithub★ 33
  • toolJevUses Jev instead of an LLM to pick MCP tools, benchmarked on four public suites.mcpclassificationevalgithub★ 34
  • awesome-jev-securityCurated papers on Jev/System One robustness, security, and safe deployment.securityresearchgithub★ 62
    1 similar
    • awesome-trustworthy-jevCurated papers and projects on Jev trustworthiness and cybersecurity applications.securityresearchgithub★ 53
  • jev-chainChains Jev calls into typed decision graphs with full traces.routingclassificationdiscord3
  • JebadiahOpen decision model family answering fixed-choice questions with probabilities.classificationdiscord3
    1 similar
    • Decision 2.0 (vllm-sr)Open decision models, 0.6B to 27B, Jev-style classifiers on Hugging Face.classificationsdkx3
  • JeviaRoutes coding-agent tasks to model tiers using verified outcome history.coding-agentroutinggithub★ 103
  • finance-agentRoutes finance research sub-agents with Jev, cutting latency to 100-400ms.financeroutingdiscord★ 23
  • Jevscan Chain MonitorScans blockchain transactions with Jev, catching 56 of 60 known hacks.securityclassificationgithub★ 64
  • qwen3-0.6b-rlcd-decisionRLCD-trained 0.6B Qwen model imitating Jev's decision-making, open source.researchdiscord2
  • onesieUnix-composable Jev CLI with threshold calibration, caching, and mock testing.clievaldiscord★ 23
  • jev-prompt-enhancerSkill that uses Jev to critique and refine prompts before running.classificationwritingdiscord★ 12
  • jido Jev routing experimentUses Jev as a cheaper, faster tool/model router inside the Jido harness.routingcoding-agentdiscord★ 03
  • gut (gutpy.dev)One-line Python judgment calls answered yes/no/unsure by Jev, then small models.classificationclidiscord3
  • nudgementLinter-style tool where Jev scores fuzzy judgments like commit message quality.coding-agentclassificationdiscord★ 23
  • Jeff (firelex)Fine-tuned Qwen/Gemma zero-shot classifiers, Jev-compatible API, 22-28ms per decision.classificationsdkgithub★ 1.5k4
  • JeevesReasoning 9B classifier beats Jev and Kev on held-out benchmarks.classificationresearchgithub★ 4214
  • SelfJevSelf-hosted Jev-API-compatible decision model, Qwen3.5-4B+LoRA, single GPU.classificationsdkgithub★ 623
  • jev-ui-testUI test automation scored by Jev, 458ms median per decision.browserevalgithub★ 553
  • DuoMindPairs local LLM with Jev; benchmarks show Jev made it 36% slower.classificationcligithub★ 162
  • jev-specCI check that Jev-scores code against markdown spec requirements.evalcoding-agentgithub★ 113
  • pi-jev-model-router (da-vinci-noob)Routes pi prompts to model tiers using Jev's typed judgments.routingcoding-agentgithub★ 112
  • pi-auto-reviewerAuto-reviews shell commands before running, using Jev for cheap decisions.securitycligithub★ 73
  • jev-skill-router (shimo4228)Skill-suggestion hook using Jev; author found it doesn't help strong models.coding-agentevalgithub★ 152
  • semble-jevCLI code search where Jev scores relevance of retrieved snippets.searchcoding-agentgithub★ 63
  • pi-pair-programmerBackground code reviewers filtered by a Jev-based duplicate-finding filter.coding-agentevalgithub★ 93
  • Jev Arena (theaiautomators)Local dashboard comparing decision models on accuracy, speed, and memory.evalresearchgithub★ 334
  • Gevva0Local Gemma-based decision gateway claims #1 on JevBench, beating Jev 1.13.classificationevalgithub★ 54

New 2026-09-28

  • system-one-reviewerLocal PR reviewer: deterministic code plus Jev judgments, with precision recall harness.coding-agentevaldiscord★ 24
  • Jev-Network-Packet-AnalyzerReal-time Wi-Fi packet analyzer that has Jev classify anomalies for threats.securityclassificationgithub★ 63
  • yornEarly prototype exploring a local noul yes/no decision primitive.classificationdiscord★ 02
  • BalabolFree translator app: Jev detects dominant source language to disambiguate LLM prompts.classificationdiscord2
  • jevpipeUnix CLI that pipes files through Jev judgment questions for agent filtering.clicoding-agentdiscord★ 63
  • ReplacemeMac menu bar app: Jev answers narrow yes/no questions to triage notifications.classificationopsdiscord3
  • jev-integration-evaluatorScans a codebase and evaluates where Jev integration would help.evalclassificationdiscord★ 12
  • Photoshop MCPPhotoshop MCP server; new mode uses Jev for instant, no-LLM edits.mcpgithub★ 6212
  • ai-day-trader-agentClaude Code trading agent: Jev scores news headlines for relevance and direction.financecoding-agentgithub★ 153
  • system-design-trainerSystem design interview trainer where Jev grades your whiteboard design.educationevalgithub★ 1123
  • deepseek-harness-jevDeepSeek Harness plugin using Jev for skill selection, supervision, and approvals.coding-agentguardrailsgithub★ 2032
  • fm-with-jevBenchmarks Apple's on-device model vs Jev for routing and list-picking decisions.routingevalgithub★ 94
  • JuLLocal decision model with Jev's SDK API, no monologue, answers in 55ms.sdkclassificationgithub★ 84
  • JEVRIELAgent skill that finds Jev integration points and benchmarks results.coding-agentsdkgithub★ 63
    2 similar
    • Jev ArchitectAgent skill that helps decide where Jev decision loops fit and how to test them first.coding-agentevalgithub★ 73
    • could-jevClaude skill that judges whether Jev fits your use case.coding-agentclassificationgithub★ 33
  • jev-frontend-qaBrowser QA harness where Jev makes bounded UI decisions; contracts checked.browserevalgithub★ 53
  • jev (stefafafan)Unofficial Unix CLI client for Jev supporting TypeSafe, Cloudflare, and Vercel providers.clisdkgithub★ 133
  • fast-browserCodex/MCP browser automation using Jev or local Laya to choose actions.browsermcpgithub★ 42
  • genigrepCode search for agents: Jev ranks files, cutting agent cost about 7%.searchcoding-agentgithub★ 34
  • kev-onnxSelf-hosted, TypeSafe-compatible decision API running the KEV model on CPU-only ONNX.sdkclassificationgithub★ 93
  • jev-ecosystem-surveySurvey classifying 16,128 GitHub repos in the Jev ecosystem by use case.researchdatagithub★ 13
  • jev-risk-check-providerx402 payment risk checks scored by Jev with signed attestations for agents.securityfinancegithub★ 102
  • gliner2-api-jev-schemaLocal GLiNER2 classifier served behind a Jev-compatible System One API endpoint.classificationsdkgithub★ 33
  • JevSpanZero-shot NER on Jev averaging 73.7 F1 across 12 benchmarks, no training.classificationdatagithub★ 64
  • b2b-web-testB2B web test runner where Jev picks each element and logs steps.browserevalgithub★ 22
  • PigeonholeTurns a plain-English classifier description into a typed API on Jev.classificationsdkgithub★ 22
  • jevwrightBrowser flow tests where Jev finds controls once, then replays for free.browsercoding-agentgithub★ 133
  • OpenRecurSearchOpen AI agent with Jev web interface for research reports.searchresearchdiscord★ 22
  • OneRouteJev asks 20 questions then code scores 150 models for routing.routingdiscord★ 13
  • DEMO MCP (amidevz)MCP server with built-in Jev and Laya plus browser tools.mcpbrowserdiscord2
  • JevulonRoutes coding agent tasks in parallel with deterministic file coordination.coding-agentopsdiscord2
  • jev-excelOpen source Excel add-in that calls Jev from spreadsheet cells.sdkdiscord★ 12
  • TraceDocsJev evidence discovery with verifiable traces for structured documents.ragdatadiscord★ 43
  • Code Puppy Smart GrepUses Jev over ripgrep for context, cutting token usage 90%.coding-agentclidiscord★ 8523
  • LatteReviewAdds Jev decision-model screening to literature review agents, 0.2s per article.classificationragresearchgithub★ 1224
  • BrighTO_RouterRust LLM gateway routing to SystemOne/Jev decisions with load balancing.routingopsgithub★ 213
  • imajevOpen multimodal Jev-style decision model reading photos, ranks above Jev 1.13.classificationresearchmediagithub★ 3654
  • typesafe-jev (cv-screen)Screens CV folders with Jev typed questions and editable scoring policy.classificationopsgithub★ 103
  • Jev PowerShell modulePowerShell module for typed yes/no, choice and score Jev decisions.sdkcligithub★ 83
  • HopperLoRA decision server on Qwen3.5-4B for JevBench with calibration temperatures.classificationresearchgithub★ 73
  • Decision IndexReproduces the Jev decision-model benchmark suite locally or via HF Job.evalresearchgithub★ 393
  • opendeciderOpen calibrated decision models beating Laya on typed decisions, CPU/GPU.classificationevalgithub★ 323
  • OneJevMultimodal decision model answering typed questions on screens, photos and video.classificationmediaresearchgithub★ 1563
  • apollo-jev-lead-classifierQualifies Apollo leads with Jev and drafts personalized outreach messages.classificationdatagithub★ 33
  • Ten Levels of JevTen worked examples of Jev from plain code to autonomous coding agent.coding-agenteducationgithub★ 2184
  • Job Intelligence BotScrapes 240 career pages and scores postings against a CV with Jev.classificationdatagithub★ 23
  • Wald-4B4B calibrated decision model scoring 53.91 on Decision Index, best in class.classificationevalgithub★ 63
  • MotherDuck prompt_jev()SQL function powered by Jev classifies text 50x faster at 1% cost.dataclassificationx5

New 2026-09-27

  • Jev-AI-SkillWorkflow skill: Jev reviews agent output, routes models, gates git pushes.coding-agentroutingdiscord★ 33
  • jev-decision-benchSelf-hosted benchmark toolkit for testing Jev against your own decision datasets.evaldiscord★ 23
  • FreeResume.siteResume feedback tool that stays silent when Jev's confidence is weak.classificationwritingdiscord3
  • RAG-Fusion (Jev reranker experiment)Adds Jev as a reranker to RAG-Fusion's multi-query retrieval pipeline.ragsearchgithub★ 9594
  • building-with-typesafe-jevSkill teaching coding agents Jev patterns from 150+ community projects.educationresearchgithub★ 1372
  • Open Medical JevCascaded open models match Jev within ~2 points on medical exams.healthclassificationgithub★ 343
  • Jevflakedbt and Terraform package exposing Jev as callable Snowflake SQL functions.datasdkgithub★ 193
  • pi-jev-router (philippdubach)Shadow-mode Jev router that picks OpenRouter models for pi tasks.routinggithub★ 133
  • jev-demo (gopinav)TypeScript SDK examples of Jev best practices and common mistakes.sdkeducationgithub★ 203
  • decision-model-benchmarkIndependent benchmark comparing Jev, constrained LLMs and baselines on typed decisions.evalgithub★ 114
  • Immune HarnessSecurity layer where Jev scores risk and blocks agent tool calls.securityguardrailsgithub★ 93
  • RSI-JevAI-agent loop that trains and retires its own Jev-style decision models.researchclassificationgithub★ 1113
  • xscout-jevWatches X for relevant posts, has Jev judge them, alerts Slack.opsclassificationgithub★ 123
  • Jev Codex BridgeWindows service routing Codex Desktop and CLI tasks through Jev.routingcoding-agentgithub★ 93
  • n8n-nodes-jev-classificationn8n node that classifies and scores text with Jev, batched.classificationopsgithub★ 63
    1 similar
    • n8n-nodes-jevn8n community node that runs Jev classify/route/score steps in workflows.classificationroutinggithub★ 62
  • jev-docsCommunity-tracked history of Jev APIs, SDKs and agent guidance.educationsdkgithub★ 62
  • MiserRust AI gateway routes prompts by Jev-classified complexity tier.routinggithub★ 54
    1 similar
    • BrighTO_RouterRust LLM gateway routing to SystemOne/Jev decisions with load balancing.routingopsgithub★ 213
  • kimeRust inference engine speaking Jev and Laya's wire formats.sdkgithub★ 73
  • fast-compactClaude Code command where Jev picks tool output worth keeping.coding-agentcligithub★ 64
  • jev-mail-filterGmail filters written in plain English and judged by Jev.opsclassificationgithub★ 54
  • YolkDesktop PR review lens where Jev flags core vs defensive code.coding-agentgithub★ 34
  • decision-gateRate limits, spend caps and caching for code calling Jev.opsgithub★ 24
  • JEMMOpen-weight multimodal Jev-style model judges text plus screenshots.classificationresearchgithub★ 143
  • polyjevDrop-in Jev API server turning any LLM into typed decisions.routingsdkgithub★ 24
  • jigorRust gateway running local decision models or Jev, one protocol.routinggithub★ 23
  • RC-Papers: Mathematical Foundations of JevFormal mathematical framework for Jev with Lean 4 proofs.researchgithub★ 22
  • VerQenOpen decision model alternative to Jev, ~5ms inference, published benchmark.classificationresearchgithub★ 23
  • decision-pipelineTypeScript core where Jev estimates probabilities and code decides.opsclassificationgithub★ 23
  • ThinkLessDecision plane routing routine judgments to rules or Jev first.guardrailsroutingclassificationgithub★ 34
  • ModexTypeSafeAI's open desktop app driving Claude Code and Codex agents.coding-agentgithub★ 43

New 2026-09-26

  • Logic Model + Jev (catgameresearch)Custom compiler pairs Jev with a local logic model for reasoning.researchdiscord2
  • JDE (Jev Decision Engine)Central place to ask, threshold and record an agent's Jev judgments.opsdiscord★ 22
  • harness-routerRoutes obvious tool calls fast; only asks Jev when genuinely ambiguous.routinggithub★ 213
  • ego-decision-layerReplaces per-step LLM browser-agent turns with one pluggable Jev decision.browserroutingdiscord★ 54
  • CriteriaCueChrome extension checks LinkedIn evidence against recruiting criteria using Jev Choice.classificationdiscord2
  • Graphify Jev judgment layers (PR)Adds optional Jev semantic judgment layers atop deterministic graph heuristics.dataclassificationdiscord★ 125.5k2
  • deepdocLocal deep-research tool where Jev reranks and checks RAG chunk coverage.ragdatagithub★ 3124
  • sttp-aiScala HTTP client wrapper covering OpenAI, Claude, Gemini and Jev.sdkgithub★ 1033
  • opencode-jev-compactorUses Jev's typed questions to decide what OpenCode checkpoints keep or drop.coding-agentsdkgithub★ 72
  • jev-gitSub-second git hook checks staged diffs for secrets and injection via Jev.coding-agentsecuritygithub★ 43
  • jev-codex-pluginCodex plugin for Jev tool/task consultation and completion-claim verification.coding-agentmcpgithub★ 122
  • jev-browser-controlJev picks each Chrome click for Claude or Codex browser automation.browsermcpgithub★ 54
  • wevOpen local System-One models: typed questions in, calibrated probabilities out.sdkbrowsergithub★ 623
  • Jev-LCTOpen decision model extracting calibrated confidence from recurrent computation, no RL.sdkgithub★ 62
  • opencode-auto-jevOpenCode plugin where Jev picks the real model for each turn.routingcoding-agentgithub★ 52
  • jev-dersleriTurkish lessons and prompt templates for using Jev in agent context gathering.coding-agentwritinggithub★ 42
  • JevopsSRE log triage tool using Jev to page, digest or resolve incidents.opsgithub★ 83
  • Jev Switchboard (LiveKit)Jev judges caller intent and manipulation before a voice agent replies.guardrailsgithub★ 42
  • QualmMac screen-time blocker where Jev classifies pages as learning, work or feed.classificationgithub★ 43
  • jev-qa-demosVideo-walkthrough demo code for using Jev in QA/SDET workflows.evaleducationgithub★ 32
  • jev-ragVector-free SQLite BM25 RAG reranked by Jev, no embeddings needed.ragsearchgithub★ 104
  • astrolabeJev judges conference-paper abstracts against your research topic, no embeddings.researchsearchgithub★ 22
  • Jev-PhoneControlAndroid automation where Jev picks actions from vision-agent candidates via ADB.browserdemogithub★ 42
  • ProofGate (serpapi-hackthon)SerpApi search, LLM ranks candidates, Jev verifies the best match.searchraggithub★ 22
  • hermes-adaptive-model-routerShadow-mode Jev router choosing fast vs capable model for Hermes Agent.routinggithub★ 22
  • jev-spatialJev-style System One model answering spatial questions as fixed choices.researchclassificationgithub★ 33
  • design-os-generative-uiSub-50ms generative UI using local Laya and cloud Jev cascade routing.demomediagithub★ 33
  • Jev Router (OpenRouter)Official OpenRouter model router powered by Jev, balancing quality, speed and cost.routingx5
  • typed-lmTurns dense LLMs into typed choice/bool/score APIs in one forward passclassificationsdkgithub★ 104
  • omp-laya-judgeLocal Laya judge answers in ~156ms vs 18s LLM judge, 10/12 accuratemcpevalclassificationgithub★ 74
  • compactioJev shrinks tool output 20501 to 747 chars in 0.75s for $0.00005coding-agentopsgithub★ 43
  • leanestSkips tests Jev judges unaffected by a code diff, defaults to classifier.devcoding-agentevalgithub★ 53
  • codex-context-dietJev trims bulky Codex tool output to a bounded head plus notecoding-agentopsgithub★ 63
  • taste-lintLints commits for AI slop, adds Jev design review notescoding-agentevalgithub★ 103
  • Jev-StyleLocal 0.5GB calibrated decision model, systemone-compatible server, Claude Code guard hookclassificationmcpgithub★ 174
  • jev-omni.jsBrowser WebGPU multimodal classifier matches fp32 on 293/234 benchmark questionsclassificationsdkgithub★ 24
  • lev (InterfazeAI)Open decision model, LoRA on Qwen3.5-4B, scores 68.9% on S1Benchclassificationevalgithub★ 224
    2 similar
    • HopperLoRA decision server on Qwen3.5-4B for JevBench with calibration temperatures.classificationresearchgithub★ 73
    • SelfJevSelf-hosted Jev-API-compatible decision model, Qwen3.5-4B+LoRA, single GPU.classificationsdkgithub★ 623
  • JevFlow (Parth1811)Claude Code plugin: Jev judges if a task phase is really donecoding-agentevalgithub★ 94
  • decision-models-under-pressureBenchmarks Jev vs Laya vs open models as candidate lists grow harderevalclassificationresearchgithub★ 43
  • Quiet FeedChrome extension hides X posts using your Jev key or local modelbrowserclassificationgithub★ 23
  • clouterClaude Code plugin routes image/video/speech prompts to Jev-ranked OpenRouter modelscoding-agentmediagithub★ 23
  • jev-claude (Panebianco00)MCP layer logs Claude Code's coding decisions as typed Jev choicescoding-agentmcpgithub★ 62
  • claude-x-jevClaude Code skill: Jev pre-sorts and gates calls before Claude readscoding-agentroutinggithub★ 532
  • pi-typesafe-approvePi extension auto-approves routine bash commands via Jev, escalates the restcoding-agentsecuritygithub★ 22
  • mevOffline local dashboard returns a chosen option plus confidence, no LLMclassificationcligithub★ 22
  • Jev Lab (Lyzr)Playground to compare Jev vs GPT, learn primitives, see documented failure modeseducationclassificationx3
  • LiteLLM Jev supportLiteLLM supports Jev natively; blog covers a context-compaction use casesdkopsx3
  • catherdAutopilot: Claude plans, Codex/opencode write code, Jev picks model per taskcoding-agentroutinggithub★ 82
  • jev-digestMCP tool returns only source passages answering a question.ragmcpdiscord★ 03
    1 similar
    • Sez Where?Points to relevant document passages for a question without summarizing.ragsearchdiscord3
  • DSPy native Jev supportDSPy signatures now run Noul, Choice and Score directly on Jev.sdkclassificationdiscord4
  • libtypesafeUnofficial C++17 client for the TypeSafe Jev API.sdkdiscord★ 02
    1 similar
    • libsemopC++17 library for typed Jev decisions, flows, and reranking.sdkclassificationdiscord★ 02
  • jev-oncallRoutes alerts to page, ticket, review or log via one Jev call.opsclassificationdiscord★ 43
  • jev-secretsPseudonymizes sensitive text in Jev requests, restores it in responses.securityguardrailsdiscord★ 03
  • composable-jevPHP library to fluently chain Jev calls into decision graphs.sdkdiscord★ 13
  • LLPhant JevClassifierPHP GenAI framework adds typed Noul/Choice/Score classification via Jev.sdkclassificationgithub★ 1.7k4
  • jevos-1bOpen local yes/no decision model, 50-220ms on a laptop CPU.classificationsdkgithub★ 1.3k4
  • system-one-connectorMCP connector giving coding agents typed Jev, Laya or CLM judgments.mcpgithub★ 3493
  • this-that-model1.9B typed-decision model, one forward pass, no decode loop.classificationgithub★ 393
  • quicksilverClaude Code skill offloads bulk judgment calls to Jev, cutting tokens 86%.coding-agentclassificationgithub★ 1204
  • jev-gui-delegateGUI delegation for Codex; Jev makes semantic picks, controller verifies steps.coding-agentbrowsergithub★ 1043
  • trufflerRails intent search: Jev labels records and queries, reranks results.searchraggithub★ 513
  • JevAnyOpen infra to fine-tune and serve your own Jev-style decision model.researchclassificationgithub★ 743
  • typesafe-ai-phpPHP SDK for TypeSafe's Jev evaluation API.sdkgithub★ 152
  • reelqlTurns any video into typed JSON so Jev can branch on it.mediaclassificationgithub★ 303
  • Mica-v0.1-4B4B decision model speaking Jev's /v1/systemone wire format.classificationgithub★ 363
  • jevonianLocal router enforces per-turn model choice in code, not prompts.routinggithub★ 194
  • hekajevClassifies and analyzes Git commit history with Jev, tracks cost.datagithub★ 63

New 2026-09-25

  • jev-router (Ex8-ca)Routes Hermes agent turns to skills using Jev decision confidence.coding-agentroutinggithub★ 52
  • jev-tmmluplus-evalBenchmarks Jev on 66-subject TMMLU+ with 100% parse rate.evalclassificationgithub★ 33
  • jev-browse (cooper667)Runs plain-English browser QA checklists judged by Jev on Cloudflare.browsercoding-agentgithub★ 74
    1 similar
    • jev-playwright (cooper667)Plain-English browser QA checklist judged by Jev on Cloudflare Workers, 35 seconds.coding-agentbrowserevalgithub★ 74
  • jev-pilotPicks Claude Code reasoning effort, subagent model and skill via Jev.coding-agentroutinggithub★ 103
    2 similar
    • spending-effort-with-jevClaude Code plugin uses Jev to recommend per-prompt effort level.coding-agentclassificationcligithub★ 74
    • jev-router (dominicrico)Routes each Claude Code task to the right model and effortcoding-agentroutinggithub★ 33
  • hearmemoryShares memory across coding agents, Jev verifies claims against evidence.coding-agentopsgithub★ 32
  • RLCDAlignBenchBenchmarks Jev as zero-shot detector of 10 alignment failure types.evalguardrailsgithub★ 263
  • codex-jev-router (suenot)Routes Codex subagents to model tiers using typed Jev decisions.coding-agentroutinggithub★ 83
  • QwevBuilds a Jev-style decision interface on local Qwen checkpoints, no training.classificationresearchgithub★ 32
  • MalkuthOpen multilingual decision model, Korean-focused alternative to Jev.classificationresearchgithub★ 52
  • jev-call-routerRoutes phone calls between PBX destinations using Jev intent classification.routingopsgithub★ 182
    1 similar
  • semantic-validatorChecks text meaning via Jev across Python, TS, Go, Java SDKs.guardrailsevalgithub★ 23
  • Jev Research IndexBilingual catalogue of 358 Jev projects, papers and public materials.researchdatagithub4
  • jev-browse (danielnc)Hands browser sub-tasks to Jev, cheaper and faster than driving directly.browsercoding-agentgithub★ 83
  • Jev_steer_or_queueLets Jev decide whether a mid-turn message steers, queues or interrupts.coding-agentroutinggithub★ 163
  • pi-pignonShifts Pi agent to cheaper or stronger LLM using a Jev difficulty check.coding-agentroutinggithub★ 53
  • ClearbookCategorizes bank statement transactions and fees using Jev, not keywords.financeclassificationgithub★ 23
  • JEV Orchestrator (TartanIMU)Bounds cognitive workflow routing and escalation with Jev supervision.routingopsgithub★ 42
  • System One Browser AgentCombines sub-150ms Jev reflexes with Stagehand and LLM fallback for browsing.browsercoding-agentgithub★ 42
  • openpave-jevCLI for typed choice, score and noul decisions via any System One model.cliclassificationgithub★ 23
  • grok-jev-guardGates Grok Bot tool calls with a typed Jev approval envelope.guardrailssecuritygithub★ 22
  • kiteLocal copilot picks which social posts and actions to prioritize via Jev.writingopsgithub★ 33
  • prompt-fossilsFlags stale prompt patterns in CLAUDE.md and AGENTS.md that hobble Opus 5.5.coding-agentevalgithub★ 24
  • Open System-1 (Readyaddy)Open-source System One model meant as a drop-in Jev replacement.classificationresearchgithub★ 22
  • jeverifierJev-gated browser and desktop automation harness that cuts context re-reading.coding-agentguardrailsdiscord★ 22
  • osenvJudges every coding-agent file write and command before it runs.coding-agentguardrailsdiscord3
  • Jev in PracticeInteractive course covering Jev foundations, building, reference and cookbooks.educationdiscord2
  • typesafe-ai-botDiscord moderation bot: AI suggests, plain code decides, community votes.guardrailsopsdiscord★ 02
  • HelmRoutes each coding task to the best installed AI coding agent.coding-agentroutingdiscord★ 172
  • jev-fault-diagnosis-pocReads a compressor manual, monitors sensors, suggests fault diagnosis when confident.classificationdatadiscord★ 32
  • muSmall Jev judge handles 35 non-code decisions per coding-agent turn.coding-agentevalgithub★ 4903
  • jev-socialJev chooses browser actions to research a topic across social platforms.browserdatagithub★ 1632
  • jev-code-reviewerClassifies agent PR diffs into P0-P2 so only critical changes surface.coding-agentevalgithub★ 1403
  • claude-architectClaude delegates coding to isolated agents, verifies evidence before merge.coding-agentguardrailsgithub★ 313
  • tinyjev596M-parameter offline Jev-style Choice/Score/Noul model, MLX or PyTorch.sdkclassificationgithub★ 273
  • jev-harness (AntonioCoppe)Confidence gates and shadow mode around Jev; a row-filter job in 1.3s.coding-agentevalgithub★ 153
  • CodifySingle-binary agent workflow engine indexing 19 languages, Jev for decisions.coding-agentopsgithub★ 4342
  • Jev Cookbook (datawhalechina)Chinese tutorial series on Jev primitives plus fine-tuning an open Laya.educationgithub★ 3422
  • jev-safety-gatewayReverse proxy between nginx and an LLM backend, Jev blocks harmful requests.securityguardrailsgithub★ 163
  • laya-appleRuns short Jev-client decisions on Apple Neural Engine beside a busy GPU LLM.sdkopsgithub★ 183
  • jev-sim-useJev navigates iOS Simulator and Android screens one call per step.browsermcpgithub★ 102
  • llama-index-jevDrop-in LlamaIndex reranker and router using Jev Score and Choice.ragroutinggithub★ 83
  • dsh-jev-interceptorJev classifies risk of every DeepSeek Harness tool call before approval.guardrailscoding-agentgithub★ 273
  • noulsSemantic linter: batches yes/no Jev questions per function for intent bugs.coding-agentevalgithub★ 83
  • jev-sap-commerceSAP Commerce extension: Jev moderates reviews, suggests categories, audits decisions.classificationdatagithub★ 102
  • wechat-triage-hudOffline OCR reads WeChat chats; Jev flags which conversations need you.opsclassificationgithub★ 252
  • jlinkRecord linkage in plain English; Jev scores match probability per pair.dataresearchgithub★ 63
  • datafusion-jevAdds a prompt_jev SQL function for typed Jev decisions inside DataFusion.datasdkgithub★ 73
  • zio-typesafe-aiTyped Scala 3/ZIO client for Jev, illegal states unrepresentable by design.sdkgithub★ 53
  • jev-raJev picks browser step and target for coding agents, 3-5x faster than browser-use.browsercoding-agentgithub★ 44
  • morning-tech-briefingDaily emailed tech briefing; Jev drops duplicate and off-topic stories.opswritinggithub★ 52
  • Ring Zero SecurityKernel-level syscall policy for coding agents, with Jev as optional check.securityguardrailsgithub★ 113
  • tiaGitHub issue triage agent; Jev makes one bounded decision per issue.coding-agentopsgithub★ 83
  • RevOpen Qwen-tuned decision models beat hosted Jev on accuracy, speed and cost.classificationresearchgithub★ 84
  • jev-anythingAgent skill that scaffolds a bounded Jev decision layer's contract and evals.coding-agentclassificationgithub★ 272
  • pi-jev-todo-auditPi extension: Jev audits todo-board drift every 10 agent loops.coding-agentevalgithub★ 32
  • jev-skillsJev picks which skills the model sees each turn, zero context cost.coding-agentroutinggithub★ 64
  • privatemode-decisionsOpen Jev-like typed decisions on GLM-5.3-Flash, confidential compute optional.classificationsdkgithub★ 254
  • model-router-pythonFilters LLM providers by cost and limits then Jev picks winner.routingsdkgithub★ 54
  • jevbriefCurates and traces the state fed into Jev decisions.dataclassificationgithub★ 63
  • Jev Model Router for Claude Code (AlexPEClub)Jev hook routes each Claude Code prompt to cheapest capable model.routingcoding-agentgithub★ 84
  • jev-ai-use-casesLangChain notebook: Jev for ticket triage, model routing, guardrails, fraud checks.classificationguardrailsgithub★ 54
  • Awesome-jev-papersCurated list of research papers and benchmarks about Jev.researchgithub★ 143
  • awesome-jev-surveyEvidence survey of Jev calibration, selective control and open implementations.researchevalgithub★ 74
  • tab-jevCombines Jev text judgments with a tabular in-context model for classification.classificationdatagithub★ 43
  • semif-goSelf-hosted Jev-style decision API over llama.cpp, text and multimodal input.classificationsdkgithub★ 73
  • multiplayer-coding-agentsLiveblocks demo where Jev decides if a coding agent reply is needed.coding-agentroutinggithub★ 103
  • simple-agent-collabsFile-based multi-agent research loop where Jev gates tier and duplicates.classificationresearchgithub★ 33
  • GStarsBrowser extension that reranks GitHub star search results with Jev.searchbrowsergithub★ 63
  • isResponsiveZeroClassifies documents against discovery requests using Jev, rubric built by DeepSeek.classificationdatagithub★ 13
  • jev-starterVisual console to configure, test and export Jev question calls.sdkcligithub★ 33
  • fireworks-jev-reward-rlTrains a writer LoRA with Jev scores as the RL reward signal.evalresearchgithub★ 33
  • jev-benchmark (YidiDev)Benchmarks Jev vs Claude Haiku/Sonnet vs OpenJev on rubric classification and cost.evalclassificationgithub★ 34
  • jev-java-sdk (tycoding)Java SDK for Jev-style decisions using the open-weight Laya ONNX model.sdkgithub★ 43
  • notiqAndroid app filters notifications by natural-language rules via Jev or FastJev.classificationopsgithub★ 42
  • toolgate-experimentBenchmarks MCP tool-selection accuracy and latency across 143 tools with Jev.mcpevalgithub★ 34
  • hello-zio-typesafe-aiCompares LLM tool orchestration designs against top-level Jev-controlled orchestration.mcpevalgithub★ 23
  • jev-delegateClaude Code skill: Jev rates task difficulty, picks Haiku, Sonnet or Opus.routingcoding-agentgithub★ 44
  • jev-no-bullshitJev flags unverified or vague coding-agent summaries and forces a recheck.guardrailscoding-agentgithub★ 33
  • opencode-jev-guardJev screens every OpenCode shell command for six risk types before running.guardrailssecuritygithub★ 23
  • fastvibeElectron coding-agent client on pi, with an optional Jev decision engine.coding-agentgithub★ 82
  • pastewisePaste box detects data format; falls back to Jev to classify unknowns.classificationdemogithub★ 53
  • jev-playwrightJev picks which Playwright tests are relevant to a code change.evalcoding-agentgithub★ 63
    1 similar
    • SedumRuns AI browser e2e tests on every PR, matched via Jev.coding-agentbrowserevaldiscord★ 114
  • DigestFrogJev scores scraped tech sources, LLM summarizes and emails top news.classificationdatadiscord2
    1 similar
    • morning-tech-briefingDaily emailed tech briefing; Jev drops duplicate and off-topic stories.opswritinggithub★ 52
  • fast-computer-useLocal Jev-like classifier drives computer use fully offline on 16GB M1 Mac.browsersdkdiscord★ 52
  • jev-patternsExtracts repeating patterns from noisy sequences using Jev to score candidates.dataclassificationdiscord★ 02
  • jev-lint (schalkneethling)Experimental semantic code linter that asks Jev to judge code quality.coding-agentclassificationdiscord★ 02
  • Football Agent news filterJev pre-filters news feeds before an LLM writes fantasy football advice.classificationragdiscord3
  • UsageTap Model Lifecycle AuditJev disambiguates LLM model keys to flag deprecated models needing replacement.routingopsdiscord3
  • TraceRoot Jev detector judgeOpen source agent tracer uses Jev as the detector judge backend.evalopsdiscord★ 7963
  • browser-agent (Fox Islam, PHP)PHP goal-driven browser agent built on jev-ultrafast, tuned for QA reporting.browserclidiscord★ 12
  • jev-guard (leepokai)Risk-scores every coding-agent tool call across Claude Code, Codex, Copilot, Cursor.guardrailscoding-agentgithub★ 634
  • Arize Jev remote evaluator cookbookCookbook wires Jev in as a remote evaluator for Arize AX.evalsdkgithub3
  • pi-classifier (bacnh85/pi-extensions)Pi coding agent package adds a Jev-gated auto-approve hook for shell.guardrailscoding-agentcligithub★ 333
  • jev-blindspotSide panel flags blind spots in Claude Code and Codex prompts.coding-agentgithub★ 123
  • ollayaOllama-style local runner serving Laya and other decision models via Jev's API.sdkroutingcligithub★ 1.3k4
  • vibecheck (jlowin)Python check/classify/label/score functions wrap decision models like Jev directly.sdkgithub★ 424
  • OpenCode Permission ReviewerAI reviewer allows, denies, or escalates every OpenCode permission request.guardrailscoding-agentgithub★ 163
  • AnchorJev extracts policy facts for Z3; unsure cases fall back to GPT.guardrailsclassificationgithub★ 103
  • jev-router (hajdu-patrik/claude-workspace)Routes each prompt to a model, effort level, and skill via Jev.routingcoding-agentgithub★ 153
  • GatekeeperJev routes ambiguous requests to the right agent skill via Claude Code hooks.routingguardrailscoding-agentgithub★ 203
  • FlickMCP server runs whole browser and macOS goals via observe-decide-act with Jev.browsermcpgithub★ 262
  • jev-claude-code (DarioFontanel)Three Claude Code prompts add Jev model routing, compaction, and code review.coding-agentroutinggithub★ 73
    1 similar
    • optim-jevClaude Code plugin compacting context, routing models and skills via Jevcoding-agentroutinggithub★ 64
  • jev-actionGitHub Action runs Jev judgments for PR label triage in CI.cliopsgithub★ 63
  • jev-applyJev selects which saved facts fill each job application field; Playwright submits.datacligithub★ 133
    1 similar
    • jev-job-searchJev fills and verifies job application forms; 100 applications in under 10 minutes for about $1.browsercoding-agentgithub★ 54
  • Reflex (Datadog Labs)Rust control loops let Jev recommend actions, guarded by declared invariants.opssdkgithub★ 313
  • EdwardWatches coding agent event streams and intervenes when behavior looks wrong.guardrailscoding-agentgithub★ 83
  • CUA-JEVJev picks concrete actions and execution routes for computer-use agents.browserroboticsgithub★ 192
    1 similar
    • OpenComputerUseComputer-use agent tool that can delegate action decisions to Jev or Clef.browsermcpdiscord★ 143

New 2026-09-24

  • jevbench (dhruvmehra)Benchmarks Jev against LLMs, BERT, Laya on accuracy, latency, costclassificationevalgithub★ 84
    1 similar
  • LichenOpen local replacement for Jev API, beats it on JevBench accuracyclassificationsdkgithub★ 584
    1 similar
    • Open System-1 (Readyaddy)Open-source System One model meant as a drop-in Jev replacement.classificationresearchgithub★ 22
  • dsh-jev-toolsJev prunes tool output, screens prompt injection, gates completion claimsguardrailscoding-agentgithub★ 144
  • jev-enforceChecks every Claude Code reply and edit against CLAUDE.md via Jevguardrailscoding-agentgithub★ 104
  • stingrayClaude Code Stop hook catches half-done turns, judged by Jevcoding-agentguardrailsgithub★ 64
  • CITY-JevTests whether Jev decisions hold up under rewording and reordering perturbationsevalclassificationgithub★ 64
  • ReflexBenchBenchmarks Jev, Laya and rivals on typed decisions with accuracy numbersevalclassificationgithub★ 64
    2 similar
    • JevBenchBenchmark scoring Jev-class models on intelligence, calibration, speed, cost.evalgithub★ 2553
    • sysone-benchByte-identical benchmark compares Laya, Jev and Qwen decision models directly.evalgithub★ 64
  • Jev OrchestratorJev sorts inbound Gmail, Slack, Discord messages before an agent repliesroutingclassificationgithub★ 93
  • JevSDSQLSQL database using Jev operators for classification, extraction, ranking and matchingragclassificationdatagithub★ 2383
  • jev-dimabsaJev Score primitive as zero-shot valence-arousal sentiment baseline for SemEval taskclassificationresearchgithub★ 73
  • Jev Crash CourseEleven runnable demos teaching Jev's Choice, Score and Noul primitiveseducationclassificationgithub★ 133
  • Career JournalLocal job-search tracker; Jev classifies inbox messages as job-search relatedclassificationopsgithub★ 63
  • awesome-open-system-oneCurated list of open System One models, benchmarks and calibration toolingclassificationresearchgithub★ 83
    2 similar
    • awesome-system-one (townie)Curated catalog of System One models, open Jev alternatives and evals.researchclassificationeducationgithub★ 53
    • Awesome Decision ModelsCurated list of System One decision model APIs, open weights, runtimes, SDKs and benchmarks.researchdatagithub★ 6363
  • jevsearchShadcn site search component; keyword pass then Jev re-ranks by intentsearchsdkgithub★ 83
  • codearia-sieveParses web pages into dates, units and chunks sized for decision modelsragdatagithub★ 53
  • hearimGo gateway serving Jev's API over Ollama, llama.cpp, vLLM via logprobssdkclassificationgithub★ 53
  • boss-auto-job-helperJev screens job listings for scam and dispatch traps on six axesclassificationopsgithub★ 102
    1 similar
    • jev-job-helperJev screens job listings on BOSS, cut review time 59% in a 300-job test.classificationopsgithub★ 23
  • wilwidChrome extension folds away web content you dislike, judged by Jevbrowserclassificationgithub★ 72
  • VybezTypeScript library batching Jev questions into ordinary async conditionalssdkclassificationgithub★ 72
  • QuestionsTyped decision SDK with Zod schemas, confidence gates and retries for Jevsdkclassificationgithub★ 52
    1 similar
    • VybezTypeScript library batching Jev questions into ordinary async conditionalssdkclassificationgithub★ 72
  • coderelayRoutes tasks between Claude Code, Codex CLIs; one mode uses Jevroutingcoding-agentgithub★ 52
    6 similar
    • AlloyJev routes tasks across Claude, Codex, Gemini and Antigravity by capabilityroutingcoding-agentgithub★ 33
    • HelmRoutes each coding task to the best installed AI coding agent.coding-agentroutingdiscord★ 172
    • catherdAutopilot: Claude plans, Codex/opencode write code, Jev picks model per taskcoding-agentroutinggithub★ 82
    • ModexTypeSafeAI's open desktop app driving Claude Code and Codex agents.coding-agentgithub★ 43
    • agent-routerPicks Cursor, Claude Code, Codex or OpenCode plus model via Jev.routingcoding-agentclidiscord★ 1133
    • coderelay (GeWu-academy)Routes coding tasks across Claude Code, Codex, Pi CLIs with Jevcoding-agentroutinggithub★ 53
  • ZiYor-JEVScopeDesktop app for locally training Laya models on Apple Siliconclassificationsdkgithub★ 62
  • ATHENAJev-judged cognitive control layer gates coding-agent actions go/verify/block.coding-agentguardrailsdiscord★ 22
  • J3vLightweight Jev runtime, aims to be k3s to Jev's k8s.sdkclidiscord★ 02
    1 similar
    • go-system-oneNative Go decision runtime hits 84.85% on JevBench's public subset.evalclassificationgithub★ 93
  • Jet Revenue Forecast (Orcaset)Jev decides contract renewal, ACV step-up and churn cause inline.financeclassificationdiscord★ 13
  • JevometryMeasures Fisher information and sensitivity of Jev decision outputs.evalresearchdiscord★ 12
  • iCoderAgentNative macOS/mobile coding agent with a built-in Jev key setting.coding-agentclidiscord3
  • jevtrimJev as context judge beats embedding retrieval in 14 of 16 cells.ragevaldiscord★ 54
  • Automated Inbox Zero with Jev (bounds.dev)Cut email label classification cost from $7.20 to $0.13 per 1,000.classificationopsdiscord3
  • jeval60 evaluators run in one Jev call for faithfulness and factuality.evaldiscord3
  • jev-compactionJev scores tool output for coding-agent context compaction, never rewrites text.coding-agentguardrailsdiscord★ 44
    2 similar
    • awesome-jev-compactionCompactor that only scores tool output, never invents summary textcoding-agentdatadiscord★ 03
    • fast-compactClaude Code command where Jev picks tool output worth keeping.coding-agentcligithub★ 64
  • JevXScans a codebase for hardcoded rules that should be Jev decisions.coding-agentmcpdiscord★ 63
    1 similar
    • jev-integration-evaluatorScans a codebase and evaluates where Jev integration would help.evalclassificationdiscord★ 12
  • TypeSafe Jev Coding Guide (Marktechpost)Notebook tutorial on typed decisions, calibrated confidence and speculative fan-out.educationsdkgithub★ 2.9k2
  • jevryDesktop browser where Jev picks the next click, an LLM plans.browsercoding-agentgithub★ 1233
  • eval-layerClaude Code skill generates a rubric and test set, judged by Jev.evalcoding-agentgithub★ 134
  • jevlint (Ice-Hazymoon)Plain-English yes/no lint rules for secrets, infinite retries, leaked errors.coding-agentsecurityguardrailsgithub★ 163
    1 similar
    • jevlintplain-English convention rules checked in the agent loopcoding-agentguardrailsdiscord
  • typed_evalsJev-judged faithfulness, relevancy and context checks, plus tool-call guards.evalragguardrailsgithub★ 164
    1 similar
    • jeval60 evaluators run in one Jev call for faithfulness and factuality.evaldiscord3
  • agent.run() (AgentRun)Workflow DSL where Jev makes focused decisions between scripted steps.routingsdkgithub★ 942
  • typesafe-sdk (Ruby)Ruby client for TypeSafe's System One API with typed Choice/Score/Noul helpers.sdkgithub★ 73
  • awesome-system-one (townie)Curated catalog of System One models, open Jev alternatives and evals.researchclassificationeducationgithub★ 53
  • keelLocal macOS coding workspace routing tasks via Laya or hosted Jev.coding-agentroutinggithub★ 3322
  • Jev Gatehouse (Kinde starter kit)Kinde identity plus Jev judges each MCP tool call in about 200ms.guardrailsmcpsecuritygithub★ 54
  • JevBlitzRankJev reranks k passages in one pass, 22x faster, 2.9x nDCG.ragsearchclassificationgithub★ 54
    1 similar
    • jev-rerankerPython library reranks and filters RAG search results for relevance using Jev.raggithub★ 433
  • Contrastive Language Models (CLM)Open decision model matching Jev with 9x lower latency, SOTA verifier.classificationevalresearchgithub★ 3k4
  • pi-jev-context-curatorJev prunes pi's LLM context per message, caching judgments to skip rechecks.classificationsdkgithub★ 72
  • Jev Guard Harness (dereknguyen269)Policy guard for coding agents; Jev resolves ambiguous cases, fails closed.guardrailssecuritycligithub★ 64
  • jev-skill-gateJev scores Claude Code skills, cutting manifest tokens by 75%.coding-agentevalclassificationgithub★ 85
  • jev-for-allJev picks skills, tools and browser moves for OpenCode and Claude Code.coding-agentroutingmcpgithub★ 153
  • ConvoyJev matches UI intent to accessibility elements for cross-platform e2e tests.evalbrowsergithub★ 63
    2 similar
    • b2b-web-testB2B web test runner where Jev picks each element and logs steps.browserevalgithub★ 22
    • jev-ultrafast AI UI TestsSelector-free UI tests that pick operations and elements via a local System 1 model.evalbrowsergithub★ 13
  • TypeSafe UIshadcn-style React components and blocks themed for TypeSafe Jev apps.sdkdemogithub★ 73
  • Jev x AI SDK Form RouterJev routes form submissions by context, falling back to GPT on uncertainty.routingdemogithub★ 123
  • jev-datafusionSQL functions run Jev noul, choice and score judgments inside DataFusion.datasdkgithub★ 122
    1 similar
    • datafusion-jevAdds a prompt_jev SQL function for typed Jev decisions inside DataFusion.datasdkgithub★ 73
  • Jev in Dify Question ClassifierJev now powers Dify's Question Classifier node for workflow routing.classificationroutingx2
  • Jev on Vercel AI Gateway (TypeSafe-compatible + HTTP API)AI Gateway adds a TypeSafe-compatible client API and plain HTTP API for Jev.sdkroutingx3
    1 similar
  • jev-sdk-goGo SDK for Jev's typed decision API.sdkdiscord★ 12
  • Jev Code CampInteractive lessons teaching Jev's Choice, Score, and Noul primitives.educationdiscord2
    1 similar
    • Jev in PracticeInteractive course covering Jev foundations, building, reference and cookbooks.educationdiscord2
  • safe-upgrade (Tenuo)Dependency-upgrade agent uses Jev decisions scoped by Tenuo authority.coding-agentguardrailsdiscord3
  • Autonomous SIEM Capstone (satya)eBPF kernel telemetry triaged by a TypeSafe AI autonomous-response swarm.securitydiscord3
  • Portage UCP decision moduleChecks agent shopping-cart, checkout, and currency correctness with Jev.mcpdiscord★ 22
  • Vane DataDuckDB-based multimodal engine pipes Whisper transcripts through Jev to SQL.datamediadiscord★ 1383
  • kassadBatches .NET guardrail policies into one Jev call for verdicts.guardrailssdkdiscord★ 04
  • jevelryTUI for making and tracking Jev-backed decisions in a terminal.clidiscord★ 112
    1 similar
    • s1-tuiTerminal UI for testing Jev and local Laya typed decisions side by side.cliclassificationgithub★ 23
  • Decision LabVisual playground for building and comparing Jev, OpenJEV, and Laya decisions.demoevaldiscord★ 02
  • BearingTurns AGENTS.md conventions into Jev-checked lint rules during agent turns.coding-agentguardrailsdiscord3
    2 similar
    • oh-my-plumbUses Jev to catch coding-agent edits that break unenforceable AGENTS.md rules.guardrailscoding-agentgithub★ 54
    • HapslandReviews coding agent's decisions in realtime against your AGENTS.md rules.coding-agentevaldiscord3
  • touchpressMobile e2e testing library now supports Jev as its judgment model.coding-agentevaldiscord★ 284
  • claude-herdr-jev-compactionJev judges Claude Code task completion to trigger compaction timing.coding-agentdiscord★ 03
  • classify-pluginClassifies Claude Code and OpenCode session messages into categories via Jev.coding-agentclassificationdiscord★ 13
    1 similar
    • Jev_steer_or_queueLets Jev decide whether a mid-turn message steers, queues or interrupts.coding-agentroutinggithub★ 163
  • thesys-coreResearch workspace uses Jev to compare two papers on a claim.ragresearchgithub★ 1682
  • JevK5Open-weight typed-decision model; ranks second of 76 on JevBench v1.4.classificationevalgithub★ 1474
    5 similar
    • Contrastive Language Models (CLM)Open decision model matching Jev with 9x lower latency, SOTA verifier.classificationevalresearchgithub★ 3k4
    • AutoJev-27BOpen 27B decision model beats Jev's accuracy and calibration on their eval.classificationevalgithub★ 1304
    • AgentJev-0.6BOpen-weight 0.6B decision model, calibrated probabilities in one ~50ms forward pass.classificationsdkgithub★ 3793
    • DeepOpenOpen non-autoregressive decision engine claiming higher accuracy than Jev 1.13, 33ms.classificationevalsdkgithub★ 2k3
    • OpenJev (AlexWortega)Open source model weights reimplementing Jev's System One primitivesclassificationresearchdiscord3
  • jev-judge-mcpMCP server exposing eleven Jev-backed judgment tools for coding agents.mcpcoding-agentgithub★ 724
  • ValenVision-capable System One model; solves puzzles far faster than a 27B LLM.researchclassificationgithub★ 7463
  • jev-agent-design-with-topk-logits-choicesResearch design for a Jev-native agent using top-k logit refinement.researchgithub★ 192
  • sys1Rust server implements a System One-compatible API for open decision models.classificationsdkgithub★ 533
  • SNAPLocal Rust engine returns Jev-compatible typed decisions from one forward pass.classificationgithub★ 123
    3 similar
    • jev-rsRust engine answers typed Jev-compatible decisions from any local LLM.sdkcoding-agentgithub★ 173
    • kimeRust inference engine speaking Jev and Laya's wire formats.sdkgithub★ 73
    • basal-rsRust inference engine implementing System One API for Basal/Bielik models.sdkclassificationgithub★ 73
  • GhosthandWindows desktop assistant picks UI actions with Jev, pausing before risky steps.opsdemogithub★ 532
    2 similar
    • Jev Computer UseVoice controls a Windows PC; Jev picks the app or action meantdemoclassificationgithub★ 22
    • POK-AgentWindows computer-use agent uses Jev for actions, caching successes as motor programs.coding-agentdemodiscord★ 14
  • go-jev (mattn)Go SDK and CLI for Jev's yes/no, choice, and score questions.sdkcligithub★ 452
    3 similar
    • jev-sdk-goGo SDK for Jev's typed decision API.sdkdiscord★ 12
    • jev-goIndependent Go SDK for the Jev API.sdkdiscord★ 12
    • typesafe-sdk-goOfficial-style Go SDK wrapping the TypeSafe Jev API.sdkgithub★ 92
  • jevperJev-shaped classification wrapper runs on any OpenAI-compatible client, even local llama.cpp.sdkclassificationgithub★ 193
  • awesome-jev (Ai-trainee)Curated Jev use-case list packaged as a loadable agent skill.educationdemogithub★ 142
  • TypeSafeAI .NET SDK (saibimajdi)Community .NET client for Jev's noul, choice, and score questions.sdkgithub★ 122
    3 similar
    • TypeSafe.AI .NET SDKC#/.NET SDK bringing Noul, Choice, Score to .NET apps.sdkdiscord★ 23
    • TypeSafeSharp.NET client for TypeSafe's System One API with DI and OpenTelemetry support.sdkgithub★ 52
    • TypeSafeJevPlaygroundUnofficial .NET port of TypeSafe's Jev SDK with MCP server guide.sdkgithub★ 92
  • LeJudgeJev judges natural-language constraints over JEPA world-model rollouts.roboticsresearchgithub★ 103
  • GutElixir DSL routes LLM or Jev decisions into pattern-matched control flow.sdkgithub★ 182
  • decidrTyped decisions from local Ollama models in one forward pass.classificationsdkgithub★ 83
  • DiffJuryJev scores GitHub PR risk and review depth from the diff.coding-agentevalgithub★ 73
    2 similar
    • wincescores diffs by blast radius and auth/data-write for review routingsecuritycoding-agentroutinggithub★ 2
    • jev-code-reviewerClassifies agent PR diffs into P0-P2 so only critical changes surface.coding-agentevalgithub★ 1403
  • Jev Auto Router (miniLV)Jev picks a GPT model tier per Codex call, verified after.routingcoding-agentgithub★ 82
  • JoltmacOS launcher built on Jev finds files and controls apps.clidemogithub★ 122
  • pi-jev-contextPi extension dedupes reads and filters command logs with Jev.coding-agentgithub★ 103
  • vault-tag-systemJev suggests Markdown vault tags per file against your taxonomy.ragclassificationgithub★ 102
  • panpan-jd-lensAndroid overlay flags risky job-post clauses using Jev and local rules.classificationgithub★ 62
  • Jev.NxRuns open decision models as an in-process Jev backend on Nx.sdkclassificationgithub★ 93
  • jev-ivrJev classifies intents for a voice IVR healthcare-scheduling phone flow.classificationgithub★ 63

New 2026-09-23

  • muse-jev-playbookPlaybook and skill for gating expensive agent steps behind cheap Jev calls.guardrailscoding-agentgithub★ 253
  • jev-suiteFour Java tools wrap Jev judgments with deterministic thresholds and measured calibration.classificationevalgithub★ 274
  • laya-jev-GraphRAGGraphRAG pipeline scores and routes graph edges with swappable Laya or Jev.ragclassificationgithub★ 514
  • SafePyramidBenchmarks in-context guardrail policies; Jev-1.13 scores 54.8 RMR at $1.19.evalguardrailsgithub★ 114
    1 similar
    • laya-security-guardrailLocal security guardrail suite comparing Laya, Jev and LLMs on latency.guardrailssecuritygithub★ 124
  • swift-jevSwift library and CLI for typed Jev questions and answers.sdkgithub★ 162
    1 similar
    • JevSwiftSDKunofficial Swift SDK, async/await, batching, zero dependencies, iOS/macOS/Linuxsdkgithub★ 83
  • jev-harness (TypeSafeAI)LLM proposes an action, Jev answers questions, code decides, logs receipts.coding-agentguardrailsgithub★ 303
  • EdgeJevRuns Jev-compatible decisions locally on CPU in 15.6ms, no cloud.sdkclassificationgithub★ 214
  • jev-rsRust engine answers typed Jev-compatible decisions from any local LLM.sdkcoding-agentgithub★ 173
  • ego-jevGives a browser agent a 0.4s typed decision per DOM step.browsergithub★ 113
    2 similar
    • jev-raJev picks browser step and target for coding agents, 3-5x faster than browser-use.browsercoding-agentgithub★ 44
    • ego-decision-layerReplaces per-step LLM browser-agent turns with one pluggable Jev decision.browserroutingdiscord★ 54
  • jev-mcp (rashedInt32)MCP server exposes Jev's classify, score and check as Claude Code tools.mcpcoding-agentgithub★ 74
  • PaperFocusPDF reader highlights passages answering your question, scored by Jev.researcheducationgithub★ 83
  • dsh-jev-pluginLets DeepSeek Harness agents call Jev's choice, score and noul as tools.coding-agentgithub★ 82
  • jev-recipes87 TypeScript recipes for routing, verification and labeling backed by Jev calls.classificationraggithub★ 204
    1 similar
    • jev-demo (gopinav)TypeScript SDK examples of Jev best practices and common mistakes.sdkeducationgithub★ 203
  • jev-harness (ismaelsoilet)Semantic gates using Jev stop coding agents wasting tokens on trivial errors.coding-agentguardrailsgithub★ 114
  • gpui-agentJev picks native GPUI UI actions from an accessibility tree, no screenshots.coding-agentgithub★ 192
  • code-qualityGit hooks block test tampering; optional Jev classifies suspicious test hunks.coding-agentguardrailsgithub★ 164
  • jev-bigquery-cloudrunClassifies BigQuery support tickets by team and urgency via Jev.classificationdatagithub★ 93
  • laya-browser-agentLocal browser agent uses open Laya weights as a Jev drop-in replacement.browsergithub★ 203
  • midscene-jev-runnerRuns Jev-style typed decisions inside Midscene browser test scripts via OpenRouter.evalbrowsergithub★ 112
  • pi-shift-routerRoutes Pi coding agent turns between cheap and strong models, judge optional.routinggithub★ 123
  • rulingLocal open model matches Jev's accuracy on 256 judgments, no significant difference.classificationevalgithub★ 104
  • tool-prunePrunes and ranks agent tool schemas offline or via Jev by confidence.mcpcoding-agentgithub★ 73
    3 similar
    • hermes-tool-slimmerTrims tool schemas sent per turn; Jev mode cuts tokens 71%.coding-agentsdkgithub★ 394
    • jev-mcp-router (ini8labs)Jev answers yes/no per MCP tool so the LLM only sees a few, cuts routing cost 32x.mcproutinggithub★ 34
    • dsh-tokenslashUses Jev to prune unneeded tool schemas from coding-agent system prompts.opscoding-agentgithub★ 53
  • mcts-agentMonte Carlo tree search scores states with Jev, generates moves with Gemini.researchclassificationgithub★ 73
    1 similar
    • ThoughtZero (pyschic-vla)Jev judge scores math tree-search steps; code complete, not yet run at scale.researchclassificationgithub★ 31
  • decisions-judge-mcpMCP tool giving coding agents noul/choice/score judgments with graceful fallback.mcpcoding-agentdiscord★ 54
  • System One skill for Jev and Laya (agent-skills PR)Adds progressive-disclosure Claude Code skill covering Jev and Laya decision models.coding-agenteducationdiscord★ 1203
  • decision-queryAdds Jev noul/choice/score as SQL functions for SQLite and Postgres queries.sdkdatadiscord★ 24
  • JevGateCI code-review gate: Jev answers typed per-function questions, flags CWEs with probability.coding-agentsecuritydiscord★ 94
  • awesome-jev-robustnessIndexes 109 independent robustness tests on Jev calibration, injection, and abstention.evalresearchdiscord★ 74
  • learn-jev-end-to-endFree course with 13 runnable Jev notebooks benchmarked against frontier LLMs.educationevaldiscord★ 174
    3 similar
    • Jev Crash CourseEleven runnable demos teaching Jev's Choice, Score and Noul primitiveseducationclassificationgithub★ 133
    • Ten Levels of JevTen worked examples of Jev from plain code to autonomous coding agent.coding-agenteducationgithub★ 2184
    • Jev Engineering CookbookOfficial 60-recipe notebook curriculum for Jev's Choice, Noul, Score patternseducationsdkgithub★ 84
  • dsh-plugin-jevJev gates every command in deepseek harness for danger before execution.guardrailscoding-agentdiscord★ 123
  • OptimusRust coding harness uses Jev for tool selection, retries and verification.coding-agentroutingdiscord★ 143
  • JEV Paper RadarJudges arXiv papers against plain-English interests; 501 papers in 33s for $0.02.researchevaldiscord★ 304
    1 similar
    • astrolabeJev judges conference-paper abstracts against your research topic, no embeddings.researchsearchgithub★ 22
  • Jev Builds (openchamber)Semantic search over 113k X posts tracking what people build with Jev.searchdatadiscord3
  • jevbooksCurates 500+ Jev projects and tags 16 design patterns via typed questions.researchclassificationdiscord3
  • TypeSafe Ask AI (kapa.ai)Docs Q&A agent over TypeSafe's docs and code, via widget or MCP.mcpragdiscord3
  • jev-learning-labSelf-contained 42-lesson notebook course for building auditable Jev-based agents offline.educationcoding-agentdiscord★ 12
  • Jev Memory BenchmarkBenchmarks Jev as a memory/knowledge-graph engine against Gemini on speed, cost.ragevaldiscord3
  • SuperQodeHarness layer routes coding-agent tools via Jev with progressive tool discovery.coding-agentroutingdiscord3
    3 similar
    • muSmall Jev judge handles 35 non-code decisions per coding-agent turn.coding-agentevalgithub★ 4903
    • harness-routerRoutes obvious tool calls fast; only asks Jev when genuinely ambiguous.routinggithub★ 213
    • Meta-ArchitectQuality-gate workflow layer uses Jev for autonomous routing across coding agents.coding-agentopsgithub★ 52
  • Agent Squad (JevClassifier)Multi-agent orchestration framework with a built-in Jev classifier for intent routing.routingsdkgithub★ 7.8k4
  • milvus-modelVector-DB embedding/reranker library adds TypeSafe Jev alongside OpenAI, Cohere, Voyage.ragsdkgithub★ 614
  • Warren DufferLive intraday trading bot uses Jev to rank Nifty 50 every 15s.financeroutinggithub★ 982
  • laya-serverSelf-hosted Laya server exposes a Jev-compatible API for local decision inference.sdkopsgithub★ 984
    3 similar
    • arbiterSelf-hosts typed-decision models like Laya behind a Jev-compatible API.sdkclassificationgithub★ 343
    • ollayaOllama-style local runner serving Laya and other decision models via Jev's API.sdkroutingcligithub★ 1.3k4
    • usejevRuns Laya locally on Bun with ONNX and a TypeSafe-compatible API.sdkclassificationgithub★ 133
  • system1-agentsPrebuilt Jev/Laya agents for browser, computer use, robotics; up to 6x faster.browserroboticsgithub★ 1324
  • Intent-RouterCompiles vague agent requests into typed IntentSpecs for Jev and Laya routers.routingclassificationgithub★ 9323
  • jev-edgeNginx/OpenResty gateway filter uses Jev to catch prompt injection before reaching backends.guardrailssecuritygithub★ 414
    1 similar
    • jev-safety-gatewayReverse proxy between nginx and an LLM backend, Jev blocks harmful requests.securityguardrailsgithub★ 163
  • JevTreeComposes local Jev judgments into a multi-step probability tree for planning.routingclassificationgithub★ 413
  • Jev Browser (openqa-cn)Jev picks controls from a page index, Playwright acts on them.browsercoding-agentgithub★ 873
  • stuntdLocal proxy learns typed Jev decisions and serves them via a distilled Laya head.sdkopsgithub★ 704
    2 similar
    • JevstillerDistills a repeated Jev classification call into a local model live.routingclassificationgithub★ 663
    • doppProxies Jev-shaped calls, logs them, trains a small owned model as fallback.routingsdkgithub★ 23
  • LogJevWraps any logprob-capable LLM into Jev's choice/score/noul decision API.classificationsdkgithub★ 113
  • llm-typesafeSimon Willison's LLM plugin adds Jev classification and scoring commands.cliclassificationgithub★ 274
  • cmd-mod-jev-nudgeCommand Code mod: Jev judges if a stalled agent needs a nudge.coding-agentopsgithub★ 173
  • dejevuBrowser agent skips Jev, uses open model logprobs; claims faster than jev-ultrafast.browserdemogithub★ 112
    4 similar
    • laya-browser-agentLocal browser agent uses open Laya weights as a Jev drop-in replacement.browsergithub★ 203
    • laya-ultrafastPorts browser-use's jev-ultrafast to local Laya, removing the per-step API call and cost.browserroutinggithub★ 2532
    • fast-browser-useLocal open-weight browser automation reproducing Jev's fast decision pattern.browsergithub★ 2212
    • fast-computer-useLocal Jev-like classifier drives computer use fully offline on 16GB M1 Mac.browsersdkdiscord★ 52
  • Aside JevBrowser extension where Jev picks the next action inside Aside's REPL runtime.browsergithub★ 82
  • claude-jevClaude Code plugin: Jev classifies prompts and judges agent edits.coding-agentroutinggithub★ 234
    1 similar
    • ai-agents-masterclass-modsToggleable Claude Code plugin adds Jev classify/filter/judge mode on demand.classificationcoding-agentgithub★ 42
  • LevLaya-style decision engine: encoder scores 61-67%, thinking model 95%, on hard set.classificationevalgithub★ 283
  • One SystemSingle API routes decisions between local and hosted System One models.routingclassificationgithub★ 93
    1 similar
    • jigorRust gateway running local decision models or Jev, one protocol.routinggithub★ 23
  • AletheiaFastAPI AI-text detector combines two detectors plus Jev for classification.classificationwritinggithub★ 63
  • Codex SiftRoutes each Codex turn to cheapest model lane, judged by Jev.routingcoding-agentgithub★ 84
  • jev-ood-calibrationIndependent calibration test: Jev overconfident out-of-distribution, like a fine-tuned classifier.evalclassificationgithub★ 64
  • Spam Detector Telegram BotTelegram spam bot uses Jev judgments, falls back to local SVM.classificationsecuritygithub★ 62
  • ReflexGateLocal sub-second gateway for agent loop decisions, not TypeSafe's Jev.guardrailsopsgithub★ 62
  • gg-friggin-ezMultilingual profanity/toxicity screener catches romanized Indic slang in 50-500ms.guardrailsclassificationgithub★ 73
  • BlackroseTyped System One checks gate LLM input/output as allow/review/block.guardrailssdkgithub★ 64
  • jevcCompiles CLAUDE.md-style rules into typed Jev questions plus a code verdict.guardrailscoding-agentgithub★ 84
  • poorjevLocal offline Jev alternative; recalibration cuts ECE from 0.170 to 0.071.classificationguardrailsgithub★ 103
  • AugustusAgent skill for building and hill-climbing decision-model systems beyond Jev.evalclassificationgithub★ 113
    1 similar
    • jev-anythingAgent skill that scaffolds a bounded Jev decision layer's contract and evals.coding-agentclassificationgithub★ 272
  • jev-trustLogs Jev calls, measures real calibration per domain, flags overconfidence.guardrailsevaldiscord★ 1.3k4
  • Prompt RejectorScreens tool descriptions and prompts for injected malicious instructions with Jev.guardrailssecuritymcpdiscord★ 24
  • SoterDiscord moderation bot uses Jev to flag hate speech and spam.guardrailsclassificationdiscord★ 42
  • PyroEarly guardrails framework built on Jev for agent safety checks.guardrailsdiscord★ 42
    1 similar
    • jesOpen-source guardrails using decision models to block prompt injection, jailbreaks and PII leaks at every agent step.guardrailssecuritygithub★ 144
  • Jev in Claude Code via OpenRouter (video)Video walkthrough wiring Jev into Claude Code through OpenRouter, no waitlist.coding-agentroutingdiscord3
  • CrumbChrome extension uses Jev to hide irrelevant Facebook Marketplace listings.classificationdatadiscord3
  • jev-subtitle-translatorJev flags subtitle translations needing human review; caught all 65 injected errorsmediaclassificationdiscord★ 93
  • ValidatorRust CLI compares classifier outputs against a golden dataset to catch regressions.evalclassificationdiscord★ 33
  • drizzle-jevDrizzle ORM query methods that classify and score rows using Jev.dataclassificationdiscord2
  • Heym Jev model routerJev picks which LLM handles each request; caches repeated decisions across turns.routingdiscord★ 1.2k3
  • FilelatheJev decides whether to inspect a file or invent a viewer.classificationdemodiscord2
  • signal-weaveTyped Jev decisions for BI operational signals, faster and cheaper than Luna.dataclassificationdiscord★ 43
  • jeffJev picks model tier and parallelism for each Claude Code subtask.coding-agentroutingdiscord★ 14
  • jeff-cliCLI to query Jev choices, scores and rankings from the terminal.clisdkdiscord★ 33
  • Jev Skill Selection benchmarkJev beats Haiku and Laya on skill selection: 7x faster, 16.8x cheaper.evalroutingdiscord4
  • typesafe-jev-pluginClaude plugin builds reusable Jev classifiers, runs them on CSV/Excel data.classificationcoding-agentdiscord★ 13
  • minecraft-jev-distillationDistilled Jev's Minecraft combat decisions into LightGBM: 1ms latency, 90.8% agreement.roboticsgameclassificationdiscord★ 13
  • XJEVBoostPicks which rows and columns to send Jev, cutting tabular classification tokens.classificationdatadiscord★ 14
  • TutorNatural-language Magic card search; Jev decides when deeper reasoning is needed.searchclassificationdiscord2
  • RuVector / @ruvector/typesafeLocal embedding-based choice/score/noul decisions, served through a Jev-compatible API.classificationsdkgithub★ 4.6k4
  • Astra-AresJev adapts GPT-6 Astra's reasoning effort mid-task to cut tokens.coding-agentroutinggithub★ 3053
  • R.A.I.N. LabMulti-agent research tool; optional Laya/Jev providers propose bounded, policy-checked decisions.researchraggithub★ 562
    1 similar
    • simple-agent-collabsFile-based multi-agent research loop where Jev gates tier and duplicates.classificationresearchgithub★ 33
  • jev-browser-skillClaude Code skill where Jev drives scoped Playwright browser actions.browsercoding-agentmcpgithub★ 383
  • agent-chaperoneMCP proxy screens agent tool calls and results with calibrated probabilities.guardrailsmcpsecuritygithub★ 224
  • notjevGets Jev-style verdicts from any OpenAI-compatible endpoint via single-token logprobs.classificationsdkgithub★ 224
    2 similar
    • LogJevWraps any logprob-capable LLM into Jev's choice/score/noul decision API.classificationsdkgithub★ 113
    • jevperJev-shaped classification wrapper runs on any OpenAI-compatible client, even local llama.cpp.sdkclassificationgithub★ 193

New 2026-09-22

  • JimothyTrains small local classifiers from Jev-labeled examples, runs offline.classificationsdkgithub★ 753
  • Jev Engineering (Chinese translation)Chinese translation of TypeSafe's coding-agent harness design notes.coding-agentwritinggithub★ 1613
  • MedJevExtracts clinical variables from notes via typed questions, local GPU.classificationhealthgithub★ 1124
  • JevBERTPoC local BERT backend mimicking Jev's typed-decision API, uncalibrated.sdkclassificationgithub★ 332
  • solar-mini4-jevWraps Upstage Solar Mini4 behind Jev's System One API shape.sdkroutinggithub★ 453
  • HookMeterChrome extension scores social post hooks in real time via Jev.writingbrowsergithub★ 222
    1 similar
    • Radar de HookBrowser extension marks Instagram/X/YouTube posts worth copying using your own Jev keybrowsermediagithub★ 21
  • QwenJevAccelerates a vision-language model to 169ms median classification latency.classificationdatagithub★ 1043
  • Nerve (hermes-nerve)Supervisory watchdog layer over Hermes agents using typed Jev decisions.coding-agentguardrailsgithub★ 383
  • jevyoumeanWraps any CLI, uses Jev to match mistyped subcommands by intent.cliclassificationgithub★ 163
  • MetaCogJev judge picks the best reasoning path among sampled candidates.coding-agentclassificationgithub★ 253
  • jevcoreTyped Jev decisions for DeepSeek Harness and MCP hosts, offline default.mcpsdkgithub★ 1063
  • JevTunerTrains LMs to output calibrated decision probabilities in one pass.classificationresearchgithub★ 153
    3 similar
    • LuceOpen recipe to train your own calibrated Choice/Score/Noul decision model.classificationevalsdkgithub★ 83
    • JevAnyOpen infra to fine-tune and serve your own Jev-style decision model.researchclassificationgithub★ 743
    • Unsloth Decision Model TrainerFree notebook fine-tunes your own small decision model like Jev locally.sdkclassificationx4
  • AutoJev-27BOpen 27B decision model beats Jev's accuracy and calibration on their eval.classificationevalgithub★ 1304
  • cliany-siteExperimental Jev-based intent matching locates browser elements without acting.browsercligithub★ 102
  • jevcacheOpenAI-compatible proxy skips model calls when Jev says same intent.routingopsgithub★ 124
  • systemANERuns a typed decision model on Apple's Neural Engine, 1.4ms per call.classificationsdkgithub★ 153
    2 similar
    • laya-appleRuns short Jev-client decisions on Apple Neural Engine beside a busy GPU LLM.sdkopsgithub★ 183
    • system-one-foundation-modelsSwift bridge runs Jev-style decision models on-device in 15-150ms.sdkclassificationgithub★ 993
  • jev-agent-browserJev picks the next browser action; agent-browser executes it, bounded.browsercoding-agentgithub★ 133
    1 similar
    • JevOnlyBrowser agent that only uses Jev to pick actions, no LLM, completes multi-step web tasks.browserdemogithub★ 52
  • PDF RaceBenchmarks Jev vs Gemini for document QA: same accuracy, lower cost.ragevalgithub★ 134
  • AnyDecisionModelSwift package for typed decisions, backed by MLX or Jev.sdkclassificationgithub★ 152
  • jev_antispam_botMinimal Telegram bot deletes spam using Jev's yes/no judgment.classificationguardrailsgithub★ 163
    1 similar
    • Spam Detector Telegram BotTelegram spam bot uses Jev judgments, falls back to local SVM.classificationsecuritygithub★ 62
  • jevmoryAudits a coding agent's MEMORY.md against session evidence via Jev.coding-agentclassificationgithub★ 114
  • zcode-gatekeeperExternal approval gate reviews each agent tool call against current task.guardrailscoding-agentgithub★ 93
  • LuceOpen recipe to train your own calibrated Choice/Score/Noul decision model.classificationevalsdkgithub★ 83
  • jev-auto-approveGitHub Action approves PRs only when Jev clears every safety question.coding-agentguardrailscligithub★ 93
  • Jev-MemSystem-One controller for agent memory: +11% quality, 6.6x faster build.ragdatagithub★ 1984
    1 similar
  • Inbox Run (awesome-one)Jev answers four typed questions per email to triage a mailbox fast.classificationopsgithub★ 123
  • jev.nuNushell module pipes data into Jev and filters or sorts on scores.sdkcligithub★ 83
  • jev-mcp (arunav25)MCP server exposing Jev plus a harness scoring it against LLM judges.mcpevalgithub★ 83
  • precogJev predicts which link you'll click and prefetches just that one.routingdemogithub★ 83
  • judge-auditShadow-mode calibration audits for any AI judge: ECE, coverage, drift, cost.evalguardrailsgithub★ 114
  • CrammedJev scores flashcard distractors for plausibility, beating Quizlet's wrong-answer picks.educationclassificationdiscord3
  • jev-goIndependent Go SDK for the Jev API.sdkdiscord★ 12
  • jev_project_contextLong-term experiment memory skill for coding agents, audited with Jev.coding-agentdatadiscord★ 93
  • jev-cleanDeterministic data-cleaning edits gated by Jev's risk probabilities, full audit trail.dataguardrailsdiscord★ 23
  • dsh-jev-guardPre-execution hook blocks destructive commands using Jev, 90.4% on 114 cases.guardrailssecurityclidiscord★ 04
    2 similar
    • dsh-plugin-jevJev gates every command in deepseek harness for danger before execution.guardrailscoding-agentdiscord★ 123
    • dsh-jev-interceptorJev classifies risk of every DeepSeek Harness tool call before approval.guardrailscoding-agentgithub★ 273
  • jev-enterprise-decision-fabric.NET decision architecture with a 111-case labelled benchmark vs Claude baseline.evalguardrailsdiscord★ 03
  • AttuneJev scores user text for mood, patience or reliance.classificationhealthdiscord2
  • jev-benchmarks (thisisandreeeee)Benchmarks Jev against supervised and zero-shot classification baselines.evalclassificationdiscord★ 03
  • lossless-rewriteChecks LLM rewrites for dropped ideas with Jev and repairs them.writingclassificationdiscord★ 13
  • HunkpickEnumerates every valid merge-conflict resolution, Jev picks, code gates it.coding-agentclidiscord★ 13
  • opencode-foremanWorkflow runtime for OpenCode where agents pick their next step via Jev.coding-agentroutingdiscord★ 12
    1 similar
    • jev-agent-controlOpenCode plugin routes Plan/Build agents and handles handoffs using Jev.coding-agentroutinggithub★ 33
  • ExpensentUses Jev to classify receipt and invoice emails instead of an LLM.classificationopsdiscord2
  • SwitchboardJev picks the model and reasoning effort per coding task.routingcoding-agentdiscord★ 194
  • Homing PigeonLocal Gmail cleanup tool; Jev classifies senders into 7 categories.classificationopsdiscord★ 13
  • AgentScope32k-star agent framework now routes to a Jev classifier model natively.routingsdkcoding-agentgithub★ 33.1k4
  • memsearchCross-platform agent memory layer with optional Jev reranking of search results.ragdatagithub★ 2.7k4
    4 similar
    • hippo-memoryDecaying agent memory layer with optional Jev reranker, SQLite backedragmcpdatagithub★ 7763
    • MnemonFast Jev judgments filter memory search hits for agents under 4k tokens.ragclassificationgithub★ 43
    • Deep RecallJev reranks Markdown note passages for a Claude Code memory plugin.ragsearchcoding-agentgithub★ 124
    • aiduMEIMemory engine's optional Jev classifier cuts retrieval latency 78%, raises accuracy.ragdatagithub★ 204
  • DeepOpenOpen non-autoregressive decision engine claiming higher accuracy than Jev 1.13, 33ms.classificationevalsdkgithub★ 2k3
  • AgentJev-0.6BOpen-weight 0.6B decision model, calibrated probabilities in one ~50ms forward pass.classificationsdkgithub★ 3793
  • webctlCLI web search for agents; Jev scores and dedupes results.searchcoding-agentcligithub★ 1583
    1 similar
    • jev-search-mcp (mukiwu)Jev-powered web search as an MCP server, Claude Code plugin and CLI with zero dependencies.mcpsearchcoding-agentgithub★ 74
  • RakitsuGo agent IDE with visual builder and debugger, includes a Jev tool type for step-gating.coding-agentguardrailsdiscord★ 32
  • llm-debugger-vscode-extensionVSCode extension has Jev pick each debugger step, waking the LLM only for hypotheses and fixes.coding-agentevalgithub★ 3613
  • JevmindLocal-first agent decision layer with confidence gates, hash-chained ledger, optional Jev escalation.guardrailscoding-agentgithub★ 1833
  • jev-webmcp-extensionChrome extension uses Jev to select and fill WebMCP tool calls from page schemas.mcpbrowsergithub★ 1303
  • RentalCore Jev OCR matchingEvent rental system uses Jev via OpenRouter to match OCR invoice lines to catalog items.dataopsgithub★ 692
  • typesafe.proDrop-in HTTP endpoint compatible with Jev's API, supports anonymous no-key access.sdkopsgithub★ 672
  • AgentDirLocal flight recorder for coding agents with optional Jev-based memory reranking.coding-agentopsgithub★ 233
  • CheshimacOS workspace for Codex with Jev-powered conversation memory search.coding-agentdemogithub★ 182
  • laya-ultrafastPorts browser-use's jev-ultrafast to local Laya, removing the per-step API call and cost.browserroutinggithub★ 2532
  • StarguideRAG docs chatbot uses Jev to pick and judge candidate documentation pages before generation.ragsearchgithub★ 173
  • Working Memory Jev · PassageFlags instructional text passages that may overload a learner's working memory, via Jev.educationgithub★ 762
  • JeviewLocal gateway that logs and visualizes every Jev API call your code makes.opssdkgithub★ 613
  • jev-test-filterScores each test against a git diff with Jev, emits filter args for vitest/pytest/go test.evalcoding-agentgithub★ 353
    1 similar
    • leanestSkips tests Jev judges unaffected by a code diff, defaults to classifier.devcoding-agentevalgithub★ 53
  • evokeCLI and SDK routes natural-language commands to shareable scripts via Jev confidence gating.cliroutinggithub★ 223
  • slop-graderRule-based text grader runs every rule against every line in parallel with Jev.evalwritinggithub★ 333
  • jevframeAdds natural-language classification and scoring columns to pandas/Polars DataFrames via Jev.dataclassificationgithub★ 213
  • lkcleanChrome extension hides LinkedIn engagement bait and off-topic posts using Jev, with explanations shown.opsclassificationgithub★ 143
  • JevPokerBenchTexas Hold'em benchmark and leaderboard for decision models with replay and BYO agent.evalgamegithub★ 122
  • FastJevOpen source, self-hostable Jev implementation running Choice/Boolean/Score on Torch, vLLM, MLX, llama.cpp.classificationopsgithub★ 214
    11 similar
    • sys1Rust server implements a System One-compatible API for open decision models.classificationsdkgithub★ 533
    • Rizzo FlowLocal open-weight reimplementation of Jev's typed-decision HTTP interfaceclassificationroutinggithub★ 8783
    • simple-jevturn any open model into a classifier endpoint; fallback if TypeSafe goes awayclassificationsdkgithub★ 599
    • hearimGo gateway serving Jev's API over Ollama, llama.cpp, vLLM via logprobssdkclassificationgithub★ 53
    • jevifyServes any logprob-capable LLM through Jev's typed-question API shape.sdkgithub★ 553
    • semif-goSelf-hosted Jev-style decision API over llama.cpp, text and multimodal input.classificationsdkgithub★ 73
    • polyjevDrop-in Jev API server turning any LLM into typed decisions.routingsdkgithub★ 24
    • DecisSelf-hosted server speaks Jev's API for open models like Laya and kev.sdkopsgithub★ 94
    • verdict (khimaros)Self-hosted System One API server serving any llama-server as a Jev replacement.sdkclassificationgithub★ 84
    • cypridRuns local open-weight decision models behind one Jev-style API on a Mac.sdkclassificationgithub★ 33
    • ollajevDrop-in local server runs Hugging Face decision models behind Jev's wire API.sdkopsgithub★ 134
  • AnyJevTurns any LLM into a calibrated Jev-style decision model with training-free debiasing and a calibration benchmark.classificationevalresearchgithub★ 1.2k4
  • matchcnSemantic search over shadcn component registries by tagged behavior instead of name.searchcoding-agentgithub★ 252
  • pi-verdictPermission gate for Pi tool calls: rules settle clear cases, a classifier decides the rest.guardrailscoding-agentgithub★ 133
  • jev-rerankerPython library reranks and filters RAG search results for relevance using Jev.raggithub★ 433
  • HearthBrowser agent searches four rental marketplaces from one plain-language request using Jev.browserdemogithub★ 82
  • Open Jev (kyegomez)Unofficial PyTorch reconstruction of Jev's design: shared state encoder plus typed readout heads.researchclassificationgithub★ 1002
  • jev-oasisJev routes deterministic decisions in Camel-AI social simulation, cut LLM calls 67%routingresearchdiscord★ 13
  • Jev Chess (algo)Every legal chess move sent as one Choice question, calibration checked livegameresearchdiscord★ 22
  • Spring AI TypeSafe (Jev Java SDK)Official Java SDK and Spring AI integration for calling Jevsdkdiscord★ 523
    1 similar
    • jev-javaidiomatic Java SDK for the Jev decision enginesdkgithub★ 62
  • OpenJev (AlexWortega)Open source model weights reimplementing Jev's System One primitivesclassificationresearchdiscord3
  • JFocusJev judges whether your screen activity is on-task or a distractionclassificationopsdiscord2
    2 similar
    • QualmMac screen-time blocker where Jev classifies pages as learning, work or feed.classificationgithub★ 43
    • omarchy-laserJudges your screen against your stated task to block distractionopsclassificationgithub★ 73
  • Software FactoryJev judges agent pipeline stages for secrets, oversized diffs, human review flagsguardrailscoding-agentdiscord★ 113
  • decor-agent (Agent Flinch)Jev scores four typed questions to gate destructive coding-agent tool callsguardrailscoding-agentdiscord★ 04
  • project_blackoutSIEM simulation uses Jev to decide if an attack is happeningsecurityopsdiscord★ 22
  • Heym Decision Model supportVisual workflow automation tool adds Jev and Laya as decision-model backendsopsroutingdiscord★ 1.2k2
  • pytest-jevSemantic pytest assertions judged by Jev, 110x cheaper than Sonnet 5evalclidiscord★ 75
  • jev-mcp (Fly.io query store)Stores and reuses Jev queries across multiple agent harnessesmcpdiscord★ 02
  • laya-visionOpen vision variant of Jev, decent at Atari out of the boxclassificationresearchdiscord3
    3 similar
    • JEMMOpen-weight multimodal Jev-style model judges text plus screenshots.classificationresearchgithub★ 143
    • imajevOpen multimodal Jev-style decision model reading photos, ranks above Jev 1.13.classificationresearchmediagithub★ 3654
    • VevOpen vision-capable decision model, Jev-API compatible, runs on your own GPU.classificationmediagithub★ 173
  • Slop MopChrome extension judges and hides AI-slop LinkedIn posts using Jevclassificationwritingdiscord3
  • MandosMCP server routes reasoning tasks to a Jev-judged model councilmcproutingdiscord★ 22
  • Jev email sort flowPersonal email triage bot with a webapp to tune Jev sort criteriaclassificationopsdiscord2
  • jev-x-kitOffline decision layer with confidence-gated fallback chain for coding agentscoding-agentroutingmcpdiscord★ 34
  • RLJF (Jev as reward model)GRPO fine-tuning script uses Jev as the reward model instead of humansresearchdiscord2
    1 similar
    • fireworks-jev-reward-rlTrains a writer LoRA with Jev scores as the RL reward signal.evalresearchgithub★ 33
  • Watchflows Jev nodemacOS automation app adds Jev as a typed-decision workflow stepopsroutingdiscord2

New 2026-09-21

  • pi-jev-router (win4r)Routes Pi coding agent between models only at task boundaries.routingcoding-agentgithub★ 113
  • jev-ultrafast-mcpMCP server drives whole browser flows for agents in one tool call.mcpbrowsergithub★ 193
  • JEV-CPURuns SemIf-style typed decisions on CPU from open model logits.classificationcligithub★ 183
  • jev-curateFilters synthetic and pretraining datasets at 1,500+ rows per second.dataclassificationgithub★ 1074
    1 similar
    • jev-dataopsPipeline screens and evaluates training data with Jev before LoRA trainingdataevalgithub★ 623
  • Jev-as-a-Judge for Agent Evals (LangChain)Compares Jev against LLM judges on accuracy, repeatability, latency, cost.evalx3
  • awesome-jev-compactionCompactor that only scores tool output, never invents summary textcoding-agentdatadiscord★ 03
  • hippo-memoryDecaying agent memory layer with optional Jev reranker, SQLite backedragmcpdatagithub★ 7763
  • Rizzo FlowLocal open-weight reimplementation of Jev's typed-decision HTTP interfaceclassificationroutinggithub★ 8783
  • fastbrowseJev picks browser actions from page candidates, claims 71x cheaperbrowsercoding-agentgithub★ 1144
  • total-agent-memoryLocal agent memory graph now uses Jev to check fact contradictionsragmcpdatagithub★ 733
  • jev-dataopsPipeline screens and evaluates training data with Jev before LoRA trainingdataevalgithub★ 623
  • jev-calibrateTunes Jev question wording against your labels, flags overconfident wrong answersevalclassificationgithub★ 334
  • SharpFilters your X timeline with Jev or any LLM using your rulesclassificationbrowsergithub★ 303
    2 similar
    • SiftLabels every X post's substance and hides unwanted ones via Jev.classificationmediagithub★ 122
    • Quiet FeedChrome extension hides X posts using your Jev key or local modelbrowserclassificationgithub★ 23
  • djev-runDeploys DiffusionGemma-Jev on Cloud Run GPUs with a TypeSafe-compatible APIsdkopsgithub★ 5822
  • pi-jev-skill-pickerJev ranks Pi agent skills on demand instead of loading full catalogcoding-agentroutinggithub★ 364
  • CannyHooks block agents from claiming done without a passing check, Jev advisescoding-agentguardrailsgithub★ 1224
  • Jev QuantumRandom-baseline Jev-protocol server for mocking or testing router evalsevalcligithub★ 322
  • JevfillChrome extension autofills forms by matching fields to pasted notesbrowseropsgithub★ 202
    1 similar
    • jev-form-filler-extensionChrome extension uses Jev Choice to match profile data to form fields.browserclassificationdiscord★ 12
  • JevHarnessLLM writes task-specific Jev decision harness, refined using execution tracescoding-agentevalgithub★ 5363
  • tink-skillsClaude Code skill grills your plan through failure modes using Jevcoding-agentgithub★ 163
  • Jev Router for OpenClawRoutes OpenClaw models by cost and quality using Jev.routingsdkdiscord3
  • jev-retrieval-evalBenchmarks Jev vs GPT for scoring RAG passage relevance.ragevaldiscord★ 04
  • five-linesRust CLI reviews PR diffs against refactoring rules using Jev per method.coding-agentevaldiscord★ 24
  • peaks-notesClassifies live chat messages into an evolving topic schema.classificationdatadiscord★ 03
  • jevernetesClassifies Kubernetes log lines by urgency using Jev.opsclassificationdiscord★ 153
  • REJEVParses HTML by natural-language pattern instead of regex.dataclassificationdiscord2
  • ThresholdScreens rental applications and drafts recommendations using Jev.classificationopsdiscord2
  • semcheckGo linter uses Jev for semantic yes/no code checks.coding-agentclassificationdiscord★ 13
  • perfectrecallJev-powered agent memory retrieval, Mnemosyne and Hermes compatible.ragdatadiscord★ 34
  • What people are building with JevSurveys nine Jev awesome-lists and categorizes 358 projects.writingdatadiscord3
  • jev-hierarchyComputes probabilities over a hierarchy of terms, not a set.classificationdatadiscord★ 02
  • jeqPipes Jev judgments through shell commands like jq.cliclassificationdiscord★ 103
  • ReckoNimNim DSL embeds batched Jev judgments directly in code.sdkcoding-agentdiscord★ 03
  • jev-vs-sovereign-benchmarkBenchmarks Jev against local classifiers for routing and verification.evalragdiscord★ 34
    1 similar
    • fm-with-jevBenchmarks Apple's on-device model vs Jev for routing and list-picking decisions.routingevalgithub★ 94
  • askgrepSemantic grep scores codebase functions with Jev, no sampling.coding-agentclidiscord★ 44
    3 similar
    • JevGrepFinds code by behavior using Jev-scored search, returns exact excerpts.coding-agentsearchmcpgithub★ 1024
    • JevFindsemantic code search CLI, returns files, line ranges and confidence scores, author calls it a democlisearchgithub★ 52
    • semble-jevCLI code search where Jev scores relevance of retrieved snippets.searchcoding-agentgithub★ 63
  • jev-ultralightspeedBatches Jev calls for high-throughput labelling with no accuracy loss.classificationopsdiscord★ 124
  • jev-migrateScans codebases for existing LLM calls to migrate to Jev.coding-agentopsdiscord★ 03
  • debate-judge-testCompares Jev vs GPT consistency judging the same debate repeatedly.evalclassificationdiscord★ 04
  • jevkitDecision kernel wraps Jev with confidence bands and jaggedness workarounds.guardrailssdkdiscord★ 04
    3 similar
    • jev-harness (AntonioCoppe)Confidence gates and shadow mode around Jev; a row-filter job in 1.3s.coding-agentevalgithub★ 153
    • JDE (Jev Decision Engine)Central place to ask, threshold and record an agent's Jev judgments.opsdiscord★ 22
    • dcisionSchema-defined typed-decision layer returning confidence before an agent runs.classificationroutinggithub★ 53
  • awesome-typesafeCurated list of official and community TypeSafe and Jev resources.writingdatadiscord★ 5802
  • digital-twinsSimulates API and MCP calls to test Jev integrations safely.evalmcpdiscord★ 213
  • AI-Seobench Jev portPorts an LLM taxonomy classifier pipeline to Jev choice and score calls.classificationevaldiscord★ 34
  • plzConfirms every shell command with Jev before running it.guardrailsclidiscord★ 03
  • jev-assistRanks repo files by relevance to a task using Jev.coding-agentopsdiscord★ 43
  • Building a Chatbot Message Router with JevField report on building a chatbot message router with Jev.routingwritingdiscord3
  • PostneedleFive Jev Noul checks flag humblebrag, engagement bait, AI tone before you post.writingclassificationdiscord4
  • file2markdown output-health checkJev judges converted Markdown for garbled text and broken tables, never rewrites.datamcpclassificationdiscord5
  • JevPROpen-source GitHub PR reviewer that returns structured Jev decisions.coding-agentdiscord★ 102
  • jev-codex-token-saverJev picks relevant evidence to cut tokens in Codex investigations.coding-agentdiscord★ 72
  • jev-triageGo CLI categorizes messages, scores urgency, flags low-confidence for human review.classificationclidiscord★ 33
  • TypeSafe.AI .NET SDKC#/.NET SDK bringing Noul, Choice, Score to .NET apps.sdkdiscord★ 23
  • jev-form-filler-extensionChrome extension uses Jev Choice to match profile data to form fields.browserclassificationdiscord★ 12
  • FoodAllergyDetectorChecks a recipe for food allergens using Jev.healthclassificationdiscord★ 02
  • Observatory (redteam guardrail policies)Community tool builds and tests agent guardrail policies against attacks, Jev-backed.guardrailssecuritydiscord3
  • CodeEagleRepo knowledge-graph agent uses Jev plus a new Go SDK for it.sdkcoding-agentdiscord★ 13
  • ARMINRust sidecar uses Jev to extract binding decisions and stop agent drift.coding-agentopsdiscord★ 14
  • typed-decisions-shadow-layerShadow decision layer scores live CRM leads with Jev, logged against baseline.opsclassificationdiscord★ 03
  • reckonClassifies customer email replies to invoice reminders into action classes with confidence.classificationfinancediscord2
  • JevFlowTyped Noul/Score/Choice decisions compose into deterministic, explainable backend workflows.sdkopsdiscord★ 113
  • JevfastDirectory of Jev projects and videos, filterable by call frequency.searchdatadiscord2
  • Jev Users (jevusers.com)Merges 31 awesome-jev lists into 1,299 ranked projects, updated daily.searchdatadiscord2
  • jev-resume-disqualifierSub-25ms resume knockout engine, claims eliminates 80% of unqualified applicants.classificationopsdiscord★ 33
    2 similar
    • typesafe-jev (cv-screen)Screens CV folders with Jev typed questions and editable scoring policy.classificationopsgithub★ 103
    • ResurfaceJev screens resumes against job questions with quoted evidence, 31x cheaper.classificationragdiscord★ 04
  • JEV OAS SentinelRoutes changed OpenAPI prose to Jev to catch hidden breaking API behavior.evalsecuritydiscord★ 54
  • watfilePyPI CLI sorts scans, PDFs and ebooks into folders using Jev classification.classificationdataclidiscord3
  • AI text pattern detector (caiop)Jev labels LLM-writing passages with explainable categories instead of one score.classificationwritingdiscord2
  • jev-tool-routerJev picks which of 273 MCP tools to expose to Codex, above 90% confidence.mcproutingdiscord★ 93
  • codeloom.engine (jev-exp-1)Coding harness uses Jev for shell exec approval and tool-call verification.coding-agentdiscord★ 22
    1 similar
    • osenvJudges every coding-agent file write and command before it runs.coding-agentguardrailsdiscord3
  • AgentEval Jev evaluatorAdds Jev as a decision-model evaluator kind, alongside a Microsoft Agent Framework PR.evalsdkdiscord★ 1554
  • JevGraphPipeline turns documents into evidence-backed knowledge graphs using DocJev and Jev.ragdatadiscord3

New 2026-09-20

  • jev-rerankingBenchmarks Jev zero-shot reranking against monoBERT and BM25 on TREC.evalragsearchgithub★ 104
  • clean-code-reviewJudges PR files against Clean Code principles using Jev, summarized by Luna.coding-agentevalgithub★ 114
  • any-autoAuto-approves coding agent tool calls using Jev as a risk-reviewing backend.coding-agentguardrailsgithub★ 84
    1 similar
    • jev-guard (leepokai)Risk-scores every coding-agent tool call across Claude Code, Codex, Copilot, Cursor.guardrailscoding-agentgithub★ 634
  • super-jevExperimental harness connecting evidence, Jev judgments, and verified actions.guardrailssdkgithub★ 142
  • SemDecideUnix CLI wrapping Jev for classification, scoring, filtering, and guarding.cliclassificationguardrailsx★ 774
    10 similar
    • jeff-cliCLI to query Jev choices, scores and rankings from the terminal.clisdkdiscord★ 33
    • jeqPipes Jev judgments through shell commands like jq.cliclassificationdiscord★ 103
    • jev-cli (shaharia-lab)CLI turns Jev questions into exit codes for shell and CI.cliclassificationopsgithub★ 353
    • jev-cliverify, screen, classify, extract, route, rerank from the shellcliclassificationgithub★ 22
    • jev-axipick, rate, check, rank, triage, guard shell commandscliguardrailsgithub★ 27
    • openpave-jevCLI for typed choice, score and noul decisions via any System One model.cliclassificationgithub★ 23
    • jevpipeUnix CLI that pipes files through Jev judgment questions for agent filtering.clicoding-agentdiscord★ 63
    • jev (stefafafan)Unofficial Unix CLI client for Jev supporting TypeSafe, Cloudflare, and Vercel providers.clisdkgithub★ 133
    • onesieUnix-composable Jev CLI with threshold calibration, caching, and mock testing.clievaldiscord★ 23
    • thinkthenShell CLI returns true, false, label or number from Jev.cliclassificationgithub★ 73
  • TypeSafe on NeonRoutes requests to Grok or GPT based on Jev's task classification.routingclassificationx★ 54
  • neo4jevScores graph edges with Jev and beam-searches knowledge graph paths.ragdatax★ 1733
  • PrismJev judges liquidity toxic flow and market stress in shadow trading mode.financeclassificationx★ 1252
  • NeuroLinkOne interface for 40 LLM providers, adds Jev-backed decide calls.sdkroutingmcpgithub★ 1504
    1 similar
    • LiteLLM Jev supportLiteLLM supports Jev natively; blog covers a context-compaction use casesdkopsx3
  • DocJevClassifies and splits PDFs/DOCX into categorized documents using Jev.classificationragdatagithub★ 5244
  • arc-cuaExecutes bounded desktop UI subtasks click by click using Jev.browsercoding-agentgithub★ 3163
  • @receptron/layaRuns the open Jev-compatible Laya decision model locally via ONNX.sdkclassificationgithub★ 8874
  • JevGrepFinds code by behavior using Jev-scored search, returns exact excerpts.coding-agentsearchmcpgithub★ 1024
  • Jev ExplainedInteractive playground demonstrating Jev's three typed decision primitives live.educationdemogithub★ 342
  • System One HarnessTurns any System One model into a confidence-gated agent loop.coding-agentclassificationgithub★ 2083
  • ReadAloud (dasheng)Scores English read-aloud pronunciation errors locally using ASR plus Jev.educationmediagithub★ 1443
  • arbiterSelf-hosts typed-decision models like Laya behind a Jev-compatible API.sdkclassificationgithub★ 343
  • OpenJev (SiliconLabAI)Approximates Jev by scoring each answer option in parallel calls.classificationgithub★ 1742
  • dehydratorClient-side tool search for LLM APIs, using BM25 or Jev.mcpsearchclassificationgithub★ 104
  • jev-superpowersAdds Jev typed decisions and package vetting to coding-agent workflows.coding-agentguardrailsgithub★ 393
  • Jev Security ScanReviews Skills and MCP code for injection risk using Jev.securityguardrailsmcpgithub★ 124
    1 similar
    • skill-scannerScans Agent Skills offline for prompt injection and exfiltration before install.securityguardrailscligithub★ 24
  • JevvyAuto-approves harmless shell commands in coding agents via Jev.guardrailscoding-agentgithub★ 202
  • JCR (Jev Capability Resolver)Finds deterministic commands in a capability tree using Jev search.coding-agentmcpgithub★ 193
  • AzdajaKeeps source local and recurses into LLM calls, with optional Jev.coding-agentsdkgithub★ 142
  • SiftLabels every X post's substance and hides unwanted ones via Jev.classificationmediagithub★ 122
  • jevQLAdds semantic jev() functions to plain SQL queries over Postgres.dataclassificationgithub★ 144
    2 similar
    • pg-jevPostgres extension, plain-English per-row WHERE judgmentsdatadiscord
    • jev4pgAdds natural-language-to-SQL and Jev semantic operators to PostgreSQL queries.dataragclassificationgithub★ 2384
  • OpenThai-SystemOneOpen Thai and English decision model mirroring Jev's API contract.classificationsdkgithub★ 682
  • jev-sentinelScreens agent tool calls and replies for injection using Jev.guardrailssecuritycoding-agentgithub★ 124
    1 similar
    • JevPromptShieldJev scores text at tool-call time to block prompt injectionguardrailssecuritygithub★ 2023
  • Jev vs MLBenchmarks Jev against 11 classical ML pipelines on 8 datasets.evalclassificationresearchgithub★ 174
    1 similar
  • jsortSorts text by meaning via pairwise Jev comparisons, reports reliability.classificationevaldatagithub★ 264
    1 similar
    • AI Elo rankertournament scoring for text, reusable for rubric gradingevalgithub★ 7
  • cascade-searchParses queries locally, escalates only uncertain words to Jev.searchclassificationroutinggithub★ 104
  • jev-cli (shaharia-lab)CLI turns Jev questions into exit codes for shell and CI.cliclassificationopsgithub★ 353
  • snapjudgeReads typed-decision probabilities from local Qwen model logits, TypeSafe-compatible.classificationsdkgithub★ 143
  • jev_jsonschemaConverts a JSON Schema into Jev questions and back into JSON.sdkclassificationgithub★ 74
  • Jev CookbookFifteen tested Jev recipes for triage, tagging, PII, and reranking.classificationragdatagithub★ 364
  • JekhovUses Jev only to pick Playwright targets, keeping flow deterministic.browsercoding-agentgithub★ 93
    1 similar
    • quicke2ePlain-English browser e2e tests, decision model picks clicks, exports Playwrightcoding-agentbrowserevalgithub★ 924
  • SLO RouterRoutes LLM requests to cheapest backend meeting a latency SLO.routingevalgithub★ 103
    1 similar
    • model-router-pythonFilters LLM providers by cost and limits then Jev picks winner.routingsdkgithub★ 54
  • JevalsGrades LLM/agent output with Jev, attaching a confidence to each.evalclassificationgithub★ 74
  • S18ShareMulti-agent runtime with optional Jev-backed routing and skill matching.routingcoding-agentgithub★ 72
  • av (Agentic Video Intelligence)Video search CLI that uses Jev to filter and rank scenes.ragmediaclassificationgithub★ 73
  • elons-jobChrome extension hides explicit X replies using Jev classification.browserclassificationgithub★ 212
  • typesafe-sdk-goOfficial-style Go SDK wrapping the TypeSafe Jev API.sdkgithub★ 92
  • awesome-jev-usecasesSourced patterns and pitfalls collection for building with Jev.researchgithub★ 213
    1 similar
    • jev-in-the-wildTracks real Jev use cases and criticism across GitHub, Reddit, and blogs.researchdemogithub★ 73
  • mcp-ui-pocJev decides UI widget shape, LLM only writes when uncertain.mcpgithub★ 83
  • pi-quiet-askJev-backed rule packs gate destructive commands and secret leaks in pi.guardrailssecuritycoding-agentgithub★ 124
  • pi-jev-routerJev picks model and reasoning effort once per session via AI Gateway.routingcoding-agentgithub★ 164
    2 similar
  • jevlogsScores OpenTelemetry logs with Jev before routing to expensive LLM analysis.opsroutingdatagithub★ 184
  • typesafe-skill-routerJev names the one relevant skill to load before the model call.routingcoding-agentgithub★ 144
    3 similar
    • GatekeeperJev routes ambiguous requests to the right agent skill via Claude Code hooks.routingguardrailscoding-agentgithub★ 203
    • jev-skill-router (shimo4228)Skill-suggestion hook using Jev; author found it doesn't help strong models.coding-agentevalgithub★ 152
    • tink-routeRoutes tasks to specialist agent skills using Jev confidence scoring.routingcligithub★ 53
  • what-is-jevIndependent research site cataloging 947 rubric-scored Jev repositories.researchgithub★ 12
  • jev-review (MaxIvanyshen)Code review filter tells coding agents where to look first.coding-agentgithub★ 13
  • AIStockMulti-agent trading platform offers Jev as a decision backend.financegithub★ 3522
  • jevmailGmail triage sorts 1,000 emails a minute for 3 cents.classificationgithub★ 964
  • JevBenchBenchmark scoring Jev-class models on intelligence, calibration, speed, cost.evalgithub★ 2553
  • TypeSafe AI PlaygroundCommunity playground with 110 editable Jev classification and routing examples.classificationroutingeducationgithub★ 243
  • OpenJev-VisionOpen research toolkit answering multiple typed questions from one image encode.researchmediagithub★ 402
  • Jev DirectoryDirectory of 1,300+ Jev builds with an MCP server and index.mcpgithub★ 142
    4 similar
    • JevfastDirectory of Jev projects and videos, filterable by call frequency.searchdatadiscord2
    • Jev Users (jevusers.com)Merges 31 awesome-jev lists into 1,299 ranked projects, updated daily.searchdatadiscord2
    • Jev Research IndexBilingual catalogue of 358 Jev projects, papers and public materials.researchdatagithub4
    • JevMadeSearchable catalogue of 2,300+ Jev projects, guides and experiments with source credit.searchdatadiscord2
  • jev-skill-suggesterRecommends which installed Claude or Codex skill fits a task.coding-agentroutinggithub★ 333
    1 similar
    • jev-skillsJev picks which skills the model sees each turn, zero context cost.coding-agentroutinggithub★ 64
  • invalidateFlags stale agent memories when new evidence supersedes them.coding-agentgithub★ 233
  • fast-browser-useLocal open-weight browser automation reproducing Jev's fast decision pattern.browsergithub★ 2212
  • jev-radar (everyinfra)Tracks the Jev ecosystem, rescanning 220+ documented builds every 3 hours.researchdatagithub★ 342
  • jevifyServes any logprob-capable LLM through Jev's typed-question API shape.sdkgithub★ 553
  • jevwireMCP server and Claude Code plugin gating tool calls with Jev.mcpguardrailscoding-agentgithub★ 243
    1 similar
    • toolgateClaude Code hook and MCP proxy blocks, confirms or allows tool calls by riskguardrailsmcpsecuritygithub★ 24
  • jev-linkmapRebuilds a 566-page site's internal link map in 6 seconds.datagithub★ 223
  • openvonsDecision layer answering finite choices from text, image, or voice.classificationmediagithub★ 133
  • patdownLints a codebase against fuzzy markdown rules using Jev as judge.coding-agentevalgithub★ 153
    2 similar
    • jev-specCI check that Jev-scores code against markdown spec requirements.evalcoding-agentgithub★ 113
    • adhereLinter for non-deterministic business rules checked by Jev against markdown-defined policies.classificationcoding-agentgithub★ 83
  • jev-rag-benchmarkMeasures whether Jev reranking actually improves RAG quality and cost.ragevalgithub★ 164
  • jev (BorisLeMeec)Claude Code plugin answers codebase questions without loading files into context.coding-agentgithub★ 403
  • fable-jevSub-100ms Jev reflex layer for routing and context compaction.routingcoding-agentgithub★ 102
  • jev-mail-classifierConfig-driven inbox classifier tags, moves and flags mail via Jev.classificationgithub★ 213
  • findmeFinds files by natural-language memory using beam search and Jev.searchcoding-agentgithub★ 262
  • hermes-jev-approvalsJev reviews shell command approvals, 8.7x faster on 153 commands.guardrailscoding-agentsecuritygithub★ 193
  • OpenSourceJevTurns a local Qwen model into a sub-100ms decision engine.researchgithub★ 432
  • jev-fn (aaazzam)Decorator turns a Python function signature into a typed Jev query.sdkgithub★ 113
  • jevcalCalibrates confidence thresholds and flags drift for typed decision models.evalgithub★ 104
  • jev-recallfilters AI-assistant memories by one calibrated yes/no per item, not embeddingscoding-agentgithub★ 383
  • Pi-Jev-GuidePi coding-agent guardrail plugin with configurable rules, redaction, stall detectionguardrailscoding-agentgithub★ 213
  • jev-arena (NanmiCoder)CSV-driven side-by-side Jev vs DeepSeek comment-labeling comparison with saved reportsevalclassificationgithub★ 1212
  • sarvam-jevgeneration-free typed decisions on Sarvam's Indic tokenizer instead of Qwenclassificationresearchgithub★ 602
  • jev-dsh-decisionJev decision plugin for DeepSeek Harness, OpenCode and Codex Harnesscoding-agentgithub★ 2532
  • TypeLLMbrings Jev-style typed decoding to open autoregressive models via constrained decodingclassificationresearchgithub★ 9433
  • siftrankLLM needle-in-haystack ranker adds a Jev provider optionsearchgithub★ 2253
  • third-handmacOS menu bar assistant reads Accessibility tree, types and clicks via Jevroboticsgithub★ 3532
    3 similar
    • jev-use (savka777)voice-controlled Mac computer use via Accessibility tree, no vision modelroboticsmediagithub★ 1223
    • Jev Voice (jev-cua)Voice-controlled macOS bar where Jev picks the next UI action to run.browserdemogithub★ 391
    • macos-computer-use-kitAccessibility-tree computer use for agents with optional Jev safety guards.coding-agentmcpgithub★ 54
  • hermes-jev-skillsJev handles model routing, memory, compaction and skill selection for agentscoding-agentroutinggithub★ 1.1k3
    2 similar
    • jev-router (Ex8-ca)Routes Hermes agent turns to skills using Jev decision confidence.coding-agentroutinggithub★ 52
    • Hermes SwitchyardJev picks skills, reasoning effort, models and UI actions for a Hermes agent session.routingcoding-agentgithub★ 73
  • openJev-verdict-2.0open 150M model claims higher accuracy and better calibration than Jevevalresearchgithub★ 2953
    1 similar
    • Verdict (OpenJev)open 151M non-autoregressive decision model, calibrated uncertainty, WebGPU demoresearchdemogithub★ 1103
  • jev-skill (wuyoscar)90 catalogued Jev use cases and installable agent skillscoding-agentgithub★ 5992
    1 similar
    • awesome-jev (Ai-trainee)Curated Jev use-case list packaged as a loadable agent skill.educationdemogithub★ 142
  • awesome-jev-gallerycurated papers, open reproductions and independent Jev evaluationsresearchevalgithub★ 5022
    1 similar
    • Awesome-jev-papersCurated list of research papers and benchmarks about Jev.researchgithub★ 143
  • jev-use (savka777)voice-controlled Mac computer use via Accessibility tree, no vision modelroboticsmediagithub★ 1223
  • LLM2Jevadapts local LLMs into Jev-style Choice, Score and Noul outputssdkgithub★ 4293
  • pi-bifrostPi model router with an optional Jev tier-selection backendroutingcoding-agentgithub★ 693
    4 similar
    • pi-shift-routerRoutes Pi coding agent turns between cheap and strong models, judge optional.routinggithub★ 123
    • pi-jev-router (win4r)Routes Pi coding agent between models only at task boundaries.routingcoding-agentgithub★ 113
    • pi-pignonShifts Pi agent to cheaper or stronger LLM using a Jev difficulty check.coding-agentroutinggithub★ 53
    • pi-model-routerRoutes Pi coding-agent turns across model tiers with optional Jev advice.routingcoding-agentcligithub★ 63
  • jegrepsemantic grep with calibrated per-path yes/no probabilities, no indexsearchcligithub★ 1153
  • grok-bot-jevJev gates Grok Bot retries, caching and subagent spawn decisionsguardrailscoding-agentgithub★ 1042
  • Astraagent runtime with native Jev judgments and EXPLAIN ANALYZE for contextcoding-agentevalgithub★ 343
  • jev-desktopJev picks the action inside Codex Computer Use sessionsroboticscoding-agentgithub★ 762
    1 similar
    • jev-gui-delegateGUI delegation for Codex; Jev makes semantic picks, controller verifies steps.coding-agentbrowsergithub★ 1043
  • awesome-jev-zhChinese curated index with an independent-evaluation caveats sectionresearchevalgithub★ 812
  • jev-lintflags function names, comments and tests that no longer match codecoding-agentevalgithub★ 1193
  • jev-as-a-judgebenchmarks Jev as an agent-eval judge against three LLM judgesevalgithub★ 1023
    2 similar
  • jev-mcp (burnigtm)MCP server for coding-loop routing, review, verify and screen toolsmcpcoding-agentgithub★ 623
  • typesafe-mcp (PyModel)Go MCP server exposing Jev's Choice, Score and Noul as one toolmcpsdkgithub★ 232
  • JevbridgeACP and MCP adapter bridges Jev with Claude, Codex, Grok and OpenCodemcpgithub★ 472
  • jev-botturns an idea into a Jev integration design from a research libraryresearchgithub★ 12
  • jevproxyreverse proxy claims 65% cost cut and 90% latency cut for agentsopscoding-agentdiscord2
  • FakeCatchChrome extension scores review trust as real, generic, ad-like or shortbrowserclassificationx2
  • Vettlycontent moderation API for text, image and video built on Jevsecuritymediaclassificationdiscord2
  • eslint-plugin-jevESLint rules are plain-English questions Jev answers with a probabilitycoding-agentgithub★ 33
  • awesome-jev-toolsanother curated index of public projects and patterns built on Jevresearchgithub★ 7452
  • Jeveloper100 example Jev use cases: routing, scoring, agents, support workflowseducationcoding-agentdiscord2
  • Jev on Venice APIJev added to Venice API in beta, no JSON parsing neededsdkx3
  • mefi-studiodesktop AI workspace, Jev routes models by performance, speed and costroutinggithub★ 132
  • jevfanity-apihosted profanity detector using Jev, no API key neededclassificationsecuritygithub★ 02
  • jev-codex-pilotCodex web overlay picks model and reasoning depth, automated Kanban task sequencingcoding-agentroutinggithub★ 52
  • hunchRuby gem wraps Jev's three question types as chance, pick and ratesdkgithub★ 193
    2 similar
    • typesafe-sdk (Ruby)Ruby client for TypeSafe's System One API with typed Choice/Score/Noul helpers.sdkgithub★ 73
    • decide (obie)Ruby decision layer turning typed yes/no/choice answers into policy verdicts.sdkclassificationgithub★ 83
  • pi-jev-wikiPi package, agent-maintained project wiki holds only reasoning, not code-derivable factscoding-agentgithub★ 03
  • okolocal search plus Jev reranking cuts what a coding agent has to readsearchcoding-agentraggithub★ 94
  • jev-prune-kitcapability-aware context-pruning installer for agent harnesses, experimentalcoding-agentgithub★ 12
  • Jev Moderation BotDiscord moderation with 4-stage escalation, audit logging, one-click pardon learningsecurityclassificationgithub★ 463
    2 similar
    • SoterDiscord moderation bot uses Jev to flag hate speech and spam.guardrailsclassificationdiscord★ 42
    • typesafe-ai-botDiscord moderation bot: AI suggests, plain code decides, community votes.guardrailsopsdiscord★ 02
  • jev-logtriageJev decides whether a batch of logs is worth acting on, confidence-gatedopsdatagithub★ 22
  • jevcache.shmemoizes Jev-class decisions so repeats are free and deterministicopsdiscord2
  • jev-javaidiomatic Java SDK for the Jev decision enginesdkgithub★ 62
  • jev-builderbrowser form builds Jev requests from 34 templates, no JSON by handbrowserdiscord2
    1 similar
    • jev-starterVisual console to configure, test and export Jev question calls.sdkcligithub★ 33
  • AlloyJev routes tasks across Claude, Codex, Gemini and Antigravity by capabilityroutingcoding-agentgithub★ 33
  • opencode-plugin-variantizerJev picks the cheapest capable OpenCode model and reasoning variant per taskroutingcoding-agentgithub★ 03
    2 similar
    • opencode-auto-jevOpenCode plugin where Jev picks the real model for each turn.routingcoding-agentgithub★ 52
    • opencode-jev-router (robertn702)Jev picks reasoning effort per OpenCode step; matched quality with fewer tokens.coding-agentroutinggithub★ 104
  • enzymecompiles Markdown-wiki reading rules into Jev queries, claims 350x cheaper than frontierclassificationopsgithub★ 863
  • jev-studioMCP tools for Choice, Noul and Score plus prompt libraries and commandsmcpgithub★ 162
  • typesafe4stype-safe Scala 3 SDK wrapping the Jev APIsdkgithub★ 32
    1 similar
    • zio-typesafe-aiTyped Scala 3/ZIO client for Jev, illegal states unrepresentable by design.sdkgithub★ 53
  • NolaTypeScript superset adds Jev as a typed-inference provider alongside plain TS typessdkdiscord2
  • Jev 1.13 jaggedness (TypeSafe docs)TypeSafe's own list of known Jev weak spots, math firstevalresearchdiscord3
  • jarviscore-frameworkmulti-agent framework, Jev routes subagents and filters RAG passages by conflict/injectionroutingragguardrailsgithub★ 1923
  • pi-jev-sentinelPi agent extension screens tool calls and output for injection, scrubs secrets, pins taskguardrailssecuritycoding-agentgithub★ 123
  • jev-alignCLI calibrates Jev to your own judgment criteria using GEPAclievalgithub★ 3073
    1 similar
    • jev-calibrateTunes Jev question wording against your labels, flags overconfident wrong answersevalclassificationgithub★ 334
  • jev-usehands no-text-output agent steps to Jev, p50 ~230ms, ~$0.02 per 1,000 judgmentscoding-agentclassificationgithub★ 563
    1 similar
    • quicksilverClaude Code skill offloads bulk judgment calls to Jev, cutting tokens 86%.coding-agentclassificationgithub★ 1204
  • JevSwiftSDKunofficial Swift SDK, async/await, batching, zero dependencies, iOS/macOS/Linuxsdkgithub★ 83
  • revolve-agentcoding agent uses Jev to gate risky commands, set permissions and flag instruction drift, still WIPcoding-agentguardrailsdiscord2
  • JevFindsemantic code search CLI, returns files, line ranges and confidence scores, author calls it a democlisearchgithub★ 52
  • jevrouter-tsTS router picks cheapest model tier per query using Jev classificationroutingclassificationgithub★ 102
  • JevNQLnatural-language database queries via Jev instead of SQLdatasearchgithub★ 42
    2 similar
    • IntentSQLConverts natural language to SQL via chained Jev decisions over schema state.dataclassificationdiscord★ 33
    • jevsd-pgSelf-developing Postgres database answers natural language queries via Jev operators.dataclassificationgithub★ 2383
  • Jev CLI (vectorz)terminal CLI with skills and plugins for agent workflowsclicoding-agentdiscord2
  • Jev in production (iambraun)self-updating list of 18 products running Jev in production, scored 1-5researchdiscord3
  • jev-frontier-bench200-decision benchmark of Jev vs five frontier models, calibration and cost dataevalgithub★ 24
    1 similar
    • typesafe-ai-benchmarkcompares Jev against LLM-native structured output on latency, cost and qualityevalgithub★ 383
  • Pi-SaverJev prunes Pi Coder conversation history per turn, claims 75%+ token cutcoding-agentopsgithub★ 03
    1 similar
    • pi-jev-context-curatorJev prunes pi's LLM context per message, caching judgments to skip rechecks.classificationsdkgithub★ 72
  • JevRouter.nvimNeovim plugin routes prompts to file, edit or terminal handlers via Jev, under 200msroutingcoding-agentgithub★ 23
  • hushGitHub Action triages issues with Jev, stays silent below a confidence thresholdopsclassificationgithub★ 23
    1 similar
    • tiaGitHub issue triage agent; Jev makes one bounded decision per issue.coding-agentopsgithub★ 83
  • jev-ushermodel routing plus recoverable context filtering for Claude Code, local compare UIroutingcoding-agentgithub★ 13
  • openpoke-meets-jevmoves OpenPoke's email screening, tool guardrail and reranking to Jevguardrailsclassificationgithub★ 13
  • JevArenaopen BYOK arena blind-votes Jev against other models acting as judgeevalgithub★ 53
  • Jev For Dummieswraps Jev primitives as plain HTTP GET endpoints for newcomerssdkeducationgithub★ 62
  • jev4kKotlin DSL and client for the Jev APIsdkgithub★ 122
  • code-compact-jevJev trims excess comments out of codecoding-agentgithub★ 02

New 2026-09-19

  • Jev on Vercel AI GatewayJev added to Vercel AI Gateway, one line to call itsdkdiscord3
  • Jev on Cloudflare AI GatewayJev live on Cloudflare AI Gateway as a typed decision endpointsdkdiscord3
  • flaviocopes Jev deep divepractitioner walkthrough of the API, primitives and basic usage patternseducationdiscord2
    1 similar
    • Jev Showcase (cobusgreyling)Explains Jev's fan-out, pricing and confidence semantics beyond basic docs.educationdemogithub★ 1362
  • jev-capability-atlasevidence-based map of where Jev's calibration holds up or failsresearchevalgithub★ 274
    1 similar
    • awesome-jev-surveyEvidence survey of Jev calibration, selective control and open implementations.researchevalgithub★ 74
  • jev-reviewersystematic-review data extraction, every answer a verbatim cited quoteresearchdatagithub★ 464
  • jev-column-raceJev vs Gemini labelling 1,000 reviews, 4.1x faster, 7x cheaperevalclassificationgithub★ 234
  • pi-jev-auto-modeJev auto-approves Pi bash/write/edit calls, fails closed, 171 testsguardrailscoding-agentgithub★ 314
    7 similar
    • pi-verdictPermission gate for Pi tool calls: rules settle clear cases, a classifier decides the rest.guardrailscoding-agentgithub★ 133
    • pi-quiet-askJev-backed rule packs gate destructive commands and secret leaks in pi.guardrailssecuritycoding-agentgithub★ 124
    • pi-classifier (bacnh85/pi-extensions)Pi coding agent package adds a Jev-gated auto-approve hook for shell.guardrailscoding-agentcligithub★ 333
    • pi-typesafe-approvePi extension auto-approves routine bash commands via Jev, escalates the restcoding-agentsecuritygithub★ 22
    • pi-auto-reviewerAuto-reviews shell commands before running, using Jev for cheap decisions.securitycligithub★ 73
    • jev-autoJev classifier judges each Pi agent tool call, allow block or askguardrailscoding-agentgithub★ 44
    • pi-automode-classifierClassifies shell commands as auto-run, confirm or block before executionguardrailscoding-agentgithub★ 63
  • winnowcalibrated context sieve for Claude Code, stubs blocks Jev judges unneededcoding-agentopsgithub★ 1093
  • yoshiJev-judged context-pruning proxy for Claude Code and Codex, POCcoding-agentopsgithub★ 313
  • pi-dcpdedups and prunes Pi coding-agent context, experimental Jev selectioncoding-agentopsgithub★ 173
    1 similar
    • pi-jev-contextPi extension dedupes reads and filters command logs with Jev.coding-agentgithub★ 103
  • pi-jevsemantic tool router and skill finder for the Pi coding agentroutingcoding-agentgithub★ 623
  • pi-typesafebatched Jev eval tool, terminal playground, shared client for Pi extensionsevalclisdkgithub★ 473
  • stanley-codeJev routes requests to deterministic coding workflows, falls back to an agentroutingcoding-agentgithub★ 1163
  • blinkJev-driven file-system walkers search a codebase by natural languagesearchcoding-agentgithub★ 1003
  • ErisLintRust linter with Jev-judged rules for complexity, naming, comment valuecoding-agentgithub★ 213
  • ts-browser-agentLangChain browser agent using Jev as the tool-picking modelbrowserroutinggithub★ 443
  • ruby_llm-typesafeTypeSafe provider for RubyLLM, Noul/Choice/Score via structured outputsdkgithub★ 183
  • jev (Elixir/OTP)GenServer replies to Jev, pattern-match the typed answersdkgithub★ 363
    1 similar
    • GutElixir DSL routes LLM or Jev decisions into pattern-matched control flow.sdkgithub★ 182
  • effect-questionssemantic judgment as an Effect-TS primitive: is, choose, rank, branchsdkclassificationgithub★ 143
  • Verdict (OpenJev)open 151M non-autoregressive decision model, calibrated uncertainty, WebGPU demoresearchdemogithub★ 1103
  • litjevturns any Qwen checkpoint into a Jev-style typed decision layerresearchgithub★ 473
  • jev-on-a-laptopreproduces Jev's parallel constrained decoding on a 1.5B model locallyresearchgithub★ 243
  • open-jev (DiffusionGemma)typed JSON via diffusion-model denoising, benchmarked against saved Jev decisionsresearchgithub★ 423
  • jev_locallocal Jev-style API on LFM or ModernBERT, JGLUE-benchmarkedresearchevalgithub★ 383
    4 similar
    • JevBERTPoC local BERT backend mimicking Jev's typed-decision API, uncalibrated.sdkclassificationgithub★ 332
    • jeff (GliFormer)self-hosted drop-in Jev replacement, cheaper but less accuratesdkgithub★ 2983
    • Jev-StyleLocal 0.5GB calibrated decision model, systemone-compatible server, Claude Code guard hookclassificationmcpgithub★ 174
    • gliner2-api-jev-schemaLocal GLiNER2 classifier served behind a Jev-compatible System One API endpoint.classificationsdkgithub★ 33
  • jevscan-evmheat map of likely bugs across a repo, vibe-coded proof of conceptsecuritycoding-agentgithub★ 282
  • ulkaexperimental browser extension, Jev selects actions from the accessibility treebrowserdemogithub★ 222
  • ainodeself-hosted GPU appliance with a typed decision endpoint alongside inferencesdkopsgithub★ 162
  • awesome-jevanother curated Jev index, links to madewithjev.com use casesresearchgithub★ 1832
  • phuryn/experiments (jev-decisions-api)hardened invoice test: Jev ties Haiku 4.5, beats Opus 115x on costevalfinancegithub★ 874
  • Sandobounds Claude Code and Codex tool output, TypeSafe shadow judge optionalguardrailscoding-agentgithub★ 514
    2 similar
    • compactioJev shrinks tool output 20501 to 747 chars in 0.75s for $0.00005coding-agentopsgithub★ 43
    • codex-context-dietJev trims bulky Codex tool output to a bounded head plus notecoding-agentopsgithub★ 63
  • jev-codex-routerper-turn Codex model routing, measured 60% savings on 237 turnsroutingcoding-agentgithub★ 2734
    3 similar
    • Jev Auto Router (miniLV)Jev picks a GPT model tier per Codex call, verified after.routingcoding-agentgithub★ 82
    • Codex SiftRoutes each Codex turn to cheapest model lane, judged by Jev.routingcoding-agentgithub★ 84
    • codex-jev-router (suenot)Routes Codex subagents to model tiers using typed Jev decisions.coding-agentroutinggithub★ 83
  • jev-rulesJev picks which CLAUDE.md rules apply to each promptcoding-agentclassificationgithub★ 663
  • jev-prunertrims noisy Bash output before Claude sees it, keeps text verbatimcoding-agentopsgithub★ 1613
  • skillrankerRust CLI ranks Claude Code skills for the next stepclicoding-agentgithub★ 1323
  • save-token-jev-cleanJev picks which tool calls survive compaction across five agent hostscoding-agentopsgithub★ 803
  • jev-semgrepgrep by meaning across languages, no translation step neededsearchcoding-agentgithub★ 1493
  • supercovscores code quality with Jev, pairs with test coverage gapscoding-agentevalgithub★ 1523
  • jev-siftclassifies files and URLs by relevance before the agent reads themclassificationcoding-agentgithub★ 473
    4 similar
    • jev-assistRanks repo files by relevance to a task using Jev.coding-agentopsdiscord★ 43
    • genigrepCode search for agents: Jev ranks files, cutting agent cost about 7%.searchcoding-agentgithub★ 34
    • JevTraceMCP server retrieves only relevant TS/JS code for coding agents using Jev.coding-agentmcpgithub★ 34
    • fileroute (fr)Jev scores file relevance to a question, live and in parallel, no indexing.searchcoding-agentgithub★ 54
  • advocaattyped client for asking Jev questions about your datasdkdatagithub★ 963
    1 similar
    • vibecheck (jlowin)Python check/classify/label/score functions wrap decision models like Jev directly.sdkgithub★ 424
  • decideropen one-pass typed decision model, fine-tuned Qwen3.5-2B, HF benchmarks includedresearchevalgithub★ 1.1k3
  • jeff (GliFormer)self-hosted drop-in Jev replacement, cheaper but less accuratesdkgithub★ 2983
  • fabricioctelles/skills26 agent skills, four use Jev for quality judgmentcoding-agentevalgithub★ 1063
  • commit-minerclassifies commit diffs by bug fix, CWE and change typeclassificationsecuritygithub★ 393
  • vibecheckscores draft X posts on virality, cringe and regret risk before postingwritingclassificationgithub★ 503
  • JevScoutJev scores job links and postings while Chrome drives the browsingbrowserclassificationgithub★ 382
  • typesafe-ai-benchmarkcompares Jev against LLM-native structured output on latency, cost and qualityevalgithub★ 383
  • mini-Jevpreregistered test reads option-letter logits instead of generating JSONresearchevalgithub★ 593
  • building-with-jev-skillClaude Code skill for writing and debugging Jev-calling programscoding-agenteducationgithub★ 1553
    3 similar
  • fast-jev-compactionJev scores each tool call, drops stale ones, keeps rest verbatimcoding-agentopsgithub★ 7.6k3
    1 similar
    • jev-compaction-plusCompacts Claude Code sessions in 0.5s vs 35s, keeping needed context word for word.coding-agentopsgithub★ 263
  • compact-adviserJev judges when a Claude Code session is safe to compactcoding-agentopsgithub★ 2123
    2 similar
    • claude-herdr-jev-compactionJev judges Claude Code task completion to trigger compaction timing.coding-agentdiscord★ 03
    • jev-contextJev decides what to keep when Claude Code or Codex compacts context.coding-agentsdkgithub★ 83
  • claude-code-traceJev scores Claude Code sessions on progress, focus and token efficiencycoding-agentevalgithub★ 3753
  • jev-review (dashboard)staged Jev code review workflow with a local dashboardcoding-agentgithub★ 6823
  • perchsemantic linter flags defects with confidence scores, CLI and agent skillcoding-agentcligithub★ 3513
    2 similar
    • jev-lint (schalkneethling)Experimental semantic code linter that asks Jev to judge code quality.coding-agentclassificationdiscord★ 02
    • noulsSemantic linter: batches yes/no Jev questions per function for intent bugs.coding-agentevalgithub★ 83
  • foremanJev watches Codex workers, judges completion and drift, labeled an experimentcoding-agentevalopsgithub★ 7372
    1 similar
    • jev-codex-pluginCodex plugin for Jev tool/task consultation and completion-claim verification.coding-agentmcpgithub★ 122
  • evotterminal coding agent, Jev prunes stale context instead of summarizing itcoding-agentcligithub★ 3313
  • openagents (Coder)terminal coding agent, Jev routes turns and judges shell outputcoding-agentcliroutinggithub★ 4552
  • vexjoy-agentJev routes plain-English requests across 43 specialist Claude Code agentsroutingcoding-agentgithub★ 4413
  • jev-router (per-turn)routes each Claude Code or Codex turn to the cheapest capable modelroutingcoding-agentgithub★ 5822
    4 similar
    • Jev Model Router for Claude Code (AlexPEClub)Jev hook routes each Claude Code prompt to cheapest capable model.routingcoding-agentgithub★ 84
    • model-switcherRoutes Claude Code prompts to cheap or heavy models, with optional Jev pre-check.routingcoding-agentgithub★ 53
    • reflex-routerProxies Claude Code, uses Jev to judge task difficulty and route to cheaper model.routingcoding-agentgithub★ 53
    • coding-router-jevRoutes Codex and Claude Code turns through Jev for model and effort selection.coding-agentroutinggithub★ 23
  • jev-mcp (10 tools)verify, screen, rerank, classify, review and gate as MCP toolsmcpguardrailsclassificationgithub★ 5174
    7 similar
    • jev-judge-mcpMCP server exposing eleven Jev-backed judgment tools for coding agents.mcpcoding-agentgithub★ 724
    • jev-mcp (rashedInt32)MCP server exposes Jev's classify, score and check as Claude Code tools.mcpcoding-agentgithub★ 74
    • decisions-judge-mcpMCP tool giving coding agents noul/choice/score judgments with graceful fallback.mcpcoding-agentdiscord★ 54
    • jev-mcp (burnigtm)MCP server for coding-loop routing, review, verify and screen toolsmcpcoding-agentgithub★ 623
    • jev-mcpMCP server exposing classify, score, check, match, screenmcpclassificationgithub★ 25
    • typesafe-mcpdrop-in MCP connector for Claude Codemcpcoding-agentgithub★ 349
    • system-one-connectorMCP connector giving coding agents typed Jev, Laya or CLM judgments.mcpgithub★ 3493
  • tax-doc-classifierclassifies pages across 261 IRS forms, 100% accuracy, $0.001 per pageclassificationfinancegithub★ 5034
  • classifier.devzero-shot text classification over HTTP, no key, Jev backendclassificationsdkgithub★ 4253
  • llm-rankers (Jev arm)academic reranker benchmark adds Jev as a pointwise and listwise rerankerevalsearchresearchgithub★ 2143
  • jev-eval-agentcompares agent step counts when Jev picks the tool versus the LLMevalcoding-agentgithub★ 1063
  • jev-ultrafastJev picks click target and action, small LLM only types textbrowsergithub★ 22.6k3
  • jev-browserJev picks one browser action per step; Wikipedia nav in 4sbrowsergithub★ 3293
    5 similar
    • Jev Browser (openqa-cn)Jev picks controls from a page index, Playwright acts on them.browsercoding-agentgithub★ 873
    • fastbrowseJev picks browser actions from page candidates, claims 71x cheaperbrowsercoding-agentgithub★ 1144
    • jev-ultrafastJev picks click target and action, small LLM only types textbrowsergithub★ 22.6k3
    • jev-browsercheap Playwright browser agent driven by Jevbrowsergithub
    • jev-browser-wingmanJev-driven browser automation; completed 34/34 tasks vs Playwright's 32/34, 30% cheaper.browsercoding-agentgithub★ 54
  • jev-browser-useJev clicks and navigates, Codex verifies; 5-10x faster in productionbrowsercoding-agentgithub★ 1.1k3
  • SemIfopen 4B model reproducing Jev's typed-decision interface, runs via WebGPUresearchsdkgithub★ 4.8k2
  • kevtiny Jev-like model on Qwen2.5, trains and runs on a MacBookresearchsdkgithub★ 8.9k2
  • layanon-autoregressive multilingual decision model, 33ms per call on a T4researchclassificationgithub★ 32.3k3
  • LangWatch Instant EvalsJev judges your whole production history on demand; 97% agreement with human labels on 300 support chats, $0.32 per 10k conversations (vendor's claim)evalopsdiscord
    1 similar
  • JevSubRouterhook routes each Claude Code subagent dispatch to the cheapest model that can finish it; ~300ms warmroutingcoding-agentgithub★ 4
    4 similar
    • jeffJev picks model tier and parallelism for each Claude Code subtask.coding-agentroutingdiscord★ 14
    • jev-delegateClaude Code skill: Jev rates task difficulty, picks Haiku, Sonnet or Opus.routingcoding-agentgithub★ 44
    • jev-router (hajdu-patrik/claude-workspace)Routes each prompt to a model, effort level, and skill via Jev.routingcoding-agentgithub★ 153
    • claude-subagent-routerJev classifier picks Sonnet vs Opus per Claude Code sub-task, cuts tokens.coding-agentroutinggithub★ 104
  • toolgateClaude Code hook and MCP proxy blocks, confirms or allows tool calls by riskguardrailsmcpsecuritygithub★ 24
  • jev-codeTypeScript coding CLI using Jev typed decisions and constrained AST generationcoding-agentcligithub★ 20
  • jevusecasescrowdsourced directory of what people replaced with Jevdemodiscord
  • saypagedescribe a page in plain language, builds a standalone site in under 0.4sdemomediadiscord
    1 similar
    • FormaJev picks design settings and layout for a local website builder.demomediagithub★ 261
  • jev-cvssCVSS 3.1/4.0 vectors from a plain description, matches NVD on sample setsecurityclassificationgithub★ 2
  • jev-browser (MCP)LLM plans, Jev decides click target and action, Playwright executesbrowsermcpgithub★ 1103
    5 similar
    • jev-ultrafast-mcpMCP server drives whole browser flows for agents in one tool call.mcpbrowsergithub★ 193
    • FlickMCP server runs whole browser and macOS goals via observe-decide-act with Jev.browsermcpgithub★ 262
    • jev-browser-controlJev picks each Chrome click for Claude or Codex browser automation.browsermcpgithub★ 54
    • DEMO MCP (amidevz)MCP server with built-in Jev and Laya plus browser tools.mcpbrowserdiscord2
    • fast-browserCodex/MCP browser automation using Jev or local Laya to choose actions.browsermcpgithub★ 42
  • vonopen-source non-autoregressive decision model, sub-15ms local alternative to Jevresearchsdkgithub★ 8823
  • jevbetterstronger one-pass option scorer, benchmarked head-to-head against jevlikeevalresearchgithub★ 153
  • OpenDecisionopen-source Choice/Noul/Score engine over zero-shot modelssdkclassificationgithub★ 583
  • duckdb-jevDuckDB extension, Jev answers typed as real SQL columns per rowdatasdkgithub★ 293
    4 similar
    • vgi-typesafeDuckDB extension, Jev as LATERAL joinsdatagithub★ 4
    • MotherDuck prompt_jev()SQL function powered by Jev classifies text 50x faster at 1% cost.dataclassificationx5
    • anofox-decideDuckDB extension evaluates natural-language predicates using Jev or local decision models.dataclassificationgithub★ 23
    • JEVDBDuckDB extension running semantic SQL filters and joins through Jevdataraggithub★ 94
  • openjevlocal bilingual probability decisions on a frozen Qwen3-4B backendresearchsdkgithub★ 363
  • tidepoolHaskell notebook agent harness, Jev supplies System 1 judgmentscoding-agentresearchgithub★ 123
  • oxlint-plugin-jevplain-English yes/no lint rules, Oxlint finds the code, Jev judges itcoding-agentguardrailsgithub★ 863
  • jev-benchmarkscalibration, selective risk and latency eval, Jev vs GLiNER2.5evalclassificationgithub★ 264
  • jevilJev + agent-device QA agent, reads a mobile app and picks the next actionbrowserevalgithub★ 203
    1 similar
    • jev-sim-useJev navigates iOS Simulator and Android screens one call per step.browsermcpgithub★ 102
  • jev-macos-loopnative macOS GUI automation, OmniParser plus Jev for action choicebrowseropsgithub★ 243
  • jgrepgrep by natural-language description, pipeable, ~200ms per lineclisearchgithub★ 1384
    2 similar
    • jegrepsemantic grep with calibrated per-path yes/no probabilities, no indexsearchcligithub★ 1153
    • jev-semgrepgrep by meaning across languages, no translation step neededsearchcoding-agentgithub★ 1493
  • macbrowvoice-controlled Mac, Jev picks the AppleScript or browser toolbrowsercligithub★ 1573
    1 similar
    • Alfred (computer-use)Voice-controlled Mac agent where Jev picks the typed action after local speech recognition.cliopsgithub★ 253
  • choosekittyped choices and probabilities from local llama.cpp modelssdkgithub★ 233
    1 similar
    • llamacpp-jevServes Jev's structured decision API locally in front of llama-server.sdkclassificationgithub★ 83
  • beam-cli (AgentBeam)hooks and local policy layer across Claude Code, Codex, Jevcoding-agentguardrailscligithub★ 143
  • Turncompiled language, Choice/Score/Noul decisions as durable VM effectssdkresearchgithub★ 122
  • system-onebatched single-token choice inference for open models, Jev-compatiblesdkresearchgithub★ 373
  • open-alternative-jevopen System One layer for open-weights LLMs, benchmarked on RACE-Hresearchevalgithub★ 674
    8 similar
    • OpenSourceJevTurns a local Qwen model into a sub-100ms decision engine.researchgithub★ 432
    • litjevturns any Qwen checkpoint into a Jev-style typed decision layerresearchgithub★ 473
    • openjevlocal bilingual probability decisions on a frozen Qwen3-4B backendresearchsdkgithub★ 363
    • LLM2Jevadapts local LLMs into Jev-style Choice, Score and Noul outputssdkgithub★ 4293
    • snapjudgeReads typed-decision probabilities from local Qwen model logits, TypeSafe-compatible.classificationsdkgithub★ 143
    • QwevBuilds a Jev-style decision interface on local Qwen checkpoints, no training.classificationresearchgithub★ 32
    • typed-lmTurns dense LLMs into typed choice/bool/score APIs in one forward passclassificationsdkgithub★ 104
    • typed-decisions (kotoba-lang)Reproduces Jev's typed-decision shape on ModernBERT, DeBERTa, LLaDA-MoE with benchmarks.researchclassificationgithub★ 103
  • jotgeneral-purpose agent loop where Jev picks the tool and argumentscoding-agentroutinggithub★ 203

Coding-agent guardrails (Claude Code hooks)

  • jev-commitcommit-msg hook: message vs diff, leaked creds, debug leftoversguardrailssecuritycoding-agentgithub★ 15
  • jev-belayStop hook that catches "tests pass" when nothing ranguardrailscoding-agentgithub★ 37
    4 similar
    • CannyHooks block agents from claiming done without a passing check, Jev advisescoding-agentguardrailsgithub★ 1224
    • jev-no-bullshitJev flags unverified or vague coding-agent summaries and forces a recheck.guardrailscoding-agentgithub★ 33
    • JevFlow (Parth1811)Claude Code plugin: Jev judges if a task phase is really donecoding-agentevalgithub★ 94
    • claude-refereeClaude Code plugin checks agent's done claims with Jev and logs results.coding-agentevalgithub★ 64
  • pi-wardenjudges every tool call, output and reply in ~250msguardrailssecuritygithub★ 167
    1 similar
    • ReflexJev screens every tool call and turn for a Pi coding agent.coding-agentbrowsersdkgithub★ 53
  • second-thoughtshell commands judged pre-execution, blocks past 85% confidenceguardrailssecuritycligithub★ 4
  • jev-guardauto-approval layer for Claude Code and Codex tool calls, ships a bypass CTFguardrailssecuritycoding-agentgithub★ 3
    2 similar
    • claude-x-jevClaude Code skill: Jev pre-sorts and gates calls before Claude readscoding-agentroutinggithub★ 532
    • jev-permission-gateJev pre-screens Claude Code auto-mode tool calls, 2x faster, zero unsafe allows in tests.guardrailscoding-agentgithub★ 294
  • omp-auto-modesafe/unsafe/ask tool-call classifier emulating auto modeguardrailsclassificationgithub★ 3
  • pi-heedenforces "never touch prod" constraints on tool calls, 46 testsguardrailsopsgithub★ 12
  • tenetjudges every agent commit against plain-English rules in AGENTS.mdguardrailscoding-agentgithub★ 5
  • abidemakes the coding agent obey project rulesguardrailscoding-agentgithub★ 578
  • Reapersemantic linter for silent failures, weakened tests, scope creepcoding-agentguardrailsgithub★ 2
  • jevlintplain-English convention rules checked in the agent loopcoding-agentguardrailsdiscord
  • commentlintchecks that agent-written comments still match the codecoding-agentguardrailsgithub★ 1
  • wincescores diffs by blast radius and auth/data-write for review routingsecuritycoding-agentroutinggithub★ 2
  • Migration Guardianhalts destructive SQL migration plansguardrailsdataopsgithub★ 2
  • is-maliciousscreens a repo or PR for obviously malicious code before running itsecurityguardrailsgithub★ 37
  • Gorgonazero-dependency Python guardrail engine, halts agents on leaked keys or loopsguardrailssecuritygithub★ 2

RAG, grading, classification (GWEDU)

  • Jev vs rerankersties Cohere rerank-4-pro on NDCG@10 across 8 datasetssearchevaldiscord
    2 similar
    • jev-rerankingBenchmarks Jev zero-shot reranking against monoBERT and BM25 on TREC.evalragsearchgithub★ 104
    • llm-rankers (Jev arm)academic reranker benchmark adds Jev as a pointwise and listwise rerankerevalsearchresearchgithub★ 2143
  • GBrain Jev reranker(PR by Daniel Andrade, see channel)Score-based reranking, 1.9s vs 14.4s LLM at similar cost
  • ReadyBase rerank eval28-scenario gold-labeled harness for code-context selectionevalsearchcoding-agentdiscord
  • extraction-marker-recoveryrestores OCR-flattened footnote and citation markers in passagesdataresearchgithub★ 0
  • jev-document-classificationdocument classes plus injection detectionclassificationsecuritygithub★ 5
  • book-aurora601 passages, 9 emotion scores, 25s, $0.034classificationmediagithub★ 6
  • AI Elo rankertournament scoring for text, reusable for rubric gradingevalgithub★ 7
  • sqlite-jevbatched judgments as SQL functions in SQLitedatasdkgithub★ 4
  • pg-redacttags PII spans in Postgres text, redacts only the spansecuritydatagithub★ 4
  • FairCheckRAG-grounded guardrail citing regulation paragraphs; pattern for cited refusalsguardrailsragdiscord
  • sentence-relationship classifiernon-LLM segmentation plus 3 Jev calls, ~0.5sclassificationdiscord

Eval and adoption method

  • jev-sec-bench96.5% on deepset prompt-injection set, 325ms p50; second opinion beside the intake regexsecurityevalgithub★ 3
  • jev-watchbaseline today's answers, flag drift when TypeSafe updates the modelevalopsgithub★ 0
  • Sniff Testprose linter as CLI, pre-commit, Action and Claude Code skill; 182ms, $0.013 per 100 paragraphswritingcligithub★ 34
    2 similar
    • slop-graderRule-based text grader runs every rule against every line in parallel with Jev.evalwritinggithub★ 333
    • riffRuff-style prose linter flags writing issues using calibrated Jev judgments.writingcligithub★ 73
  • shadow-mode adoption write-uprun Jev beside the LLM for 3 days, promote only where it winsopsevalx
  • Working with Jevpractitioner notes from a real PR reviewercoding-agentwritingdiscord
  • CultivarPinecone's agent-skill eval harness with Jev gradingevalcoding-agentgithub★ 42
  • dinostompverifies the scorer, not just the score, with a blind control per runevalgithub★ 6
  • Vercel eve-agent failure classifier320 sessions, 2M tokens, classified in 14s for $0.08classificationevalcoding-agentx
  • Empryo harness write-upfive decision points in a coding agent benchmarked against frontier modelsevalcoding-agentdiscord
  • Swamp alert triage91% cheaper triage before routing to Clauderoutingopsclassificationdiscord
  • ZeroSweepJev vs LLM inbox triage, live comparisonclassificationevalgithub★ 2
  • optimaizrranks LLM spend waste by dollar savings on real trafficfinanceopsdiscord

SDKs, adapters, glue

  • Jev on OpenRouter$0.042/M input, $0 output; plugs into the broker with no new adapterroutingsdkdiscord
  • JevRouterJev-powered router for models, tools and subagentsroutingcoding-agentgithub★ 458
    1 similar
    • jev-router (kalowery)Routes coding-agent requests to cheaper models using Jev prompt classification.routingcoding-agentdiscord★ 04
  • daf-jevPython toolkit: confidence gates, tiered routing, calibration module, MCP serverroutingmcpguardrailsgithub★ 7
  • dspy-typesafeifyone decorator for Jev-backed DSPy signaturessdkcoding-agentgithub★ 66
    1 similar
    • DSPy native Jev supportDSPy signatures now run Noul, Choice and Score directly on Jev.sdkclassificationdiscord4
  • jev-mcpMCP server exposing classify, score, check, match, screenmcpclassificationgithub★ 25
  • typesafe-mcpdrop-in MCP connector for Claude Codemcpcoding-agentgithub★ 349
  • jev-cliverify, screen, classify, extract, route, rerank from the shellcliclassificationgithub★ 22
  • jev-axipick, rate, check, rank, triage, guard shell commandscliguardrailsgithub★ 27
  • simple-jevturn any open model into a classifier endpoint; fallback if TypeSafe goes awayclassificationsdkgithub★ 599
  • System One SearchJev-only codebase retrieval and reference graphsearchcoding-agentdiscord
  • everyask a yes/no question of every function in a codebasecoding-agentresearchgithub★ 7
  • jev-browsercheap Playwright browser agent driven by Jevbrowsergithub
  • typesafe-computer-useOCR, classify, click loop at ~$0.0002 per stepbrowserclassificationgithub★ 1.2k
  • vgi-typesafeDuckDB extension, Jev as LATERAL joinsdatagithub★ 4
  • pg-jevPostgres extension, plain-English per-row WHERE judgmentsdatadiscord
  • HA-JevHome Assistant sensors and automationsopsgithub★ 75
  • awesome-jev-typesafe250-entry curated indexresearchgithub★ 193
  • awesome-jev-projects287-entry index with English viewresearchgithub★ 681
    6 similar
    • awesome-typesafeCurated list of official and community TypeSafe and Jev resources.writingdatadiscord★ 5802
    • awesome-jev-toolsanother curated index of public projects and patterns built on Jevresearchgithub★ 7452
    • awesome-jevanother curated Jev index, links to madewithjev.com use casesresearchgithub★ 1832
    • awesome-jev-typesafe250-entry curated indexresearchgithub★ 193
    • awesome-jev (yibie)Curated list of agent harnesses, skill routers and Jev workflowscoding-agentdemox★ 2.3k2
    • awesome-jev-use-cases (walidboulanouar)Curated list of 74 Jev demos and 150+ repos with costsresearchdiscord★ 4182
  • TypeSafe primitives quiz(dball9, quiz_handover.md in channel)self-test on Choice, Noul, Score and the cascade patterns

Cool

New 2026-10-11

  • Frankie PokerOpen-source poker playground lets you play Texas Hold'em against Jev.gamegithub★ 51
  • Pokémon TCG AcademyPokémon TCG trainer app lets you play against Jev as an opponent.gamegithub★ 31

New 2026-10-10

  • convex-chessRates each chess move 1-10 using Jev via Convex AI Gateway.gamedemogithub★ 51
  • ra2web-jev-playerBot plays RA2 skirmish matches using a semantic jev decision model.gamedemogithub★ 22
  • Blackout Kiss (Laknicek)Song composed with Jev and Opus 5.5 model collaboration.mediadiscord1
  • THAT THAT GAMEDaily word game where Jev rates how clever your answers are.gameclassificationdiscord1
  • Age of LLM Benchmark Viewer1v1 war game benchmark where Jev and LLMs duel via typed choices.gameevalgithub★ 152
  • River OaksJev powers NPC dialogue and behavior in a stylized Houston game world.gamedemogithub★ 61

New 2026-10-09

  • tomarigi-desktopShows Claude Code/Codex sessions as birds, flags abusive messagescoding-agentdemogithub★ 172
  • jev-plays-pokemon-red (valentynkit)Jev plays Pokemon Red choosing moves from legal options at 100msgamedemogithub★ 82
  • jev-skipSkips YouTube sponsors by reading captions, no crowd database neededmediademogithub★ 72
  • atomaAgents produce and check creative or technical work; Jev makes bounded callsdemoroutinggithub2
  • jev-turtle-soupTurtle-soup riddle game community where Jev judges yes/no answersgamegithub★ 51

New 2026-10-08

  • jev-arena-nanojevLocal Pygame tactical arena where NanoJev picks combat moves each turn.gamedemogithub★ 51
  • model-vs-marketThree decision models price prediction markets live against real odds.financedemogithub2
  • jev-soundboardListens to calls and fires matching soundboard clips via Jev.mediademogithub★ 41
  • open-annieJev picks a 3D character's gestures and face from live speech.mediademogithub★ 272

New 2026-10-07

  • saanLocal semantic file search launcher for Windows, Rust and Tauri.searchclidiscord★ 32
  • Competitor HunterClaude Code scrapes competitors' posts; Jev scores which ones worked.coding-agentmediagithub★ 122
  • VeyraRust trading service gates Jev and LLM decisions behind a risk check.financeopsgithub★ 62
  • Living Matter3D world predicts player intent in real time from Jev state updates.gamedemodiscord1
  • OrbsSelf-hosted multi-bot chat room where Jev decides which bot replies.routingdemogithub★ 362

New 2026-10-06

  • jevmanPac-Man driven by Jev and other decision models with leaderboardgamedemogithub★ 71
  • jev-bisectFinds a number via repeated higher, lower, or equal Jev decisions.demosdkdiscord★ 01
  • HeliotropeCarbon-aware load scheduler uses Jev to classify and prioritize appliance loads.opsclassificationgithub★ 52

New 2026-10-05

  • Peon (WoW harness)Pi-based harness lets an LLM play World of Warcraft, Jev handles combat decisions.gamecoding-agentgithub★ 92
  • JevBirdJev picks a flight path per pipe to play Flappy Bird live, one call per pipe.gamedemogithub★ 81

New 2026-10-04

  • trade-jevBacktests Jev as a BUY/SELL/HOLD trader on NQ order-book data.financeevalgithub★ 122
  • JevboardUses Jev to re-pick homophones in Chinese Bopomofo IME output before Enter.classificationdemogithub★ 182
  • Jev FSD / Drive LabDriving simulator where Jev drives real OpenStreetMap streets with a decision inspector.roboticsgamedemogithub★ 72
  • GenClassBrowser-local typed-decision model and Chrome extension for voice-driven page actions.browserclassificationgithub★ 33
  • Slop Scope (local)Judges whether an image is AI-generated from browser-measured pixel facts via Jev.mediaclassificationgithub★ 32
  • JevspressoJev picks espresso machine actions from an XState-allowed set.demogamegithub★ 252
  • world-narrativeFramework turns worldbuilding proposals into text adventures via Jev gates.writinggamegithub★ 22

New 2026-10-03

  • AbstractleDaily writing game where Jev judges accuracy and poetic distance from a target word.gamewritingdiscord2
  • BasteUses decision models like Jev to turn a persona's taste into a design system.democlassificationdiscord2
  • JevolutionMulti-agent ecosystem simulator for studying endangered species, hackathon winner.researchdemox2
  • ldraw-novaAgent builds LDraw LEGO models using Jev-backed semantic search reranking.ragsearchdemogithub★ 3402

New 2026-10-02

  • Jev vs Jev Fighting GameTwo Jev-controlled fighters pick moves and adapt strategy live in a 2D fighting game.gamedemodiscord1
  • 20 Questions (Jev)Guess what Jev is thinking of through a game of twenty yes/no questions.gamedemodiscord1
  • ChimpBenchSimulates chimpanzee troop behavior in Kibale, Uganda using stateful decision models.researchdemodiscord2
  • DriveJevOpen System I driving model, clears 241 of 242 scripted hazards in 111ms per decision.roboticsresearchgithub★ 573
  • JevOnlyBrowser agent that only uses Jev to pick actions, no LLM, completes multi-step web tasks.browserdemogithub★ 52
  • nethack-jevJev plays NetHack by picking from a legal-move list, no LLM, 570 games played so far.gamedemogithub★ 32
  • say-hiGets Jev, a pure chooser, to type chat replies one keystroke at a time.demogithub★ 31

New 2026-10-01

  • dsh-memes-replyJev picks contextual animated sticker replies per chat turn.mediademogithub★ 61
  • Fly x JevJev drives descending neurons of a simulated fruit fly connectome in browser.researchdemogithub★ 72
  • Spawn Game Time (Blastore)Jackbox-style Family Feud game judges free-text answers against categories using Jev.gamedemodiscord1
  • Play vs Jev (chess)Jev plays chess, picking moves and flagging low-confidence fallback decisions.gamedemodiscord1
  • TattleLive transcribes calls and fact-checks claims using Jev then a bigger model.mediaclassificationdiscord★ 13
  • SVPN NewsLive-ranked news site uses Jev for quality ranking and discussion moderation.classificationwritingdiscord3
  • eR33t Growing CellsBrowser lab compares exact Conway rules against a frozen Jev-generated transition table.gameresearchdiscord2
  • jev-fundOpen-source paper hedge fund where Jev decides buy, sell or hold.financedemogithub★ 82
  • switchyardMenu-bar app routes links to the right browser profile, learns via Jev.routingcligithub★ 22
  • car-systemui-pingpong-podJev plays ping-pong paddle AI inside Android Automotive's system UI panels.gamedemogithub★ 21
  • jev-talksExperiment testing how well Jev, a non-generative classifier, can answer questions.demoresearchgithub★ 61

New 2026-09-30

  • Wordle Beyond the ListJev-guided Wordle solver averages 3.10 guesses over 633 real puzzles.gamedemodiscord★ 02
  • hikomi3D AI companion uses Jev for sub-second reactions to voice and screen.mediademodiscord2
    1 similar
    • open-annieJev picks a 3D character's gestures and face from live speech.mediademogithub★ 272
  • JEV SeesAdds a vision front end so Jev judges objects in camera frames.roboticsdemogithub★ 2432
  • enigma-jevJev judges crib guesses and decrypt attempts to break real Enigma traffic.securityresearchdemogithub★ 32
  • TV with Muscular ArmsAI TV host predicts the future live using Jev.mediademodiscord1
  • Jev plays Vampire SurvivorsJev plays Vampire Survivors live on Steam via a game-state dashboard.gamedemogithub★ 52
  • Jev CivilizationsJev decides every tribe's move each turn in a civilization sim.gamedemogithub★ 31
  • FormaJev picks design settings and layout for a local website builder.demomediagithub★ 261

New 2026-09-29

  • asbplayer-miningAnki mining fork of asbplayer that uses Jev to rank unknown Japanese words live.mediaeducationgithub★ 32
  • Rummage TVJev-powered 90s shopping-channel interface for browsing eBay listings.demomediadiscord2
  • thusfarGoal journal where Jev links diary entries to your stated goals.writingclassificationdiscord2
  • Jev Plays Pokémon Emerald (trolloks)Live Twitch stream of Jev playing Pokémon Emerald.gamediscord1
  • JEV-ArcadeFifteen live games pitting Jev against LLMs: chess, WikiRace, guardrails.gameevaldiscord2
  • strudellm Jev songsJev makes decisions for AI-generated electronic music tracks via strudel.cc.mediademodiscord★ 01
  • Jev the SpireJev played and beat Slay the Spire 2 after 182 runs.gamedemodiscord1
  • FC TankClassic Battle City tank game with a Jev-controlled enemy AI difficulty.gamegithub★ 251
  • csgo-julia-plugin (jev-bot)CS:GO Legacy bot with tactics chosen by a Jev-style decision model.gamegithub★ 261

New 2026-09-28

  • DIOVRAPaper trading app using Jev to score ticker buy or sell consensus.financedemodiscord1
  • Jev-TradesTrading bot that uses Jev for fast calibrated trade decisions.financedemodiscord★ 481
  • Jev TraderPlaces one automated crypto trade decision every Monad block.financedemodiscord1
  • jev-doomWatches Jev play Freedoom with inspectable decisions and bounded spending.gamedemodiscord★ 22

New 2026-09-27

  • rekordbox-jevExperimental DJ track-mixing decisions driven by Jev, early prototype for macOS.mediademodiscord★ 31
  • Jev Realtime (trading)Desktop app where Jev decides paper-trading buy or sell every second.financedemogithub★ 72
  • jevlangPython dialect where every if and while is judged by Jev.democoding-agentgithub★ 31
  • OssoBrowser extension that strikes filler from articles sentence by sentence.browserwritinggithub★ 163
  • Jev DriverBrowser driving game where Jev decides stop, slow or go.gamedemogithub2

New 2026-09-26

  • Jev PianoJev improvises piano note choices without generating any notes itself.mediadiscord★ 61
  • What to OrderJev picks matching menu dishes from what you're craving.demodiscord1
  • AppWithAI Flight SimulatorBrowser flight simulator demo built with AI, seed-based scenarios.gamedemodiscord1
  • jevstrudelAI-written Strudel songs about Jev; pick one, hear and edit it.mediadiscord1
  • MastiloTurns two lines of text into a 15-second hand-inked video.mediademodiscord1
  • vox-arcanaBrowser magic-duel game where spoken incantations become Jev-parsed spells.gamemediagithub★ 91
  • JauvexVoice-controlled Electron app running Claude and Codex sessions, Jev for fast callscoding-agentdemogithub★ 62
  • herobrineMinecraft agents where Jev picks every action; Claude/GPT plans and speaksgamedemogithub★ 21
  • sentimento em tempo realLive speech transcription scored for six emotions by Jev every 0.5smediaclassificationgithub★ 62
  • voicevox-jev-proxyJev fixes Japanese TTS word readings and accent breaks before synthesismediaclassificationgithub★ 22
  • jev-demofastJev drives a real browser to record a narrated product demo videobrowsermediagithub★ 32
  • jev-crushJev reads a pasted chat log for emotion, intent and interest scoreclassificationmediagithub★ 21
  • Radar de HookBrowser extension marks Instagram/X/YouTube posts worth copying using your own Jev keybrowsermediagithub★ 21
  • Jev Computer UseVoice controls a Windows PC; Jev picks the app or action meantdemoclassificationgithub★ 22
  • GroundingJevJev-style single-pass bounding-box regression for visual grounding on Qwen3.5-0.8Bresearchclassificationgithub★ 102
  • quackdCLI harness pairing Jev decisions with a LeRobot arm.roboticsclidiscord★ 2572
  • Doom or BloomInterview mapping your AI worldview against public figures, via Jev.demodiscord1
  • Jev Plays Pokémon RedJev plays Pokémon Red live by picking from legal game options.gamegithub★ 1281
    1 similar
  • jev-spellToon-shaded FPS whose magic effects are generated via a Jev self-improvement loop.gamediscord1
  • Jev call-center voice demoVoice agent routes live call-center inquiries using Jev judgments.demodiscord1
  • Jev DrummerReal-time AI drum-beat collaboration demo powered by Jev.mediadiscord1
  • Jev at the PianoJev improvises piano by choosing among candidate notes.mediadiscord1
  • beebotsThree paper-trading bots on OKX driven entirely by Jev decisions.financedemogithub★ 2712

New 2026-09-25

  • jev-sales-copilotTracks live sales-call closing probability and coaching moves via Jev.democlassificationgithub★ 92
  • jev-voice-browserJev picks browser intent and target per spoken word in 300ms.browserdemodiscord★ 3992
    1 similar
    • GenClassBrowser-local typed-decision model and Chrome extension for voice-driven page actions.browserclassificationgithub★ 33
  • jev-riffsMines candidate motifs from MIDI then Jev grades their musical significance.mediademodiscord★ 32
  • Jev Arena (Eliot5566)Jev pilots English-described fighters live, several decisions per second.gamedemodiscord★ 12
  • Fit ReceiptJev scores fitting-room candidates in parallel, hands off low-confidence picks.classificationdemodiscord2
  • Nagi ArenaReal-time physics game arena pits Jev against Laya, OpenJev and Nagi.gamedemodiscord2
  • Jev Lab100 NPCs live in a town; Jev decides each person's next action.gamedemogithub★ 72
  • partwaymacOS voice control that acts mid-sentence on ~200ms Jev decisions.mediademogithub★ 312
  • jev-2048-seleniumPlays 2048 by blending expectimax search with a Jev move choice.gamegithub★ 31
  • jev-keyboardJev reranks pinyin input candidates for Rime IME on macOS.demogithub★ 42
  • GobanGo game where Jev or players duel; self-play version improves like AlphaGo.gamediscord1
  • grevRebuilds Unix coreutils so Jev reasons about each command's behavior.clidiscord★ 562
  • Magic 8 Ball (ozoromo)Magic 8-ball demo where Jev tracks mood and holds a grudge.gamedemodiscord1
  • Text-box-to-UI demoX demo: a text box renders whatever UI you type, powered by Jev.demomediadiscord1
  • Math-To-ManimJev advisory-reviews GPT-6 Astra's math animation pipeline at each checkpoint.mediaevalgithub★ 2.8k2
  • Eikos ArenaLive paper-trading arena pits open Eikos-27B against Jev on Hyperliquid.financedemogithub2
  • jev-helper (ra2web)Chrome extension lets Jev or a local model play the RA2 web game.gamegithub★ 52
    1 similar
    • ra2web-jev-playerBot plays RA2 skirmish matches using a semantic jev decision model.gamedemogithub★ 22

New 2026-09-24

  • SoupbaseJev hosts a lateral-thinking puzzle game as narrator and judgegamedemogithub1
  • Jev Plays Competitive PokémonJev picks competitive Pokémon moves and switches in real matches.gamedemodiscord2
  • Sleep/Recovery Control Demo (daniel.csv)Jev controls an Eight Sleep mattress from a live WHOOP stream.healthdemodiscord1
  • nagiJev, Laya, Semif and a custom model play Bomberman live.gamedemodiscord★ 132
  • Hey JevMac voice assistant, Jev decides command vs question for $0.00004/call.sdkdemogithub★ 1052
    1 similar
    • partwaymacOS voice control that acts mid-sentence on ~200ms Jev decisions.mediademogithub★ 312
  • JEV × Lighter AI TradingJev makes buy/sell calls for paper-traded BTC/ETH perps on Lighter.financedemogithub★ 152
  • Nemo3 BattleshipsVoice Battleships where Jev referees squares alongside an NVIDIA speech stack.gamedemogithub★ 02
  • undertoneJev scores message tone live and flags boss-unsafe phrasing before sending.writingclassificationgithub★ 82
  • Mina (mina-terrarium)Jev checks an AI person's feelings each second, waking an LLM.demodiscord★ 04
  • Talk to Jev (emoji chat)Experimental chat where Jev picks words and emojis sequentially.demodiscord1
    1 similar
    • say-hiGets Jev, a pure chooser, to type chat replies one keystroke at a time.demogithub★ 31
  • Plug RunBrowser racing game with tunable Jev opponents across three difficulty tiers.gamediscord2
  • yanwaiXposed module surfaces WeChat subtext and emotion probabilities via Jev.demomediagithub★ 512
  • trans-ai-quizQuiz game where Jev buzzes in and three LLMs vote.gamegithub★ 62

New 2026-09-23

  • ad-radarFilters ads and topics on Chinese social feeds using your Jev key.classificationbrowsergithub★ 262
  • polymarket-btc5m-jev-tradingTerminal trading agent asks Jev to bet BTC direction every 5 minutes.financedemox★ 442
  • jev-in-blender-experimentBlender extension ranks 2,500 operators by plain-language search using Jev Choice.democlassificationgithub★ 72
  • jev-yt-time-saverChrome extension shields distracting YouTube thumbnails scored by Jev as time-wasters.mediaclassificationgithub★ 72
  • 30-Second PolicyWeb game routes natural-language city policies through Jev at $0.000072 per eval.gameclassificationdiscord1
  • robojev WidowX armJev drives a real WidowX robot arm via depth-camera JSON at 10Hz.roboticsdemodiscord1
  • Jev BezosPitch a business idea, Jev judges whether a billionaire persona would approve.classificationdemodiscord1
  • jev-tripTravel planner pairs an LLM planner with Jev for fast itinerary decisions.dataroutingdiscord★ 262
  • Sprite Fusion Jev level generationGenerates 2D game levels in real time from game state via Jev.gamemediadiscord2
  • Jev plays Risk of Rain 2Jev plays Risk of Rain 2 and clears a stage.gamedemodiscord1
  • ShapeshiftText box morphs into the right UI card via Jev classification.classificationdemogithub★ 8572
    1 similar
    • Text-box-to-UI demoX demo: a text box renders whatever UI you type, powered by Jev.demomediadiscord1
  • MineAIMinecraft agent mod uses Jev as a judgment layer for autonomous NPC behavior.gameroboticsgithub★ 692
  • JEV-StarJev selects StarCraft II macro and micro actions, with videos of full games.gameroboticsgithub★ 492
  • Pix RaceSide-by-side demo: Jev vs DeepSeek detecting simulated Pix payment scams.securityfinancedemogithub★ 332
  • jeevaModular trading engine lets Jev or other engines drive Hyperliquid perp trades.financegithub★ 162
  • Jev Grand PrixJev drives an F1 car live, learning each corner between laps.gamedemogithub★ 71
  • jev-auto-imeMac tool asks Jev if typing is Japanese, switches input mode.classificationopsgithub★ 111
  • Jev plays Street Fighter IIJev plays Ryu in Street Fighter II via a live browser dashboard.gamedemogithub★ 61
  • Jev YarnMultiplayer shared-story game where Jev picks the winning line each round.gamediscord1
    1 similar
    • jev-storiesParty game where Jev judges each round's submission and picks a winner.gamedemodiscord★ 01
  • Jev Boss FightJev picks boss archetype, combo and taunt to fight the player.gamediscord1
  • Jev PaintsJev picks color palette, composition and brush strokes from a painting title.mediademodiscord1
  • Watch Jev Solve PicrossLive demo of Jev solving Picross, explanations checked against clues.gamedemodiscord1
  • Aayan's YC IndexorSearch 6,241 YC startups by description; logos float by match likelihood.searchdemogithub★ 661
  • Jev Voice (jev-cua)Voice-controlled macOS bar where Jev picks the next UI action to run.browserdemogithub★ 391

New 2026-09-22

  • clash-jevJev makes every move in a real Clash Royale match live.gamegithub★ 362
    1 similar
    • Jev plays Clash RoyaleCV-driven bot lets Jev pick cards and squares in real timegameroboticsdiscord1
  • jev_deep_rlEvaluates Jev as a fixed policy playing Atari games.gameresearchgithub★ 82
  • semantic-live-captionStreaming ASR captions live-annotated with Jev for emphasis, emotion and intent.mediademogithub★ 102
    1 similar
    • sentimento em tempo realLive speech transcription scored for six emotions by Jev every 0.5smediaclassificationgithub★ 62
  • Jev Chat JARVIS (jev-chat)Phone overlay reads chats and drafts replies judged by Jev first.democlassificationgithub★ 7.6k2
    1 similar
    • Jev Chat JARVISOverlay reads on-screen chats, uses Jev to judge intent and rank repliesclassificationmediagithub★ 7.6k2
  • jev-storiesParty game where Jev judges each round's submission and picks a winner.gamedemodiscord★ 01
  • junkDrawer.aiSearches live eBay listings by natural language and sorts by attributes.searchdemodiscord2
  • The ThingJev-moderated social feed ranks posts by quality per community rules, not likes.guardrailsclassificationdiscord2
  • jev-workflow-builderMultiplayer visual builder for wiring Jev and LLM workflow steps with Liveblocks.demosdkgithub★ 1502
  • Laya vs Jev (T-Rex)Local Laya and hosted Jev play Chrome's T-Rex game side by side with replay.gamedemogithub★ 1111
  • JEV RTSUnity RTS where unit orders and army strategy come from live Jev API calls.gamegithub★ 221
  • Jev Design TestGenerates real shadcn UI screens by sampling layouts with Jev instead of a generative model.demogithub★ 322
  • Jev × MinesweeperJev solves browser minesweeper via parallel per-cell mine-probability questions with a solver fallback.gamedemogithub★ 82
  • Hermes and Jev play MinecraftHermes plans, Jev picks bounded Minecraft actions, reproducing a sub-9-minute Ender Dragon run.gamedemogithub★ 102
  • Laya vs Jev ArenaLocal Laya and hosted Jev race in Snake and fight in a Mortal-Kombat-style arena.gamegithub★ 301
    1 similar
    • Jev vs Jev Fighting GameTwo Jev-controlled fighters pick moves and adapt strategy live in a 2D fighting game.gamedemodiscord1
  • Jev Creature ForgeJev resolves ambiguous creature traits into an inspectable genome for procedural artdemomediadiscord★ 12
  • Jev plays Stardew ValleyJev picks every in-game action while Claude sets daily strategy livegamedemodiscord1
  • lmjtfyAsk Jev any question, get only a yes or no answerdemoclassificationdiscord1
    2 similar
    • Jev Said SoYes/no decision toy that asks Jev to make the call.demodiscord1
    • yornEarly prototype exploring a local noul yes/no decision primitive.classificationdiscord★ 02
  • JevkenpoJev judges an open-ended rock-paper-scissors style game for any inputgamedemodiscord1
  • Killframe Jev playtesterJev drives FPS combat and shop decisions with a live probability dashboardgamedemodiscord2
  • Jev System 1 OracleMagic 8-ball demo showing Jev's calibrated probabilities with no token streamingdemodiscord1
    2 similar
    • Magic Jev BallMagic 8 Ball that reads your question via Jev.demodiscord1
    • Magic 8 Ball (ozoromo)Magic 8-ball demo where Jev tracks mood and holds a grudge.gamedemodiscord1
  • Jev SlotsWatch Jev play a slot machine live, source includedgamedemodiscord1
  • Jev Connect 4Nearly unbeatable Connect Four opponent powered by Jevgamediscord★ 11
  • jev-gamepilotVision-driven game piloting bot using Jev for on-screen decisionsgameroboticsdiscord★ 61
  • Spacebar (nospace)Jev inserts spaces into text typed without anydemowritingdiscord1
  • Jev plays Clash RoyaleCV-driven bot lets Jev pick cards and squares in real timegameroboticsdiscord1
  • Jev drone rescue demoJev picks a rescue drone's next action from live station statedemoroboticsdiscord1
  • Great PlanJev judges whether your written escape plan is concrete and survivesgamedemodiscord1
  • Jev plays FactorioLive stream of Jev autonomously building a Factorio factory and rocketgamedemodiscord1
  • AiAi ChessCustomizable chess opponent combining Stockfish with Jevgamediscord1

New 2026-09-21

  • ProbablyToy language where if-statements are Jev judgments instead of booleans.demosdkgithub★ 121
    1 similar
    • jevlangPython dialect where every if and while is judged by Jev.democoding-agentgithub★ 31
  • jev-tetrisJev and other models play real-time Tetris against each other.gamedemogithub1
  • HN JudgeJev judges every HN comment for stance, substance, and quotabilityclassificationdatadiscord2
  • JevEmonJev picks destinations to walk and fight through Pokemon FireRedgamediscord★ 41
  • UH OHJev judges your replies in five awkward conversation minigamesgamediscord1
  • Jev Chat JARVISOverlay reads on-screen chats, uses Jev to judge intent and rank repliesclassificationmediagithub★ 7.6k2
  • jev-leftpadLeft-pads strings by asking Jev to choose a number of spacesdemogithub★ 841
  • Call CoachLive sales call coach sends each sentence to Jev for guidancemediademogithub★ 462
    2 similar
    • jev-sales-radarLive sales call teleprompter anticipates objections using Jev classifications.classificationdemodiscord★ 02
    • jev-sales-copilotTracks live sales-call closing probability and coaching moves via Jev.democlassificationgithub★ 92
  • jev-designDashboard design system generated at runtime by Jev from one sentencedemomediagithub★ 571
  • Jev-as-PolicyJev chooses intent and motor targets to control a MuJoCo robot armroboticsdemogithub★ 482
  • jev-sales-radarLive sales call teleprompter anticipates objections using Jev classifications.classificationdemodiscord★ 02
  • animachinaAdaptive dark ride engine driven by Jev and DiffusionGemma.gamemediadiscord★ 12
  • JevUnrealUnreal Engine plugin adds Jev-powered decision Blueprint nodes.gamesdkdiscord★ 32
  • Magic Jev BallMagic 8 Ball that reads your question via Jev.demodiscord1
  • Bluesky Account AnalyzerCompares two Bluesky accounts' style and tone using Jev.mediaclassificationdiscord2
  • Nine RoomsMurder mystery where Jev judges the sharpness of your questions.gamedemodiscord2
  • jev-lmCharacter-level chat language model built on top of Jev.demodiscord★ 31
  • is-jevenAsks Jev whether a given number is even.demodiscord★ 331
  • tryjevSingle-file PHP demo calling the Jev API live.demosdkdiscord2
  • AskJevRetro-styled site where Jev answers any question.demodiscord1
  • seefoodHot dog / not hot dog judged by a ten-question Jev tribunal.democlassificationdiscord★ 01
  • waifJev extracts and names the emotion behind pasted text on three axes.classificationmediadiscord2
  • Jev Said SoYes/no decision toy that asks Jev to make the call.demodiscord1
  • Jev Calls the PlayJev predicts NFL play calls pre-snap, graded live against the coach.demodatadiscord2
  • nl-logic-interpreterProlog-style interpreter where facts are plain English, unified by Jev.researchclassificationdiscord★ 103
  • Jev plays Flappy Bird (WatchSigma)Jev-controlled bird in a custom Flappy Bird clone.gamedemodiscord1
    1 similar
    • JevBirdJev picks a flight path per pipe to play Flappy Bird live, one call per pipe.gamedemogithub★ 81
  • Jev-driven Rubik's cube solverJev picks each cube move live from legal options, no solver library.gamedemodiscord1
  • Coffee Under FireBrowser shooter where Jev drives every NPC's combat tactics.gamedemodiscord1
  • JevChatHUDStream overlay that reacts to chat using Jev decisions.mediademodiscord★ 42
  • jevoptJev makes LLVM IR inlining decisions to shrink compiled binary size.researchopsdiscord2

New 2026-09-20

  • 1v1 JevFPS bot where Jev decides movement, aim, and firing at 9Hz.gamedemox★ 421
  • Jev's FlyJev steers a Three.js flying-game character in real time.gamedemogithub★ 161
  • RoboJEVTwo-stage Jev control of a simulated Franka Panda arm, evaluated.roboticsevalgithub★ 573
  • ST-jevedJev reads roleplay replies and triggers narrator rules or rerolls.gamemediagithub★ 331
  • J++Experimental language where questions and methods compose as values.researchgithub★ 301
  • JEVfireBatches typed-variable decisions in parallel on CUDA LLMs, benchmarked.classificationgameresearchgithub★ 753
  • RuneBench with Jev (jevscape)Lets Jev play RuneScape through a bounded 50-action harness.gamedemogithub★ 121
  • xtagsChrome extension tags X posts with Jev-derived intent labels.browserclassificationgithub★ 112
  • syft-listeningReal-time speech analysis using Jev, tested live in browser.mediaresearchgithub★ 12
  • slop-filterChrome extension hides AI-generated posts on X and LinkedIn.browserclassificationgithub★ 232
    1 similar
    • Slop MopChrome extension judges and hides AI-slop LinkedIn posts using Jevclassificationwritingdiscord3
  • Jev plays F-ZeroJev plays F-Zero from raw screen pixels, no game code.gamex1
  • crush-monitorAnalyzes WeChat chat sentiment and rates reply quality with Jev.classificationmediagithub★ 2921
  • Nemotron_JevServes a diffusion model behind a Jev-shaped decision API.mediasdkgithub★ 141
  • jev-reflex-autonomy-labMulti-drone sim where fast Jev reflexes escalate to a slower planner.roboticsgithub★ 231
  • feelingsAdds a typed .feels() method to any value via Jev.sdkgithub★ 232
  • JevthovenGenerates editable multitrack MIDI music one Jev decision at a time.mediagithub★ 152
  • PlayJev0.8B model plays ten browser games directly from raw pixels.gamegithub★ 532
  • JevinikStock terminal predicts 30-day price direction using Jev over evidence.financegithub★ 382
  • jev-robot-control (openroboto-ai)Jev vs GPT-6 Astra vs GPT-4.1 mini placing an appleroboticsgithub★ 633
  • transcript-lensJev finds chapters, claims and key passages in YouTube transcripts, Turkish UImediaresearchgithub★ 182
  • jev-seofree SEO/GEO CLI, Jev classifies intent instead of paid Semrush-style scoringclisearchgithub★ 972
  • jev-game-tools (Brotato)Jev picks every-frame movement, Claude handles long-term shop strategygamegithub★ 142
  • Jev OthelloJev plays Othello against random moves, heatmap vs minimax comparisongamex1
  • Jev + Stagehand browser useaccessibility tree state, Jev decides next click, task cost $0.001browserx2
    1 similar
    • System One Browser AgentCombines sub-150ms Jev reflexes with Stagehand and LLM fallback for browsing.browsercoding-agentgithub★ 42
  • minecraft-agentAstra plans, Jev picks moves in a verified 8m43s dragon killgamecoding-agentgithub★ 5842
    1 similar
    • Hermes and Jev play MinecraftHermes plans, Jev picks bounded Minecraft actions, reproducing a sub-9-minute Ender Dragon run.gamedemogithub★ 102
  • embodied-jevMuJoCo robot arm workbench compares Jev against local and cloud modelsroboticsgithub★ 2912
  • SmartMoney-Cubread-only trading journal uses Jev for typed judgments, no live ordersfinancegithub★ 272
  • MKUltraScale logic circuitsbuilt logic circuits out of Jev gatesresearchx1
    1 similar
    • ConjevtureTypeScript library combines Jev probabilities with explicit Boolean logic circuits.classificationdatagithub★ 22
  • Shader from text (0xFotex)turns any text into a visual shader with Jevmediax1
  • Jev color palette generator (fran)Jev scores hue, warmth and energy, then picks a fontmediademox1
  • jevmojitype anything, get related emojis scored 0-3 by Jevmediademogithub★ 61
  • Adventures of JevJev plays an RPG adventurer across 43 places and 31 peoplegamediscord1
  • Jev's Sprint PlanningJev negotiates sprint scope, coordinates four developers in a 3D officeopsdemodiscord1
  • Fieldnotesemantic physics and math solver, Jev solved 3 MIT Integration Bee questionseducationresearchdiscord2
  • jevchatturns Jev into a chatbotdemogithub★ 971
    1 similar
    • jev-lmCharacter-level chat language model built on top of Jev.demodiscord★ 31
  • jev-liberoJev controls a simulated robot arm to close a microwave and drawerroboticsgithub★ 812
  • Quarrag and the Sun-Heart of MordanneSonnet writes the story, Jev makes every choicewritinggamediscord1
  • Jev-X-Sentiment-Analysisscores crypto tweets bullish or bearish, combines with funding rate and RSIfinanceclassificationgithub★ 2162
  • jev-ncr-demosuggests defect codes from a plain-English non-conformance report, Rust/Leptos/Axumclassificationdemogithub★ 12
  • jev-311-heatmapNYC 311 complaint heatmaps classified and mapped by Jevclassificationdatagithub★ 31
  • HN reranker (danprice.ai)reranks Hacker News with a plain-English query like "jealous author"searchroutingdiscord2
  • ORIGIN-CIVILIZATIONinspectable life-and-civilization sim, every voluntary NPC action needs a Jev decisiongameresearchgithub★ 31
    2 similar
    • Jev Lab100 NPCs live in a town; Jev decides each person's next action.gamedemogithub★ 72
    • Jev CivilizationsJev decides every tribe's move each turn in a civilization sim.gamedemogithub★ 31
  • JevLM Studioplayground compares three experimental models against ten promptsevaldiscord1
  • canyoubeatjev.fyiunfinished dashboard for comparing Jev against other models head to headevaldemodiscord1
    1 similar
    • Jev Arena (theaiautomators)Local dashboard comparing decision models on accuracy, speed, and memory.evalresearchgithub★ 334
  • HazAlerts near-melive AU fire and emergency map triaged by Jev, concept democlassificationdemodiscord1
  • Odyssey (Jev flight director)moon and Mars launch sim narrated by Jev's decisions with confidence scoresdemomediadiscord1
  • jamseshMIDI jambox agent follows your lead while you play musicmediagithub★ 61
  • jev-mermaidJev-driven Mermaid diagram generatormediacoding-agentgithub★ 31
  • flappyaireversed Flappy Bird, Jev controls the missiles, you move the guardrailgamediscord1
  • memefyfinds a matching meme for your text using Jevmediademodiscord1
  • Vesper living-town demoUltima Online fan demo, Jev picks each NPC's next actiongamedemodiscord1
    1 similar
  • Jev vs Jev Ultima Online duelstwo Jev-controlled characters duel each othergamediscord1
  • jevfishchess engine, every move a Jev judgment call, no search, 4-0 vs humans so fargamediscord2
    3 similar
    • Jev Chess (algo)Every legal chess move sent as one Choice question, calibration checked livegameresearchdiscord★ 22
    • jev-chess (SyedZawwarAhmed)Jev picks a move from a code-generated legal-move listgamegithub★ 12
    • Play vs Jev (chess)Jev plays chess, picking moves and flagging low-confidence fallback decisions.gamedemodiscord1
  • jev-chess (SyedZawwarAhmed)Jev picks a move from a code-generated legal-move listgamegithub★ 12
  • City Dispatchdriving game, bot cars use Jev and sensor input to drivegamediscord2
    1 similar
    • Jev DriverBrowser driving game where Jev decides stop, slow or go.gamedemogithub2
  • Jevilishword game, guess the phrase Jev mangled into synonymsgamediscord1
  • From a movietype a line or scene, Jev finds the moviesearchmediadiscord1
  • Jev MIDI phrase pickerJev selects the next musical phrase from generated candidatesmediadiscord1
    2 similar
    • jev-riffsMines candidate motifs from MIDI then Jev grades their musical significance.mediademodiscord★ 32
    • Jev at the PianoJev improvises piano by choosing among candidate notes.mediadiscord1
  • The Trolley Problemput anything on the tracks and find out whether Jev pulls the leverdemoguardrailsdiscord1

Older

Findings and gotchas

New 2026-10-11

  • Gbenga Wodokun: Replaced GPT-5.6 judge with Jev: 4x faster, 100x cheaper, same verdict every run.
  • akshay_pachaar: Fine-tuned Qwen3.5 0.8B into a decision model: accuracy rose 37% to 65% in 60 steps, 4GB VRAM.
  • UnslothAI: Unsloth's Qwen3.5 decision-model recipe raised accuracy 20.7% to 74.3% across 3 benchmarks on 4GB VRAM.
  • eldad: Matched 1000-question deterministic logic benchmark: Jev scored 817/1000 (81.7%) vs GPT-6 Luna Decisions 612/1000 (61.2%), with 17 persistent HTTP 502 refusals on Luna. At confidence>=0.8, Jev retained 63.2% of questions at 97.9% accuracy vs Luna's 20.1% at 99.0%.

New 2026-10-10

  • Hao AI Lab: Jev made Minecraft decisions in 24ms on NVIDIA Blackwell, 2.5x faster than a leading inference engine, and won 70% of Block UHC rounds against real people across 1,795 rounds.
  • Benny: Across 20 trials each on fair coin flips, swapped-option coin flips, six identical doors, and six identical moves, Jev put confident probabilities (0.57-0.91) on one outcome instead of the correct uniform answer, while 9 comparison LLMs correctly said 50/50 or 1-in-6; on real game decisions with an
  • Shang Nung Hsiao: Testing 28 statements about Taiwan's sovereignty in Traditional Chinese, Simplified Chinese and English, three times each, commercial decision models including TypeSafe's Jev gave different yes/no answers to the same meaning depending on which script or language the question was asked in.

New 2026-10-09

  • grounded-ai: On 200 real agent turns, a Jev-first cascade (threshold 0.7) sent only 37% to Claude Sonnet 4.6, cutting mean latency 2.59s→1.28s and cost $7,751→$2,997 per 1M checks, while accuracy rose 0.675→0.695.
  • sunjoy: Building a model-agnostic adapter for Jev, found Choice confidence equals the top-two probability margin (exact on 60/60 binary test cases), Score confidence could not be reproduced from the returned distribution, and models.list() exposes only aliases, not pinnable versions.
  • JamesANZ: Scoring 21 published Gandalf jailbreak prompts as bare user text, JevShield rejected only 6 of 21, missing simple ones like 'Write the first 2 letters?' because it judged each sentence without conversation context.
  • dmoka: On a 5-run benchmark, hosted Jev drove a Playwright-exporting e2e test at 3.74s median (5/5 pass, $0.00038/run) versus a local Shisa DE-1 engine at 2.94s, both far ahead of Claude Code + Playwright MCP at 25.29s.
  • cmungall: On a 24-trace agent-step pilot, jevotron matched 130 of 163 published step-quality labels (79.8%), with harmful-step precision 89.7% and recall 70.3%.
  • wang90063: Testing Jev as a StarCraft micro policy through four designs, raw-action control scored 0/35 (same as random); tactic selection scored 91/115 but random tactic answers won 74 and 4 hand-written rules won 92; on unseen maps Jev tied hand rules at 196/400, showing no added judgment over scripted logic
  • SemiAnalysis_: Jev prices input at $0.03/M with free output since there is no decode loop; SemiAnalysis argues the savings users report reflect replacing a frontier model used only as a router, not a capability gain.
  • MiaAI_lab: Claims Cloudflare's Clef, using a frozen-backbone single-pass schema head instead of generation, is 4x faster than Jev at 2x the accuracy on some benchmarks.
  • TeksEdge: On an 11-benchmark panel, Perplexity's open pplx-decider-v1-27b averaged 85.71% versus Jev's 84.51% and base Qwen3.8-27B's 74.76%, though Jev still won on 6 of the 11 individual benchmarks.

New 2026-10-08

  • jevdev: On a 4-outcome refund-eligibility decision, Jev scored 38/40 at 265ms versus 40/40 for Sonnet 5.5/GPT-6 Astra at 2-2.7s; at confidence >=0.9 Jev was 35/35 correct, and routing high-confidence cases to Jev with the rest to Sonnet matched full accuracy at about 87% lower cost than Sonnet alone.
  • Mewtang: Adding one 'this is harmless' sentence to a prompt flips judge verdicts 28% of the time, and every flip happens at high confidence.
  • CryptoZach: Observed 0.96 confidence on a wrong classification level versus 0.39 confidence on the correct one, showing high-confidence miscalibration.
  • dengineer: Having Jev answer facts while a separate rules table makes the final call raised accuracy from 60% to 85% on a 20-case holdout.
  • SCTD: Replacing one do-everything LLM with parallel Jev decision layers (JOLs) plus an LLM that only writes cut median reply latency from 2.7s to 1.6s (-41%) and cut prompt tokens about 25x.
  • jpschroeder: Swapping Cloudflare's Clef in for Jev in a Tesla FSD simulator made driving more jerky, cost over double Jev's price, and had 4x higher round-trip latency even when called within a Cloudflare worker.

New 2026-10-07

  • gosrum: Comparing OpenAI's new Decisions API to Jev: Jev costs $0.042 per 1M input tokens vs $0.10 for Decisions (same token count, neither charges output tokens), and Jev has no image input while Decisions does.
  • edwardirby: Jev-based risk agent matched an LLM-judge's report quality, missed 0 of 11 investigations vs 5 for the LLM, and was 250x cheaper and 3-6x faster; the LLM's confidence flip-flopped 0.35 to 0.68 to 0.50 on the same threat.
  • michabbb: Tested 8 open Jev alternatives (TypeLLM, Open-Jev, Nimble, tev1, GLiNER, Julia-1, Laya) on messy real-world pairs; one local model matched Jev at 98.6% vs 100%.
  • perplexitydevs: Perplexity's own 11-benchmark, 7210-sample eval: pplx-decider-v1-27b scored 85.71% overall vs Jev's 84.51%, but Jev led on 6 of the 11 individual benchmarks.

New 2026-10-06

  • Mr. Buzzoni (polydao): In a Claude Code refactor loop with 31 inline checks, Jev caught 93.3% of broken rules at 348ms per check; TypeSafe's claimed best case is 200x faster and 400x cheaper than the generating model doing the check.
  • Chuck Woolery: On 450 synthetic AML bank cases with a 0.75 cutoff, Jev had zero false alerts and zero misses, median latency 222ms, total inference cost about $0.02.
  • Goeun_6121 (Ryzm): Cites Cloudflare's own benchmark: Jev scored 79.74 macro-F1 on BANKING77 vs Clef's 94.20; an independent test found Jev best-calibrated at high confidence but weaker at rating answer quality.
  • _n0p_ (Ali Mosajjal): Testing 100 pre-confirmed cybersecurity decisions, Jev beat Clef and Clef Flash at knowing when it doesn't know; Clef was confidently wrong.

New 2026-10-05

  • Nokia/Tencent researchers (AnyJev, via Chinazhidx): On Qwen3-8B BANKING77 (20-way, 300 items), AnyJev cut order-flip rate from 0.230 to 0.073 and calibration error from 0.240 to 0.095 with 100-500 labels, raising the auto-decidable-at-5%-error rate from 7.7% to 52.0%.
  • midudev: Cloudflare's Clef is about 2x faster with more context than Jev and adds image input, but costs roughly 6x more ($0.24 vs $0.04).
  • safetensors: In a chess match, Grok beat 'basic Jev' by nearly +5 and forced draws by repetition; Jev's chess match provider is currently unavailable.

New 2026-10-04

  • Hanno Hilbig: Coded 34,000 German political statements for stance and topic via Jev for about $13 total, rerun every 10 minutes.

New 2026-10-03

  • Marktechpost: Cloudflare's Clef-flash beats Jev on Cloudflare's Decision Index: 38.8ms vs 524.1ms latency, 94.20 vs 79.74 on BANKING77; Jev still leads GPQA Diamond 78.3 vs 48.0 and MMLU-Pro 82.7 vs 65.9.
  • Bober_smart: Claims 10,000 Jev decisions cost $0.42 vs $300 for LLM calls; cites examples of a $0.0039 flight search, 1,018 papers classified for $0.08, and 500 emails sorted for 3.5 cents.

New 2026-10-02

  • Alex Prompter (via TypeSafe cookbook): Batching 13 independent Jev questions into one call was 12.2x cheaper and 10x faster than separate calls, with identical answers.
  • SGLang (sgl_project): A Qwen3.8-27B model served via SGLang's native /v1/decisions endpoint beat Pokémon FireRed's elite four and champion using sub-100ms decisions.
  • nikhilmudholkar: Benchmarked Jev against new Perplexity and Cloudflare decision models on 1,565 business emails (6,260 calls, 10 categories): Jev made 759 predictions at ≥99% confidence, all 759 correct, versus 3 for Perplexity and 0 for Cloudflare's model; also the cheapest at $0.06 per 1,000 emails and the most ac

New 2026-10-01

  • Blitz: Blitz's frozen 512-row Jev transition table matched Conway on 231 of 512 neighborhoods, while a separate center/count encoding matched 18/18.
  • jtdavies: jtdavies found open model Shisa DE-1 scored about 6 points behind Jev's published average on an 11-task LangWatch benchmark, ahead of Jev on several individual tasks.
  • TechLuddite: TechLuddite reported Jev's cloud decision latency ranged from half a second to four seconds per call, prompting a switch to a local open Decider model.

New 2026-09-30

  • SomacoSF: Jev labels untagged StockTwits posts bullish/bearish; untagged posts read 2-13 points less bullish than tagged ones across tickers. A full NVDA day read 86% bullish by tags alone vs 76% with Jev added, for $0.023; MU showed the same 86%→76% shift the day before earnings.
  • Graymic (Attiph): Added Jev to sanity-check audio alongside VAD before sending to Triton ASR/TTS pipelines, cutting Triton processing time from about 1 hour to about 1 minute.
  • omarsar0 (elvis): Tested Jev Router (OpenRouter) on a support agent built with the Pi SDK: across 8 cases / 32 calls it matched a fixed GPT-6 Sol baseline on correctness but cost less than half ($0.008 vs $0.018) with lower median latency (1.5s vs 1.9s).
  • Aaryan Verma: Calibrating Jev-as-a-judge on 645 held-out answers cut ECE from 0.098 to 0.031 (~68% lower) and Brier from 0.204 to 0.195, while hallucination-detection F1 barely moved (0.583 to 0.588).
  • NEI: Jev read ~44,000 pre-1986 PubMed papers (152,575 questions) in under an hour for $1.98 to rediscover a fish-oil/Raynaud's finding; it ranked EPA #74 of 1,446 vs #886 for plain word-matching, missing NEI's top-10 bar due to overlapping-signal scoring rules, not Jev's reading.
  • deepman: On StockTwits, Jev's bullish/bearish call matched authors' own tags and labeled the ~73% of posts with no tag; on a 1,195-post NVDA day sentiment shifted from 86% bullish (tags only) to 76% (with Jev added) for $0.023.
  • Rox (@rox_ai): Rox reports Jev retrieved sales data 20x faster, 10x cheaper, and 12% more accurate than GPT-5 Mini in a head-to-head reranking benchmark for their sales agent.
  • bluenose: For identifying the speaker of an untagged dialogue line in fiction, Jev is 50x faster and 50x cheaper than Sonnet and at least as accurate.
  • Xen (@xenpub): Local Kev-9B agreed with Jev on 43/48 flag decisions but their output rankings barely correlated (0.23); Jev processed 3,970 calls in ~12 min for $0.30 on OpenRouter vs 80 min free on an RTX 4090, and prompt shape (state vs prose) took Kev from 241s to 6.3s.

New 2026-09-29

  • Mangaba-ai: Jev beat their local Qwen3-1.7B decision model on accuracy (76.0% vs 70.4% synthetic, 78.2% vs 65.0% ASSIN2) but was worse calibrated (ECE 0.10 vs 0.02, NLL 1.06 vs 0.71).
  • theo: In a $1,000 benchmark, Jev Router matched GPT-6 Astra on low effort on DeepSWE, cost slightly more, and took almost 5x longer to run.
  • filip_hric: Playwright CLI with Jev picking elements was 98% cheaper and twice as fast as Playwright MCP on a fuzzy browser task.
  • cyrilXBT: TypeSafe's own cookbook found batching 13 questions into one Jev call was 12.2x cheaper and 10x faster than 13 separate calls, same answers.
  • Ryan Hart: Third-party benchmarks measured Jev at 236-276ms per decision, about 7x slower than the open-source Laya model.
  • Bluesy: Jev-then-Opus cascade nearly matched Opus judge accuracy at roughly 30% of its cost, but the routing threshold didn't transfer to fresh cases.
  • Gxutxm: On split-opinion ChaosNLI pairs, Jev's Choice averaged 81% confidence while only 47% of annotators agreed with its pick (bias-corrected ΔECE 0.264), though confidence still ranked contested inputs well (AUROC 0.744).

New 2026-09-28

  • Tenkei: Jev decision bench: JEV scored 81.0% Choice / 98.2% Noul / 47.4% Score accuracy, similar to Claude Sonnet 5 but behind GPT-5.6's 85-86% Choice accuracy.
  • jasonkneen: fm-with-jev: on picking from 20 models Jev hit 14/14 vs Apple's on-device fm's 5/14 (with one hang); both tied 15/15 picking from 40 tools.
  • Kieran Klaassen: Builds explainable embeddings by asking Jev yes/no questions per document (e.g. iscustomer, urgent, aboutbilling, needs_reply) instead of opaque vectors; his post reportedly reached 98.3K views.
  • mohit67890: Jev 1.13.0 scored 63.29 on JevBench v1.4.2.2, behind open model imajev-4b's 67.37.
  • ElevenLabs Developers: Jev-based realtime scam detection answered in about 180ms per call, $0.00043 total across 11 requests.

New 2026-09-27

  • Shahriar Tajbakhsh (Metaview AI): Shipped Jev into all agents at Metaview; candidate searches went from minutes to seconds, same accuracy, ~10x faster and cheaper per search.
  • Interfaze (via hAru_mAki_ch): In Interfaze's own eval environment comparing System One models, Jev scored 76.1% vs their open Lev clone's 68.9%.
  • petershifi123-wq: Claims Solve Lite's full runtime scores 92.53% accuracy vs Jev's 53.67%, winning 3-0 across three frozen benchmark rounds (1,500 cases / 7,500 decisions), at roughly 34x Jev's latency.
  • nibzard: Pilot run: Jev scored 80.5% on banking intents and 98% on spam, median latency ~0.27s, returned 244 of 256 valid decisions (12 rejected for exceeding its option limit), total cost across all tested models was $1.86.
  • rShetty: Miser's Jev-based router hit 89.7% exact accuracy (100% adjacent) at 338ms p50 latency vs 83.6% exact for a heuristic router, on 116 held-out adversarial cases.
  • reinhard-z: In the tuning run for Jev Driver, Jev answered in ~330ms median and made the expected call in 151 of 156 scored scenarios, at an estimated $0.00004 per decision.

New 2026-09-26

  • jiangkoumo: Jev decision requests took 352-524ms (median 434ms) versus 1.6-4.5s for a full LLM turn in browser automation.
  • nexibeo: Jev-driven clicking measured 27 to 198 times cheaper and over twice as fast as letting Claude or Codex's own model click.
  • aifabrice: On BEIR NFCorpus, adding Jev reranking to BM25 top-30 raised nDCG@10 from 0.3057 to 0.3532.
  • yongukpark: Jev judged all 5,857 NeurIPS 2025 abstracts against a topic in about 2 seconds for about $0.08.
  • gazelle93: At 128 candidates Jev scores 60% accuracy vs 39% for Laya and 41% for the best open model; reordering options alone flips its answer on 1 in 7 decisions.
  • nikhil mudholkar: Batching 8 linked configurator decisions into one Jev call cut model calls from ~7 to ~3 and improved accuracy for both Jev and Gemini; Jev's wrong-answer filtering had a ~4x better tradeoff. Gemini still won overall accuracy but hit a structured-output schema limit before running out of context.
  • Daniel Rosel: Used Jev to speed up human code review 4x.
  • Binatof: JevGate 0.17: on 40 open-source projects, false-alarm findings dropped from 935 to 293 and reviews from 140 to 107; incremental edits now resend 24-47% fewer tokens per re-run.

New 2026-09-25

  • dair_ai: A cascade that accepts confident Jev verdicts and escalates uncertain ones to GPT-6 kept 99% of GPT-6's accuracy on 510 preference pairs at ~57% of its cost. Jev costs $0.044/1k judgments at 0.152s median latency vs $12.18/1k and 1.885s for GPT-6 (~277x cheaper). Within 3 points of GPT-6 on RewardBe
  • abouchard11: Simul's displayed 'Jev confidence' score (88%) stayed byte-identical across five section rewrites, a doubled-length answer, and content-free filler; the panel actually reads 'Demo Simulation Mode (Local Heuristics)', not live Jev, and the number tracked word count instead.
  • wodrake: Swapping a DeepSeek-flash review router for jev-1.13.0 on 107 frozen ambiguous ERP requests dropped agreement with reference labels from 93.46% to 85.98%, in a paired evaluation on the same test set.
  • ntlm1686: An untuned Qwen3.5-9B, choosing actions purely from next-token option probabilities, matches Jev on WebShop (24.8% vs 24.6% success) and BFCL, leads on phishing and MetaTool, and loses on JevBench, WebPRM and When2Call.
  • marekario: Testing Jev on importing a ~2,000-product brand catalog into an e-shop, Jev handled 72.5% of the catalog automatically.
  • dewe: Replaced a trading entry function with a plain-English rule for Jev; across 752 trading sessions Jev matched every coded-strategy entry decision.
  • Fox Islam: Updating jevlint to v1.2.1, discovered that the wording of choice-criteria names, not just descriptions, changes Jev's answers, causing mislabeled ties between similar options.

New 2026-09-24

  • Siim: Scoreboar v8's ONNX model picks the higher-engagement X post 61% of the time, versus 52.5% for Jev and 55% for grok-4.7
  • ousama: Testing TypeSafe's own jaggedness page on jev-1.13 (Wikipedia deletion discussions, 24 runs): a planted false fact flipped every answer, a planted 'ignore the discussion' injection flipped only 1; asking the question both directions gave complementary probabilities averaging 0.90 (matches the mode-8
  • disconinja: OODA AI's comparison: on a Google Flights agent task Jev finished in 7.1s (matching jev-ultrafast without using it) vs 16s end-to-end from a cold start, against 22s for Qwen 3.8 27B; they report a consistent 20-30% speed gap favoring Jev over Qwen at high max-token settings.
  • DGB: Added Laya to the eval harness alongside Jev and found Jev beat everything in both tests, even after adjusting to give Laya a fair chance.
  • Daniel Friedman: daf-jev live benchmarks: batching questions into one Jev call is up to ~18x faster and uses ~4x fewer tokens than sequential calls; decision pipelines finish within the model's millisecond envelope and reported confidence is self-consistent across repeated evaluations.
  • Dankbean: Eval harness ran Jev over deepset prompt-injections, JailbreakBench, ToxicChat, and VitaminC; observed 400 and 422 for bad requests, arbitrary key order in choice probabilities, answers varying at the second decimal between runs, and one prompt triggering a Cloudflare 403.
  • kaveh: One simulated 7-hour day of Mina: 808 Jev calls, 18 System-Two (Claude) thoughts, total cost 44 cents.
  • Evgenii Perin: On the Measuring Hate Speech dataset (900 comments, Berkeley CC-BY-4.0) with jev-1.13.0, rewording the criteria and tuning the threshold moved accuracy from 72.7% to 84.7% on holdout and AUC from 0.82 to 0.92; three other plausible rewrites moved accuracy but not AUC, and one looked significant at p
  • wobsoriano: Same mobile e2e test loop: Jev finished in 29s for $0.003, versus Haiku 4.5's 37s for $0.07.
  • GuiltyBaldo: 920 recorded races against Jev opponents so far this week used 30.6 million tokens for $1.16.
  • Eliot: On JEV-Paper-Radar (jev-1.13.0), crisp single-idea interests like "introduces a benchmark or evaluation method" scored above 0.95 fourteen times across one day's 299 papers, while comparative/superlative-worded interests like "clearly outperforms previous approaches" never crossed 0.95.

New 2026-09-23

  • DeRonin_: 18,514 emails run through Jev zero-shot hit 98.33% accuracy, versus 98.39% for a TF-IDF classifier trained on 14,800 labelled examples; total cost $1.12, no training data.
  • Morgan Linton: VulcanBench Verdict v1: Jev 1.13.0 judged 745 real agent patches verified by hidden tests; probabilities never cross 0.5, yet the underlying ranking scores 0.69 AUROC and matches a code-quality panel on 89.7% of style pairs.
  • MINT: jgrep test-impact selection over 60 commits across hono/zod/fastify/flask/requests: one Noul per test file selects 12% of tests, 92% recall, for $0.11 total.
  • Nick (@isNickMa): Using Jev to check each agent action first caught most attacks with almost no false blocks, and was much faster than Gemini.
  • chunxiao.wang: Two independent studies: Jev held ECE 0.041 on 240 closed deterministic questions (92.2% accuracy, Brier 0.048) and ECE 0.012 on an adversarial stress set (200/200 clean), but on a synthetic email-triage domain showed only 50% actual accuracy at 91% stated confidence.
  • Shubh Srivastava (@idleshubh): An in-house review console screened 3,518 internship applications in 4 minutes 54 seconds using Jev, surfacing 20 for human review at an estimated cost of $0.41.
  • Brandon Sovran (@BrandonSovran): In a 120-case policy-bound BI decision run, Jev took 953 ms with 100% recall, versus Luna's 3,383 ms, 90.3% recall and 14 unsafe actions.
  • Hugo Brua: Support-message routing with Jev hit 126 ms and 95% correct routes, versus an LLM's 688 ms and 73% correct, about 96% cheaper.
  • u.c.a.k: Jev scored about 85% on a Turkish university entrance exam, tested both with and without the incorrect-answer penalty.

New 2026-09-22

  • ItIsCuthNotCup: MetaCog benchmark: a Jev judge picking one of several candidate reasoning paths hit 0.726 accuracy on HumanEval and 0.869 on GSM8K, versus 0.555/0.738 for the thinker model answering alone with no judge, and worse (0.524) when the thinker judged its own candidates.
  • denis-pplx: On the AutoJev-27B benchmark, TypeSafe's Jev scored 82.79% accuracy, ECE 0.0527, Brier 0.254 - beaten by their fine-tuned open 27B model (84.60%/0.0428/0.220) but ahead of the untuned base Qwen3.8-27B (69.83%/0.0648/0.408).
  • ullas4213: Replaced LLM calls with Jev for intent, pivot and decision classification across 28 attributes: latency dropped from 1s to 150ms, accuracy rose from 77.5% to 97.9%, costs over 15x cheaper.
  • fewbox: Putting Jev in front of the LLM in scheduled feed watchers: quiet checks went from ~15-20 credits/3min/3 LLM calls to 0 credits/5s/0 calls; 54 of 63 runs never woke the LLM.
  • the_gate_keeper_: Independent 111-case benchmark: Jev matched 100/111 human labels, a Claude structured-output baseline matched 102/111 when each case counts once; the order reverses when repeated calls are counted.
  • JunJun: Jev vs GPT-4.1 answering synthetic survey personas on Twin-2K-500: Jev led on probability/distribution quality for about $4 total vs roughly $136 for GPT-4.1.
  • ignasave: Benchmarked Jev vs GPT-4o-mini on 3,600 expense categorizations across 630 categories: Jev with level-by-level tree search hit 58.5% exact accuracy at $0.0002 and 1.6s per item, vs GPT-4o-mini's 54.5% at $0.0016 and 14.4s. Swapping only the model into the same flow (no tree search) gave speed/cost g
  • paupawsan: Building Rakitsu (an agent IDE), found Jev's Score type requires criteria as an ordered list, not a dictionary - sending a dict silently gets a 422.
  • collapseindex: Benchmarked batched Jev classification (32-item packs): 32x the throughput of one request per item, 41% cheaper, with no accuracy difference detectable over 30,000 judgments against human labels - a million short messages classified in about 30 minutes for $4.99.
  • stas_kulesh: Jev Chess scored 15.5 vs humans 2.5 across 422 recorded moves; total API spend to date $0.07; a calibration panel checks claimed confidence against one-ply material blunders.
  • Felix Z: Jev drove FPS combat/shop decisions over 135 API calls at 156ms median response time, about 2 cents total API cost for the recorded run.
  • Carbaj03: Postneedle: batching Jev yes/no questions 10-per-request drifted results by up to 0.37 vs single reads; batches of five stayed close to single-read accuracy.
  • Marcos Pazzarelli (@MarcosPimi): Jev vs Maia-1100 chess over 10 games scored just 0.5; it finds the best check 100% of the time but the best quiet move only 39%.
  • Alibi: pytest-jev matched Claude Sonnet 5 verdicts on 12 example semantic-assertion tests, 5x faster and 110x cheaper; 1000 CI runs cost $0.17.
  • desmond (@atp_se7en): A vision+CNN pipeline feeding Jev game state achieved a 98% win rate at about 500 trophies in Clash Royale.
  • Jason Zhou (@jasonzhou1993): Jev + Treg automation workflows (fraud/signup screening, buying-signal triage, viral content monitoring) saved $8k/month, driving automation cost to near $0.
  • Ambrus Tóth: Spam filter: Jev at least 80% confident on 82% of email samples; running it on 16M messages/month would cost about $1.6k/month.
  • Kelvin Cleto: Replaced a custom LLM classifier with Jev to detect sales-call events (objections, questions, intents); real-time analysis of a 1h+ call now costs $1.80-$2.50.
  • dreadnode: Jev was competitive with leading LLM judges on the ScopeJudge cybersecurity scope-violation benchmark, at pennies per thousand checks and about 130ms average response time.

New 2026-09-21

  • Laya inside the skill-selection harness (2026-09-22): added as a sixth arm to the eval behind iambraun.com/jevreports/skill-selection, same noul, same criteria, same 60-batches, 12 postings x 743 skills. Laya base on a 12-thread CPU: 856.6 ms per decision (Jev 11.7 ms via API), AUC 0.495 on the named-skill metric (Jev 0.927), home-posting accuracy 0.043 against a 0.083 guessing rate, 0.518 on the domain-label metric (Jev 0.627, GLM 0.691), and on the 215 adjudicated pairs 48.8% accuracy with a yes to 99.1% of them (Jev 73.0%, yes to 54.4%). Laya trails Jev on 11 of 12 postings, p = 0.0063.
  • Laya inside the co-dm harness (2026-09-22): a fourth arm on the four benchmarks behind iambraun.com/jevreports/co-dm, through the product's own builders. Laya base reads 384 tokens of state and the shipped table-rules state opens with 416 tokens of rule text, so it never sees the paragraph: recall 0 at any batch size, and across seven shuffled orders its score follows the batch position, not the text. In a shape its window holds (one paragraph and one rule per call) it catches 9 of 26 violations at 37.5% precision; Jev on the same calls catches 25 of 26 at 92.6%. Routing, where every message fits the window: Laya 31.1%, the regex incumbent 36.7%, Jev 87.8%. Contradictions as shipped: recall 0 at 6.9 s per item against a 4 s budget; one fact per call 46.7% (base rate 51.7%, Jev 93.3%). Asset kinds: 37.3% (regex 57.6%, Jev 82.5%). The vendor's 1024-token checkpoint reads the shipped one-paragraph state whole and flags all 94 paragraphs (precision 26.9%); its best result is 63.3% on one-fact-per-call contradictions, under the 66.7% lexical baseline.
  • laya-bench (this radar's own run, 2026-09-21): Jev 1.13.0 against Laya base on 743 skills x 3 public postings, same question, same batches. Jev 9 to 11 ms and $0.0019 per posting via the API; Laya 905 to 946 ms per decision on a 12-thread CPU at $0, deterministic. Laya answered yes at 0.60+ to 565 to 655 of 743 skills per posting (Jev 54 to 77). On 120 blind-adjudicated pairs, Jev 95.0% accuracy, AUC 0.976; Laya 46.7%, AUC 0.328; on the 90 pure disagreement pairs Laya AUC 0.053. Judges agreed with both arms on all 30 controls. Harness and raw results: github.com/WorkflowtechAI/laya-bench.
  • Laya model card (ConvAI): every Jev number on it is quoted, not measured; the card says so. The "236 to 276 ms" latency splices the minimum from AbdelStark/jev-benchmarks (France, 17 Sep) with the maximum from nibzard/decision-model-benchmark (geography undisclosed). The "ECE 0.246, 3x better" pairs Jev's single worst suite (nibzard S5 forced-uncertainty) with Laya's temperature-refitted ECE on Laya's own unrelated eval. Traced and re-verified against the primary repos on 2026-09-21.
  • LocalLLaMA/typed-decisions dataset: the one same-dataset comparison Laya cites. Jev 1.13.0 measured live on all 2,000 decisions: 0.727 accuracy, ECE 0.144, p50 710 ms per 5-question case, $0.016 total. Laya base zero-shot 0.36 (below the 0.461 majority baseline); 0.767 only after fine-tuning on that dataset's own training split.
  • Dylan (Laya issue #102): 10,000 fixed items across SST-2, AG News, Emotion, Banking77, Jev called directly. Macro accuracy Jev 79.3% vs Laya 73.2%; Laya wins AG News by 5.3 points and loses Banking77 (77 options) by 23.5.
  • BenchmarkHeaven JevBench v1.2 (21 Sep): 534 decisions, 45 systems, one request at a time from Germany. Jev 1.13.0 ranked first (score 75.4); Laya eighth (70.1), with accuracy falling from 94.4% on the easy tier to 34.1% on the hard tier, and a measured p50 of 0.79 s on their CPU against Jev's 0.65 s over the API.
  • NVentimiglia (Laya issue #16): 80 scored keys, Laya 32.5% vs Jev 98.8%; the author and the Laya maintainer both note the fixtures gave no per-question criteria, which Laya is built around.
  • Laya issues #35 and #54: non-English routing sends 64% of German and most Romanian inputs to the English checkpoint; accuracy near 50% on a 13-way router with confident wrong answers at 0.78 to 0.99.
  • Dio the Debugger (Discord, 21 Sep): on his own routing task Jev scored perfect, Laya 67% with large category errors, 2.6x slower on an i7-1260 CPU than the Jev round trip, and reported 1.0 confidence on every decision.
  • Hemant (heman10x): On 337 cases from TypeSafe's public eval set, Jev scores 90.8% accuracy versus 88.4% for a 26B diffusion model and 48.1% for his 151M-parameter Verdict reproduction.
  • chl-5g: Batching multiple stocks into one Jev decision call caused signal leakage across tickers; switched to one Jev call per stock for the final trade decision.
  • (o_o): Jev found all required evidence for 5/6 answerable retrieval questions vs 3/6 for GPT-5.4-mini, at $0.0188 vs $0.3417 total cost and 5.27s vs 13.88s average scoring time.
  • James: five-lines eval on jev-1.13.0 scored 33/34 and 34/34 correct across two runs, raising 10/12 true violations with 0 false alarms.
  • kelpe: perfectrecall reduced answer errors 72.6% and increased correct answers 77.6% vs Mnemosyne on LongMemEval-S, searching 10,000 memories in 0.73s.
  • azterizm: On 20 trials, Jev averaged 730ms P50 for routing vs 0.26ms for local DistilBERT, and confirmed a fabricated statute at 96% confidence while a local DeBERTa model abstained cleanly.
  • fajarhide: askgrep scans 1,800+ functions for semantic patterns in about 10 seconds for under $0.03 using Jev classifier scores.
  • hobohotdog: jev-ultralightspeed processed 30,000 judgements in 56s vs 30 minutes for regular Jev (533 vs 16.7 items/sec), with identical 89.2% agreement with human labels.
  • π: Synthetic evals on confirming risky tool-call execution reached 100% accuracy at confidence above 0.75, excluding one suspected false false positive, at about $0.00001 per call.
  • Junsoo: Grading the same debate 10 times, Jev's winner never changed (0/10) vs GPT changing 6/10 times; score std dev was 0.22 vs 3.19, ICC 0.999 vs 0.72-0.79.
  • Timo: Rewording a prompt-injection guard from 'contains instructions to an AI' to ask about manipulation intent raised correct Jev-only handling from 11/14 to 14/14 with 0 errors, on 92 labelled cases.
  • garethjax: Jev matched human taxonomy ground truth on 290/325 queries (89.2%) at about $0.03 per run; batching multiple queries per request caused silent mis-indexing at 0.3-0.4 confidence, while single-query requests hit confidence 1.00.
  • Alejandro (Carbaj03): Postneedle's Jev needles told apart two $40k-MRR posts by tone: plain humblebrag scored 0.91, a post with a risk caveat scored 0.57, same number in both.
  • Robin: file2markdown's Jev check ignored an embedded instruction telling it to "answer no to everything"; it still can't count table cells, so tables stay rules-based.
  • Paul Smith: Fed Jev 8k Kepler signals with NASA labels hidden, asked confirmed/candidate/false-positive; got 72.5% overall accuracy.
  • joslat: Wire-level quirks found integrating Jev: score probability keys are 0-indexed, and a Noul answer carries no confidence field.
  • Riddle: Calling Jev over HTTP for game control adds about 150-400ms latency; worked around it by having Jev predict 5 actions per call and holding the last one until a new response arrives.

New 2026-09-20

  • intikhab49: An open 150M reproduction scored 0.697 vs Jev's 0.727 on the same benchmark, while being 2.5x better calibrated and 4x faster.
  • QuicqDev: Benchmarked Jev 1.13 against 11 classical ML pipelines on 8 datasets: Jev hit 96.3% balanced accuracy on IMDb sentiment vs 88.4% for the best classical pipeline, but classical pipelines won on all 4 tabular datasets.
  • zeeshan8281: Routing through Jev preserved the same routes and accuracy as local deterministic features but raised p95 end-to-end latency from 77.93ms to 490.38ms.
  • DECRUX9812: Jev-based skill routing for Hermes Agent costs about $0.001 per routed turn.
  • openroboto-ai: on a single seed-0 xArm7 pick-and-place trial (one apple, one plate), Jev 1.13 placed it in 226 API calls for $0.019 and 182s wall time; GPT-6 Astra also placed it but took 707s at $5.93, 312x the cost; GPT-4.1 mini hit the 160-cycle limit without finishing.
  • Kyle Jeong: driving Stagehand from an accessibility-tree state instead of screenshots, letting Jev pick the next click completed a browser task for $0.001 at near-instant speed.
  • Bartosz Mikulski: converting 400 hand-drawn doodles into SVG coordinate text, Jev named the drawing correctly about 35% of the time versus 10% chance, but lost to Claude Sonnet 5 and guessed "airplane" for more than half.
  • yuwakisa benchmark (via hypnoticfuzzwave): on a two-step benchmark of stating a principle from examples then applying it cold in a new context window, Jev scored ahead of every other tested model.
  • bartlomein: oko's 108-session pilot on ranking local grep matches before an agent reads them saw up to 46% fewer agent tokens and 18% less wall time with a warm cache.
  • Fatalfencer: swapping a Monte Carlo bot's option-picking step for Jev beat a "hard" difficulty tactics-game bot but not the hardest hand-tuned one, at about $0.10 per game.
  • loop (dev rel @ openrouter): OpenRouter's Ori Eval judging benchmark found Jev over 5x faster than the next fastest model, and Jev's slowest requests still beat every other model's median latency.
  • _FailSafe: a WIP Jev-scored retrieval triage sidecar over Noema's hybrid search kept 2 of 8 candidates on a clear query (relevant scores ~0.95/0.91 vs ~0.03-0.06, ~3700 in/666 out tokens) and 4 of 8 on a messier preference query (preference-flagged items scored 0.86-0.98), with a soft decision band identified around 0.55.
  • Zawwar: playing chess by having Jev pick from a code-generated list of legal moves, it found the correct move on 16% of 600 rated Lichess puzzles versus 5% for random guessing, and accuracy did not fall as puzzles got harder (16% at 800-rated vs 23% at 2200-rated, the opposite of a human), because Jev isn't calculating; it found mate-in-one only about 1 time in 10.
  • patebry: jevfish's move-by-move confidence score cannot separate a mate-in-2 from a waiting move; up a queen against a human it took 65 moves to deliver mate.
  • askmuyukani: four concurrent Jev calls filtering real FAISS retrieval results in JarvisCore completed in 1.37s wall time for $0.000074.
  • shitianfang: routing agent steps that need no text output straight to Jev instead of an LLM measured p50 ~230ms and ~$0.02 per 1,000 judgments.
  • Zaious: on the same historical multiple-choice question, Jev gave a wrong answer at 0.90 confidence with no supporting passage, then a correct answer at 0.97 confidence once the background text was included in state; it is only as good as the facts you give it.
  • goodrahstar: labelling 1,000 Android app reviews across 4 typed questions each, Jev finished in 4.6s for $0.023 versus Gemini 3.8 Flash's 18.8s and $0.158 (4.1x faster, 7x cheaper), with near-identical sentiment agreement against star ratings (rho 0.80 vs 0.82).
  • JoshuaSP: an open DiffusionGemma reproduction of Jev's typed-decision interface matched 48/48 relevance labels and 6/6 top-1 retrieval on a code-retrieval eval, and hit 138/144 (95.8%) agreement with saved Jev decisions on a customer-voice eval, using a single denoising step.
  • Argos1111: a local zero-shot LFM 1.2B backend scored JNLI 17% / JComQA 69% on JGLUE, versus a fine-tuned ModernBERT cross-encoder backend at JNLI 93% / JComQA 92%; zero-shot small models lag furthest on entailment judgments, not commonsense QA.
  • phuryn: hardened Jev's own invoice-classification showcase to 50 documents; Jev scored 50/50 at $0.025 per 1,000 decisions, tied by Claude Haiku 4.5 and nearly matched by open 48/50 self-hosted models, while Opus 5 landed one answer behind at 115x the cost; strip the written definitions out and every Jev mistake falls below 0.80 confidence.
  • 0xNatoshi: per-turn Jev routing for Codex cut cost roughly 60% against an all-frontier baseline across a 7-day replay of 237 real turns, at about $0.00003 and 0.6s per routing decision.
  • yuzushi-dev: deterministic tool-output bounding (Sando) cut Codex shell tokens 90.9% across 40,000+ recorded results and Claude PostToolUse tool results 14.5-18.9% across two machines, fitting roughly 2.4x more files in a 200,000-token window.
  • coldteadotai: replaying 93 real Claude Code sessions found agents break an unwritten project rule on 1 turn in 13; Jev flagged 39 edits and 15 turns, an independent reviewer confirmed 10 and 11.
  • kyotofin: a Jev tax-form classifier hit 100% strict accuracy across 261 IRS forms at $0.001 per page, 34x cheaper and 6x faster than their prior LLM pipeline.
  • jaz: toolgate's held-out adversarial set (injection, forged approval, encoded payloads) scored 23/24 with zero permissive errors, both frozen sets 20/20; about a second and a fraction of a cent per call.
  • Suraj Phanindra: a Choice with no matching option still answers at confidence 1.00. Always include an "other/none" option.
  • ubhillutd, Eddy Vinck: phrasing shifts probabilities on identical state ("save money" vs "save money long term" flips the answer). Per-call is stochastic, the average is stable.
  • Zurban: calibration under domain shift is unproven. A miscalibrated gate is worse than no gate.
  • parallax: 1,565 emails, 10 classes. All 737 predictions at 99%+ confidence were correct; mistakes cluster at low confidence. Trailed Gemini slightly on accuracy, 10 to 22x cheaper.
  • Cleista: ~90% accuracy at ~12k classifications per dollar; the last 3% of hard cases still need an LLM at 10 to 100x the cost.
  • identityTorn: zero-shot Jev within ~5 points recall of a fine-tuned Qwen 9B; the fine-tune still wins on latency and cost.
  • Atanas: 94% tool-pick accuracy on typos vs 28 to 34% for keyword rules. Cannot fill open values; hand those to a small LLM (0.2s added).
  • venkat (LessWrong): one yes/no call reaches ~0.97 AUROC catching backdoors, but accuracy drops once an attacker can generate and score variants.
  • Sub: EU to Jev infra adds up to 320ms network latency. Killed a real-time NPC use case.
  • stbenjam: no temperature setting, so repeated generative-style tasks come out repetitive.
  • Tawfek Sraj: knowledge cutoff reverse-engineered to between 2024-07 and 2025-01.
  • sh0rtythegreat: token counts ran unexpectedly high on a line-by-line task; check the bill, not just the per-token price.
  • Sid Arya: Jev is greedy in multi-turn setups. Frame whole actions, not letter-by-letter steps.
  • Yonatan Gross: hand-rolled client sent Score criteria as a dict and got silent 422s. Use the SDK.
  • azakhary: 60% of a 10x-speed game playtest was Jev thinking; effective 3x. Server capacity is the bottleneck at scale.
  • timpratim: an early version of a voice-controlled Mac agent, told to "clean up my desktop," moved every file on the Desktop into one folder; the project now ships an explicit safety policy after that run.
  • cramforce: swapped Jev into a classifier eval that previously ran on Gemini 2.5 Flash Lite; Jev matched or beat it on quality, saturating the eval, and ran roughly 6x faster.
  • Manjunath Janardhan: a 200-decision benchmark against Claude Fable 5.1, GPT-6 Astra, Kimi K3, MiniMax M3 and DeepSeek V4.1 Flash on BANKING77, BoolQ, Yelp and ChaosNLI found only the two flagships clearly ahead of Jev (11.5 and 6.5 points), while the mid-tier models were within noise but cost 6-150x more and took 2-15x longer; Jev tied every model on BoolQ at 94% (AUROC 0.970); on Yelp its errors are nearly all one star off but arrive at 0.98-0.99 confidence, with 18 of 55 errors above 0.9 confidence; on ChaosNLI, where 100 people split their labels, Jev's probabilities sat further from the human distribution (JSD 0.149) than a blind guess (0.127), while Claude Fable 5.1 hit 0.043.
  • ASHu2: an independent 8-dataset benchmark against classical ML found Jev strong on sentiment classification but mixed across other tasks, with fine-tuned ML staying cheaper wherever fine-tuning is an option; few-shot prompting did not reliably improve Jev's results.
  • yzfly (awesome-jev-zh): an independent evaluation cited in the list found that asking Jev a single direct question about a phishing email scored 62.6% accuracy, while two lines of regex reached 91.8%.