The complete index

Every public Jev record, organized

147 projects and examples that demonstrate what Jev can do, each one checked against its public artifact and re-read on a rolling basis. Search the ledger below, or browse by domain — every cluster carries an editorial note on what it proves and where the evidence stops.

  • 147 records
  • 147 artifact-verified or reproduced
  • 3 official · 144 community
  • 13 domains
  • Last re-check 2026-10-08

Public evidence ledger

Verified capabilities

147 of 147 records

  1. E2 · artifact verifiedCommunityFeatured

    Assess a simulated company under attack

    A local cybersecurity lab replays synthetic telemetry and asks Jev for compromise probability, classification, severity, and an advisory response as evidence accumulates.

    choicescorenoulsecurity

    Artifactreproducible repo

    Checked

  2. E2 · artifact verifiedOfficialFeatured

    Batch a regulatory briefing into one call

    An official cookbook asks 13 regulatory questions over a pinned GDPR article in one Jev request and compares that batch with 13 separate requests.

    choicescorenoulautomation

    Reported12.2 × cheaper

    Checked

  3. E2 · artifact verifiedOfficialFeatured

    Check LLM citations against source context

    An official cookbook combines exact quote matching with a Jev relation judgment to label citations verified, unsupported, contradicted, or fabricated.

    choiceagentsdeveloper tools

    Artifactofficial example

    Checked

  4. E2 · artifact verifiedCommunityFeatured

    Choose drone tactics from camera-derived state

    A simulated quadrotor uses Jev for low-frequency tactical judgments while classical vision, flight control, and safety reflexes remain in code.

    choicescorenoulrobotics

    Reported77.5 meters

    Checked

  5. E2 · artifact verifiedCommunityFeatured

    Complete a StarCraft shareware mission

    A reproducible harness lets Jev direct combat, exploration, and economy actions in the original StarCraft shareware campaign.

    choicegames

    Reported1 mission completed

    Checked

  6. E2 · artifact verifiedCommunityFeatured

    Drive macOS from structured screen state

    A computer-use loop combines OCR and accessibility data, then asks Jev which bounded action should move the Mac toward a plain-English goal.

    choicenoulagentsautomation

    Reported0.0002 USD per step

    Checked

  7. E2 · artifact verifiedCommunityFeatured

    Pick Pokémon turns from live ROM state

    A battle harness reads FireRed state from RAM and lets Jev choose the next legal move or switch while ordinary code advances the fight.

    choicegames

    Reported0.03 USD per run

    Checked

  8. E2 · artifact verifiedCommunityFeatured

    Replace Claude Code compaction with Jev decisions

    A Claude Code plugin and npm library asks Jev which old tool calls and results still matter, then drops or truncates the stale ones, so /compact replaces the lossy built-in summary with the kept messages verbatim.

    noulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  9. E2 · artifact verifiedOfficialFeatured

    Route smart-home assistant requests

    TypeSafe's smart-home demo evaluates a request against many typed questions in parallel, then lets code use only the answers relevant to that request.

    choicenoulhomeautomation

    Artifactofficial example

    Checked

  10. E2 · artifact verifiedCommunityFeatured

    Search Google Flights in seconds

    A browser agent turns the visible DOM into an indexed action space and uses Jev to choose the next operation and compatible target.

    choiceagentsautomation

    Reported7.073 seconds

    Checked

  11. E2 · artifact verifiedCommunityFeatured

    Search the web with typed intent judgments

    A metasearch front end lets Jev choose the query, sources, and time range, then score every result for relevance while code fans out to engines.

    choicescoresearchagents

    Artifactworking demo

    Checked

  12. E2 · artifact verifiedCommunityFeatured

    Turn Home Assistant state into automation signals

    A Home Assistant integration exposes Jev probabilities, choices, and scores as entities and action responses that automations can use.

    choicescorenoulhomeautomation

    Artifactreproducible repo

    Checked

  13. E2 · artifact verifiedCommunity

    Answer finance questions from two judged passages

    A finance RAG benchmark where Jev picks the passages: the agentic baseline reads whole SEC filings - 82,908 tokens for 46 of 50 right - while Jev-judged retrieval reads about 840 tokens and answers all 50; benchmark code and write-up are public.

    choicenoulresearchdeveloper tools

    Reported50 of 50 questions right at ~840 tokens per question

    Checked

  14. E2 · artifact verifiedCommunity

    Answer with one of five words

    A Turkish chat toy where whatever you type is answered by one of five fixed phrases - Jev picks which word, and two more answers decide the punctuation, in a single call.

    choicescorenoulresearch

    Artifactworking demo

    Checked

  15. E2 · artifact verifiedCommunity

    Ask Jev inside SQL and get real types back

    A DuckDB extension where jev_choice, jev_score, and jev_noul are scalar functions over table rows: the criteria literal is both the set of permitted answers and the column's SQL type.

    choicescorenouldeveloper tools

    Artifactreproducible repo

    Checked

  16. E2 · artifact verifiedCommunity

    Authorize agent tool calls before they run

    A starter kit from Kinde: identity and permissions decide what an agent may do, and Jev judges each call in about 200 milliseconds - does it match the request, is it destructive, does it follow planted text, does it exfiltrate - before anything runs.

    choicenoulscoresecurityagents

    Artifactworking demo

    Checked

  17. E2 · artifact verifiedCommunity

    Benchmark Wordle solvers with Jev word priors

    A reproducible experiment beyond the 'optimal' Wordle solver: Jev scores how answer-like each of 12,972 accepted words is, the prior feeds entropy search over 1,925 days of NYT answers, and every claim is archived with pinned inputs and a reproduce script.

    scoreresearchgames

    Reported1925 days of NYT answers in the benchmark

    Checked

  18. E2 · artifact verifiedCommunity

    Call Jev from Flutter apps

    A Flutter plugin that brings typed Jev calls to Dart apps: question builders for Noul and Choice with option-count validation, the systemone wire protocol on the native endpoint, and a Python-side test harness for the plugin's bridge.

    choicenoulmobiledeveloper tools

    Artifactreproducible repo

    Checked

  19. E2 · artifact verifiedCommunity

    Catch contract drift hidden in OpenAPI prose

    OAS Sentinel compares two OpenAPI documents in two layers: deterministic checks find structural breaks, and Jev answers bounded semantic questions about changed prose - retries, ordering, pagination, error meaning - that schema diffs cannot see.

    choicenouldeveloper toolssecurity

    Artifactreproducible repo

    Checked

  20. E2 · artifact verifiedCommunity

    Chat without generating any free text

    An experiment that makes a model which cannot write text answer anyway: for every word of the reply Jev picks 1 of 254 meaning-based word groups, then the word inside that group - and the README reports exactly where that stops working.

    choiceresearch

    Artifactworking demo

    Checked

  21. E2 · artifact verifiedCommunity

    Check rewrites for lost ideas

    Lossless Rewrite closes the loop on AI shortening your report: your model rewrites, and Jev checks every protected idea survived - exact wording, meaning, or a reviewed checklist - then helps repair what went missing.

    nouldeveloper toolsresearch

    Artifactreproducible repo

    Checked

  22. E2 · artifact verifiedCommunity

    Check suspicious messages for scams with typed verdicts

    Paste a suspicious SMS, email, DM or listing into ScamCheck - web app, API or browser extension - and get a scam verdict, risk score, plain-English reasons and next steps, with all wording from the project's own templates rather than a model.

    noulscoresecurity

    Artifactreproducible repo

    Checked

  23. E2 · artifact verifiedCommunity

    Choose a market quote every Monad block

    A trading loop reads the Kuru MON-USDC order book and asks Jev for a buy-or-sell judgment before code places a post-only limit order.

    choicetradingautomation

    Artifactreproducible repo

    Checked

  24. E2 · artifact verifiedCommunity

    Classify CI failures as flaky or real

    A GitHub Action that classifies test failures as regression, flaky, environment, or unknown: deterministic signals first, one structured Jev choice second, and a local policy that never auto-reruns tests or masks failures.

    choicenouldeveloper tools

    Artifactreproducible repo

    Checked

  25. E2 · artifact verifiedCommunity

    Classify dino obstacles, let physics time them

    An autonomous bot that plays the Chrome T-Rex Runner to a thousand points by asking Jev which action each obstacle requires - jump, duck, or run - and letting a measured physics model decide exactly when to press.

    choicegames

    Artifactreproducible repo

    Checked

  26. E2 · artifact verifiedCommunity

    Classify SVGs by weighting typed responses

    A research task treating Jev's answer distributions as classifier features: many small typed questions about an SVG, responses weighted and combined until the signal classifies the image - a study of whether decision calls can stand in for maths on pixels.

    noulchoiceresearch

    Artifactreproducible repo

    Checked

  27. E2 · artifact verifiedCommunity

    Compact conversation context by calibrated judgment

    jevtrim is a comparative analysis of context compaction driven by calibrated judgments instead of summarization: Jev scores every chunk for relevance, ordinary Python keeps what fits the token budget, and the result is auditable and replayable offline.

    noulscoreagentsresearch

    Artifactreproducible repo

    Checked

  28. E2 · artifact verifiedCommunity

    Control a robot arm through layered choices

    Jev drives a LIBERO robot through 27 control inputs with layered decisions - intent, then motion family, then input - while reversible physics previews evaluate candidate effects locally before anything executes.

    choicenoulscorerobotics

    Artifactworking demo

    Checked

  29. E2 · artifact verifiedCommunity

    Control Chrome by voice with typed commands

    A Chrome MV3 extension that turns speech into browser actions: Jev routes each spoken command to open a site, search, click a link, fill a form field or go back, with typed answers instead of parsed free text; ships with tests, CI and a side-panel command log.

    choicenoulautomationdeveloper tools

    Artifactreproducible repo

    Checked

  30. E2 · artifact verifiedCommunity

    Cover YouTube spoilers with one Noul judgment per comment

    A Chrome extension covers every YouTube comment the moment it appears, asks Jev one Noul question per comment, and keeps it covered whenever the probability says it discloses a concrete plot event.

    noulautomation

    Artifactreproducible repo

    Checked

  31. E2 · artifact verifiedCommunity

    Curate a daily science edition with typed labels

    Pipette reads arXiv, bioRxiv, medRxiv and 58 journals each morning and publishes a short, diverse daily edition: Jev labels and ranks, quoted sentences are the authors' own abstract lines, method and probabilities are public, and output is CC0 open data.

    noulresearchsearch

    Artifactreproducible repo

    Checked

  32. E2 · artifact verifiedCommunity

    Decide Home Assistant commands with typed answers

    A Home Assistant conversation agent - installable via HACS - that decides with a TypeSafe System One model instead of an LLM: your spoken or typed commands become typed choices and Nouls that drive devices, with metered costs documented.

    choicenoulhomeautomation

    Artifactreproducible repo

    Checked

  33. E2 · artifact verifiedCommunity

    Dodge Terraria bosses with typed reflexes

    A tModLoader mod whose boss fights are driven by Jev: every 200ms one request asks a nine-way intent, a five-band danger score, and whether to dash or jump; a per-frame reflex layer turns intents into keypresses.

    choicescorenoulgames

    Artifactreproducible repo

    Checked

  34. E2 · artifact verifiedCommunity

    Drive a browser with typed page judgments

    Give jev-browser a task and a URL: Jev picks one action per step from the page's clickable, typeable, and selectable elements and scores goal-met and stuck likelihood, while code owns budgets, recovery, and stop gates.

    choicenoulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  35. E2 · artifact verifiedCommunity

    Drive an Android phone from the A11Y tree

    An Android automation agent with a two-tier brain: Jev decides fast from the accessibility tree, a vision agent takes over only when the structural view is ambiguous, and ADB executes - cheap judge on the hot path, expensive one on the exceptions.

    choicenoulmobileautomation

    Artifactreproducible repo

    Checked

  36. E2 · artifact verifiedCommunity

    Drive Android from semantic UI state

    A proof-of-concept Android loop stabilizes the screen, builds a short list of valid actions, and lets Jev pick one while code executes it.

    choicemobileagentsautomation

    Artifactreproducible repo

    Checked

  37. E2 · artifact verifiedCommunity

    Drive the Aside browser with typed decisions

    bside pilots the Aside browser with Jev instead of a chat LLM: each tick answers which action, which element, and whether the goal is met, over an action schema the pilot cannot hallucinate outside of.

    choicenoulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  38. E2 · artifact verifiedCommunity

    Evaluate Jev on the SNIPS NLU benchmark

    An evaluation of Jev on the SNIPS natural-language-understanding benchmark - intent detection and slot filling - asking how far a model that never generates text gets on a task normally solved by a trained tagger, using label names alone.

    choicenoulresearch

    Artifactreproducible repo

    Checked

  39. E2 · artifact verifiedCommunity

    Expose Jev decisions to MCP clients

    JDE, the Jev Decision Engine, wraps Jev as an MCP server: any MCP client - Claude Code, Cursor, your own agent - gets typed decision tools, with a policy layer, a decision ledger, and recorded evals comparing the hosted jev-1.13.0 against local alternatives.

    choicenoulscoreagentsdeveloper tools

    Artifactreproducible repo

    Checked

  40. E2 · artifact verifiedCommunity

    Expose typed Jev judgments to MCP agents

    An MCP server gives compatible agents tools for classification, scoring, checking, matching, screening, and custom typed Jev questions.

    choicescorenoulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  41. E2 · artifact verifiedCommunity

    Extract clinical review data as verbatim quotes

    A browser app for systematic-review extraction: Jev never writes the answer - it points at line ids in trial reports and supplements, and code copies the quote out with its file, page, row, or slide, highlighted where it sits.

    choicenoulresearch

    Artifactworking demo

    Checked

  42. E2 · artifact verifiedCommunity

    Fact-check a link claim by claim against its own sources

    A live link-checker extracts each claim from a submitted page, asks Jev whether the cited excerpts support, contradict, or fail to establish it, and reports REAL or FAKE only when enough evidence agrees.

    choiceresearch

    Artifactreproducible repo

    Checked

  43. E2 · artifact verifiedCommunity

    Fade low-value web text with sentence-level Nouls

    Osso fades the parts of a page that are not what the reader came for: each sentence of the main text gets a Jev probability, workspace sites are covered by default, and password fields, account pages and reviews are left alone by construction.

    noulresearchautomation

    Artifactreproducible repo

    Checked

  44. E2 · artifact verifiedCommunity

    Feel Jev speed and cost in a 25-second game

    A tiny Japanese game with no send button: type IT buzzwords and each keystroke gets a Jev judgment that stretches a meter, so 25 seconds of play answers what curl never does - how fast and how cheap the model feels inside a real app.

    choicegamesdeveloper tools

    Artifactreproducible repo

    Checked

  45. E2 · artifact verifiedCommunity

    File local documents with typed safety lanes

    A local-first document filer: text is extracted locally, Jev decides category, confidentiality and prompt-injection risk as typed Choices, and low-confidence or suspicious files land in review lanes - never overwritten, with audit preview and undo.

    choiceautomationsecurity

    Artifactreproducible repo

    Checked

  46. E2 · artifact verifiedCommunity

    Filter admin tables by vibe with typed decisions

    A Filament plugin for Laravel admin panels that filters tables by natural language instead of SQL: type "the customer is angry" and Jev's typed decisions drive the where-clauses - shipped as a Packagist package with a driver system, tests and a live demo.

    choicenouldeveloper tools

    Artifactreproducible repo

    Checked

  47. E2 · artifact verifiedCommunity

    Filter news and YouTube feeds by interest

    A single-file Python portal that filters news and YouTube feeds by interests you describe in plain English: the official typesafe-sdk scores every item, routine business hides separately, and the result is one static page with News and YouTube tabs.

    noulscoresearchautomation

    Artifactreproducible repo

    Checked

  48. E2 · artifact verifiedCommunity

    Filter streams of text with typed batch judgments

    jevpipe pipes thousands of lines, files or records through one question and gets a typed judgment per item - grep-style filtering where Jev decides - shipped as a Rust binary on PyPI with an agent skill that teaches coding agents when to reach for it.

    choicenouldeveloper toolsagents

    Artifactreproducible repo

    Checked

  49. E2 · artifact verifiedCommunity

    Filter your X timeline with five typed signals

    An open-source Chrome extension asks Jev to judge each visible X post for relevance, substance, practical value, promotion, and engagement bait, then dims, collapses, or hides it under weights the reader owns.

    noulautomation

    Artifactreproducible repo

    Checked

  50. E2 · artifact verifiedCommunity

    Find city permits from a free-text plan

    Describe what you want to do in San Francisco and a rules engine works out which permits you need - Jev answers each rules question as a typed choice, falls back to asking you when unknown, and summarizes fees and deadlines per permit.

    choicelegaldeveloper tools

    Artifactreproducible repo

    Checked

  51. E2 · artifact verifiedCommunity

    Find the word aphasia can't say

    For people with aphasia who know what they mean but can't get the word out: describe it any way you can, and Jev picks the best guesses from a fixed 900-word list as big tap-to-hear picture tiles - never inventing a word.

    choicemobile

    Artifactworking demo

    Checked

  52. E2 · artifact verifiedCommunity

    Fit confidence thresholds on your own data

    jevcal measures a typed decision model on private labeled data, fits per-question thresholds to a target accuracy, and fails CI when a model update drifts.

    choicescorenouldeveloper tools

    Artifactreproducible repo

    Checked

  53. E2 · artifact verifiedCommunity

    Gate agent commits with plain-English rules

    tenet makes agents fix rule violations before you ever see the diff: rules live in a YAML file in plain language, and Jev answers each one with a calibrated probability that becomes a pass-or-fail cutoff.

    nouldeveloper toolsagents

    Artifactreproducible repo

    Checked

  54. E2 · artifact verifiedCommunity

    Gate agent posts for leaks by audience

    A guard for OpenClaw agents that reads an outgoing message and where it is going: Jev judges whether a client name, credential or internal hostname is about to reach the wrong readers, then confirms, blocks or rewrites per channel.

    noulchoicesecurityagents

    Artifactreproducible repo

    Checked

  55. E2 · artifact verifiedCommunity

    Gate LLM-generated requirements with 22 questions

    A requirements quality gate: one batched call asks 22 atomic questions across five MECE facets about an AI-generated requirement, and thresholds route it pass, human review, or reject - never rewriting, only judging.

    choicescorenouldeveloper toolslegal

    Reported0.0022 USD

    Checked

  56. E2 · artifact verifiedCommunity

    Gate paid-API spend on relevance and budget

    Four proof-of-concepts wiring Jev in front of a pay-per-call API marketplace: relevance below 0.7 confidence skips the paid call entirely, a typed choice routes to exactly one endpoint, and a second independent gate enforces per-team budget caps.

    choicenoulautomationdeveloper tools

    Artifactreproducible repo

    Checked

  57. E2 · artifact verifiedCommunity

    Gate risky coding-agent tool calls

    A Pi extension asks Jev to flag destructive, exfiltrating, or out-of-scope tool calls and to classify failures in command output.

    choicescorenoulagentsdeveloper tools

    Reported203 requests

    Checked

  58. E2 · artifact verifiedCommunity

    Give a companion character sub-second typed reflexes

    A Japanese companion game where one Jev pass decides the character's true feeling - Choice, affection Score, dislike Noul, topic Choice - in 0.2 to 0.5 seconds, so her face and a one-liner land before Claude-written dialogue and a Gemini TTS voice.

    choicescorenoulgames

    Reported0.5 seconds

    Checked

  59. E2 · artifact verifiedCommunity

    Give a simulated person fast feelings

    Mina, a simulated 34-year-old librarian, runs on two systems: Jev reads her body, senses and clock every second and accumulates feelings; only when a feeling crosses its line does an LLM stop and think.

    choicenoulscoreagentsautomation

    Reported0.53 USD

    Checked

  60. E2 · artifact verifiedCommunity

    Give Jev eyes over cameras and streams

    A bridge connecting images, video streams and RGB-D cameras to Jev's judgment engine: identify what matters in a frame, estimate risk, score a situation or judge many visible objects at once - typed answers over the visual world, 46 stars in its first day.

    noulchoiceresearchdeveloper tools

    Artifactreproducible repo

    Checked

  61. E2 · artifact verifiedCommunity

    Grep for what code does, not what it's called

    jgrep answers semantic queries like catches an error and silently ignores it over a whole source tree in about two seconds for a cent: one typed yes-or-no judgment per code chunk, sixteen chunks per request, no index.

    nouldeveloper tools

    Artifactreproducible repo

    Checked

  62. E2 · artifact verifiedCommunity

    Ground agent answers in cited document blocks

    TraceDocs structures documents into source-linked blocks; Jev judges each with four Nouls - relevant, evidence, contradicts-premise, prompt-injection - returning a cited evidence set with a trace; the LLM writes from evidence and refuses when none exists.

    noulagentsresearch

    Artifactreproducible repo

    Checked

  63. E2 · artifact verifiedCommunity

    Guard .NET LLM apps with typed checks

    Kassad brings calibrated guardrails to .NET: every prompt, completion, tool call and citation passes narrow typed checks answered by Jev, batched one round trip per stage, thresholded in code into Allow, Flag, Review, or Block.

    noulscoresecuritydeveloper tools

    Artifactreproducible repo

    Checked

  64. E2 · artifact verifiedCommunity

    Improvise at the piano with typed mood choices

    Describe a mood - a slow, sad waltz - and a live piano plays it, with Jev deciding continuously as it goes: every musical choice is a typed question answered with probabilities, streamed to a public site in real time.

    choicenoulmusic

    Artifactreproducible repo

    Checked

  65. E2 · artifact verifiedCommunity

    Interview your AI worldview with typed questions

    Doom or Bloom maps where you stand between AI doom and bloom: a dynamic interview where the engine picks the next curated question by where your answers are thinnest, with Jev interpreting answers and scoring candidate follow-ups.

    choicescorenoulresearch

    Artifactworking demo

    Checked

  66. E2 · artifact verifiedCommunity

    Judge Azure Foundry agents with typed evaluators

    Four drop-in evaluators for Azure AI Foundry, rebuilt on Jev: Intent Resolution, Task Adherence, Tool Call Accuracy, Groundedness - each metric becomes small typed questions answered in one call, combined into a 1-5 score listing its checks.

    noulchoiceresearchdeveloper tools

    Artifactreproducible repo

    Checked

  67. E2 · artifact verifiedCommunity

    Judge every new arXiv paper each morning

    Paper Radar reads all of arXiv so you read the few that matter: every new paper is judged against plain-English interests with calibrated per-interest probabilities - about six cents a day for everything, no pre-filtering.

    noulscoreresearchsearch

    Reported0.002 USD

    Checked

  68. E2 · artifact verifiedCommunity

    Judge player persuasion by each character's values

    An npm library that lets game characters argue back: give a name, a persona and a goal, pass what the player typed, and Jev judges whether that character - with those values - was convinced; the same line can win a greedy merchant and offend an honest guard.

    choicescorenoulgamesdeveloper tools

    Artifactreproducible repo

    Checked

  69. E2 · artifact verifiedCommunity

    Jump to the video moment that answers you

    A Chrome extension for YouTube: ask the video a question in plain text, and Jev's typed judgments locate the moment that answers it - the player seeks straight there instead of you scrubbing.

    noulchoicesearchdeveloper tools

    Artifactreproducible repo

    Checked

  70. E2 · artifact verifiedCommunity

    Let Jev paint a sketch region by region

    A generative painting instrument: select part of a sketch and Jev chooses its paint material - one typed Choice per region with probabilities and certainty bands - and the material flies in and paints itself; demo film and live gallery included.

    choicegames

    Artifactreproducible repo

    Checked

  71. E2 · artifact verifiedCommunity

    Lint agent writes in 300 ms with team rules

    A fuzzy linter that watches a coding agent write and speaks up 0.3 seconds later: Jev checks the file against your team's rules - race conditions in effects, missing cleanup, house style - so the agent fixes them before any human reviews the code.

    noulchoiceagentsdeveloper tools

    Artifactreproducible repo

    Checked

  72. E2 · artifact verifiedCommunity

    Make Jev talk one word at a time

    A demonstration pushing the judgment-only model past its envelope: Jev 'writes' by answering which-word-comes-next in 250-word batches - each option shown as the whole reply so far plus the word - top-3 shortlist, final pick, until sentence end.

    choiceresearch

    Artifactreproducible repo

    Checked

  73. E2 · artifact verifiedCommunity

    Measure confidence-gated routing on labelled data

    An independent study routes Jev confidence into a larger model on two labelled datasets and shows the winning settings do not transfer between them.

    choiceresearch

    Reported80.2 %

    Checked

  74. E2 · artifact verifiedCommunity

    Measure Jev against LLMs on real workflows, weekly

    A weekly measured series pitting Jev against frontier LLMs on the same real workflow steps: week one routed inbound leads - Jev 90 percent correct at 366 milliseconds and four cents per thousand, against Sonnet 5's 78 percent at 2.6 seconds and three dollars.

    choicenoulscoreresearch

    Reported90 %

    Checked

  75. E2 · artifact verifiedCommunity

    Mine repeating patterns from noisy sequences

    Finding and extracting repeating patterns from noisy sequences with Jev: instead of a hand-tuned distance metric, typed questions decide what counts as the same pattern, and matches come back with probabilities instead of thresholds you guess.

    choicescoreresearchdeveloper tools

    Artifactreproducible repo

    Checked

  76. E2 · artifact verifiedCommunity

    Moderate community messages with typed probabilities

    jevmod scores every community message for spam, scam, harassment, NSFW, self-harm, doxxing, off-topic and custom plain-English rules; operators set thresholds, decisions log their numbers, and bots ship for Discord, Twitch, YouTube and Reddit.

    noulsecurityautomation

    Reported0.04 USD per 1,000 messages

    Checked

  77. E2 · artifact verifiedCommunity

    Moderate SAP Commerce reviews with four questions

    An SAP Commerce extension where Jev answers four yes-or-no questions per product review - abusive, spam, personal data, on-topic - and code turns the probabilities into approve, reject, or pending, with a dry-run mode that judges against human decisions first.

    noulautomation

    Artifactreproducible repo

    Checked

  78. E2 · artifact verifiedCommunity

    Morph one input into the right UI with typed intent

    One text box that becomes the right UI as you type - an event card, checklist, timer, color picker, bill splitter or poll - with Jev Nouls deciding what the input means and code rendering the component; live demo included.

    nouldeveloper tools

    Artifactreproducible repo

    Checked

  79. E2 · artifact verifiedCommunity

    Navigate a Neo4j graph one hop at a time

    At each node of a Neo4j graph the outgoing relationships become Choice options; Jev returns a full probability distribution over which one to follow, with a goal-reached Noul riding in the same call.

    choicenoulsearch

    Artifactreproducible repo

    Checked

  80. E2 · artifact verifiedCommunity

    Paint alongside Jev on a shared emoji canvas

    A shared 1,000-by-1,000 emoji canvas where humans place strokes and Jev paints with them: after each stroke one typed call picks a contextually relevant emoji and where to put it, and a yes/no decides whether your stroke was finished.

    choicenoulgamesresearch

    Artifactreproducible repo

    Checked

  81. E2 · artifact verifiedCommunity

    Paper-trade crypto on typed market state

    A Next.js dashboard that streams live crypto data, computes fifteen moving averages and eleven oscillators into one typed market state, and lets a Jev agent trade a simulated 100k portfolio - paper only, no broker connected.

    choicenoulscoretrading

    Artifactreproducible repo

    Checked

  82. E2 · artifact verifiedCommunity

    Pick the best line of a shared story

    Jev Yarn is a party game where everyone writes the next sentence and Jev picks the winner: one taste request scores every line on four dimensions, and Nouls handle room filters and whether the story feels finished.

    choicescorenoulgames

    Artifactworking demo

    Checked

  83. E2 · artifact verifiedCommunity

    Pick the right viewer for any file

    Drop a file or paste a URL and get a floating window with the right tool: Jev decides and composes a viewer from the registry, and when the format is unknown, Haiku invents a spec for a mini-app on the spot.

    choicenouldeveloper tools

    Artifactworking demo

    Checked

  84. E2 · artifact verifiedCommunity

    Plan day trips with a typed checker

    Jev Trip is an explainable day-trip planner where the LLM plans ahead and Jev chooses and checks: scope choices keep one day in one city, per-place choices rank every candidate with probabilities, and code owns routes, times and validation.

    choiceagentsautomation

    Artifactreproducible repo

    Checked

  85. E2 · artifact verifiedCommunity

    Play League of Legends with three Jev heads

    Jev plays Yasuo in League of Legends through three decision heads - strategy once a second, tactics six to seven times a second while units are on screen, and build checks every twenty seconds - while code reads the game and executes.

    choicenoulscoregames

    Artifactreproducible repo

    Checked

  86. E2 · artifact verifiedCommunity

    Play Pac-Man on typed decisions in the browser

    A single-file Pac-Man that asks the System One endpoint for typed decisions per game tick, playable live without setup or locally with your own key pointed at the official endpoint.

    choicegames

    Artifactworking demo

    Checked

  87. E2 · artifact verifiedCommunity

    Play Super Mario from emulator state

    An experimental controller translates NES telemetry into object-centric JSON and lets Jev choose the next legal controller macro.

    choicescorenoulgames

    Artifactreproducible repo

    Checked

  88. E2 · artifact verifiedCommunity

    Plug a fail-closed decision layer into your agent

    A pluggable decision layer for the ego agent: System One (Jev) by default, swappable to local or other OpenAI-compatible backends, fail-closed guardrails - and the README's whole argument is that it is measured: 16 suites, 429 checks, rerun in full.

    choicenoulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  89. E2 · artifact verifiedCommunity

    Predict the window you want next on macOS

    A macOS menu-bar switcher asks Jev which of the ten most recent apps you intend on a hotkey press and falls back to the last-used app on any failure.

    choiceautomationdeveloper tools

    Artifactreproducible repo

    Checked

  90. E2 · artifact verifiedCommunity

    Propose support decisions a guard can check

    A Kosovo electronics shop’s support agent reads a unified Albanian and English inbox and lands every message on auto-resolve, verification, or escalate - with Jev only proposing toward caution and deterministic code owning facts, access, and the final call.

    choicenoulautomation

    Artifactreproducible repo

    Checked

  91. E2 · artifact verifiedCommunity

    Prune stale coding-agent tool history

    A Pi extension uses Jev to decide which old tool calls and results still matter while keeping conversation text verbatim.

    noulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  92. E2 · artifact verifiedCommunity

    Rank zsh history suggestions while typing

    A zsh plugin asks Jev which recent command you are completing and shows the best match with its probability, while code owns gating and acceptance.

    choicenouldeveloper tools

    Reported0.8 seconds per request

    Checked

  93. E2 · artifact verifiedCommunity

    Rate LinkedIn posts for AI slop and bait

    A browser extension where Jev answers 17 typed questions about every LinkedIn post - thirteen AI-tell questions plus four bait questions in the same request - and fixed weights turn the answers into a badge and a bait chip.

    choicescoreautomation

    Reported67 %

    Checked

  94. E2 · artifact verifiedCommunity

    Rate pending shell commands on a color-coded safety rubric

    The kamchatka terminal agent asks Jev to place each pending shell command on a three-level safety rubric - reads and reports, changes something reversibly, destroys or sends something out - drawn green, yellow, or red beside the permission prompt.

    scoreagentsdeveloper tools

    Artifactreproducible repo

    Checked

  95. E2 · artifact verifiedCommunity

    Read customer-inquiry emotion in six questions

    A Japanese-language experiment scoring 100 labeled customer inquiries with the official TypeSafe SDK: one request per inquiry answers six questions at once - sentiment, emotion, anger intensity, urgency, churn risk, and sarcasm.

    choicescorenoulresearch

    Artifactreproducible repo

    Checked

  96. E2 · artifact verifiedCommunity

    Recover OCR marker formatting with typed choices

    OCR flattens superscripts: a footnote star, an endnote number and a unit power land in the stream as look-alike tokens. A regex over-finds the suspects, then Jev classifies each - footnote, citation or unit - as a typed choice so formatting can be restored.

    choicedeveloper toolsresearch

    Artifactreproducible repo

    Checked

  97. E2 · artifact verifiedCommunity

    Redact PII inside Postgres, by content

    A support inbox where Jev decides per span whether text is personal data and of what kind, and a redact() SQL function enforces the masking in Postgres by the viewer's clearance - content-aware, not pattern-based.

    choicesecurity

    Artifactreproducible repo

    Checked

  98. E2 · artifact verifiedCommunity

    Remove ad-like DOM elements in Chrome

    A Chrome extension finds ad-shaped DOM candidates and asks Jev whether each candidate is a paid advertisement before code removes it.

    noulautomation

    Artifactreproducible repo

    Checked

  99. E2 · artifact verifiedCommunity

    Rerank Postgres product search with two questions

    Product search inside PostgreSQL - typos, barcodes, typeahead, facet counts, all in SQL - with an optional second stage asking Jev two questions per search to rerank the shortlist; the README is itself the report, three failed versions included.

    choicenoulsearchdeveloper tools

    Artifactreproducible repo

    Checked

  100. E2 · artifact verifiedCommunity

    Rerank search results by natural-language criteria

    A Pinecone official examples repository: full-text search retrieves 200 candidates, and one Jev judgment pass reranks them to 10 by natural-language criteria, with a Claude baseline column for comparison.

    scorechoicesearch

    Artifactworking demo

    Checked

  101. E2 · artifact verifiedCommunity

    Research social media through real browser operations

    Jev Social pairs the decision model with socai, a CLI that drives your real Chrome across Instagram, TikTok, and LinkedIn: Jev chooses each next read-only operation - search, open a post or profile, read comments - and socai executes it.

    choicesearchautomation

    Artifactreproducible repo

    Checked

  102. E2 · artifact verifiedCommunity

    Resolve merge conflicts by choosing candidates

    Hunkpick resolves git conflicts by computing every plausible resolution itself - ours, theirs, union, line merge, token merge - discarding the ones that fail to parse, and asking Jev only to choose.

    choicedeveloper tools

    Artifactreproducible repo

    Checked

  103. E2 · artifact verifiedCommunity

    Restyle websites with typed catalog choices

    A Chrome MV3 extension that restyles any site from a plain-English prompt: one Jev call picks palette, fonts, spacing and intent from a fixed catalog, and deterministic code compiles role-stamped CSS - no LLM ever writes CSS that can break.

    choicenouldeveloper tools

    Artifactreproducible repo

    Checked

  104. E2 · artifact verifiedCommunity

    Review commits and copy with typed questions

    nudgement reviews commit messages, code, comments, tests and UI copy before you commit: exact checks plus focused Jev questions catch what formatters cannot - a message that misrepresents the diff, or a test that passes even when behavior broke.

    nouldeveloper tools

    Artifactreproducible repo

    Checked

  105. E2 · artifact verifiedCommunity

    Review pull requests against Five Lines of Code

    A single binary that reviews a diff against the ten refactoring rules of Clausen's Five Lines of Code: the countable rules run on a real parser, and Jev answers the judgment rules as typed questions whose probabilities set the bar.

    noulscoredeveloper tools

    Artifactreproducible repo

    Checked

  106. E2 · artifact verifiedCommunity

    Review pull requests with calibrated bug risk

    A single-file local pull-request reviewer: deterministic code does the plumbing while Jev judges each hunk with a real-issue Noul, scores severity, and returns a PR-level risk with a needs-human probability - no agent loop, no prompts to tune.

    noulscoredeveloper tools

    Artifactreproducible repo

    Checked

  107. E2 · artifact verifiedCommunity

    Route a multilingual helpdesk with typed triage

    FastGate fronts an English/Uzbek/Russian university helpdesk with four narrow Jev judgments per message plus per-passage grounding, routes deterministically in code, and ships an independent benchmark of Jev on low-resource languages.

    choicenoulautomationresearch

    Artifactreproducible repo

    Checked

  108. E2 · artifact verifiedCommunity

    Route focused code-review investigations

    A staged review workflow uses Jev to identify risky areas, select evidence, classify mechanisms, score severity, and route follow-up checks.

    choicescorenouldeveloper tools

    Artifactreproducible repo

    Checked

  109. E2 · artifact verifiedCommunity

    Route messy music requests before the LLM

    A DJ chatbot's pre-LLM router, built as a learning study: free-form /play requests - English, Spanish, typos, pasted lyrics, emojis - classified by Jev in one call (~500 ms, ~$0.000036) into a structured hint that tells the expensive LLM what it is handling.

    choicenoulmusicdeveloper tools

    Artifactreproducible repo

    Checked

  110. E2 · artifact verifiedCommunity

    Run a live crypto desk with Jev as the brain

    A multi-bot OKX trading desk with Jev as the brain and code as the body: each tick reads market and account state, computes features in TypeScript, asks Jev typed questions, then holds or places and cancels real orders - dashboard as the glass.

    choicenoultrading

    Artifactreproducible repo

    Checked

  111. E2 · artifact verifiedCommunity

    Run an e-commerce store on typed decisions alone

    Jev Mart is a Japanese e-commerce demo where every operational decision - listings, inquiries, reviews, triage - is a typed Jev call and an LLM is never invoked; it ships with a twelve-chapter lecture set that teaches the pattern.

    noulchoicedeveloper toolsresearch

    Artifactreproducible repo

    Checked

  112. E2 · artifact verifiedCommunity

    Run Jev judgments from the shell before an agent acts

    jev-axi is a CLI that puts a half-second Jev opinion in front of every command an agent runs: pick, rate, check, rank, triage, and guard, plus a PreToolUse hook that blocks risky Bash calls before they execute.

    choicescorenoulsecurityagentsdeveloper tools

    Reported375 ms

    Checked

  113. E2 · artifact verifiedCommunity

    Run typed judgments as eleven shell commands

    A CLI named jev turns Jev's typed questions into pipable commands - verify, screen, classify, extract, find, rerank, match, route, ask, compact, batch - with keychain auth and exit codes a script can gate on.

    choicenoulscoredeveloper tools

    Artifactreproducible repo

    Checked

  114. E2 · artifact verifiedCommunity

    Score a whole novel by emotion, one passage at a time

    Book Aurora sends every ~90-word passage of Frankenstein to Jev as ten parallel score questions - nine emotions plus overall intensity - and draws each answer as one feathered row of a full-book aurora strip.

    scoreresearch

    Reported0.0337 USD

    Checked

  115. E2 · artifact verifiedCommunity

    Score code quality for coding agents

    Supercov asks Jev yes-or-no questions about every source file, does the arithmetic in code, and turns weak spots and coverage gaps into agent tasks.

    nouldeveloper tools

    Reported0.01 USD per megabyte of source

    Checked

  116. E2 · artifact verifiedCommunity

    Score video statements by viewpoint

    jevmeter turns a video into a BS meter: every sentence scored and every dodge flagged under a preset viewpoint - debates, earnings calls, podcasts, pitches - rendered as a 16:9 edit you can post.

    scorenoulresearch

    Reported0.05 USD

    Checked

  117. E2 · artifact verifiedCommunity

    Screen Spain's official gazette every morning

    A daily pipeline reads the BOE with one Jev call per provision, scoring impact, tagging topics, and selecting an original paragraph as the summary.

    choicescorenoullegalautomation

    Reported0.01 USD per daily gazette

    Checked

  118. E2 · artifact verifiedCommunity

    Select which tests a change needs

    A sibling GitHub Action that decides which allowlisted test groups a pull request needs and which can safely be skipped: the group list stays in your config, Jev picks among it, and deterministic rules a model cannot bypass have the final word.

    choicedeveloper tools

    Artifactreproducible repo

    Checked

  119. E2 · artifact verifiedCommunity

    Shadow a live CRM with typed decisions

    A working example of a shadow decision layer under a production CRM: one call returns intention, business division, human-handover probability, urgency, and lead quality - decided and logged, never sent.

    choicenoulscoreautomationagents

    Artifactreproducible repo

    Checked

  120. E2 · artifact verifiedCommunity

    Show every search decision as it happens

    OpenRecurSearch is an agentic web-search interface where nothing stays hidden: each research layer and each Jev decision appears live in the chat - option probabilities, score and latency included - while the markdown report streams into the side panel.

    choicescoresearchagents

    Artifactreproducible repo

    Checked

  121. E2 · artifact verifiedCommunity

    Sift emoji out of the pile by description

    Type a description and matching emoji fly out of the pile: one request carries the query as shared state with one score question per emoji, answered in parallel, and code ranks and floors the results.

    scoresearch

    Artifactreproducible repo

    Checked

  122. E2 · artifact verifiedCommunity

    Silence ad noise on Android, failing open

    An Android noise gate for notifications and SMS that asks Jev whether each message is an ad - and suppresses only what is explicitly flagged. Every uncertain path resolves to allow, because swallowing a verification code costs more than leaking an ad.

    choicenoulmobileautomation

    Artifactreproducible repo

    Checked

  123. E2 · artifact verifiedCommunity

    Sort Gmail and grade replies with typed choices

    A light terminal Gmail client for Omarchy Linux: Jev files incoming mail into Gmail labels only above a confidence threshold, and grades your reply as you write - clarity, tone, length, next step, and which questions you haven't answered yet.

    choiceautomationdeveloper tools

    Artifactreproducible repo

    Checked

  124. E2 · artifact verifiedCommunity

    Speak in sentences compiled from decisions

    JevSpeak holds a conversation without any generative model in the loop: Jev answers about thirteen parallel questions per turn, the decisions normalize into a semantic IR, and a deterministic compiler writes the sentence.

    choicescorenoulresearch

    Artifactworking demo

    Checked

  125. E2 · artifact verifiedCommunity

    Steer a hospital delivery robot past four hundred people

    A hospital delivery robot judged about four times a second: code samples twelve candidate paths and removes predicted collisions, Jev picks one among the survivors, and a deterministic safety brake guarantees no contact - built at the TypeSafe JEVATHON.

    choicerobotics

    Reported84 s

    Checked

  126. E2 · artifact verifiedCommunity

    Steer a music composition with enum labels

    A playground asks Jev to decide only closed-vocabulary labels for character, key, meter, and phrasing while deterministic code writes, engraves, and plays the notes.

    choicescorenoulmusic

    Artifactworking demo

    Checked

  127. E2 · artifact verifiedCommunity

    Steer a simulated fly brain with typed decisions

    Spikecast puts Jev at the controls of a fly's simulated brain: the insect walks a road, and every turn, stop or dash is a typed decision with the synapse memory and neural firing visualised beside it - with docs separating what is real from what is modelled.

    choiceresearchgames

    Artifactreproducible repo

    Checked

  128. E2 · artifact verifiedCommunity

    Steer a town of fifty citizens with one broadcast

    You are the Town Crier of a 3D town: write one broadcast and all fifty citizens decide in parallel - investigate, join, flee, warn, or ignore - in a single batched choice request.

    choicegames

    Artifactworking demo

    Checked

  129. E2 · artifact verifiedCommunity

    Steer hydroponic crops with typed judges

    A Home Assistant add-on runs a crop-steering irrigation engine with Jev judging what fixed rules got wrong - ramp timing, probe trust, shot landing, EC moves, alerts - inside code-enforced physical limits; the README replays real Jev answers from 27 Sep 2026.

    choicenoulscorehomeautomation

    Reported0.92 probability

    Checked

  130. E2 · artifact verifiedCommunity

    Steer matter paths with live decisions

    An interactive installation - explore a flooded observatory where glowing matter builds paths from your movement, gaze and actions, with real-time Jev decisions deciding how the world responds, shipped as a playable web app with live-decision tests.

    choicenoulgamesresearch

    Artifactreproducible repo

    Checked

  131. E2 · artifact verifiedCommunity

    Steer RAG chunking and index hygiene with Jev

    Chunk documents where meaning changes, not at a character count: Jev judges sentence continuation, planted instructions are quarantined before embedding, and retrieved passages classified as evidence, conflict or noise - on your existing vector database.

    noulchoicesearchsecurity

    Artifactreproducible repo

    Checked

  132. E2 · artifact verifiedCommunity

    Supervise a coding agent while it works

    Foreman runs an independent observation loop that asks Jev whether a coding worker is progressing, stuck, complete, or ready for verification.

    noulagentsdeveloper tools

    Artifactreproducible repo

    Checked

  133. E2 · artifact verifiedCommunity

    Tag a calibre library with typed suggestions

    A calibre plugin that suggests subject and genre tags for your books with Jev: suggestions arrive for review before anything is applied, existing tags are preserved, and the plugin is independent GPL software with a live demo and production notes.

    choicedeveloper toolsresearch

    Artifactreproducible repo

    Checked

  134. E2 · artifact verifiedCommunity

    Triage a store inbox into four lanes

    A quiet clerk for a small online store's inbox: a decision model reads every customer email first and sorts it into four lanes - needs a person now, can wait, template could answer, sales pitch - packaged as an n8n workflow with a shadow mode.

    noulchoiceautomationdeveloper tools

    Artifactreproducible repo

    Checked

  135. E2 · artifact verifiedCommunity

    Triage customer messages with seven parallel questions

    A Portuguese customer-service demo where each client message makes one Jev call with seven typed questions in parallel - intent, urgency, sentiment and more - and a cost table prices it against routing everything to a large LLM instead.

    choicenoulscoreautomationdeveloper tools

    Artifactreproducible repo

    Checked

  136. E2 · artifact verifiedCommunity

    Triage GitHub issues with calibrated labels

    An issue classifier whose decision logic is ordinary Python gated on returned numbers: eight typed questions in one call, labels written only where confidence clears a threshold, and everything else escalated.

    choicescorenouldeveloper tools

    Artifactreproducible repo

    Checked

  137. E2 · artifact verifiedCommunity

    Triage industrial faults from live telemetry

    Fault triage for an industrial compressed-air unit with Jev: live telemetry in, classified alerts and tickets out, every judgment grounded in the unit's own service manual - read-only by design, with an eval harness that replays recorded decision cassettes.

    noulchoiceautomationdeveloper tools

    Artifactreproducible repo

    Checked

  138. E2 · artifact verifiedCommunity

    Triage Kubernetes incidents when rules run out

    A read-only Kubernetes controller that turns workload state and Events into stable incident decisions: deterministic rules decide the clear cases, and a Jev decision provider is called only for the ambiguous remainder.

    choicenoulautomationdeveloper tools

    Artifactreproducible repo

    Checked

  139. E2 · artifact verifiedCommunity

    Triage production alerts with four typed questions

    Open-source alert triage where each production alert gets one Jev call with four typed questions - actionable, severity, team and paging - and auditable code turns the probabilities into routing; Jev never pages anyone, it only judges.

    noulscorechoiceautomationdeveloper tools

    Artifactreproducible repo

    Checked

  140. E2 · artifact verifiedCommunity

    Triage security alerts as a UNIX filter

    A single Rust binary pipes JSONL security alerts through five typed Jev questions and emits validated dispositions that code, not the model, enforces.

    choicescorenoulsecuritydeveloper tools

    Reported200 ms per alert

    Checked

  141. E2 · artifact verifiedCommunity

    Triage YouTube comments by reply worthiness

    Paste a YouTube URL and every comment in the section is classified on four typed axes - then ranked so the ones actually worth a reply float to the top, for creators drowning in comment volume.

    noulchoicescoreautomationdeveloper tools

    Artifactreproducible repo

    Checked

  142. E2 · artifact verifiedCommunity

    Turn vulnerability texts into CVSS vectors

    Python scripts that ask Jev to select every CVSS metric from a vulnerability description, then compute the numeric score in code exactly per the FIRST specification - for CVSS 3.0, 3.1, and 4.0.

    choicenoulsecurity

    Artifactreproducible repo

    Checked

  143. E2 · artifact verifiedCommunity

    Use a decision model as an agent world model

    A bilingual study asking whether JEV can serve as a world model for LLM agents - predicting what happens next from typed state questions - benchmarked against generative LLMs on identical predictions, with run manifests and a phase-0 API check archived.

    noulchoiceresearchagents

    Artifactreproducible repo

    Checked

  144. E2 · artifact verifiedCommunity

    Watch novel characters take shape as you read

    A reader for public-domain Japanese novels where every paragraph is judged by Jev and each character accumulates a trait radar for the page in front of you, plus a running ranking of traits for the story so far.

    noulscoreresearch

    Artifactreproducible repo

    Checked

  145. E2 · artifact verifiedCommunity

    Watch your screen time with local and hosted judges

    Qualm is a screen-time app for Apple silicon Macs, local-first, built on two judges: Kev running on your machine and TypeSafe's hosted Jev - typed choices decide what counts as a distracting session without your activity leaving the machine unless you opt in.

    choicedeveloper toolshome

    Artifactreproducible repo

    Checked

  146. E2 · artifact verifiedCommunity

    Write a morning research note per topic

    For each topic you follow, jrp runs a pipeline: your questions steer the web search, Jev decides which findings count, and an LLM you choose writes the morning note - in English, Chinese or Japanese - with parallel stages and recorded claim-fidelity checks.

    choicenoulresearchsearch

    Artifactreproducible repo

    Checked

  147. E2 · artifact verifiedCommunity

    Write songs by choosing every note

    Jev cannot write a single note, so code lays out the bars and the legal notes with musical facts attached, and Jev chooses: mode, tempo, form, every chord, every note - one call per decision, every call replayable.

    choicemusic

    Reported0.005 USD

    Checked

Curated paths through the ledger

Browse by domain

13 clusters

Developer tools63 records

The largest cluster in the index: coding agents under supervision, code-review routing, citation checks, and typed tool gates. The recurring shape is an LLM-powered pipeline that would otherwise prompt a chat model at a decision point and instead asks Jev for a bounded judgment the surrounding code can branch on.

Most records here are reproducible repositories, and several ship as installable CLI or MCP integrations. What the cluster does not yet show is a large-scale production deployment with published reliability numbers — treat every reported metric as single-artifact evidence.

  1. E2Official

    Check LLM citations against source context

    An official cookbook combines exact quote matching with a Jev relation judgment to label citations verified, unsupported, contradicted, or fabricated.

    Artifact: official example · checked 2026-09-18
  2. E2Community

    Replace Claude Code compaction with Jev decisions

    A Claude Code plugin and npm library asks Jev which old tool calls and results still matter, then drops or truncates the stale ones, so /compact replaces the lossy built-in summary with the kept messages verbatim.

    Artifact: reproducible repo · checked 2026-09-18
  3. E2Community

    Answer finance questions from two judged passages

    A finance RAG benchmark where Jev picks the passages: the agentic baseline reads whole SEC filings - 82,908 tokens for 46 of 50 right - while Jev-judged retrieval reads about 840 tokens and answers all 50; benchmark code and write-up are public.

    Reported: 50 of 50 questions right at ~840 tokens per question · checked 2026-09-29
  4. E2Community

    Ask Jev inside SQL and get real types back

    A DuckDB extension where jev_choice, jev_score, and jev_noul are scalar functions over table rows: the criteria literal is both the set of permitted answers and the column's SQL type.

    Artifact: reproducible repo · checked 2026-09-20
  5. E2Community

    Call Jev from Flutter apps

    A Flutter plugin that brings typed Jev calls to Dart apps: question builders for Noul and Choice with option-count validation, the systemone wire protocol on the native endpoint, and a Python-side test harness for the plugin's bridge.

    Artifact: reproducible repo · checked 2026-09-30
  6. E2Community

    Catch contract drift hidden in OpenAPI prose

    OAS Sentinel compares two OpenAPI documents in two layers: deterministic checks find structural breaks, and Jev answers bounded semantic questions about changed prose - retries, ordering, pagination, error meaning - that schema diffs cannot see.

    Artifact: reproducible repo · checked 2026-09-23
  7. E2Community

    Check rewrites for lost ideas

    Lossless Rewrite closes the loop on AI shortening your report: your model rewrites, and Jev checks every protected idea survived - exact wording, meaning, or a reviewed checklist - then helps repair what went missing.

    Artifact: reproducible repo · checked 2026-09-24
  8. E2Community

    Classify CI failures as flaky or real

    A GitHub Action that classifies test failures as regression, flaky, environment, or unknown: deterministic signals first, one structured Jev choice second, and a local policy that never auto-reruns tests or masks failures.

    Artifact: reproducible repo · checked 2026-09-25
  9. E2Community

    Control Chrome by voice with typed commands

    A Chrome MV3 extension that turns speech into browser actions: Jev routes each spoken command to open a site, search, click a link, fill a form field or go back, with typed answers instead of parsed free text; ships with tests, CI and a side-panel command log.

    Artifact: reproducible repo · checked 2026-09-28
  10. E2Community

    Drive a browser with typed page judgments

    Give jev-browser a task and a URL: Jev picks one action per step from the page's clickable, typeable, and selectable elements and scores goal-met and stuck likelihood, while code owns budgets, recovery, and stop gates.

    Artifact: reproducible repo · checked 2026-09-20
  11. E2Community

    Drive the Aside browser with typed decisions

    bside pilots the Aside browser with Jev instead of a chat LLM: each tick answers which action, which element, and whether the goal is met, over an action schema the pilot cannot hallucinate outside of.

    Artifact: reproducible repo · checked 2026-09-23
  12. E2Community

    Expose Jev decisions to MCP clients

    JDE, the Jev Decision Engine, wraps Jev as an MCP server: any MCP client - Claude Code, Cursor, your own agent - gets typed decision tools, with a policy layer, a decision ledger, and recorded evals comparing the hosted jev-1.13.0 against local alternatives.

    Artifact: reproducible repo · checked 2026-10-03
  13. E2Community

    Expose typed Jev judgments to MCP agents

    An MCP server gives compatible agents tools for classification, scoring, checking, matching, screening, and custom typed Jev questions.

    Artifact: reproducible repo · checked 2026-09-18
  14. E2Community

    Feel Jev speed and cost in a 25-second game

    A tiny Japanese game with no send button: type IT buzzwords and each keystroke gets a Jev judgment that stretches a meter, so 25 seconds of play answers what curl never does - how fast and how cheap the model feels inside a real app.

    Artifact: reproducible repo · checked 2026-10-01
  15. E2Community

    Filter admin tables by vibe with typed decisions

    A Filament plugin for Laravel admin panels that filters tables by natural language instead of SQL: type "the customer is angry" and Jev's typed decisions drive the where-clauses - shipped as a Packagist package with a driver system, tests and a live demo.

    Artifact: reproducible repo · checked 2026-10-02
  16. E2Community

    Filter streams of text with typed batch judgments

    jevpipe pipes thousands of lines, files or records through one question and gets a typed judgment per item - grep-style filtering where Jev decides - shipped as a Rust binary on PyPI with an agent skill that teaches coding agents when to reach for it.

    Artifact: reproducible repo · checked 2026-09-28
  17. E2Community

    Find city permits from a free-text plan

    Describe what you want to do in San Francisco and a rules engine works out which permits you need - Jev answers each rules question as a typed choice, falls back to asking you when unknown, and summarizes fees and deadlines per permit.

    Artifact: reproducible repo · checked 2026-10-03
  18. E2Community

    Fit confidence thresholds on your own data

    jevcal measures a typed decision model on private labeled data, fits per-question thresholds to a target accuracy, and fails CI when a model update drifts.

    Artifact: reproducible repo · checked 2026-09-18
  19. E2Community

    Gate agent commits with plain-English rules

    tenet makes agents fix rule violations before you ever see the diff: rules live in a YAML file in plain language, and Jev answers each one with a calibrated probability that becomes a pass-or-fail cutoff.

    Artifact: reproducible repo · checked 2026-09-23
  20. E2Community

    Gate LLM-generated requirements with 22 questions

    A requirements quality gate: one batched call asks 22 atomic questions across five MECE facets about an AI-generated requirement, and thresholds route it pass, human review, or reject - never rewriting, only judging.

    Reported: 0.0022 USD · checked 2026-09-26
  21. E2Community

    Gate paid-API spend on relevance and budget

    Four proof-of-concepts wiring Jev in front of a pay-per-call API marketplace: relevance below 0.7 confidence skips the paid call entirely, a typed choice routes to exactly one endpoint, and a second independent gate enforces per-team budget caps.

    Artifact: reproducible repo · checked 2026-09-27
  22. E2Community

    Gate risky coding-agent tool calls

    A Pi extension asks Jev to flag destructive, exfiltrating, or out-of-scope tool calls and to classify failures in command output.

    Reported: 203 requests · checked 2026-09-18
  23. E2Community

    Give Jev eyes over cameras and streams

    A bridge connecting images, video streams and RGB-D cameras to Jev's judgment engine: identify what matters in a frame, estimate risk, score a situation or judge many visible objects at once - typed answers over the visual world, 46 stars in its first day.

    Artifact: reproducible repo · checked 2026-10-01
  24. E2Community

    Grep for what code does, not what it's called

    jgrep answers semantic queries like catches an error and silently ignores it over a whole source tree in about two seconds for a cent: one typed yes-or-no judgment per code chunk, sixteen chunks per request, no index.

    Artifact: reproducible repo · checked 2026-09-24
  25. E2Community

    Guard .NET LLM apps with typed checks

    Kassad brings calibrated guardrails to .NET: every prompt, completion, tool call and citation passes narrow typed checks answered by Jev, batched one round trip per stage, thresholded in code into Allow, Flag, Review, or Block.

    Artifact: reproducible repo · checked 2026-09-24
  26. E2Community

    Judge Azure Foundry agents with typed evaluators

    Four drop-in evaluators for Azure AI Foundry, rebuilt on Jev: Intent Resolution, Task Adherence, Tool Call Accuracy, Groundedness - each metric becomes small typed questions answered in one call, combined into a 1-5 score listing its checks.

    Artifact: reproducible repo · checked 2026-10-03
  27. E2Community

    Judge player persuasion by each character's values

    An npm library that lets game characters argue back: give a name, a persona and a goal, pass what the player typed, and Jev judges whether that character - with those values - was convinced; the same line can win a greedy merchant and offend an honest guard.

    Artifact: reproducible repo · checked 2026-09-29
  28. E2Community

    Jump to the video moment that answers you

    A Chrome extension for YouTube: ask the video a question in plain text, and Jev's typed judgments locate the moment that answers it - the player seeks straight there instead of you scrubbing.

    Artifact: reproducible repo · checked 2026-10-07
  29. E2Community

    Lint agent writes in 300 ms with team rules

    A fuzzy linter that watches a coding agent write and speaks up 0.3 seconds later: Jev checks the file against your team's rules - race conditions in effects, missing cleanup, house style - so the agent fixes them before any human reviews the code.

    Artifact: reproducible repo · checked 2026-10-04
  30. E2Community

    Mine repeating patterns from noisy sequences

    Finding and extracting repeating patterns from noisy sequences with Jev: instead of a hand-tuned distance metric, typed questions decide what counts as the same pattern, and matches come back with probabilities instead of thresholds you guess.

    Artifact: reproducible repo · checked 2026-10-03
  31. E2Community

    Morph one input into the right UI with typed intent

    One text box that becomes the right UI as you type - an event card, checklist, timer, color picker, bill splitter or poll - with Jev Nouls deciding what the input means and code rendering the component; live demo included.

    Artifact: reproducible repo · checked 2026-09-29
  32. E2Community

    Pick the right viewer for any file

    Drop a file or paste a URL and get a floating window with the right tool: Jev decides and composes a viewer from the registry, and when the format is unknown, Haiku invents a spec for a mini-app on the spot.

    Artifact: working demo · checked 2026-09-24
  33. E2Community

    Plug a fail-closed decision layer into your agent

    A pluggable decision layer for the ego agent: System One (Jev) by default, swappable to local or other OpenAI-compatible backends, fail-closed guardrails - and the README's whole argument is that it is measured: 16 suites, 429 checks, rerun in full.

    Artifact: reproducible repo · checked 2026-10-03
  34. E2Community

    Predict the window you want next on macOS

    A macOS menu-bar switcher asks Jev which of the ten most recent apps you intend on a hotkey press and falls back to the last-used app on any failure.

    Artifact: reproducible repo · checked 2026-09-18
  35. E2Community

    Prune stale coding-agent tool history

    A Pi extension uses Jev to decide which old tool calls and results still matter while keeping conversation text verbatim.

    Artifact: reproducible repo · checked 2026-09-18
  36. E2Community

    Rank zsh history suggestions while typing

    A zsh plugin asks Jev which recent command you are completing and shows the best match with its probability, while code owns gating and acceptance.

    Reported: 0.8 seconds per request · checked 2026-09-18
  37. E2Community

    Rate pending shell commands on a color-coded safety rubric

    The kamchatka terminal agent asks Jev to place each pending shell command on a three-level safety rubric - reads and reports, changes something reversibly, destroys or sends something out - drawn green, yellow, or red beside the permission prompt.

    Artifact: reproducible repo · checked 2026-09-18
  38. E2Community

    Recover OCR marker formatting with typed choices

    OCR flattens superscripts: a footnote star, an endnote number and a unit power land in the stream as look-alike tokens. A regex over-finds the suspects, then Jev classifies each - footnote, citation or unit - as a typed choice so formatting can be restored.

    Artifact: reproducible repo · checked 2026-10-03
  39. E2Community

    Rerank Postgres product search with two questions

    Product search inside PostgreSQL - typos, barcodes, typeahead, facet counts, all in SQL - with an optional second stage asking Jev two questions per search to rerank the shortlist; the README is itself the report, three failed versions included.

    Artifact: reproducible repo · checked 2026-10-08
  40. E2Community

    Resolve merge conflicts by choosing candidates

    Hunkpick resolves git conflicts by computing every plausible resolution itself - ours, theirs, union, line merge, token merge - discarding the ones that fail to parse, and asking Jev only to choose.

    Artifact: reproducible repo · checked 2026-09-24
  41. E2Community

    Restyle websites with typed catalog choices

    A Chrome MV3 extension that restyles any site from a plain-English prompt: one Jev call picks palette, fonts, spacing and intent from a fixed catalog, and deterministic code compiles role-stamped CSS - no LLM ever writes CSS that can break.

    Artifact: reproducible repo · checked 2026-09-29
  42. E2Community

    Review commits and copy with typed questions

    nudgement reviews commit messages, code, comments, tests and UI copy before you commit: exact checks plus focused Jev questions catch what formatters cannot - a message that misrepresents the diff, or a test that passes even when behavior broke.

    Artifact: reproducible repo · checked 2026-09-30
  43. E2Community

    Review pull requests against Five Lines of Code

    A single binary that reviews a diff against the ten refactoring rules of Clausen's Five Lines of Code: the countable rules run on a real parser, and Jev answers the judgment rules as typed questions whose probabilities set the bar.

    Artifact: reproducible repo · checked 2026-09-23
  44. E2Community

    Review pull requests with calibrated bug risk

    A single-file local pull-request reviewer: deterministic code does the plumbing while Jev judges each hunk with a real-issue Noul, scores severity, and returns a PR-level risk with a needs-human probability - no agent loop, no prompts to tune.

    Artifact: reproducible repo · checked 2026-09-28
  45. E2Community

    Route focused code-review investigations

    A staged review workflow uses Jev to identify risky areas, select evidence, classify mechanisms, score severity, and route follow-up checks.

    Artifact: reproducible repo · checked 2026-09-18
  46. E2Community

    Route messy music requests before the LLM

    A DJ chatbot's pre-LLM router, built as a learning study: free-form /play requests - English, Spanish, typos, pasted lyrics, emojis - classified by Jev in one call (~500 ms, ~$0.000036) into a structured hint that tells the expensive LLM what it is handling.

    Artifact: reproducible repo · checked 2026-10-07
  47. E2Community

    Run an e-commerce store on typed decisions alone

    Jev Mart is a Japanese e-commerce demo where every operational decision - listings, inquiries, reviews, triage - is a typed Jev call and an LLM is never invoked; it ships with a twelve-chapter lecture set that teaches the pattern.

    Artifact: reproducible repo · checked 2026-09-30
  48. E2Community

    Run Jev judgments from the shell before an agent acts

    jev-axi is a CLI that puts a half-second Jev opinion in front of every command an agent runs: pick, rate, check, rank, triage, and guard, plus a PreToolUse hook that blocks risky Bash calls before they execute.

    Reported: 375 ms · checked 2026-09-20
  49. E2Community

    Run typed judgments as eleven shell commands

    A CLI named jev turns Jev's typed questions into pipable commands - verify, screen, classify, extract, find, rerank, match, route, ask, compact, batch - with keychain auth and exit codes a script can gate on.

    Artifact: reproducible repo · checked 2026-09-20
  50. E2Community

    Score code quality for coding agents

    Supercov asks Jev yes-or-no questions about every source file, does the arithmetic in code, and turns weak spots and coverage gaps into agent tasks.

    Reported: 0.01 USD per megabyte of source · checked 2026-09-18
  51. E2Community

    Select which tests a change needs

    A sibling GitHub Action that decides which allowlisted test groups a pull request needs and which can safely be skipped: the group list stays in your config, Jev picks among it, and deterministic rules a model cannot bypass have the final word.

    Artifact: reproducible repo · checked 2026-09-25
  52. E2Community

    Sort Gmail and grade replies with typed choices

    A light terminal Gmail client for Omarchy Linux: Jev files incoming mail into Gmail labels only above a confidence threshold, and grades your reply as you write - clarity, tone, length, next step, and which questions you haven't answered yet.

    Artifact: reproducible repo · checked 2026-09-30
  53. E2Community

    Supervise a coding agent while it works

    Foreman runs an independent observation loop that asks Jev whether a coding worker is progressing, stuck, complete, or ready for verification.

    Artifact: reproducible repo · checked 2026-09-18
  54. E2Community

    Tag a calibre library with typed suggestions

    A calibre plugin that suggests subject and genre tags for your books with Jev: suggestions arrive for review before anything is applied, existing tags are preserved, and the plugin is independent GPL software with a live demo and production notes.

    Artifact: reproducible repo · checked 2026-10-01
  55. E2Community

    Triage a store inbox into four lanes

    A quiet clerk for a small online store's inbox: a decision model reads every customer email first and sorts it into four lanes - needs a person now, can wait, template could answer, sales pitch - packaged as an n8n workflow with a shadow mode.

    Artifact: reproducible repo · checked 2026-10-08
  56. E2Community

    Triage customer messages with seven parallel questions

    A Portuguese customer-service demo where each client message makes one Jev call with seven typed questions in parallel - intent, urgency, sentiment and more - and a cost table prices it against routing everything to a large LLM instead.

    Artifact: reproducible repo · checked 2026-10-01
  57. E2Community

    Triage GitHub issues with calibrated labels

    An issue classifier whose decision logic is ordinary Python gated on returned numbers: eight typed questions in one call, labels written only where confidence clears a threshold, and everything else escalated.

    Artifact: reproducible repo · checked 2026-09-22
  58. E2Community

    Triage industrial faults from live telemetry

    Fault triage for an industrial compressed-air unit with Jev: live telemetry in, classified alerts and tickets out, every judgment grounded in the unit's own service manual - read-only by design, with an eval harness that replays recorded decision cassettes.

    Artifact: reproducible repo · checked 2026-10-03
  59. E2Community

    Triage Kubernetes incidents when rules run out

    A read-only Kubernetes controller that turns workload state and Events into stable incident decisions: deterministic rules decide the clear cases, and a Jev decision provider is called only for the ambiguous remainder.

    Artifact: reproducible repo · checked 2026-09-23
  60. E2Community

    Triage production alerts with four typed questions

    Open-source alert triage where each production alert gets one Jev call with four typed questions - actionable, severity, team and paging - and auditable code turns the probabilities into routing; Jev never pages anyone, it only judges.

    Artifact: reproducible repo · checked 2026-09-29
  61. E2Community

    Triage security alerts as a UNIX filter

    A single Rust binary pipes JSONL security alerts through five typed Jev questions and emits validated dispositions that code, not the model, enforces.

    Reported: 200 ms per alert · checked 2026-09-18
  62. E2Community

    Triage YouTube comments by reply worthiness

    Paste a YouTube URL and every comment in the section is classified on four typed axes - then ranked so the ones actually worth a reply float to the top, for creators drowning in comment volume.

    Artifact: reproducible repo · checked 2026-10-07
  63. E2Community

    Watch your screen time with local and hosted judges

    Qualm is a screen-time app for Apple silicon Macs, local-first, built on two judges: Kev running on your machine and TypeSafe's hosted Jev - typed choices decide what counts as a distracting session without your activity leaving the machine unless you opt in.

    Artifact: reproducible repo · checked 2026-10-03

Automation37 records

Browser, desktop, and smart-home control driven from observed state: the visible DOM becomes a table of allowed operations, Home Assistant state becomes automation signals, and Jev picks the operation while deterministic code executes it. This is the clearest demonstration of the split the whole index is about — model chooses, code acts.

Because these records touch real interfaces, the interesting evidence is in the failure boundaries: frames, occlusion, timing windows, and devices that change state mid-command. The limitations field on each record is where authors document what their automation still cannot do.

  1. E2Official

    Batch a regulatory briefing into one call

    An official cookbook asks 13 regulatory questions over a pinned GDPR article in one Jev request and compares that batch with 13 separate requests.

    Reported: 12.2 × cheaper · checked 2026-09-18
  2. E2Community

    Drive macOS from structured screen state

    A computer-use loop combines OCR and accessibility data, then asks Jev which bounded action should move the Mac toward a plain-English goal.

    Reported: 0.0002 USD per step · checked 2026-09-18
  3. E2Official

    Route smart-home assistant requests

    TypeSafe's smart-home demo evaluates a request against many typed questions in parallel, then lets code use only the answers relevant to that request.

    Artifact: official example · checked 2026-09-18
  4. E2Community

    Search Google Flights in seconds

    A browser agent turns the visible DOM into an indexed action space and uses Jev to choose the next operation and compatible target.

    Reported: 7.073 seconds · checked 2026-09-18
  5. E2Community

    Turn Home Assistant state into automation signals

    A Home Assistant integration exposes Jev probabilities, choices, and scores as entities and action responses that automations can use.

    Artifact: reproducible repo · checked 2026-09-18
  6. E2Community

    Choose a market quote every Monad block

    A trading loop reads the Kuru MON-USDC order book and asks Jev for a buy-or-sell judgment before code places a post-only limit order.

    Artifact: reproducible repo · checked 2026-09-18
  7. E2Community

    Control Chrome by voice with typed commands

    A Chrome MV3 extension that turns speech into browser actions: Jev routes each spoken command to open a site, search, click a link, fill a form field or go back, with typed answers instead of parsed free text; ships with tests, CI and a side-panel command log.

    Artifact: reproducible repo · checked 2026-09-28
  8. E2Community

    Cover YouTube spoilers with one Noul judgment per comment

    A Chrome extension covers every YouTube comment the moment it appears, asks Jev one Noul question per comment, and keeps it covered whenever the probability says it discloses a concrete plot event.

    Artifact: reproducible repo · checked 2026-09-22
  9. E2Community

    Decide Home Assistant commands with typed answers

    A Home Assistant conversation agent - installable via HACS - that decides with a TypeSafe System One model instead of an LLM: your spoken or typed commands become typed choices and Nouls that drive devices, with metered costs documented.

    Artifact: reproducible repo · checked 2026-09-30
  10. E2Community

    Drive an Android phone from the A11Y tree

    An Android automation agent with a two-tier brain: Jev decides fast from the accessibility tree, a vision agent takes over only when the structural view is ambiguous, and ADB executes - cheap judge on the hot path, expensive one on the exceptions.

    Artifact: reproducible repo · checked 2026-10-06
  11. E2Community

    Drive Android from semantic UI state

    A proof-of-concept Android loop stabilizes the screen, builds a short list of valid actions, and lets Jev pick one while code executes it.

    Artifact: reproducible repo · checked 2026-09-18
  12. E2Community

    Fade low-value web text with sentence-level Nouls

    Osso fades the parts of a page that are not what the reader came for: each sentence of the main text gets a Jev probability, workspace sites are covered by default, and password fields, account pages and reviews are left alone by construction.

    Artifact: reproducible repo · checked 2026-09-29
  13. E2Community

    File local documents with typed safety lanes

    A local-first document filer: text is extracted locally, Jev decides category, confidentiality and prompt-injection risk as typed Choices, and low-confidence or suspicious files land in review lanes - never overwritten, with audit preview and undo.

    Artifact: reproducible repo · checked 2026-09-30
  14. E2Community

    Filter news and YouTube feeds by interest

    A single-file Python portal that filters news and YouTube feeds by interests you describe in plain English: the official typesafe-sdk scores every item, routine business hides separately, and the result is one static page with News and YouTube tabs.

    Artifact: reproducible repo · checked 2026-10-04
  15. E2Community

    Filter your X timeline with five typed signals

    An open-source Chrome extension asks Jev to judge each visible X post for relevance, substance, practical value, promotion, and engagement bait, then dims, collapses, or hides it under weights the reader owns.

    Artifact: reproducible repo · checked 2026-09-20
  16. E2Community

    Gate paid-API spend on relevance and budget

    Four proof-of-concepts wiring Jev in front of a pay-per-call API marketplace: relevance below 0.7 confidence skips the paid call entirely, a typed choice routes to exactly one endpoint, and a second independent gate enforces per-team budget caps.

    Artifact: reproducible repo · checked 2026-09-27
  17. E2Community

    Give a simulated person fast feelings

    Mina, a simulated 34-year-old librarian, runs on two systems: Jev reads her body, senses and clock every second and accumulates feelings; only when a feeling crosses its line does an LLM stop and think.

    Reported: 0.53 USD · checked 2026-09-24
  18. E2Community

    Moderate community messages with typed probabilities

    jevmod scores every community message for spam, scam, harassment, NSFW, self-harm, doxxing, off-topic and custom plain-English rules; operators set thresholds, decisions log their numbers, and bots ship for Discord, Twitch, YouTube and Reddit.

    Reported: 0.04 USD per 1,000 messages · checked 2026-09-28
  19. E2Community

    Moderate SAP Commerce reviews with four questions

    An SAP Commerce extension where Jev answers four yes-or-no questions per product review - abusive, spam, personal data, on-topic - and code turns the probabilities into approve, reject, or pending, with a dry-run mode that judges against human decisions first.

    Artifact: reproducible repo · checked 2026-09-24
  20. E2Community

    Plan day trips with a typed checker

    Jev Trip is an explainable day-trip planner where the LLM plans ahead and Jev chooses and checks: scope choices keep one day in one city, per-place choices rank every candidate with probabilities, and code owns routes, times and validation.

    Artifact: reproducible repo · checked 2026-09-28
  21. E2Community

    Predict the window you want next on macOS

    A macOS menu-bar switcher asks Jev which of the ten most recent apps you intend on a hotkey press and falls back to the last-used app on any failure.

    Artifact: reproducible repo · checked 2026-09-18
  22. E2Community

    Propose support decisions a guard can check

    A Kosovo electronics shop’s support agent reads a unified Albanian and English inbox and lands every message on auto-resolve, verification, or escalate - with Jev only proposing toward caution and deterministic code owning facts, access, and the final call.

    Artifact: reproducible repo · checked 2026-09-25
  23. E2Community

    Rate LinkedIn posts for AI slop and bait

    A browser extension where Jev answers 17 typed questions about every LinkedIn post - thirteen AI-tell questions plus four bait questions in the same request - and fixed weights turn the answers into a badge and a bait chip.

    Reported: 67 % · checked 2026-09-20
  24. E2Community

    Remove ad-like DOM elements in Chrome

    A Chrome extension finds ad-shaped DOM candidates and asks Jev whether each candidate is a paid advertisement before code removes it.

    Artifact: reproducible repo · checked 2026-09-18
  25. E2Community

    Research social media through real browser operations

    Jev Social pairs the decision model with socai, a CLI that drives your real Chrome across Instagram, TikTok, and LinkedIn: Jev chooses each next read-only operation - search, open a post or profile, read comments - and socai executes it.

    Artifact: reproducible repo · checked 2026-09-25
  26. E2Community

    Route a multilingual helpdesk with typed triage

    FastGate fronts an English/Uzbek/Russian university helpdesk with four narrow Jev judgments per message plus per-passage grounding, routes deterministically in code, and ships an independent benchmark of Jev on low-resource languages.

    Artifact: reproducible repo · checked 2026-09-27
  27. E2Community

    Screen Spain's official gazette every morning

    A daily pipeline reads the BOE with one Jev call per provision, scoring impact, tagging topics, and selecting an original paragraph as the summary.

    Reported: 0.01 USD per daily gazette · checked 2026-09-18
  28. E2Community

    Shadow a live CRM with typed decisions

    A working example of a shadow decision layer under a production CRM: one call returns intention, business division, human-handover probability, urgency, and lead quality - decided and logged, never sent.

    Artifact: reproducible repo · checked 2026-09-23
  29. E2Community

    Silence ad noise on Android, failing open

    An Android noise gate for notifications and SMS that asks Jev whether each message is an ad - and suppresses only what is explicitly flagged. Every uncertain path resolves to allow, because swallowing a verification code costs more than leaking an ad.

    Artifact: reproducible repo · checked 2026-09-23
  30. E2Community

    Sort Gmail and grade replies with typed choices

    A light terminal Gmail client for Omarchy Linux: Jev files incoming mail into Gmail labels only above a confidence threshold, and grades your reply as you write - clarity, tone, length, next step, and which questions you haven't answered yet.

    Artifact: reproducible repo · checked 2026-09-30
  31. E2Community

    Steer hydroponic crops with typed judges

    A Home Assistant add-on runs a crop-steering irrigation engine with Jev judging what fixed rules got wrong - ramp timing, probe trust, shot landing, EC moves, alerts - inside code-enforced physical limits; the README replays real Jev answers from 27 Sep 2026.

    Reported: 0.92 probability · checked 2026-09-27
  32. E2Community

    Triage a store inbox into four lanes

    A quiet clerk for a small online store's inbox: a decision model reads every customer email first and sorts it into four lanes - needs a person now, can wait, template could answer, sales pitch - packaged as an n8n workflow with a shadow mode.

    Artifact: reproducible repo · checked 2026-10-08
  33. E2Community

    Triage customer messages with seven parallel questions

    A Portuguese customer-service demo where each client message makes one Jev call with seven typed questions in parallel - intent, urgency, sentiment and more - and a cost table prices it against routing everything to a large LLM instead.

    Artifact: reproducible repo · checked 2026-10-01
  34. E2Community

    Triage industrial faults from live telemetry

    Fault triage for an industrial compressed-air unit with Jev: live telemetry in, classified alerts and tickets out, every judgment grounded in the unit's own service manual - read-only by design, with an eval harness that replays recorded decision cassettes.

    Artifact: reproducible repo · checked 2026-10-03
  35. E2Community

    Triage Kubernetes incidents when rules run out

    A read-only Kubernetes controller that turns workload state and Events into stable incident decisions: deterministic rules decide the clear cases, and a Jev decision provider is called only for the ambiguous remainder.

    Artifact: reproducible repo · checked 2026-09-23
  36. E2Community

    Triage production alerts with four typed questions

    Open-source alert triage where each production alert gets one Jev call with four typed questions - actionable, severity, team and paging - and auditable code turns the probabilities into routing; Jev never pages anyone, it only judges.

    Artifact: reproducible repo · checked 2026-09-29
  37. E2Community

    Triage YouTube comments by reply worthiness

    Paste a YouTube URL and every comment in the section is classified on four typed axes - then ranked so the ones actually worth a reply float to the top, for creators drowning in comment volume.

    Artifact: reproducible repo · checked 2026-10-07

Research and measurement35 records

Measurement-first projects: routing strategies scored on labelled data, novels scored passage by passage, clinical reviews extracted as verbatim quotes, and chat interfaces that emit decisions instead of free text. These records are where the directory's evidence habits matter most, because the authors are usually testing a hypothesis rather than shipping a product.

Read the metric conditions closely in this cluster. Several studies report numbers from one dataset, one run, or one author-supplied baseline, and the record says so explicitly rather than averaging it away.

  1. E2Community

    Answer finance questions from two judged passages

    A finance RAG benchmark where Jev picks the passages: the agentic baseline reads whole SEC filings - 82,908 tokens for 46 of 50 right - while Jev-judged retrieval reads about 840 tokens and answers all 50; benchmark code and write-up are public.

    Reported: 50 of 50 questions right at ~840 tokens per question · checked 2026-09-29
  2. E2Community

    Answer with one of five words

    A Turkish chat toy where whatever you type is answered by one of five fixed phrases - Jev picks which word, and two more answers decide the punctuation, in a single call.

    Artifact: working demo · checked 2026-09-22
  3. E2Community

    Benchmark Wordle solvers with Jev word priors

    A reproducible experiment beyond the 'optimal' Wordle solver: Jev scores how answer-like each of 12,972 accepted words is, the prior feeds entropy search over 1,925 days of NYT answers, and every claim is archived with pinned inputs and a reproduce script.

    Reported: 1925 days of NYT answers in the benchmark · checked 2026-10-02
  4. E2Community

    Chat without generating any free text

    An experiment that makes a model which cannot write text answer anyway: for every word of the reply Jev picks 1 of 254 meaning-based word groups, then the word inside that group - and the README reports exactly where that stops working.

    Artifact: working demo · checked 2026-09-20
  5. E2Community

    Check rewrites for lost ideas

    Lossless Rewrite closes the loop on AI shortening your report: your model rewrites, and Jev checks every protected idea survived - exact wording, meaning, or a reviewed checklist - then helps repair what went missing.

    Artifact: reproducible repo · checked 2026-09-24
  6. E2Community

    Classify SVGs by weighting typed responses

    A research task treating Jev's answer distributions as classifier features: many small typed questions about an SVG, responses weighted and combined until the signal classifies the image - a study of whether decision calls can stand in for maths on pixels.

    Artifact: reproducible repo · checked 2026-10-03
  7. E2Community

    Compact conversation context by calibrated judgment

    jevtrim is a comparative analysis of context compaction driven by calibrated judgments instead of summarization: Jev scores every chunk for relevance, ordinary Python keeps what fits the token budget, and the result is auditable and replayable offline.

    Artifact: reproducible repo · checked 2026-09-24
  8. E2Community

    Curate a daily science edition with typed labels

    Pipette reads arXiv, bioRxiv, medRxiv and 58 journals each morning and publishes a short, diverse daily edition: Jev labels and ranks, quoted sentences are the authors' own abstract lines, method and probabilities are public, and output is CC0 open data.

    Artifact: reproducible repo · checked 2026-09-28
  9. E2Community

    Evaluate Jev on the SNIPS NLU benchmark

    An evaluation of Jev on the SNIPS natural-language-understanding benchmark - intent detection and slot filling - asking how far a model that never generates text gets on a task normally solved by a trained tagger, using label names alone.

    Artifact: reproducible repo · checked 2026-10-05
  10. E2Community

    Extract clinical review data as verbatim quotes

    A browser app for systematic-review extraction: Jev never writes the answer - it points at line ids in trial reports and supplements, and code copies the quote out with its file, page, row, or slide, highlighted where it sits.

    Artifact: working demo · checked 2026-09-20
  11. E2Community

    Fact-check a link claim by claim against its own sources

    A live link-checker extracts each claim from a submitted page, asks Jev whether the cited excerpts support, contradict, or fail to establish it, and reports REAL or FAKE only when enough evidence agrees.

    Artifact: reproducible repo · checked 2026-09-27
  12. E2Community

    Fade low-value web text with sentence-level Nouls

    Osso fades the parts of a page that are not what the reader came for: each sentence of the main text gets a Jev probability, workspace sites are covered by default, and password fields, account pages and reviews are left alone by construction.

    Artifact: reproducible repo · checked 2026-09-29
  13. E2Community

    Give Jev eyes over cameras and streams

    A bridge connecting images, video streams and RGB-D cameras to Jev's judgment engine: identify what matters in a frame, estimate risk, score a situation or judge many visible objects at once - typed answers over the visual world, 46 stars in its first day.

    Artifact: reproducible repo · checked 2026-10-01
  14. E2Community

    Ground agent answers in cited document blocks

    TraceDocs structures documents into source-linked blocks; Jev judges each with four Nouls - relevant, evidence, contradicts-premise, prompt-injection - returning a cited evidence set with a trace; the LLM writes from evidence and refuses when none exists.

    Artifact: reproducible repo · checked 2026-09-28
  15. E2Community

    Interview your AI worldview with typed questions

    Doom or Bloom maps where you stand between AI doom and bloom: a dynamic interview where the engine picks the next curated question by where your answers are thinnest, with Jev interpreting answers and scoring candidate follow-ups.

    Artifact: working demo · checked 2026-09-24
  16. E2Community

    Judge Azure Foundry agents with typed evaluators

    Four drop-in evaluators for Azure AI Foundry, rebuilt on Jev: Intent Resolution, Task Adherence, Tool Call Accuracy, Groundedness - each metric becomes small typed questions answered in one call, combined into a 1-5 score listing its checks.

    Artifact: reproducible repo · checked 2026-10-03
  17. E2Community

    Judge every new arXiv paper each morning

    Paper Radar reads all of arXiv so you read the few that matter: every new paper is judged against plain-English interests with calibrated per-interest probabilities - about six cents a day for everything, no pre-filtering.

    Reported: 0.002 USD · checked 2026-09-24
  18. E2Community

    Make Jev talk one word at a time

    A demonstration pushing the judgment-only model past its envelope: Jev 'writes' by answering which-word-comes-next in 250-word batches - each option shown as the whole reply so far plus the word - top-3 shortlist, final pick, until sentence end.

    Artifact: reproducible repo · checked 2026-10-08
  19. E2Community

    Measure confidence-gated routing on labelled data

    An independent study routes Jev confidence into a larger model on two labelled datasets and shows the winning settings do not transfer between them.

    Reported: 80.2 % · checked 2026-09-18
  20. E2Community

    Measure Jev against LLMs on real workflows, weekly

    A weekly measured series pitting Jev against frontier LLMs on the same real workflow steps: week one routed inbound leads - Jev 90 percent correct at 366 milliseconds and four cents per thousand, against Sonnet 5's 78 percent at 2.6 seconds and three dollars.

    Reported: 90 % · checked 2026-09-25
  21. E2Community

    Mine repeating patterns from noisy sequences

    Finding and extracting repeating patterns from noisy sequences with Jev: instead of a hand-tuned distance metric, typed questions decide what counts as the same pattern, and matches come back with probabilities instead of thresholds you guess.

    Artifact: reproducible repo · checked 2026-10-03
  22. E2Community

    Paint alongside Jev on a shared emoji canvas

    A shared 1,000-by-1,000 emoji canvas where humans place strokes and Jev paints with them: after each stroke one typed call picks a contextually relevant emoji and where to put it, and a yes/no decides whether your stroke was finished.

    Artifact: reproducible repo · checked 2026-09-29
  23. E2Community

    Read customer-inquiry emotion in six questions

    A Japanese-language experiment scoring 100 labeled customer inquiries with the official TypeSafe SDK: one request per inquiry answers six questions at once - sentiment, emotion, anger intensity, urgency, churn risk, and sarcasm.

    Artifact: reproducible repo · checked 2026-09-24
  24. E2Community

    Recover OCR marker formatting with typed choices

    OCR flattens superscripts: a footnote star, an endnote number and a unit power land in the stream as look-alike tokens. A regex over-finds the suspects, then Jev classifies each - footnote, citation or unit - as a typed choice so formatting can be restored.

    Artifact: reproducible repo · checked 2026-10-03
  25. E2Community

    Route a multilingual helpdesk with typed triage

    FastGate fronts an English/Uzbek/Russian university helpdesk with four narrow Jev judgments per message plus per-passage grounding, routes deterministically in code, and ships an independent benchmark of Jev on low-resource languages.

    Artifact: reproducible repo · checked 2026-09-27
  26. E2Community

    Run an e-commerce store on typed decisions alone

    Jev Mart is a Japanese e-commerce demo where every operational decision - listings, inquiries, reviews, triage - is a typed Jev call and an LLM is never invoked; it ships with a twelve-chapter lecture set that teaches the pattern.

    Artifact: reproducible repo · checked 2026-09-30
  27. E2Community

    Score a whole novel by emotion, one passage at a time

    Book Aurora sends every ~90-word passage of Frankenstein to Jev as ten parallel score questions - nine emotions plus overall intensity - and draws each answer as one feathered row of a full-book aurora strip.

    Reported: 0.0337 USD · checked 2026-09-20
  28. E2Community

    Score video statements by viewpoint

    jevmeter turns a video into a BS meter: every sentence scored and every dodge flagged under a preset viewpoint - debates, earnings calls, podcasts, pitches - rendered as a 16:9 edit you can post.

    Reported: 0.05 USD · checked 2026-09-23
  29. E2Community

    Speak in sentences compiled from decisions

    JevSpeak holds a conversation without any generative model in the loop: Jev answers about thirteen parallel questions per turn, the decisions normalize into a semantic IR, and a deterministic compiler writes the sentence.

    Artifact: working demo · checked 2026-09-20
  30. E2Community

    Steer a simulated fly brain with typed decisions

    Spikecast puts Jev at the controls of a fly's simulated brain: the insect walks a road, and every turn, stop or dash is a typed decision with the synapse memory and neural firing visualised beside it - with docs separating what is real from what is modelled.

    Artifact: reproducible repo · checked 2026-10-02
  31. E2Community

    Steer matter paths with live decisions

    An interactive installation - explore a flooded observatory where glowing matter builds paths from your movement, gaze and actions, with real-time Jev decisions deciding how the world responds, shipped as a playable web app with live-decision tests.

    Artifact: reproducible repo · checked 2026-10-08
  32. E2Community

    Tag a calibre library with typed suggestions

    A calibre plugin that suggests subject and genre tags for your books with Jev: suggestions arrive for review before anything is applied, existing tags are preserved, and the plugin is independent GPL software with a live demo and production notes.

    Artifact: reproducible repo · checked 2026-10-01
  33. E2Community

    Use a decision model as an agent world model

    A bilingual study asking whether JEV can serve as a world model for LLM agents - predicting what happens next from typed state questions - benchmarked against generative LLMs on identical predictions, with run manifests and a phase-0 API check archived.

    Artifact: reproducible repo · checked 2026-10-08
  34. E2Community

    Watch novel characters take shape as you read

    A reader for public-domain Japanese novels where every paragraph is judged by Jev and each character accumulates a trait radar for the page in front of you, plus a running ranking of traits for the story so far.

    Artifact: reproducible repo · checked 2026-09-29
  35. E2Community

    Write a morning research note per topic

    For each topic you follow, jrp runs a pipeline: your questions steer the web search, Jev decides which findings count, and an LLM you choose writes the morning note - in English, Chinese or Japanese - with parallel stages and recorded claim-fidelity checks.

    Artifact: reproducible repo · checked 2026-10-04

Agent control28 records

Supervision records: watching a coding agent work, gating risky tool calls, pruning stale tool history, and exposing typed judgments to MCP clients. The cluster treats the agent as the system under control and Jev as the referee that decides what the agent may do next.

All of these are community-published repositories verified at the artifact level. None of them claim production uptime, which is why each record's verification note matters more than its demo video.

  1. E2Official

    Check LLM citations against source context

    An official cookbook combines exact quote matching with a Jev relation judgment to label citations verified, unsupported, contradicted, or fabricated.

    Artifact: official example · checked 2026-09-18
  2. E2Community

    Drive macOS from structured screen state

    A computer-use loop combines OCR and accessibility data, then asks Jev which bounded action should move the Mac toward a plain-English goal.

    Reported: 0.0002 USD per step · checked 2026-09-18
  3. E2Community

    Replace Claude Code compaction with Jev decisions

    A Claude Code plugin and npm library asks Jev which old tool calls and results still matter, then drops or truncates the stale ones, so /compact replaces the lossy built-in summary with the kept messages verbatim.

    Artifact: reproducible repo · checked 2026-09-18
  4. E2Community

    Search Google Flights in seconds

    A browser agent turns the visible DOM into an indexed action space and uses Jev to choose the next operation and compatible target.

    Reported: 7.073 seconds · checked 2026-09-18
  5. E2Community

    Search the web with typed intent judgments

    A metasearch front end lets Jev choose the query, sources, and time range, then score every result for relevance while code fans out to engines.

    Artifact: working demo · checked 2026-09-18
  6. E2Community

    Authorize agent tool calls before they run

    A starter kit from Kinde: identity and permissions decide what an agent may do, and Jev judges each call in about 200 milliseconds - does it match the request, is it destructive, does it follow planted text, does it exfiltrate - before anything runs.

    Artifact: working demo · checked 2026-09-24
  7. E2Community

    Compact conversation context by calibrated judgment

    jevtrim is a comparative analysis of context compaction driven by calibrated judgments instead of summarization: Jev scores every chunk for relevance, ordinary Python keeps what fits the token budget, and the result is auditable and replayable offline.

    Artifact: reproducible repo · checked 2026-09-24
  8. E2Community

    Drive a browser with typed page judgments

    Give jev-browser a task and a URL: Jev picks one action per step from the page's clickable, typeable, and selectable elements and scores goal-met and stuck likelihood, while code owns budgets, recovery, and stop gates.

    Artifact: reproducible repo · checked 2026-09-20
  9. E2Community

    Drive Android from semantic UI state

    A proof-of-concept Android loop stabilizes the screen, builds a short list of valid actions, and lets Jev pick one while code executes it.

    Artifact: reproducible repo · checked 2026-09-18
  10. E2Community

    Drive the Aside browser with typed decisions

    bside pilots the Aside browser with Jev instead of a chat LLM: each tick answers which action, which element, and whether the goal is met, over an action schema the pilot cannot hallucinate outside of.

    Artifact: reproducible repo · checked 2026-09-23
  11. E2Community

    Expose Jev decisions to MCP clients

    JDE, the Jev Decision Engine, wraps Jev as an MCP server: any MCP client - Claude Code, Cursor, your own agent - gets typed decision tools, with a policy layer, a decision ledger, and recorded evals comparing the hosted jev-1.13.0 against local alternatives.

    Artifact: reproducible repo · checked 2026-10-03
  12. E2Community

    Expose typed Jev judgments to MCP agents

    An MCP server gives compatible agents tools for classification, scoring, checking, matching, screening, and custom typed Jev questions.

    Artifact: reproducible repo · checked 2026-09-18
  13. E2Community

    Filter streams of text with typed batch judgments

    jevpipe pipes thousands of lines, files or records through one question and gets a typed judgment per item - grep-style filtering where Jev decides - shipped as a Rust binary on PyPI with an agent skill that teaches coding agents when to reach for it.

    Artifact: reproducible repo · checked 2026-09-28
  14. E2Community

    Gate agent commits with plain-English rules

    tenet makes agents fix rule violations before you ever see the diff: rules live in a YAML file in plain language, and Jev answers each one with a calibrated probability that becomes a pass-or-fail cutoff.

    Artifact: reproducible repo · checked 2026-09-23
  15. E2Community

    Gate agent posts for leaks by audience

    A guard for OpenClaw agents that reads an outgoing message and where it is going: Jev judges whether a client name, credential or internal hostname is about to reach the wrong readers, then confirms, blocks or rewrites per channel.

    Artifact: reproducible repo · checked 2026-09-30
  16. E2Community

    Gate risky coding-agent tool calls

    A Pi extension asks Jev to flag destructive, exfiltrating, or out-of-scope tool calls and to classify failures in command output.

    Reported: 203 requests · checked 2026-09-18
  17. E2Community

    Give a simulated person fast feelings

    Mina, a simulated 34-year-old librarian, runs on two systems: Jev reads her body, senses and clock every second and accumulates feelings; only when a feeling crosses its line does an LLM stop and think.

    Reported: 0.53 USD · checked 2026-09-24
  18. E2Community

    Ground agent answers in cited document blocks

    TraceDocs structures documents into source-linked blocks; Jev judges each with four Nouls - relevant, evidence, contradicts-premise, prompt-injection - returning a cited evidence set with a trace; the LLM writes from evidence and refuses when none exists.

    Artifact: reproducible repo · checked 2026-09-28
  19. E2Community

    Lint agent writes in 300 ms with team rules

    A fuzzy linter that watches a coding agent write and speaks up 0.3 seconds later: Jev checks the file against your team's rules - race conditions in effects, missing cleanup, house style - so the agent fixes them before any human reviews the code.

    Artifact: reproducible repo · checked 2026-10-04
  20. E2Community

    Plan day trips with a typed checker

    Jev Trip is an explainable day-trip planner where the LLM plans ahead and Jev chooses and checks: scope choices keep one day in one city, per-place choices rank every candidate with probabilities, and code owns routes, times and validation.

    Artifact: reproducible repo · checked 2026-09-28
  21. E2Community

    Plug a fail-closed decision layer into your agent

    A pluggable decision layer for the ego agent: System One (Jev) by default, swappable to local or other OpenAI-compatible backends, fail-closed guardrails - and the README's whole argument is that it is measured: 16 suites, 429 checks, rerun in full.

    Artifact: reproducible repo · checked 2026-10-03
  22. E2Community

    Prune stale coding-agent tool history

    A Pi extension uses Jev to decide which old tool calls and results still matter while keeping conversation text verbatim.

    Artifact: reproducible repo · checked 2026-09-18
  23. E2Community

    Rate pending shell commands on a color-coded safety rubric

    The kamchatka terminal agent asks Jev to place each pending shell command on a three-level safety rubric - reads and reports, changes something reversibly, destroys or sends something out - drawn green, yellow, or red beside the permission prompt.

    Artifact: reproducible repo · checked 2026-09-18
  24. E2Community

    Run Jev judgments from the shell before an agent acts

    jev-axi is a CLI that puts a half-second Jev opinion in front of every command an agent runs: pick, rate, check, rank, triage, and guard, plus a PreToolUse hook that blocks risky Bash calls before they execute.

    Reported: 375 ms · checked 2026-09-20
  25. E2Community

    Shadow a live CRM with typed decisions

    A working example of a shadow decision layer under a production CRM: one call returns intention, business division, human-handover probability, urgency, and lead quality - decided and logged, never sent.

    Artifact: reproducible repo · checked 2026-09-23
  26. E2Community

    Show every search decision as it happens

    OpenRecurSearch is an agentic web-search interface where nothing stays hidden: each research layer and each Jev decision appears live in the chat - option probabilities, score and latency included - while the markdown report streams into the side panel.

    Artifact: reproducible repo · checked 2026-10-04
  27. E2Community

    Supervise a coding agent while it works

    Foreman runs an independent observation loop that asks Jev whether a coding worker is progressing, stuck, complete, or ready for verification.

    Artifact: reproducible repo · checked 2026-09-18
  28. E2Community

    Use a decision model as an agent world model

    A bilingual study asking whether JEV can serve as a world model for LLM agents - predicting what happens next from typed state questions - benchmarked against generative LLMs on identical predictions, with run manifests and a phase-0 API check archived.

    Artifact: reproducible repo · checked 2026-10-08

Games and interactive fiction17 records

Emulator and game-state experiments — Mario, StarCraft, Pokémon, Terraria, shared stories — where Jev reads structured game state and picks the next move. Games are the lowest-risk place to demonstrate tight control loops, so the cluster works as a proving ground for latency and decision quality under continuous state change.

The evidence ceiling here is honest: these are demos and reproducible repos, not benchmarks. Their value is showing typed decisions holding up at interactive frame rates, which is hard to fake.

  1. E2Community

    Complete a StarCraft shareware mission

    A reproducible harness lets Jev direct combat, exploration, and economy actions in the original StarCraft shareware campaign.

    Reported: 1 mission completed · checked 2026-09-18
  2. E2Community

    Pick Pokémon turns from live ROM state

    A battle harness reads FireRed state from RAM and lets Jev choose the next legal move or switch while ordinary code advances the fight.

    Reported: 0.03 USD per run · checked 2026-09-18
  3. E2Community

    Benchmark Wordle solvers with Jev word priors

    A reproducible experiment beyond the 'optimal' Wordle solver: Jev scores how answer-like each of 12,972 accepted words is, the prior feeds entropy search over 1,925 days of NYT answers, and every claim is archived with pinned inputs and a reproduce script.

    Reported: 1925 days of NYT answers in the benchmark · checked 2026-10-02
  4. E2Community

    Classify dino obstacles, let physics time them

    An autonomous bot that plays the Chrome T-Rex Runner to a thousand points by asking Jev which action each obstacle requires - jump, duck, or run - and letting a measured physics model decide exactly when to press.

    Artifact: reproducible repo · checked 2026-09-25
  5. E2Community

    Dodge Terraria bosses with typed reflexes

    A tModLoader mod whose boss fights are driven by Jev: every 200ms one request asks a nine-way intent, a five-band danger score, and whether to dash or jump; a per-frame reflex layer turns intents into keypresses.

    Artifact: reproducible repo · checked 2026-09-20
  6. E2Community

    Feel Jev speed and cost in a 25-second game

    A tiny Japanese game with no send button: type IT buzzwords and each keystroke gets a Jev judgment that stretches a meter, so 25 seconds of play answers what curl never does - how fast and how cheap the model feels inside a real app.

    Artifact: reproducible repo · checked 2026-10-01
  7. E2Community

    Give a companion character sub-second typed reflexes

    A Japanese companion game where one Jev pass decides the character's true feeling - Choice, affection Score, dislike Noul, topic Choice - in 0.2 to 0.5 seconds, so her face and a one-liner land before Claude-written dialogue and a Gemini TTS voice.

    Reported: 0.5 seconds · checked 2026-09-27
  8. E2Community

    Judge player persuasion by each character's values

    An npm library that lets game characters argue back: give a name, a persona and a goal, pass what the player typed, and Jev judges whether that character - with those values - was convinced; the same line can win a greedy merchant and offend an honest guard.

    Artifact: reproducible repo · checked 2026-09-29
  9. E2Community

    Let Jev paint a sketch region by region

    A generative painting instrument: select part of a sketch and Jev chooses its paint material - one typed Choice per region with probabilities and certainty bands - and the material flies in and paints itself; demo film and live gallery included.

    Artifact: reproducible repo · checked 2026-09-28
  10. E2Community

    Paint alongside Jev on a shared emoji canvas

    A shared 1,000-by-1,000 emoji canvas where humans place strokes and Jev paints with them: after each stroke one typed call picks a contextually relevant emoji and where to put it, and a yes/no decides whether your stroke was finished.

    Artifact: reproducible repo · checked 2026-09-29
  11. E2Community

    Pick the best line of a shared story

    Jev Yarn is a party game where everyone writes the next sentence and Jev picks the winner: one taste request scores every line on four dimensions, and Nouls handle room filters and whether the story feels finished.

    Artifact: working demo · checked 2026-09-22
  12. E2Community

    Play League of Legends with three Jev heads

    Jev plays Yasuo in League of Legends through three decision heads - strategy once a second, tactics six to seven times a second while units are on screen, and build checks every twenty seconds - while code reads the game and executes.

    Artifact: reproducible repo · checked 2026-09-25
  13. E2Community

    Play Pac-Man on typed decisions in the browser

    A single-file Pac-Man that asks the System One endpoint for typed decisions per game tick, playable live without setup or locally with your own key pointed at the official endpoint.

    Artifact: working demo · checked 2026-09-25
  14. E2Community

    Play Super Mario from emulator state

    An experimental controller translates NES telemetry into object-centric JSON and lets Jev choose the next legal controller macro.

    Artifact: reproducible repo · checked 2026-09-18
  15. E2Community

    Steer a simulated fly brain with typed decisions

    Spikecast puts Jev at the controls of a fly's simulated brain: the insect walks a road, and every turn, stop or dash is a typed decision with the synapse memory and neural firing visualised beside it - with docs separating what is real from what is modelled.

    Artifact: reproducible repo · checked 2026-10-02
  16. E2Community

    Steer a town of fifty citizens with one broadcast

    You are the Town Crier of a 3D town: write one broadcast and all fifty citizens decide in parallel - investigate, join, flee, warn, or ignore - in a single batched choice request.

    Artifact: working demo · checked 2026-09-22
  17. E2Community

    Steer matter paths with live decisions

    An interactive installation - explore a flooded observatory where glowing matter builds paths from your movement, gaze and actions, with real-time Jev decisions deciding how the world responds, shipped as a playable web app with live-decision tests.

    Artifact: reproducible repo · checked 2026-10-08

Security13 records

Alert triage as UNIX filters, vulnerability text mapped to CVSS vectors, contract drift caught in OpenAPI prose, and PII redaction inside Postgres. The pattern across the cluster is narrow, checkable judgments inserted before a human or an agent acts on a security signal.

Security records carry an extra burden the directory cannot resolve: correctness of the judgment matters more than fluency, and no record here includes an independent audit. The limitations fields are worth reading before trusting any of these in a real pipeline.

  1. E2Community

    Assess a simulated company under attack

    A local cybersecurity lab replays synthetic telemetry and asks Jev for compromise probability, classification, severity, and an advisory response as evidence accumulates.

    Artifact: reproducible repo · checked 2026-09-18
  2. E2Community

    Authorize agent tool calls before they run

    A starter kit from Kinde: identity and permissions decide what an agent may do, and Jev judges each call in about 200 milliseconds - does it match the request, is it destructive, does it follow planted text, does it exfiltrate - before anything runs.

    Artifact: working demo · checked 2026-09-24
  3. E2Community

    Catch contract drift hidden in OpenAPI prose

    OAS Sentinel compares two OpenAPI documents in two layers: deterministic checks find structural breaks, and Jev answers bounded semantic questions about changed prose - retries, ordering, pagination, error meaning - that schema diffs cannot see.

    Artifact: reproducible repo · checked 2026-09-23
  4. E2Community

    Check suspicious messages for scams with typed verdicts

    Paste a suspicious SMS, email, DM or listing into ScamCheck - web app, API or browser extension - and get a scam verdict, risk score, plain-English reasons and next steps, with all wording from the project's own templates rather than a model.

    Artifact: reproducible repo · checked 2026-10-02
  5. E2Community

    File local documents with typed safety lanes

    A local-first document filer: text is extracted locally, Jev decides category, confidentiality and prompt-injection risk as typed Choices, and low-confidence or suspicious files land in review lanes - never overwritten, with audit preview and undo.

    Artifact: reproducible repo · checked 2026-09-30
  6. E2Community

    Gate agent posts for leaks by audience

    A guard for OpenClaw agents that reads an outgoing message and where it is going: Jev judges whether a client name, credential or internal hostname is about to reach the wrong readers, then confirms, blocks or rewrites per channel.

    Artifact: reproducible repo · checked 2026-09-30
  7. E2Community

    Guard .NET LLM apps with typed checks

    Kassad brings calibrated guardrails to .NET: every prompt, completion, tool call and citation passes narrow typed checks answered by Jev, batched one round trip per stage, thresholded in code into Allow, Flag, Review, or Block.

    Artifact: reproducible repo · checked 2026-09-24
  8. E2Community

    Moderate community messages with typed probabilities

    jevmod scores every community message for spam, scam, harassment, NSFW, self-harm, doxxing, off-topic and custom plain-English rules; operators set thresholds, decisions log their numbers, and bots ship for Discord, Twitch, YouTube and Reddit.

    Reported: 0.04 USD per 1,000 messages · checked 2026-09-28
  9. E2Community

    Redact PII inside Postgres, by content

    A support inbox where Jev decides per span whether text is personal data and of what kind, and a redact() SQL function enforces the masking in Postgres by the viewer's clearance - content-aware, not pattern-based.

    Artifact: reproducible repo · checked 2026-09-23
  10. E2Community

    Run Jev judgments from the shell before an agent acts

    jev-axi is a CLI that puts a half-second Jev opinion in front of every command an agent runs: pick, rate, check, rank, triage, and guard, plus a PreToolUse hook that blocks risky Bash calls before they execute.

    Reported: 375 ms · checked 2026-09-20
  11. E2Community

    Steer RAG chunking and index hygiene with Jev

    Chunk documents where meaning changes, not at a character count: Jev judges sentence continuation, planted instructions are quarantined before embedding, and retrieved passages classified as evidence, conflict or noise - on your existing vector database.

    Artifact: reproducible repo · checked 2026-10-01
  12. E2Community

    Triage security alerts as a UNIX filter

    A single Rust binary pipes JSONL security alerts through five typed Jev questions and emits validated dispositions that code, not the model, enforces.

    Reported: 200 ms per alert · checked 2026-09-18
  13. E2Community

    Turn vulnerability texts into CVSS vectors

    Python scripts that ask Jev to select every CVSS metric from a vulnerability description, then compute the numeric score in code exactly per the FIRST specification - for CVSS 3.0, 3.1, and 4.0.

    Artifact: reproducible repo · checked 2026-09-20

Smart home5 records

Home Assistant integrations, hydroponic control, and screen-time judging — small, high-variance domestic systems where a typed decision maps to one safe action. The official TypeSafe smart-home routing record sits here alongside community equivalents, which makes the cluster a natural side-by-side of official and community practice.

Household devices make failure visible, so authors tend to document fail-open behavior and manual overrides carefully. Both are recorded in the limitations of each entry.

  1. E2Official

    Route smart-home assistant requests

    TypeSafe's smart-home demo evaluates a request against many typed questions in parallel, then lets code use only the answers relevant to that request.

    Artifact: official example · checked 2026-09-18
  2. E2Community

    Turn Home Assistant state into automation signals

    A Home Assistant integration exposes Jev probabilities, choices, and scores as entities and action responses that automations can use.

    Artifact: reproducible repo · checked 2026-09-18
  3. E2Community

    Decide Home Assistant commands with typed answers

    A Home Assistant conversation agent - installable via HACS - that decides with a TypeSafe System One model instead of an LLM: your spoken or typed commands become typed choices and Nouls that drive devices, with metered costs documented.

    Artifact: reproducible repo · checked 2026-09-30
  4. E2Community

    Steer hydroponic crops with typed judges

    A Home Assistant add-on runs a crop-steering irrigation engine with Jev judging what fixed rules got wrong - ramp timing, probe trust, shot landing, EC moves, alerts - inside code-enforced physical limits; the README replays real Jev answers from 27 Sep 2026.

    Reported: 0.92 probability · checked 2026-09-27
  5. E2Community

    Watch your screen time with local and hosted judges

    Qualm is a screen-time app for Apple silicon Macs, local-first, built on two judges: Kev running on your machine and TypeSafe's hosted Jev - typed choices decide what counts as a distracting session without your activity leaving the machine unless you opt in.

    Artifact: reproducible repo · checked 2026-10-03

Mobile5 records

Android automation from the accessibility tree, ad silencing that fails open, aphasia assistance, and Flutter bindings. The cluster demonstrates that typed decisions work from semantic UI state on constrained devices, not just from desktop DOM dumps.

Every record here is community-published and artifact-verified against a specific device or emulator. Portability across Android versions is exactly the kind of claim none of them make, so none should be assumed.

  1. E2Community

    Call Jev from Flutter apps

    A Flutter plugin that brings typed Jev calls to Dart apps: question builders for Noul and Choice with option-count validation, the systemone wire protocol on the native endpoint, and a Python-side test harness for the plugin's bridge.

    Artifact: reproducible repo · checked 2026-09-30
  2. E2Community

    Drive an Android phone from the A11Y tree

    An Android automation agent with a two-tier brain: Jev decides fast from the accessibility tree, a vision agent takes over only when the structural view is ambiguous, and ADB executes - cheap judge on the hot path, expensive one on the exceptions.

    Artifact: reproducible repo · checked 2026-10-06
  3. E2Community

    Drive Android from semantic UI state

    A proof-of-concept Android loop stabilizes the screen, builds a short list of valid actions, and lets Jev pick one while code executes it.

    Artifact: reproducible repo · checked 2026-09-18
  4. E2Community

    Find the word aphasia can't say

    For people with aphasia who know what they mean but can't get the word out: describe it any way you can, and Jev picks the best guesses from a fixed 900-word list as big tap-to-hear picture tiles - never inventing a word.

    Artifact: working demo · checked 2026-09-26
  5. E2Community

    Silence ad noise on Android, failing open

    An Android noise gate for notifications and SMS that asks Jev whether each message is an ad - and suppresses only what is explicitly flagged. Every uncertain path resolves to allow, because swallowing a verification code costs more than leaking an ad.

    Artifact: reproducible repo · checked 2026-09-23

Music4 records

Composition steered by enum labels, songs written note by note, and piano improvisation from typed mood choices. A small cluster, but a useful one: creative domains show Jev making aesthetic judgments under explicit human direction rather than operational decisions under automation.

The records are demos and reproducible repos. They evidence that the decision interface works for creative control, not anything about musical quality, which no published metric here attempts to score.

  1. E2Community

    Improvise at the piano with typed mood choices

    Describe a mood - a slow, sad waltz - and a live piano plays it, with Jev deciding continuously as it goes: every musical choice is a typed question answered with probabilities, streamed to a public site in real time.

    Artifact: reproducible repo · checked 2026-10-02
  2. E2Community

    Route messy music requests before the LLM

    A DJ chatbot's pre-LLM router, built as a learning study: free-form /play requests - English, Spanish, typos, pasted lyrics, emojis - classified by Jev in one call (~500 ms, ~$0.000036) into a structured hint that tells the expensive LLM what it is handling.

    Artifact: reproducible repo · checked 2026-10-07
  3. E2Community

    Steer a music composition with enum labels

    A playground asks Jev to decide only closed-vocabulary labels for character, key, meter, and phrasing while deterministic code writes, engraves, and plays the notes.

    Artifact: working demo · checked 2026-09-18
  4. E2Community

    Write songs by choosing every note

    Jev cannot write a single note, so code lays out the bars and the legal notes with musical facts attached, and Jev chooses: mode, tempo, form, every chord, every note - one call per decision, every call replayable.

    Reported: 0.005 USD · checked 2026-09-22

Trading3 records

Per-block market quotes, paper trading on typed market state, and one live crypto desk with Jev as the decision core. The cluster is small and the stakes are the highest in the directory, since a misjudged control decision maps directly to money.

Read these records with the strictest posture in the index: reported performance comes from single artifacts, market conditions are not controlled, and no record has been independently verified. They demonstrate architecture, not profitability.

  1. E2Community

    Choose a market quote every Monad block

    A trading loop reads the Kuru MON-USDC order book and asks Jev for a buy-or-sell judgment before code places a post-only limit order.

    Artifact: reproducible repo · checked 2026-09-18
  2. E2Community

    Paper-trade crypto on typed market state

    A Next.js dashboard that streams live crypto data, computes fifteen moving averages and eleven oscillators into one typed market state, and lets a Jev agent trade a simulated 100k portfolio - paper only, no broker connected.

    Artifact: reproducible repo · checked 2026-09-25
  3. E2Community

    Run a live crypto desk with Jev as the brain

    A multi-bot OKX trading desk with Jev as the brain and code as the body: each tick reads market and account state, computes features in TypeScript, asks Jev typed questions, then holds or places and cancels real orders - dashboard as the glass.

    Artifact: reproducible repo · checked 2026-10-04

Robotics3 records

Drone tactics from camera-derived state, a robot arm driven through layered choices, and a hospital delivery robot steering past four hundred people. Physical control is where bounded judgments meet the least forgiving failure modes, and the cluster shows authors responding with layered choice hierarchies and conservative action sets.

All three records are community-published demonstrations. None reports long-run reliability on hardware, which is the number that would actually matter for deployment.

  1. E2Community

    Choose drone tactics from camera-derived state

    A simulated quadrotor uses Jev for low-frequency tactical judgments while classical vision, flight control, and safety reflexes remain in code.

    Reported: 77.5 meters · checked 2026-09-18
  2. E2Community

    Control a robot arm through layered choices

    Jev drives a LIBERO robot through 27 control inputs with layered decisions - intent, then motion family, then input - while reversible physics previews evaluate candidate effects locally before anything executes.

    Artifact: working demo · checked 2026-09-23
  3. E2Community

    Steer a hospital delivery robot past four hundred people

    A hospital delivery robot judged about four times a second: code samples twelve candidate paths and removes predicted collisions, Jev picks one among the survivors, and a deterministic safety brake guarantees no contact - built at the TypeSafe JEVATHON.

    Reported: 84 s · checked 2026-09-27

Add to the record

Missing a public Jev project?

Send the artifact, what Jev decided, and any measured outcome. Nothing appears in the index before editorial review.

Submit evidence