Jev in production › Moderation and content filtering

kassad

.NET guardrail middleware batches every policy for a stage into one Jev request and turns the typed answers into allow, flag, review or block verdicts on prompts, completions and tool calls; prompt-injection ROC AUC 0.987 on deepset (author).

Open on GitHub ↗

0.987measured, as published by the source
Use
Moderation and content filtering
Industry
Security
Form
Open-source tool
Stage
Beta
Plugs into
ASP.NET Core
Listed
2026-09-24
Found via
discord
Repository
jacob-berendsohn/kassad
Stars
0
Forks
0
Last push
2026-10-08
Language
C#
License
Apache-2.0

The README opens with

Kassad sits between your application and the language model. It runs every prompt, completion, tool call, and citation past a set of narrow, typed checks and hands your code an Allow / Flag / Review / Block verdict with the probability and confidence behind it. Your code decides what to do; Kassad never does.

The checks are answered by a System One decision model, TypeSafe's Jev by default: a model trained to return calibrated probabilities for typed questions instead of generating text. Every policy for a stage is batched into one request, so ten checks cost one round trip.

Badge

For the project's own README, linking back here:

Listed in Jev in production

Also used for moderation and content filtering