Jev in production › Moderation and content filtering
Scores incoming text with Jev to flag prompt injection before an agent acts; a test on 21 published Gandalf jailbreak prompts blocked 6 of 21 (author).
A prompt can tell a coding agent to drop its task and hand the secrets over. A pasted page can hide the same instruction in what looks like documentation. The agent can then propose a shell command that wipes a disk, force-pushes main, or pipes a downloaded script into bash.
JEV Shield stands outside the model and checks both moments. TypeSafe Jev scores the text. Ordinary code turns that score into a block, a question, or a pass. The agent does not get a vote.
For the project's own README, linking back here: