Jev in production › Evaluation and testing
Reviews every checkpoint of GPT-6-authored math animation scripts via a separate Jev API call, advisory by default with an optional gated mode (Math-To-Manim).

Six real renders from the Astra Morse film and the existing showcase. Explore the films →
A question becomes a mathematical world you can move through.
GPT-6 Astra builds the explanation. Jev challenges every step. Manim makes it visible.
The primary pipeline now uses the official Codex SDK, your Codex ChatGPT login, and gpt-6-astra for authors and evidence auditors. Real TypeSafe Jev (jev-1.13.0) evaluates each checkpoint through its separate API. Astra develops the learning brief, verifies the mathematics, directs the visual argument, writes the scene, and reviews the actual render. Jev reviews are advisory by default: their scores are retained, but do not trigger regeneration or extra investigations. Strict gates remain available with...
For the project's own README, linking back here: