Jev in production › Model and agent routing
Local OpenAI-compatible proxy in which Jev judges whether a new prompt has the same intent as a cached one and serves the stored answer on a match, failing open to the model. A Jev-confirmed hit returned in 394ms, against 3244ms for a miss (author).

MorrowCache — routers pick a model. This proxy decides whether to call one.
npm package and CLI stay @kushalicious/jevcache / jevcache (download history unchanged). Site: morrowcache.vercel.app.
A local OpenAI-compatible chat response cache. 2.0 (Agent Turn Cache) skips the model in two cases: the same question in different words, and an agent retrying the same job (same prior conversation, same-intent latest ask). It stores final assistant text only. A later stream of that answer can HIT. A response that contains toolcalls is returned and never stored.
For the project's own README, linking back here: