Jev by @typesafeai is now on OpenRouter, in beta. Jev is a System One model. Instead of generating text, it takes your app's state plus a typed question and returns a typed decision with a probability attached. There is no JSON prompting, parsing layer, and nothing to validate Show more
ASSAY-001
A pre-registered, independently rescored audit of Jev's calibration and type safety: calibrated on CLINC150 (ECE 0.0204), systematically overconfident on Banking77 (ECE 0.0936), and zero type errors across 8,576 responses.
- Category
- Practices & Patterns
- Published by
- Community
- Author
- jourdanlabs
- Added
- 2026-09-22
Highlights
- On CLINC150 Jev's chosen-option probabilities were calibrated at ECE 0.0204; on Banking77 they were not, at ECE 0.0936 and systematically overconfident.
- Across 8,576 responses there were zero type errors.
- The protocol was frozen on 2026-09-17 before any query, with corpus hashes for Banking77 and CLINC150 sealed before the first API call.
- Every request and response is logged verbatim and sealed with SHA-256; smoke-test connectivity runs are excluded from scoring.
- A blind independent re-score written from a spec on a different base model matched every field of the original scoring.
Quickstart
python3 harness/controls.py
python3 harness/score.py banking77 --sum-tol 0.02
python3 harness/score.py clinc150 --sum-tol 0.02Watch out
No license file, so reuse terms are unclear; re-running the model needs a TypeSafe key at ~/.config/typesafe/api_key and produces a new run, not the recorded one.
More like this
From the community
Posts from builders shipping with Jev right now.
Jev lands on OpenRouter
700 leads scored for $0.09
JEV is INSANE. We gave it 700 high-intent leads and personalised outreach messages. In 40 seconds, it predicted how each message would perform, assigned a confidence score and detected lead-message mismatches. All for just $0.09. JEV can also score leads, analyse buying Show more
Beating Gemini Flash Lite on an eval
Ran @typesafeai's Jev against an existing classifier eval that previously used Gemini 2.5 Flash Lite. It won both on quality (saturated the eval) and speed (6x)
Browser Use Ultrafast, powered by Jev
Breaking: Browser Use + Jev = Ultrafast ⚡ Findings flights took 7s and cost only $0.0039 🤯 > new action space every step > DOM state space > small LLM fallback to type (this video is at 1x speed btw) Built a tiny open source browser agent. try it below ↓
A really smart switch statement
hype-free explanation of jev: jev does not replace gpt / claude jev is just a *really* smart switch statement like if 2016 ml classifiers got 2026 levels of intelligence it's a new* type of tool that will make a lot of workloads insanely fast, cheap, and accurate * = and by Show more
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x
When a designer gets Jev
when a designer gets access to Jev



