I made a DuckDB extension where you can use @typesafeai 's Jev to do quick classification of rows in any csv/parquet file or duckdb table about 10sec for 1k rows ~ better than using an LLM, way more ergonomic than a classifier game-changing for data analysis!
Jev Review
A local-first MCP server that gives Claude Code, Codex, Cursor, and OpenCode structured quality scores from Jev across correctness, complexity, tests, security, and other dimensions while the agent writes code.
- Category
- Tools & Integrations
- Published by
- Community
- Author
- NiazMorshed2007
- Added
- 2026-09-22
Highlights
- One focused MCP tool, jev_review, runs as a local Node.js process over stdio; there is no hosted backend, database, or telemetry service.
- Jev returns typed Score, Choice, and Noul decisions, which the server converts into dimension metrics, confidence levels, and previous-evaluation comparisons.
- There is deliberately no blended 82/100 score; changes are reported per dimension, such as Readability 6.3 to 8.1.
- The coding agent stays responsible for diagnosing weaknesses and editing code; Jev only supplies the scalar signal.
- Install from GitHub with npx plugins add; manual MCP configuration is documented for all four clients.
Quickstart
export JEV_API_KEY="YOUR_API_KEY"
npx plugins add NiazMorshed2007/jev-reviewWatch out
MIT-licensed. Needs Node.js 20+, a Jev API key, and one of the four supported clients; review context (task, diff, files, repositoryContext) is sent to TypeSafe's API, which enforces roughly a 32,768-token state ceiling.
More like this
From the community
Posts from builders shipping with Jev right now.
Classifying rows in DuckDB
A playable 16-judgment demo
typesafe's jev is fun! live demo you can play with: typesafe-demo.val.run
AI multiple choice, not essay writing
WTF is Jev by @typesafeai? Here’s the tl;dr ELI5: Think AI multiple choice, not AI essay writing. It doesn’t chat. It makes decisions your software can act on: “Spam or not?” “Which tool should this agent use?” “Does this need a human?” The exciting part: roughly 200x faster Show more
Screening agent actions with Jev
Tested TypeSafe’s Jev (no-text, probability-only model) as an AI agent safety monitor. Checking each action first worked well caught most attacks with almost no false blocks, and much faster than Gemini.
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x
Cua's small System One models
1/ Introducing CUA-S1: a family of System One Models, small, specialized, and built for computer use. Today we're open-sourcing CUA-S1-FORMS, the first in the family: github.com/trycua/cua
A 706K-parameter form filler
cua open sourced a 706k param model that fills a whole form in one 50ms pass the llm agent doing the same form took 23 turns and 39.6 seconds the specialists are going to eat the generalists from the bottom
1/ Introducing CUA-S1: a family of System One Models, small, specialized, and built for computer use. Today we're open-sourcing CUA-S1-FORMS, the first in the family: github.com/trycua/cua
