Skip to content
JevDirectory.org
CommunityRepos & SDKs234 starsVerified 2026-09-22

Jev Visual

An educational Jev-like inference project for Apple Silicon: Qwen3.5-0.8B scores candidate answers from shared multimodal context instead of generating text, with browser, CLI, and HTTP interfaces.

Category
Repos & SDKs
Published by
Community
Author
hr98w
Added
2026-09-22
Tagscommunitypythonopen-modelsmlxapple-silicondemo

Highlights

  • Reuses one shared image prefill and forks the cache to score candidate outputs from model logits, avoiding autoregressive structured generation.
  • Supports 1 to 64 questions with 2 to 26 options each; probabilities are relative to the supplied options, not correctness estimates.
  • Demos cover an AI sorting factory, Breakout, and camera gestures; a recorded Breakout test cleared 9 bricks and returned 6 balls over 80 decisions.
  • Needs an Apple Silicon Mac with Metal and downloads about 596 MiB of weights; setup was tested on an M4 with 16 GB running macOS 15.1.
  • The README is explicit that it is independent of TypeSafe and does not reproduce Jev's architecture, RLCD training, calibration, or serving.

Quickstart

bash
uv venv --python 3.13
source .venv/bin/activate
uv pip install -r requirements-lock.txt
uv pip install --no-deps -e .
jev-visual-download
JEV_VISUAL_MODEL_PATH=.models/Qwen3.5-0.8B-4bit uvicorn jev_visual.server:app --host 127.0.0.1 --port 8788

Watch out

MIT-licensed for the original code with third-party notices. Needs an Apple Silicon Mac with Metal, Python 3.13, uv, and the downloaded Qwen3.5-0.8B-4bit weights; only the Qwen adapter is verified.

More like this

57GitHub stars
An Apple Silicon decision layer that scores every allowed option of a schema (booleans, enums, multi-selects) from MLX model logits in one prefill, assembles schema-valid JSON itself, and exposes a System One-compatible endpoint.
Repos & SDKs#community#python#open-models
Communityjevmlx
14.3kGitHub stars
Open multilingual System 1 decision models with published checkpoints for choice, score and noul questions, plus a router that dispatches each request to the right checkpoint in one forward pass.
Repos & SDKs#community#python#open-models
Communitylaya
1.9kGitHub stars
An open 0.6B replica of Jev that turns states and questions into full probability distributions without decoding answer tokens, trained and evaluated on Maze, Snake, and ViZDoom.
Repos & SDKs#community#python#open-models
CommunityNanoJev
Back to all resources

From the community

Posts from builders shipping with Jev right now.

Follow @typesafeai

Classifying 1,500 real emails

this model is actually insane at email classification i tested it on 1500 of my own emails to see how well it works and I am blown away

Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

Fast browser use with Stagehand

we built blazing fast computer/browser use with Jev + @Stagehanddev. this task cost $0.001 and executed at near instant speed (in a remote browser btw) the loop: observe the page, send a11y tree as state + actions as questions, Jev decides the next action, then Stagehand Show more

Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

LLM-as-a-judge, sped up

Jev has spoken. It picked which model is AGI. 20–200x faster. 40–400x cheaper. This could make things like LLM-as-a-judge insanely fast and nearly free. (I tried a bunch of prompts and still didn’t burn through $0.10.)

Image
Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

Instant compaction with Jev

A Claude session from 1M to 86K tokens

This is actually insane. This uses @typesafeai Jev model, as a plugin in Claude to review all the un-nesseasary tool calls, and it takes 1s to run! Like, literally, 1 second to take my Claude session from nearly 1M to ... 86K tokens! 😮 Ask your claude to install it and be  Show more

Image
Image
tamara
tamara
@tamarajtran

found the perfect use case for @typesafeai Jev: instant compaction in 2026, why is compaction still a summarization prompt? Jev can make it instant by scoring every tool call and dropping what’s irrelevant

Reply

Vercel's fx safety reviewer, 18x faster

We're seeing extraordinary results from @typesafeai. Default mode in 𝚏𝚡 is auto, with a safety reviewer analyzing every command. That reviewer runs on GPT Luna today. Jev is up to 18x faster (p95) *and* more accurate. It's coming to @vercel AI Gateway and likely new default.

Pranit
Pranit
Vercel
@fazxes

We benchmarked fx auto mode (safety) classifier with @typesafeai's Jev. tl;dr: ~5-18x faster and more accurate than 𝚐𝚙𝚝-𝟻.𝟼-𝚕𝚞𝚗𝚊, our current top choice

Image
Reply