Skip to content
JevDirectory.org
CommunityTools & Integrations4 starsVerified 2026-09-22

typesafe-jev

A local screening workbench for a folder of CVs: it asks Jev a small fixed set of typed questions about each candidate, scores the answers with plain arithmetic you can read, and shows a sortable shortlist.

Category
Tools & Integrations
Published by
Community
Author
gtaras7
Added
2026-09-22
Tagscommunitytypescriptclassificationdataevaluation

Highlights

  • Point it at one CV or a folder of hundreds, and each candidate gets one verdict from typed Jev answers scored by readable arithmetic.
  • The role policy - weights, caps, and questions - is edited in the app, and a change re-scores every stored candidate in about 20 ms and costs nothing.
  • Judgments are kept separate from the arithmetic, so policy edits never re-call Jev for candidates that were already judged.
  • Each project in the repository is self-contained with its own README, tests, and measured results; cv-screen is the first.
  • The cv-screen README documents the results it produced, including the two bugs its test data caught.

Watch out

MIT-licensed; needs a TypeSafe API key to judge candidates, CV contents are sent to Jev, and this is a self-contained experiment rather than a production hiring system.

More like this

811GitHub stars
Vercel Labs' terminal CLI for the AI SDK: its `ai evaluate` command answers named Boolean, Choice, and Score questions, with Jev as the default evaluation model and one unchanged input shared across all questions.
Tools & Integrations#community#typescript#cli
Communityai-cli
273GitHub stars
An experimental Sutro CLI that evaluates CSV, Parquet, and JSONL data with Jev, asks you to label ambiguous and random audit rows, and uses GEPA to propose improved function definitions.
Tools & Integrations#community#python#cli
Communityjev-align
210GitHub stars
An open-source, commercially hosted generative-engine-optimization platform that uses Jev judgments within a broader application to track brand mentions and placement in AI answers across ChatGPT, Claude, Gemini, and Perplexity.
Tools & Integrations#community#typescript#postgres
Communitynotra
Back to all resources

From the community

Posts from builders shipping with Jev right now.

Follow @typesafeai

Classifying 1,500 real emails

this model is actually insane at email classification i tested it on 1500 of my own emails to see how well it works and I am blown away

Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

Fast browser use with Stagehand

we built blazing fast computer/browser use with Jev + @Stagehanddev. this task cost $0.001 and executed at near instant speed (in a remote browser btw) the loop: observe the page, send a11y tree as state + actions as questions, Jev decides the next action, then Stagehand Show more

Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

LLM-as-a-judge, sped up

Jev has spoken. It picked which model is AGI. 20–200x faster. 40–400x cheaper. This could make things like LLM-as-a-judge insanely fast and nearly free. (I tried a bunch of prompts and still didn’t burn through $0.10.)

Image
Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

Instant compaction with Jev

A Claude session from 1M to 86K tokens

This is actually insane. This uses @typesafeai Jev model, as a plugin in Claude to review all the un-nesseasary tool calls, and it takes 1s to run! Like, literally, 1 second to take my Claude session from nearly 1M to ... 86K tokens! 😮 Ask your claude to install it and be  Show more

Image
Image
tamara
tamara
@tamarajtran

found the perfect use case for @typesafeai Jev: instant compaction in 2026, why is compaction still a summarization prompt? Jev can make it instant by scoring every tool call and dropping what’s irrelevant

Reply

Vercel's fx safety reviewer, 18x faster

We're seeing extraordinary results from @typesafeai. Default mode in 𝚏𝚡 is auto, with a safety reviewer analyzing every command. That reviewer runs on GPT Luna today. Jev is up to 18x faster (p95) *and* more accurate. It's coming to @vercel AI Gateway and likely new default.

Pranit
Pranit
Vercel
@fazxes

We benchmarked fx auto mode (safety) classifier with @typesafeai's Jev. tl;dr: ~5-18x faster and more accurate than 𝚐𝚙𝚝-𝟻.𝟼-𝚕𝚞𝚗𝚊, our current top choice

Image
Reply