I made a DuckDB extension where you can use @typesafeai 's Jev to do quick classification of rows in any csv/parquet file or duckdb table about 10sec for 1k rows ~ better than using an LLM, way more ergonomic than a classifier game-changing for data analysis!
Jev Agent Skill Router
An optional routing layer for Hermes Agent, shipped as a Python CLI and library, that sends a request and a public skill catalogue to Jev and returns route, no_skill, or review plus confidence and candidate distributions.
- Category
- Tools & Integrations
- Published by
- Community
- Author
- GodsBoy
- Added
- 2026-09-22
Highlights
- Reported 68 of 72 synthetic requests routed correctly (94.4%) versus 51 of 72 (70.8%) for a lexical baseline, with zero wrong routes.
- Sends up to three parallel Jev Choice batches and a final request; two candidates survive each batch and the final distribution is conditional on the shortlist.
- Returns a three-valued specialist-needed Noul, final Choice confidence and winner probability, per-call candidates, latency, and API token usage.
- The public benchmark uses 24 synthetic skills and 72 synthetic requests, and offline tests also exercise a catalogue of 1,000 skills.
- A full live run sends roughly 288 API requests, bounded to four concurrent calls within a case, with no automatic retries and a 600-call ceiling.
Quickstart
python3 -m venv .venv
. .venv/bin/activate
python -m pip install -e '.[dev]'
export TYPESAFE_API_KEY=YOUR_API_KEY
jev-router route --model jev-1.13.0 --catalogue examples/catalogue.jsonWatch out
MIT-licensed and not affiliated with TypeSafe; needs Python 3.11+ and TYPESAFE_API_KEY, results are exploratory reused-data numbers with no repeated-run analysis, and typed output guarantees shape rather than correctness.
More like this
From the community
Posts from builders shipping with Jev right now.
Classifying rows in DuckDB
A playable 16-judgment demo
typesafe's jev is fun! live demo you can play with: typesafe-demo.val.run
AI multiple choice, not essay writing
WTF is Jev by @typesafeai? Here’s the tl;dr ELI5: Think AI multiple choice, not AI essay writing. It doesn’t chat. It makes decisions your software can act on: “Spam or not?” “Which tool should this agent use?” “Does this need a human?” The exciting part: roughly 200x faster Show more
Screening agent actions with Jev
Tested TypeSafe’s Jev (no-text, probability-only model) as an AI agent safety monitor. Checking each action first worked well caught most attacks with almost no false blocks, and much faster than Gemini.
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x
Cua's small System One models
1/ Introducing CUA-S1: a family of System One Models, small, specialized, and built for computer use. Today we're open-sourcing CUA-S1-FORMS, the first in the family: github.com/trycua/cua
A 706K-parameter form filler
cua open sourced a 706k param model that fills a whole form in one 50ms pass the llm agent doing the same form took 23 turns and 39.6 seconds the specialists are going to eat the generalists from the bottom
1/ Introducing CUA-S1: a family of System One Models, small, specialized, and built for computer use. Today we're open-sourcing CUA-S1-FORMS, the first in the family: github.com/trycua/cua
