Skip to content
JevDirectory.org
Patterns

Confidence-gated routing

Use the answer to decide what to do and confidence to decide whether to act, with per-action thresholds set by the cost of being wrong.

Category
Patterns
Also known as
—
Related terms
5
Directory entries
36
Docs
docs.typesafe.ai
Added
2026-09-24

Definition

The voice-banking example routes any intent below 0.6 confidence to a support agent before the action is considered. Above that, check_balance is safe at 0.6 because the worst case is a wrong read-out, while approve_transfer needs more than 0.85 or the user confirms first.

Confidence is not permission: a high-confidence answer can still be wrong, so irreversible actions keep hard limits in code next to the model.

Tagspatternsconfidencerouting

From the directory

Use the answer to decide what to do and confidence to decide whether to act: the voice-banking example routes below 0.6 to a human and needs 0.85 or more to auto-approve a transfer.
Practices & PatternsDocs#official#patterns#confidence
Official
How TypeSafe reports certainty, how it differs from probability, and how to use it architecturally to gate and route decisions.
Practices & PatternsDocs#official#confidence#routing
Official
A 3D robot that understands without generating text: one TypeSafe call per turn with nine typed questions, then code decides whether to act, ask, or shrug.
Cookbooks & DemosPlayground#community#playground#embodied-agent
Community
11GitHub stars
CLI that fits a per-question confidence threshold to a target accuracy on your own labeled data, verifies it on a held-out split, estimates how much traffic still needs an LLM, and re-checks locked thresholds in CI. It publishes no Jev results of its own.
Practices & Patterns#community#python#calibration
A Ruby client for TypeSafe's System One API that returns typed Noul, Choice, and Score answers, validates obvious input mistakes before spending a request, and pools HTTP connections across threads.
Repos & SDKs#community#ruby#sdk
An independent calibration study of Jev over three public benchmarks and 900 rule-generated support tickets, publishing every raw Gateway response and the quantization limits of returned probabilities.
Practices & Patterns#community#python#calibration

30 more matching entries in the full directory.

From the community

Posts from builders shipping with Jev right now.

Follow @typesafeai

Inferring Jev's internals from 1,000 calls

Jevの内部アーキテクチャを推測している技術記事(Jev’s Architecture Unmasked)からメモ。 ・本記事はJevのAPIを約1万回の呼び出して、内部構造を推測したもの ・従来の言語モデルを用いた分類やルーティングでは、トークンを1文字ずつ逐次生成するために膨大な無駄な計算コストが発生していた。 Show more

Reply

The open System One roundup

Jev 发布没几天,开源社区已经开始疯狂复刻了🔥 最值得推荐的五个模型: 1、Laya 421M:原生决策模型,支持 Mac 2、Decider-2B:最像 Jev,基于 Qwen3.5 3、NanoJev 0.6B:专门的 Decision Head 4、Reflex:Qwen3.5 + Direct Logits 5、System-One 4B:专门做概率校准 Show more

小墨同学
小墨同学
@xiaomovps

Jev 刚发布没几天,开源社区就出现了同款🔥 Decider-2B模型,是基于 Qwen3.5-2B 做了特殊调整 它和 Jev 模型是一样的 只做选择 评分和判断 不是文本类的 LLM 模型 但两者还是有几个明显区别: 1、模型 Jev:闭源 System One Model Decider:Qwen3.5-2B,约 1.9B 参数,Apache 2.0 开源 2、价格

Image
Reply

Jev lands on the Vercel AI Gateway

Vercel ships the AI SDK provider for Jev

Computer use at 155× cheaper than Opus 5

i built computer use using @typesafeai ! it is 155x cheaper than opus 5, ~20x faster, and generalizes across OS's more on how it works in the vid & thread below:

Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply

Foreman keeps coding agents on task