Jev just landed and the agent stack moved in three days. TypeSafe AI released Jev on September 15. It is built for the work agents actually do most of the time: choose the next tool, score a risk, decide whether to retry, or answer yes or no. Latency is under half a second. Show more
Jev vs. GPT-6 Astra
Vercel's guide for when a typed Jev decision is enough and when a job still needs GPT-6 Astra to write, see images, or call tools.
- Category
- Guides & Articles
- Format
- Article
- Published by
- Community
- Author
- Ben Sabic
- Added
- 2026-09-28
- Last verified
- 2026-09-28
Highlights
- Ben Sabic, 21 September 2026: use Jev for focused decisions with defined answers and native probabilities.
- Use GPT-6 Astra when the task also has to generate content, read an image, or work through a problem with tools.
- Jev accepts text state, including structured records, and returns choices, rubric scores, or yes-or-no probabilities.
- The post says Jev does not generate replies, explanations, or media. A screenshot has to be turned into text first.
- Astra Structured Outputs can restrict a label, but the post notes a valid category can still be the wrong assignment.
Watch out
This is a product guide, not a measured bake-off. It does not publish accuracy, latency, or cost on a shared dataset.
Reactions & coverage
Posts, threads, and videos about this entry from around the web.
Hacker News: Vercel on using Jev for defined decisions and GPT-6 Astra when the job also has to write or use tools.
flashbrew4 pointsTypeSafe AI Jev vs. GPT-6 AstraRead the thread on Hacker NewsX: How the agent stack moved in Jev's first three days
YouTube: LangChain walks through Jev as a System 1 layer in an agent loop: model routing, auto mode for risky tools, and judged evals.
X: Giving your agents a decision brain
Jev could become the control layer AI agents have been missing. Instead of spending 5–20 seconds and expensive LLM calls deciding every next step, it can route actions in milliseconds at near-zero cost. In this article, I break down how x.com/i/article/2101…
More like this
From the community
Posts from builders shipping with Jev right now.
Vercel ships the AI SDK provider for Jev
Jev from @typesafeai is on AI Gateway. Build agents that decide, route, score, and stop in milliseconds: 𝚊𝚠𝚊𝚒𝚝 𝚎𝚟𝚊𝚕𝚞𝚊𝚝𝚎({ 𝚖𝚘𝚍𝚎𝚕: '𝚝𝚢𝚙𝚎𝚜𝚊𝚏𝚎-𝚊𝚒/𝚓𝚎𝚟', 𝚜𝚝𝚊𝚝𝚎, 𝚚𝚞𝚎𝚜𝚝𝚒𝚘𝚗𝚜, }); vercel.com/changelog/type…
Computer use at 155× cheaper than Opus 5
i built computer use using @typesafeai ! it is 155x cheaper than opus 5, ~20x faster, and generalizes across OS's more on how it works in the vid & thread below:
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x
Foreman keeps coding agents on task
I just open sourced Foreman: a software factory foreman built with @typesafeai's Jev. Coding agents work the factory floor. Foreman watches them, continuously assessing progress, completeness, tests, drift, and verification, and intervenes when needed. GitHub: Show more
Unclutter: an ad and slop blocker that runs on Jev
introducing Unclutter: a smart ad + slop blocker with Jev 🤓 it auto cleans up pages from slop elements: ⬖ ads ⬖ cookie banners ⬖ upsells ⬖ bs dialogs BYOK. open source + free, download below 👇
An open-source BS meter for debates and investor calls
🚨 Open Source Jev BS meter you can use this to analyze any debate / investor call / interview / sales pitch / podcast video fact check live , for example this dario interview cost 60 Jev calls / 111K tokens / $0.0047 github.com/ChetasLua/jevm…
🚨 I gave the Trump vs Kamala debate a live BS meter using Jev every sentence, both candidates, 5 yes/no questions each 1,191 Jev calls / 1.18M tokens / 415 ms median total cost : $0.0497 same questions for both, clips picked by one fixed rule, not a fact-check
A Jev-shaped model on Cerebras and Qwen
Built an alternative version of @typesafeai but on @cerebras with Qwen 3.8 27b. Similar quality, similar performance, but vastly different cost. TypeSafe was way cheaper, and did beat Qwen on performance. Closest we can get using LLMs I think. Source: github.com/iammrduncan/ty…
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x


