Skip to content

Lichen

A local /v1/systemone server that reads label probabilities from a GGUF chat model, with prompt repetition and a confidence shrink toward uniform.

Category
Models & Reimplementations
Format
—
Published by
Community
Author
Mushroom-Systems
Added
2026-09-28
Last verified
2026-09-28

Mushroom-Systems/lichen

Organization repository on GitHub

View on GitHub

A local, API-compatible replacement for Jev, TypeSafe's System One model

GitHub stars
39
Forks
2
Primary language
Python
License
mit
Last pushed
Updated Sep 2026

Repo stats from the GitHub API, cached Sep 2026.

Highlights

  • On 231 public JevBench v1.4 items, gemma-4-26B-A4B with the default options scored 207 correct against 200 for Jev 1.13.
  • The same table lists median latency of 93 ms on one laptop GPU, against 665 ms for hosted Jev including the network.
  • The README says the two systems disagree on 13 items and the edge is not significant (p = 0.09).
  • A choice takes 2 to 62 options. Questions in one request are answered one after another.
  • The server does not check API keys and is meant for a trusted network.

Quickstart

bash
docker run --rm --gpus all -p 8765:8765 -v "$PWD/models:/models:ro" \
  lichen --model /models/gemma-4-26B_q4_0-it.gguf

Watch out

MIT. Needs Docker and an NVIDIA GPU for the published setup. Models keep their own licenses. Jev's score confidence formula is unpublished, so Lichen reuses the choice formula.

Reactions & coverage

Posts, threads, and videos about this entry from around the web.

X: deepseek-v4.1-flash-jev

Related terms

Glossary definitions related to this entry.

More like this

7GitHub stars
A Python server in front of llama-server that reads one-token label probabilities and exposes them as POST /v1/systemone.
Models & Reimplementations#community#python#open-models
689GitHub stars
A local server that reads next-token probabilities from a fine-tuned Spark model and returns Choice, Score, and Noul answers with zero generated tokens.
Models & Reimplementations#community#python#open-models
1.1kGitHub stars
A local yes/no decision model that speaks Jev's wire format on a laptop CPU, refusing Choice and Score until those question types ship.
Models & Reimplementations#community#python#open-models

From the guides

Original write-ups that draw on this entry.

A base URL is the whole client change. What still works when Ollaya, Lichen, verdict, Rizzo Flow, or jevos answers POST /v1/systemone on your machine.
Models & Reimplementations#open-models#local#api
Read article
Back to all resources

From the community

Posts from builders shipping with Jev right now.

Follow @typesafeai

Vercel's fx safety reviewer, 18x faster

We're seeing extraordinary results from @typesafeai. Default mode in 𝚏𝚡 is auto, with a safety reviewer analyzing every command. That reviewer runs on GPT Luna today. Jev is up to 18x faster (p95) *and* more accurate. It's coming to @vercel AI Gateway and likely new default.

Pranit
Pranit
Vercel
@fazxes

We benchmarked fx auto mode (safety) classifier with @typesafeai's Jev. tl;dr: ~5-18x faster and more accurate than 𝚐𝚙𝚝-𝟻.𝟼-𝚕𝚞𝚗𝚊, our current top choice

Image
Reply

Jev lands on OpenRouter

700 leads scored for $0.09

Beating Gemini Flash Lite on an eval

Browser Use Ultrafast, powered by Jev

A really smart switch statement

hype-free explanation of jev: jev does not replace gpt / claude jev is just a *really* smart switch statement like if 2016 ml classifiers got 2026 levels of intelligence it's a new* type of tool that will make a lot of workloads insanely fast, cheap, and accurate * = and by Show more

Diogo Almeida
Diogo Almeida
TypeSafe AI
@CompleteSkeptic

After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x

Reply