48 of 561 catalog entries
Articles & threads — built on Jev
Longer-form and conversational material: blog posts, engineering notes, and social threads where builders describe what they tried. Useful for judging where the approach fits and where it does not.
Threads are linked at their source. Where a post contains a performance claim, check whether the author measured it or is repeating marketing copy.
AI that does not talk
Practical guide: playground, Python and JS SDKs, raw HTTP, and the agent skill.
AIAvatarKit turn-end gate
Voice-dialog turn-end detection using Jev scores after speech.
AINews: Jev, a System One Model that only decides
Latent Space's launch-day roundup: over 100x faster and 200x cheaper than small frontier LLMs.
Browser Use + Jev
Gregor Zunic's flight-search demo with a dynamic DOM action space.
Computer use built on Jev
Aaron Levin: 155x cheaper than Opus 5, about 20x faster, and it generalizes across operating systems.
ConsoleChaosRacing driven by Jev
Racing UI wired to Jev driving decisions.
DuckDB Jev classifier
DuckDB extension that classifies rows in CSV, Parquet, or DuckDB tables with Jev, about 10 seconds per 1,000 rows.
Early Jev tools roundup
Thread cataloguing the first wave of Jev tools: MCP servers, routers, reviewers, and browser agents.
Ground Truth news-framing extension
Browser extension that classifies an article's framing, type, topic, and loaded language with Jev.
Hacker News launch thread
1,800-point thread debating whether typed decisions replace LLM calls for classification, routing, and scoring.
He says he co-invented ChatGPT. His new AI will not write a word
Dev.to walkthrough of the Vercel AI SDK evaluate integration.
Hide posts on X with natural language
Marcel Pociot's browser extension that collapses posts based on a Jev judgment.
Hook panel A/B tester
Near-real-time scoring of TikTok and Instagram hooks against about 100 personas.
Internal classifier field note
Matched-precision comparison against a private fine-tuned classifier.
Jev as an agent safety monitor
Test report using Jev to check each agent action first: most attacks caught, almost no false blocks.
Jev gomoku harness
Local tactics shrink 225 moves to about 40 candidates, then Jev picks among tiered options.
Jev in 34 seconds
Short video explainer of how Jev's typed-decision loop works.
Jev in a Grammarly-style Mac app
Desktop writing app using Jev for fast structured writing judgments.
Jev plays Minecraft (r/accelerate)
Work-in-progress demo of Jev driving Minecraft, including fleeing zombies at night.
Jev Typewriter launch
Steve Krouse's playable 16-judgment demo and video.
Jev vs Mistral and Gemini for event validation
Head-to-head test at validating local event listings, with cost and latency.
Jev vs Qwen on Cerebras
Video comparison against a structured-output LLM baseline.
jev 同士に五目並べで対戦させた
Jev vs Jev gomoku with source and timing logs.
jev-rabbit PR review bot
Work-in-progress PR reviewer with plain-English Jev rules.
Jev: System One models explained
The Neuron's explainer on AI decisions without a chatbot.
Jev: System One models explained (DataCamp)
Third-party write-up of the System One primitives, pricing, and vendor workflow evals.
Jev: The Language Model That Will Not Talk
Anthony Maio's essay on what a model that cannot generate text is for.
Kalshi prediction-market bot
Jev trades 15-minute and 1-hour BTC, ETH, and SOL markets on Kalshi.
Launch thread by Diogo Almeida
TypeSafe's founder on why RLCD-trained decision models are a shorter path to value than chat models.
Mini-Vibe Check: Jev judged everything I have written in 0.7 seconds
Every's Mike Taylor runs his whole archive through Jev.
Model router built with Jev
Ephraim Duncan's demo where Jev decides which model should serve a request.
Model router CLI
Task plus subscription list in, Jev picks which model or agent should handle it.
One judge call vs twelve dimension scores
One direct Jev question per row against 12–14 Jev-scored dimensions with locally fitted weights on three classification tasks: 5,477 test rows, 25,174 Jev calls, $1.43. Decomposition wins on Japanese NLI (0.9076 vs 0.8373) but flags about 25× more hard benign rows as attacks (37.2% vs 1.5%).
OpenCode browser use powered by Jev
Preview of fast browser use with Jev and OpenCode's browser CLI.
skillbox + Jev skill routing
MCP skill router where Jev picks the relevant skills instead of a long agent search.
Spanish AEPD corpus test
Jev versus a hand-built regex on 544 public data-protection resolutions: 98.2% agreement for about five cents.
Stagehand + Jev browser use
Observe the accessibility tree, Jev chooses the next action, Stagehand executes.
StarCraft Brood War WASM MCP demo
Brood War in WASM exposed as an MCP server, with Jev playing and still losing to a Zerg rush.
Support ticket classifier
Jev labels category, urgency, and human-versus-auto handling for support tickets.
Tabletop MMORPG action mapper
Eval of Jev turning free-text player intent into typed server actions: 96% agreement, 317 ms median.
Typed decisions, not chat
Independent walkthrough separating TypeSafe's published claims from public evidence.
TypeSafe AI debuts model for machines that plays Doom
The Register on the launch, the Doom demo, and the $40M seed round.
TypeSafe Jev: the first decision-only model class
Release-week technical roundup: API, evals, adapter, and skill.
TypeSafeのJevを正しく驚く
Japanese walkthrough of what Jev is and is not.
Vercel fx: Jev as a command safety reviewer
Guillermo Rauch: Jev reviews every fx command, faster and more accurate than a chat model.
ViZDoom Jev agent
Two decision channels on ViZDoom, navigation at 5 Hz and combat at 12 Hz, with an 18-kill test run.
What is Jev?
Short practical intro with a Python ticket-triage example.
Wiki-link clicker demo
Page-level demo where Jev picks which candidate link to click toward a goal.