Hacker News launch thread
1,800-point thread debating whether typed decisions replace LLM calls for classification, routing, and scoring.
Coverage, write-ups, and X threads. Browse the collection in Jev Library.
48 entries in the current community snapshot1,800-point thread debating whether typed decisions replace LLM calls for classification, routing, and scoring.
Latent Space's launch-day roundup: over 100x faster and 200x cheaper than small frontier LLMs.
The Register on the launch, the Doom demo, and the $40M seed round.
The Neuron's explainer on AI decisions without a chatbot.
Every's Mike Taylor runs his whole archive through Jev.
Anthony Maio's essay on what a model that cannot generate text is for.
Release-week technical roundup: API, evals, adapter, and skill.
dev.to walkthrough of the Vercel AI SDK evaluate integration.
Independent walkthrough separating TypeSafe's published claims from public evidence.
Short practical intro with a Python ticket-triage example.
Practical guide: playground, Python and JS SDKs, raw HTTP, and the agent skill.
Head-to-head test at validating local event listings, with cost and latency.
Japanese walkthrough of what Jev is and is not.
Jev vs Jev gomoku with source and timing logs.
TypeSafe's founder on why RLCD-trained decision models are a shorter path to value than chat models.
Aaron Levin: 155x cheaper than Opus 5, about 20x faster, and it generalizes across operating systems.
Gregor Zunic's flight-search demo with a dynamic DOM action space.
Steve Krouse's playable 16-judgment demo and video.
Ephraim Duncan's demo where Jev decides which model should serve a request.
Video comparison against a structured-output LLM baseline.
Matched-precision comparison against a private fine-tuned classifier.
Test report using Jev to check each agent action first: most attacks caught, almost no false blocks.
Marcel Pociot's browser extension that collapses posts based on a Jev judgment.
Work-in-progress demo of Jev driving Minecraft, including fleeing zombies at night.
Guillermo Rauch: Jev reviews every fx command, faster and more accurate than a chat model.
Observe the accessibility tree, Jev chooses the next action, Stagehand executes.
Preview of fast browser use with Jev and OpenCode's browser CLI.
Browser extension that classifies an article's framing, type, topic, and loaded language with Jev.
Desktop writing app using Jev for fast structured writing judgments.
Jev trades 15-minute and 1-hour BTC, ETH, and SOL markets on Kalshi.
Task plus subscription list in, Jev picks which model or agent should handle it.
Jev labels category, urgency, and human-versus-auto handling for support tickets.
Racing UI wired to Jev driving decisions.
MCP skill router where Jev picks the relevant skills instead of a long agent search.
Work-in-progress PR reviewer with plain-English Jev rules.
Voice-dialog turn-end detection using Jev scores after speech.
Thread cataloguing the first wave of Jev tools: MCP servers, routers, reviewers, and browser agents.
Third-party write-up of the System One primitives, pricing, and vendor workflow evals.
DuckDB extension that classifies rows in CSV, Parquet, or DuckDB tables with Jev, about 10 seconds per 1,000 rows.
Near-real-time scoring of TikTok and Instagram hooks against about 100 personas.
Local tactics shrink 225 moves to about 40 candidates, then Jev picks among tiered options.
Two decision channels on ViZDoom, navigation at 5 Hz and combat at 12 Hz, with an 18-kill test run.
Brood War in WASM exposed as an MCP server, with Jev playing and still losing to a Zerg rush.
Jev versus a hand-built regex on 544 public data-protection resolutions: 98.2% agreement for about five cents.
Eval of Jev turning free-text player intent into typed server actions: 96% agreement, 317 ms median.
Short video explainer of how Jev's typed-decision loop works.
Page-level demo where Jev picks which candidate link to click toward a goal.
One direct Jev question per row against 12–14 Jev-scored dimensions with locally fitted weights on three classification tasks: 5,477 test rows, 25,174 Jev calls, $1.43. Decomposition wins on Japanese NLI (0.9076 vs 0.8373) but flags about 25× more hard benign rows as attacks (37.2% vs 1.5%).