Hacker News launch thread
1,800-point thread debating whether typed decisions replace LLM calls for classification, routing, and scoring.
EXPLORE JEV
阅读发布介绍、技术文章与社区讨论;作者经验与性能主张需要结合原始材料判断。
THE RESOURCE INDEX
项目、工具、教程与文章
原始链接,按用途整理。
支持中文与多关键词,例如:浏览器 DOM。
全部资源
1,800-point thread debating whether typed decisions replace LLM calls for classification, routing, and scoring.
Latent Space's launch-day roundup: over 100x faster and 200x cheaper than small frontier LLMs.
The Register on the launch, the Doom demo, and the $40M seed round.
小样本使用评测,不能作为通用准确率保证。
Every's Mike Taylor runs his whole archive through Jev.
Anthony Maio's essay on what a model that cannot generate text is for.
Release-week technical roundup: API, evals, adapter, and skill.
dev.to walkthrough of the Vercel AI SDK evaluate integration.
Independent walkthrough separating TypeSafe's published claims from public evidence.
Practical guide: playground, Python and JS SDKs, raw HTTP, and the agent skill.
Head-to-head test at validating local event listings, with cost and latency.
TypeSafe's founder on why RLCD-trained decision models are a shorter path to value than chat models.
Aaron Levin: 155x cheaper than Opus 5, about 20x faster, and it generalizes across operating systems.
Ephraim Duncan's demo where Jev decides which model should serve a request.
Matched-precision comparison against a private fine-tuned classifier.
Test report using Jev to check each agent action first: most attacks caught, almost no false blocks.
Marcel Pociot's browser extension that collapses posts based on a Jev judgment.
Work-in-progress demo of Jev driving Minecraft, including fleeing zombies at night.
Guillermo Rauch: Jev reviews every fx command, faster and more accurate than a chat model.
Observe the accessibility tree, Jev chooses the next action, Stagehand executes.
Preview of fast browser use with Jev and OpenCode's browser CLI.
Browser extension that classifies an article's framing, type, topic, and loaded language with Jev.
Desktop writing app using Jev for fast structured writing judgments.
Jev trades 15-minute and 1-hour BTC, ETH, and SOL markets on Kalshi.
Task plus subscription list in, Jev picks which model or agent should handle it.
Jev labels category, urgency, and human-versus-auto handling for support tickets.
MCP skill router where Jev picks the relevant skills instead of a long agent search.
Thread cataloguing the first wave of Jev tools: MCP servers, routers, reviewers, and browser agents.
Third-party write-up of the System One primitives, pricing, and vendor workflow evals.
DuckDB extension that classifies rows in CSV, Parquet, or DuckDB tables with Jev, about 10 seconds per 1,000 rows.
Near-real-time scoring of TikTok and Instagram hooks against about 100 personas.
Local tactics shrink 225 moves to about 40 candidates, then Jev picks among tiered options.
Two decision channels on ViZDoom, navigation at 5 Hz and combat at 12 Hz, with an 18-kill test run.
Brood War in WASM exposed as an MCP server, with Jev playing and still losing to a Zerg rush.
Jev versus a hand-built regex on 544 public data-protection resolutions: 98.2% agreement for about five cents.
Eval of Jev turning free-text player intent into typed server actions: 96% agreement, 317 ms median.
25,174 Jev calls for $1.43 across template-trap, multi-dimension, and real-ledger account-coding tasks.
Page-level demo where Jev picks which candidate link to click toward a goal.
试试项目英文名、相关用途,或重置筛选。
部分项目补充了中文用途说明,原始描述保留供对照。在 GitHub 阅读完整目录 ↗