AI Resonance · 阅读最新日报 · AI 入门推荐

2026-09-22 · America/Los_Angeles · 行业新闻 · #9

Show HN: JevBench, a reproducible benchmark for typed decision models

Hi HN! I built JevBench because Jev kicks ass, and the world deserves to know how the serious open source and fake lookalike projects really perform in comparison. Jev-class models return bounded choices and probabilities instead of text, and are disruptively faster and cheaper than LLMs, while being similarly intelligent on the text input they operate on. JevBench allows looking at accuracy, latency and price all at once, in a weighted way - you can even configure the weighting. A full run asks 534 English decisions. The v1.3 score combines chance-corrected Intelligence, Calibration, Speed…

Show HN: JevBench, a reproducible benchmark for typed decision models

热度 49.5 / 100;排名与评分保留该期记录。