HunterAlphaHub
OpenRouter model reference Facts from the public catalogue, dated and labelled
Jev Laya Verified
2026-09-22

JevBench published its first leaderboard

Introducing JevBench. The first benchmark for Jev class models. Original Jev by @typesafeai in the lead at 75.3. SemIf #2 at 74.6. All results at https://t.co/hFQ5fX4JEb https://t.co/nXYfCrZnwO

Why it is here

The launch of JevBench, a reproducible benchmark for this class of model, with Jev's first score on it. Whatever one thinks of a benchmark written by the community it measures, this is the artefact that made comparison possible — before it, every number in this topic came from a different ruler.

What we checked

  • the post read through X's public syndication endpoint, 2026-09-25
  • like count and date read from the same endpoint the same day

What we did not check. We did not re-run the thing the post describes, so every figure on this page is the author's own and every claim is theirs — the note above is what we make of it after reading, not a measurement. The like count is a snapshot read on 2026-09-23 and it has moved since; 226 is what it said when we looked.

Read the post on X ↗ All X posts

Filed under Research and data The Jev topic

Send us a link

A project built with Jev, a post, a video, a correction, a tip. We open the link, check it says what you said it says, and write the entry ourselves.

Required a link, and an email to reply to. Optional everything else.

Add context — all optional

One or two sentences about what it does, in your words. We write the entry ourselves.

Cost, latency, a benchmark — anything you measured. We attribute these to you.

We store what you type and email it to ourselves. No IP address, no user agent, no referrer — the same rule as the mailing list.

What happens to what you send →