HunterAlphaHub
OpenRouter model reference Facts from the public catalogue, dated and labelled
Jev Laya Verified
2026-09-22

A 2.5-million-user product benchmarked Jev on résumés

Stanford CS PhD here 👋. I rigorously benchmark Jev/Typesafe on my consumer AI app serving 2.5 million monthly active users https://t.co/b52LE8B3yN. At HiringCafe we score user resume x job description relevance. Here's how it does 🧵

Why it is here

A Stanford PhD running a hiring product with 2.5 million monthly users benchmarked the model on the task his company actually pays for: scoring how well a résumé matches a job. Production numbers on a real workload are rare in this topic, and he shows the method.

What we checked

  • the post read through X's public syndication endpoint, 2026-09-25
  • like count and date read from the same endpoint the same day

What we did not check. We did not re-run the thing the post describes, so every figure on this page is the author's own and every claim is theirs — the note above is what we make of it after reading, not a measurement. The like count is a snapshot read on 2026-09-23 and it has moved since; 809 is what it said when we looked.

Read the post on X ↗ All X posts

Filed under Triage and routing The Jev topic

Send us a link

A project built with Jev, a post, a video, a correction, a tip. We open the link, check it says what you said it says, and write the entry ourselves.

Required a link, and an email to reply to. Optional everything else.

Add context — all optional

One or two sentences about what it does, in your words. We write the entry ourselves.

Cost, latency, a benchmark — anything you measured. We attribute these to you.

We store what you type and email it to ourselves. No IP address, no user agent, no referrer — the same rule as the mailing list.

What happens to what you send →