Benchmarking a safety classifier against Jev
We benchmarked fx auto mode (safety) classifier with @typesafeai's Jev. tl;dr: ~5-18x faster and more accurate than 𝚐𝚙𝚝-𝟻.𝟼-𝚕𝚞𝚗𝚊, our current top choice https://t.co/3G8tpRz7AG
Why it is here
A benchmark post with a concrete result: roughly 5-18x faster on a safety classifier. Their numbers, their harness; it is listed because they published the setup.
What we checked
- post resolved through the public syndication endpoint on 2026-09-22
- like count read from that response on 2026-09-22
- poster image stays on X's CDN; nothing downloaded or rehosted
What we did not check. We did not re-run the thing the post describes, so every figure on this page is the author's own and every claim is theirs — the note above is what we make of it after reading, not a measurement. The like count is a snapshot read on 2026-09-23 and it has moved since; 650 is what it said when we looked.