Somebody tested the Terra-tier claim, and it held
Tested @typesafeai's claim that their new model Jev delivered "comparable... intelligence" to GPT-5.6 Terra on "System 1" tasks. To do this, I compare both models on multiple-choice benchmarks (MMLU, GPQA, etc.). Set reasoning=none for Terra for sys 1. Result: Jev is Terra-tier.…
Why it is here
Somebody finally tested the sentence rather than the demo. He took Jev's claim of intelligence comparable to a frontier model on System One tasks and ran both on multiple-choice benchmarks with reasoning turned off — and came back with "Jev is Terra-tier", which is a stronger result than the claim it was checking.
What we checked
- the post read through X's public syndication endpoint, 2026-09-25
- like count and date read from the same endpoint the same day
What we did not check. We did not re-run the thing the post describes, so every figure on this page is the author's own and every claim is theirs — the note above is what we make of it after reading, not a measurement. The like count is a snapshot read on 2026-09-23 and it has moved since; 910 is what it said when we looked.