ZTC Models: JEV ecosystems
Answer verification from one forward pass, zero generated tokens. Measured against JEV and 10 others on 2,018 identical items.
-
Text Generation • 404B • Updated • 1.04k • 31 -
FINAL-Bench/ZTC-Judge-27B
Text Classification • 28B • Updated • 263 • 37
FINAL-Bench/ZTC-Judge-9B
Text Classification • 10B • Updated • 1.35k • 31Note Size ladder rung. Leaderboard AUC 0.6506 — above the surface baseline overall, and clearly so on professional exams.
FINAL-Bench/ZTC-Judge-4B
Text Classification • 5B • Updated • 80 • 32Note Size ladder rung. Leaderboard AUC 0.6360 — above the surface baseline overall, and clearly so on professional exams.
-
Gate Arcade
🚦31Plain vs JEV vs ZTC: a gate decides execute or hold
Typed Decision Leaderboard
🎯35Answer verifiers on one identical test set
Note Independent leaderboard — 13 answer verifiers on 2,018 identical items. ZTC (397B) 0.7364 leads; the gap to second place is not statistically distinguishable.
Verifier Playground
🧪31Run ZTC, JEV and Laya on the same question and answer
Note Paste a question and an answer; ZTC, JEV and Laya score it side by side on the same input.