Claude Opus 5 leads OpenRouter’s classification task ranking by spend. We sent the same 3,080 Banking77 utterances to it and to Jev 1.13 through the Decisions API. Opus scored 84.4% to Jev’s 81.0%, and Jev answered in 175 ms at $0.11 per thousand requests against 2.3 seconds and $2.42 for Opus.
Is Jev as Accurate as Frontier Models at Classification?
calendar_today
September 22, 2026
domain
openrouter