Is Jev as Accurate as Frontier Models at Classification?
Artikelbild eller återanvändbart omslag för OpenRouter
Claude Opus 5 leads OpenRouter's classification task ranking by spend.
We sent the same 3,080 Banking77 utterances to it and to Jev 1.13 through the Decisions API. Opus scored 84.4% to Jev's 81.0%, and Jev answered in 175 ms at $0.11 per thousand requests against 2.3 seconds and $2.42 for Opus.
Texten är källans egen beskrivning av publiceringen. Innehållet tillhör OpenRouter.
Mer att läsa
Server-Side Code Execution Tools for AI Agents, Compared
OpenRouter för 10 tim sedan
v0.40.0
Ollama för 10 tim sedan
Google froze its open source bug bounty program due to a ‘significant rise’ in AI submissions
TechCrunch AI för 14 tim sedan
Can ‘super intelligence’ and a non-binding safety pact solve AI’s image problem?
TechCrunch AI för 14 tim sedan