OpenRouter benchmarks Jev vs LLMs: 100x cheaper for decisions, but it can't write replies
OpenRouter 实测 Jev 决策模型对比 LLM,何时该用决策模型替代生成文本
OpenRouter benchmarked TypeSafe's Jev 1.13 against GPT Luna and Claude Opus on 60 support tickets. Jev classified and flagged escalation at $0.025 per 1,000 tickets with 194 ms median latency, versus $0.09 for Luna and $2.88 for Opus. It returns typed probabilities directly—no JSON parsing needed. Jev is text-only, weak at arithmetic, and can't generate prose. The post recommends routing with Jev first, then handing off to an LLM for replies.