Mercury 2.5 hits 770 tokens/s, but ranks #91 in intelligence
Mercury 2.5 LLM hits 770 tokens per second
Inception's Mercury 2.5 hits 770 output tokens per second on Artificial Analysis, ranking #2 overall. But its intelligence score is 12 (the post doesn't specify the max), ranking #91 out of 175 models, below the median of 13. Input costs $0.25/M tokens, output $0.75, with a 90% cache discount. Context window is 260k tokens, text-only, with reasoning. Bottom line: very fast, average smarts — good for latency-sensitive, low-cognition tasks.