Skip to content
AI HOT (Curated Pool)

OpenRouter's model fusion panel beats GPT-5.5 and Claude Opus 4.8 on deep research benchmark

OpenRouter融合预算模型面板超越GPT-5.5和Claude Opus 4.8

OpenRouter launched Fusion, which sends a prompt to multiple models in parallel and has a judge model synthesize the final answer. On 100 DRACO deep research tasks, Fable 5 + GPT-5.5 fused scored 69.0%, beating Fable 5 alone at 65.3%. A budget panel of Gemini 3 Flash, Kimi K2.6, and DeepSeek V4 Pro hit 64.7%—close to Fable 5 at roughly half the cost. The post doesn't disclose added latency or the exact per-call price for the budget panel.

Why it matters: OpenRouter's Fusion lets budget model panels beat solo frontier models on deep research via multi-model deliberation + judge. Concrete DRACO benchmark data and anti-cheat design make it worth reading. Score capped at 78 because it's a platform feature launch, not a model break...

Read the original ↗Export Markdown