Skip to content
r/LocalLLaMA

Qwen 3.8 27B Still Holds Up in New Benchmarks

New Artificial Analysis scores but Qwen 3.8 27b Still holds up

Artificial Analysis released v4.2 of its Intelligence Index, and Qwen 3.8 27B still holds its ground. The update adds an agentic knowledge work eval and 4,592-page long-context reasoning, while dropping the saturated GPQA Diamond. Some users say Muse 1.3 doesn't match DeepSeek Flash in practice; others prefer Muse 1.3 over any previous DeepSeek. The post doesn't spell out exact score changes or rankings.

Read the original ↗Export Markdown