AI Forecast Ledger

Third-party dispatches, not graded receipts

Qwen 3.8 27B matches GPT-5.6 Luna's score on Artificial Analysis index

simonwillison.net · 2026-08-22

An open-weights 27B model scored 52 on the Artificial Analysis Intelligence Index, equal to GPT-5.6 Luna (max) and one point behind GLM-5.2 and DeepSeek V4 Pro, both far larger. Benchmark figures are vendor/aggregator-reported, not independently verified.

“That's the same score as GPT-5.6 Luna (max), and just one point behind GLM-5.2 (max) and DeepSeek V4 Pro 0813 (max) - that GLM is [753B](https://huggingface.co/zai-org/GLM-5.2) and that DeepSeek is [1.7T parameters](https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-0813), and Luna is size unknown but presumably a whole lot bigger than 27B.”

source↗ · archived copy↗

← all dispatches