Grok 4.7 xHigh 在該基準測試中得分58%,Claude Fable 5.1 Max 得分59%。Grok 4.7 xHigh 超越了 GPT-6 Astra、GPT-5.6、Gemini、Kimi、GLM 以及幾乎所有其他前沿模型。