Grok 4.6이 61점으로 비용 효율 부문에서 주요 경쟁자와 어깨를 나란히 했습니다.
Grok 4.6이 61점으로 성과 평가에서 GPT-5.6 Sol과 동점에 올라섰습니다. 이는 Opus 5와 Fable 5에 이어 상위권에 해당하며, Grok 4.5에 비해 한 달 만에 5점 상승한 성과입니다. 또한, 에이전트 성능에서도 Elo 1753으로 Claude Opus 5 다음에 위치하며, Qwen3.8 Max와 Fable 5와 신뢰구간이 겹치는 수준에 도달했습니다.
Grok 4.6 scores 61, placing it among top competitors in cost efficiency.
Grok 4.6 has reached a score of 61, tying with GPT-5.6 Sol in performance evaluation. This places it among the top performers, just below Opus 5 and Fable 5. It has improved by 5 points from Grok 4.5 in just one month, and in agent performance, it achieved an Elo rating of 1753, ranking just after Claude Opus 5 and overlapping with Qwen3.8 Max and Fable 5 in confidence intervals.