LLM Leaderboard · Proprietary

Grok 4.6 leaderboard — benchmarks, pricing, and comparisons.

Compare Grok 4.6 vs GPT, Claude, Gemini, DeepSeek, open-weight, and frontier AI models using public benchmark scores, token pricing, context window, and access details.

Rank #5AskClash overall score: 76.6
$2.00 / $6.00Input and output token price, when published. Context: 500K.
Visit websiteVisit the model provider's website.

Grok 4.6 benchmark snapshot

AskClash combines public LLM benchmark cells into a weighted percentile score and penalizes missing coverage so narrow rows do not dominate better-measured models.

Overall76.6
Benchmark cells13
Context500K
CreatorxAI

Grok 4.6 public benchmark scores

Cached benchmark values can include HLE, GPQA, SWE-bench, SWE-Pro, SWE-Atlas, Terminal-Bench, MCP Atlas, MMMU-Pro, ARC-AGI-2, Tau2, and model-specific coding or agent scores.

RWT

9.0 score

Coding Agent Index

76.4 score

HLE

40.3 score

GPQA

94.9 score

SWE-Pro

64.7 score

SWE-Atlas

83.9 score

Terminal-Bench

88.4 score

DeepSWE

65.9 score

Finance Agent

48.3 score

MMMU-Pro

80.4 score

ARC-AGI 2

52.6 score

Grok 4.6 vs other AI models

Use these comparison links to evaluate Grok 4.6 against nearby LLMs by benchmark score, price, context window, and provider.

Related AI and tech coverage

Cached AskClash article matches that can provide release, provider, benchmark, pricing, or market context around this model.

Congrats to the @SpaceXAI team on the release of Grok 4.6.

Congrats to the @SpaceXAI team on the release of Grok 4.6. Grok 4.6 brings frontier intelligence, running and trained on NVIDIA GB300 NVL72 with NVLink to deliver exceptional performance, reliability and lowest token cost. https://t.co/WIp3iQM35T

Last cached leaderboard date: August 12, 2026. This model page is generated from the AskClash LLM Leaderboard cache and linked from the live leaderboard.