LLM Leaderboard · Proprietary

Grok 4.6 benchmarks — leaderboard rank, pricing, and comparisons.

Grok 4.6 ranks #17 on the AskClash LLM leaderboard with an overall score of 64.4. It scores 47.0 on the Coding Agent Index and 88.4 on Terminal-Bench. Compare Grok 4.6 vs GPT, Claude, Gemini, DeepSeek, open-weight, and frontier AI models using public benchmark scores, token pricing, context window, and access details.

Rank #17AskClash overall score: 64.4
$2.00 / $6.00Input and output token price, when published. Context: 500K.
Visit websiteVisit the model provider's website.

Grok 4.6 benchmark snapshot

AskClash combines public LLM benchmark cells into a weighted percentile score and penalizes missing coverage so narrow rows do not dominate better-measured models.

Overall64.4
Benchmark cells13
Context500K
CreatorxAI

Grok 4.6 public benchmark scores

Cached benchmark values can include HLE, GPQA, SWE-bench, SWE-Pro, SWE-Atlas, Terminal-Bench, MCP Atlas, MMMU-Pro, ARC-AGI-2, Tau2, and model-specific coding or agent scores.

ACB

69.0 score

RWT

9.0 score

Coding Agent Index

47.0 score

HLE

42.9 score

GPQA

94.9 score

SWE-Atlas

58.0 score

Terminal-Bench

88.4 score

DeepSWE

65.0 score

GDPval-AA

1605.5 score

Finance Agent

53.7 score

ARC-AGI 2

67.1 score

Grok 4.6 vs other AI models

Use these comparison links to evaluate Grok 4.6 against nearby LLMs by benchmark score, price, context window, and provider.

Related AI and tech coverage

Cached AskClash article matches that can provide release, provider, benchmark, pricing, or market context around this model.

AI/tech coverage

AskClash will attach cached AI and tech articles here as relevant coverage is collected.

Last cached leaderboard date: . This model page is generated from the AskClash LLM Leaderboard cache and linked from the live leaderboard.