ACB
66.4 score
Compare Grok 4.7 vs GPT, Claude, Gemini, DeepSeek, open-weight, and frontier AI models using public benchmark scores, token pricing, context window, and access details.
AskClash combines public LLM benchmark cells into a weighted percentile score and penalizes missing coverage so narrow rows do not dominate better-measured models.
Cached benchmark values can include HLE, GPQA, SWE-bench, SWE-Pro, SWE-Atlas, Terminal-Bench, MCP Atlas, MMMU-Pro, ARC-AGI-2, Tau2, and model-specific coding or agent scores.
66.4 score
56.0 score
43.1 score
63.0 score
76.0 score
73.0 score
1695.2 score
49.2 score
Use these comparison links to evaluate Grok 4.7 against nearby LLMs by benchmark score, price, context window, and provider.
Cached AskClash article matches that can provide release, provider, benchmark, pricing, or market context around this model.

xAI has released Grok 4.7, its most capable model yet. But on the Artificial Analysis Intelligence Index, it scores just 46 points, landing mid-pack and well behind Claude Fable 5.1 and GPT-6 at 53 each. The gap grows even wider in agentic coding. The upside i

xAI : SpaceXAI releases Grok 4.7, which it says is better at verifying its own work and managing longer context, available for $2/1M input and $6/1M output tokens — SpaceXAI's most powerful model for coding and knowledge work. Twice as fast, at half the price
I’ve been testing Grok 4.7… it’s a really great daily driver, and a huge step up over 4.6. Definitely worth trying!
Grok 4.7 has landed. 🚀 Congrats to @SpaceXAI on its most capable model yet for coding and knowledge work. Proud to support the team with NVIDIA accelerated computing.
Last cached leaderboard date: . This model page is generated from the AskClash LLM Leaderboard cache and linked from the live leaderboard.