LLM Leaderboard · Open Weight

DeepSeek V4 Pro leaderboard — benchmarks, pricing, and comparisons.

Compare DeepSeek V4 Pro vs GPT, Claude, Gemini, DeepSeek, open-weight, and frontier AI models using public benchmark scores, token pricing, context window, and access details.

Rank #53AskClash overall score: 20.0
$1.74 / $3.48Input and output token price, when published. Context: 1M.
Visit websiteVisit the model provider's website.

DeepSeek V4 Pro benchmark snapshot

AskClash combines public LLM benchmark cells into a weighted percentile score and penalizes missing coverage so narrow rows do not dominate better-measured models.

Overall20.0
Benchmark cells8
Context1M
CreatorDeepSeek

DeepSeek V4 Pro public benchmark scores

Cached benchmark values can include HLE, GPQA, SWE-bench, SWE-Pro, SWE-Atlas, Terminal-Bench, MCP Atlas, MMMU-Pro, ARC-AGI-2, Tau2, and model-specific coding or agent scores.

RWT

7.0 score

HLE

7.7 score

GPQA

72.9 score

MATH-500

64.5 score

SWE-bench

73.6 score

SWE-Pro

52.1 score

Terminal-Bench

59.1 score

MCP Atlas

69.4 score

Finance Agent

44.1 score

MRCR

44.7 score

DeepSeek V4 Pro vs other AI models

Use these comparison links to evaluate DeepSeek V4 Pro against nearby LLMs by benchmark score, price, context window, and provider.

Related AI and tech coverage

Cached AskClash article matches that can provide release, provider, benchmark, pricing, or market context around this model.

Kimi K3, and what we can still learn from the pelican benchmark

Chinese AI lab Moonshot AI announced Kimi K3 this morning, describing it as their "most capable model to date, with 2.8 trillion parameters". It's currently available via their website and API, but an open weight release is promised "by July 27, 2026". Moonsho

Last cached leaderboard date: April 24, 2026. This model page is generated from the AskClash LLM Leaderboard cache and linked from the live leaderboard.