LLM Leaderboard · Open Weight

Kimi K3 leaderboard — benchmarks, pricing, and comparisons.

Compare Kimi K3 vs GPT, Claude, Gemini, DeepSeek, open-weight, and frontier AI models using public benchmark scores, token pricing, context window, and access details.

Rank #4AskClash overall score: 77.3
$3.00 / $15.0Input and output token price, when published. Context: 1M.
Visit websiteVisit the model provider's website.

Kimi K3 benchmark snapshot

AskClash combines public LLM benchmark cells into a weighted percentile score and penalizes missing coverage so narrow rows do not dominate better-measured models.

Overall77.3
Benchmark cells11
Context1M
CreatorMoonshot AI

Kimi K3 public benchmark scores

Cached benchmark values can include HLE, GPQA, SWE-bench, SWE-Pro, SWE-Atlas, Terminal-Bench, MCP Atlas, MMMU-Pro, ARC-AGI-2, Tau2, and model-specific coding or agent scores.

RWT

8.0 score

HLE

43.5 score

GPQA

93.5 score

Terminal-Bench

88.3 score

DeepSWE

68.5 score

MCP Atlas

84.2 score

Finance Agent

54.4 score

CharXiv

84.8 score

MMMU-Pro

81.6 score

Kimi K3 vs other AI models

Use these comparison links to evaluate Kimi K3 against nearby LLMs by benchmark score, price, context window, and provider.

Related AI and tech coverage

Cached AskClash article matches that can provide release, provider, benchmark, pricing, or market context around this model.

How a Frontier Model Gets Built, Read from the Kimi K3 Report

An open, 2.8-trillion-parameter model shipped with 47 pages of its own recipe. Reading it tells you what building a frontier model now involves, and how little of it is the model. The post How a Frontier Model Gets Built, Read from the Kimi K3 Report appeared

Last cached leaderboard date: July 16, 2026. This model page is generated from the AskClash LLM Leaderboard cache and linked from the live leaderboard.