#11
Kimi K3New
kimi-k3Helpfulness
Instruction
Following
Comprehension
Empathy
Creative
Writing
Helpfulness
0.0
Instruction Following
0.0
Comprehension
0.0
Empathy
0.0
Creative Writing
0.0
Speed
Avg 58 tok/s
Release Date
July 16, 2026
Lab
Moonshot AI
Type
Open Source
Context Size
1.0M
Max Output Tokens
1.0M
Cost per 1 million tokens
$3.00 / $15.00
Model Inputs*
Text, Images, Video
Model Outputs*
Text
Tool Calling*
Enabled
Overall Assistant Score
An average score combining the 5 main categories.
90.60 pts
Rank #11
86th Percentile
0
Novice
33
Capable
66
Proficient
100
Expert
Kimi K3 is Moonshot's July 2026 open frontier flagship (2.8T MoE, 1M context, always-on reasoning) for long-horizon work—coding agents, research, and messy multi-step help. Our tests put it near the top of open assistants on helpfulness, instruction following, and comprehension, with strong tool follow-through. Empathy is capable but cooler than Claude or Grok. Creative writing pays the reasoning tax: default max thinking favors agent scratchpads over sustained voice, so our tests still prefer Claude Fable and the old GPT-4.5 peak for prose. Moonshot itself notes everyday UX still trails Fable 5 and GPT-5.6 Sol. Best open pick when you want an agent that finishes hard work; not the warmest writing partner. Scored at cost-efficient reasoning (not max showcase).
Overall Score vs Price
Price axis is logarithmic (USD per 1M input tokens) · Higher score and lower price is better
Ringed point is this model
Anthropic
DeepSeek
Google
Meta
Moonshot
OpenAI
SpaceX AI
Zhipu AI
Intelligence
Overall Score · Higher is better
Speed
Output tokens per second · Higher is better
Price
USD per 1M input tokens · Lower is better
Moonshot Models
Overall Score · Same provider
Moonshot Family
Overall Score · Same model family
Closest Rivals
Overall Score · Nearest by overall rank