MoonshotAI: Kimi K3

Text input Image input Text output
Author's Description

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

Key Specifications
Cost
$$$$$
Context
1M
Released
Jul 16, 2026
Speed
Ability
Reliability
Supported Parameters

This model supports the following parameters:

Tool Choice Reasoning Include Reasoning Frequency Penalty Tools Stop Presence Penalty Response Format Max Tokens Structured Outputs
Features

This model supports the following features:

Structured Outputs Response Format Reasoning Tools
Performance Summary

MoonshotAI's Kimi K3, an ultra-large-scale multimodal reasoning model, demonstrates exceptional reliability with a perfect 100% success rate across all benchmarks, indicating consistent and usable responses. However, it tends to have longer response times, ranking in the 13th percentile for speed, and is positioned at premium pricing levels, ranking in the 5th percentile for cost. Despite its higher cost and slower speed, Kimi K3 exhibits strong performance across several critical areas. It achieves perfect accuracy in both General Knowledge and Ethics, with the latter also being the most accurate model at its price point and speed. The model excels in Coding (96.0% accuracy, 93rd percentile) and Instruction Following (82.0% accuracy, 86th percentile), making it well-suited for complex coding and agentic workflows. Its Reasoning capabilities are also robust at 98.0% accuracy (84th percentile). While its Hallucinations accuracy is good at 98.0%, its Email Classification (97.0% accuracy, 35th percentile) and Mathematics (93.9% accuracy, 69th percentile) performance, while solid, are not as standout compared to its other strengths. Kimi K3's primary strengths lie in its reliability and high accuracy in knowledge-intensive and complex problem-solving tasks, despite its premium cost and slower processing.

Model Pricing

Current Pricing

Feature Price (per 1M tokens)
Prompt $3
Completion $15
Input Cache Read $0.3

Price History

Available Endpoints
Provider Endpoint Name Context Length Pricing (Input) Pricing (Output)
Moonshot AI
Moonshot AI | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Nebius
Nebius | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Fireworks
Fireworks | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
BaseTen
BaseTen | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
GMICloud
GMICloud | moonshotai/kimi-k3-20260715 262K $3 / 1M tokens $15 / 1M tokens
DigitalOcean
DigitalOcean | moonshotai/kimi-k3-20260715 1M $2.85 / 1M tokens $14.3 / 1M tokens
DeepInfra
DeepInfra | moonshotai/kimi-k3-20260715 1M $2.85 / 1M tokens $14.3 / 1M tokens
Together
Together | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Moonshot AI
Moonshot AI | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Fireworks
Fireworks | moonshotai/kimi-k3-20260715 1M $4.5 / 1M tokens $22.5 / 1M tokens
Morph
Morph | moonshotai/kimi-k3-20260715 1M $2.8 / 1M tokens $14 / 1M tokens
Modal
Modal | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Modal
Modal | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Parasail
Parasail | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Wafer
Wafer | moonshotai/kimi-k3-20260715 912K $3 / 1M tokens $15 / 1M tokens
Wafer
Wafer | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Wafer
Wafer | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Morph
Morph | moonshotai/kimi-k3-20260715 1M $6 / 1M tokens $22.5 / 1M tokens
Chutes
Chutes | moonshotai/kimi-k3-20260715 1M $3 / 1M tokens $15 / 1M tokens
Benchmark Results
Benchmark Category Reasoning Strategy Free Executions Accuracy Cost Duration
Other Models by moonshotai