← All Models

Kimi K3

Kimi K3 is Moonshot AI’s 2.8T-parameter flagship model for long-horizon agentic work, with native vision, always-on reasoning, and a 1M-token context window.

Run with LM Studio Cloud
Input$3Cached$0.30Output$15

About Kimi K3

Kimi K3 is Moonshot AI's most capable flagship model to date, built for long-horizon agentic work in LM Studio Bionic. The 2.8-trillion-parameter model is the first open model to reach the 3T-parameter class.

At a glance

  • 1M-token context window for large codebases, deep research, and large document collections
  • Native vision support for understanding images, interfaces, and screenshots
  • Always-on reasoning with low, high, and max effort levels
  • Long-horizon agentic performance for complex work spanning many files and steps

Coding and visual creation

Kimi K3 is designed for repository-scale coding, architecture work, and complex debugging. Its native vision support enables visual work such as frontend development: it can inspect screenshots, modify a codebase, and reason over the resulting interface.

Moonshot AI coding benchmark comparison for Kimi K3

Coding benchmark results as reported by Moonshot AI. Source.

General and visual agents

For research and knowledge work, Kimi K3 can analyze and synthesize across large collections of files, extract findings, create reports and presentations, and sustain reasoning across long-running tasks.

Moonshot AI general-agent and visual-agent benchmark comparisons for Kimi K3

General-agent and visual-agent benchmark results as reported by Moonshot AI. Source.

Full Benchmark Table

The following results are reproduced from Moonshot AI's Kimi K3: Open Frontier Intelligence. Scores, harnesses, comparison settings, and model names are reported by Moonshot AI; see the source article's footnotes for complete methodology and caveats.

CategoryBenchmarkHarness / noteKimi K3
(max)
Claude Fable 5
(max, with fallback)
GPT 5.6 Sol
(max)
Claude Opus 4.8
(max)
GPT 5.5
(xhigh)
GLM-5.2
(max)
CodingDeepSWEKimi Code67.570.073.059.067.046.2
CodingProgram BenchKimi Code77.876.877.671.970.863.7
CodingTerminal Bench 2.1Kimi Code88.384.688.884.683.482.7
CodingFrontierSWEDominance as of 26/7/16 · Kimi Code81.286.671.366.764.967.3
CodingSWE MarathonClaude Code42.035.039.040.014.013.0
CodingPostTrain BenchClaude Code36.641.434.634.128.434.3
CodingMLS BenchKimi Code48.349.946.242.835.540.4
CodingKimi Code Bench 2.0 (Internal)72.976.964.871.769.064.2
AgenticGDPval-AA v2 (Elo-score)166817601748160014941514
AgenticBrowseComp91.288.090.484.384.4
AgenticDeepSearchQA (f1-score)95.094.293.1
AgenticToolathlon-Verified73.277.974.976.273.559.9
AgenticMCP Atlas84.284.783.683.682.882.6
AgenticAutomation Bench30.829.129.727.222.712.9
AgenticJob Bench52.957.446.548.438.343.4
AgenticAA-Briefcase (Elo-score)154815831495135411581260
AgenticAPEX-Agents41.043.339.939.438.535.6
AgenticOffice QA Pro63.369.9*63.2*63.9*60.9*41.4
AgenticSpreadsheetBench 234.834.7*32.4*31.55*29.05*28.12
AgenticDECK-Bench (Internal)73.573.074.766.968.268.6
Reasoning & KnowledgeGPQA-Diamond93.592.694.191.093.591.2
Reasoning & KnowledgeHLE-Full43.553.344.549.8*41.4*
Reasoning & KnowledgeHLE-Full w/ tools56.063.058.057.9*52.2*
VisionMMMU-Pro81.681.283.078.981.2
VisionMMMU-Pro w/ python83.486.584.682.783.2
VisionCharXiv (RQ)84.888.984.680.584.1
VisionCharXiv (RQ) w/ python91.393.589.189.989.0
VisionMathVision94.394.895.886.792.2
VisionMathVision w/ python97.898.697.897.196.8
VisionBabyVision w/ python85.790.588.981.283.6
VisionZeroBench_main (pass@5)23.023.017.017.022.0
VisionZeroBench_main w/ python (pass@5)41.046.035.034.041.0
VisionWorldVQA ForceAnswer51.056.741.839.138.5
VisionOmniDocBench91.189.885.887.989.4
VisionPerceptionBench58.557.259.747.255.8

Architecture

Kimi K3 combines Kimi Delta Attention and Attention Residuals with a highly sparse Stable LatentMoE architecture, activating 16 of 896 experts. Moonshot AI also uses quantization-aware training with MXFP4 weights and MXFP8 activations for broad hardware compatibility.

Use Kimi K3 in LM Studio Bionic

Choose Kimi K3 from the Cloud model picker in Bionic. LM Studio Cloud inference servers are US-based, and Zero Data Retention is enabled by default: your prompts and outputs are not retained or used for training.

Read LM Studio's Kimi K3 announcement · Read Moonshot AI's Kimi K3 technical blog