Kimi

Excellent

by Moonshot AI · founded 2023 · updated Jul 2026

Kimi is Moonshot AI's assistant and open-weight model family (the K2 line), built as a trillion-parameter mixture-of-experts that activates a fraction of its weights per token. It punches at the frontier on agentic tool use and coding while pricing far below Western closed labs, and the open weights let teams self-host or fine-tune for full data control.

LLMs & AssistantsOpen SourceEstablishedFrom Free chat / from $0.60 per 1M tokens (API)
84.0TIP Score
Excellent

The score, taken apart

Fixed weights, sources attached, the formula

Cap 88Val 93Sec 68Int 80Mat 78Mom 91
Capability(30%)88
Value for Money(15%)93
Security & Compliance(15%)68
Integrations & Ecosystem(15%)80
Maturity & Reliability(10%)78
Momentum(15%)91

Best for

  • ◆Cost-sensitive frontier LLM workloads
  • ◆Self-hosted agentic coding
  • ◆Teams fine-tuning open models

Key integrations

OpenAI-compatible APIOpenRouterOpen weights (Hugging Face)MCP / agentic tool usevLLM

Strengths

  • +Frontier-class agentic coding — K2.7-Code leads Opus 4.8 on MCP-Mark Verified
  • +Trillion-parameter MoE with open weights for self-hosting and fine-tuning
  • +256K context and API pricing that undercuts most closed competitors many times over

Limitations

  • -China-based hosting raises data-residency and compliance questions for regulated buyers
  • -Trails the very top on some pure-reasoning benchmarks (GPQA-Diamond, AIME) versus GPT-5.4
  • -English-language ecosystem, docs and enterprise support thinner than US labs

Evidence behind this score

Each score is built from the claims below. This is a snapshot, so check the current terms with the vendor before you commit to anything.

1 of 3 claims below are past their re-check window. Pricing, availability and security details move fast, so confirm those with the vendor.

ClaimSourceChecked
Kimi K2.6 priced at $0.60 input / $2.50 output per 1M tokens on the official API· re-checkOpenRouter / Moonshot pricingJun 2026
K2.7-Code (1T total / 32B active MoE, 256K context) beats Opus 4.8 on MCP-Mark Verified, 81.1 vs 76.4OpenRouter model card & MarkTechPostJun 2026
K2.6 trails GPT-5.4 on GPQA-Diamond (90.5% vs 92.8%) and AIME 2026 (96.4% vs 99.2%)llm-stats.com benchmarksJun 2026

Intelligence on Kimi

Releasehigh impactJul 26, 2026

Kimi K3's 2.8T open weights land — the largest open model yet, if you can host it

Moonshot released the full Kimi K3 weights for free download a day ahead of its 27 July target: 2.8 trillion parameters with 104B active, a ~1M-token context window, and roughly 1.4TB of fast memory required even at four-bit MXFP4 precision, for which Moonshot recommends at least 64 accelerators. Together AI and Modal had day-0 hosting. Near-frontier coding quality is now self-hostable, which removes the objection for anyone who was blocked on sending data to a Chinese API.

KimiMoonshot AI release on Hugging Face; Quartz and TechTimes coverage
Releasehigh impactJul 16, 2026

Moonshot's Kimi K3 pushes open weights into 3T-parameter territory

Kimi K3 is a 2.8 trillion-parameter mixture-of-experts that trails only the newest Claude and GPT flagships on benchmarks while undercutting them sharply on price, with open weights promised days after launch. For anyone weighing self-hosting against an API bill, the gap that made that choice easy is narrowing.

KimiMoonshot AI announcement
Adoptionmedium impactJul 11, 2026

Kimi K2.7 Code becomes first open-weight model in Copilot's picker

GitHub added Moonshot's Kimi K2.7 Code to Copilot's model picker — the first open-weight model offered there. Copilot's model line-up now spans OpenAI, Anthropic, Google and open weights. That is a cheaper agentic-coding option inside a tool most teams already have.

GitHub CopilotKimiGitHub changelog
Releasemedium impactJul 10, 2026

Moonshot's Kimi K2.7-Code beats Opus 4.8 on agentic coding at open-weight prices

Moonshot AI's K2 line — a trillion-parameter mixture-of-experts with 256K context and open weights — is rattling the coding-model market: K2.7-Code leads Opus 4.8 on MCP-Mark Verified (81.1 vs 76.4) while the API prices from $0.60 per 1M tokens. A strong option for cost-sensitive and self-hosted agentic coding, with data-residency caveats for regulated buyers.

KimiDeepSeekClaude CodeOpenRouter, MarkTechPost & Moonshot pricing

Alternatives in LLMs & Assistants

Compare them →

Anthropic

91.0Exceptional

Frontier assistant known for reasoning depth, long context and reliability.

LLMs & AssistantsFreemiumMature
From Free / $20 per month (Pro)

OpenAI

90.7Exceptional

The most widely adopted general-purpose AI assistant.

LLMs & AssistantsFreemiumMature
From Free / $20 per month (Plus)

Google

90.4Exceptional

Google's multimodal frontier models woven through Workspace and Android.

LLMs & AssistantsFreemiumMature
From Free / $19.99 per month (AI Pro)