# Mistral AI vs Moonshot AI

> Kimi K3 is the most capable open-weight model, at $3 in and $15 out. Mistral's lineup is far cheaper and far easier to self-host, but tops out at 256K context.

Canonical: https://www.subconscious.dev/compare/mistral-ai-vs-moonshot-ai · By The Subconscious Team · Updated September 30, 2026

## How they compare

Moonshot's Kimi K3 is a 2.8 trillion parameter mixture-of-experts model with native vision and a 1M context. Vals AI scored it 93.4% on SWE-bench Verified with a neutral harness, fourth overall behind closed frontier models. It costs $3 in and $15 out, with cached input at $0.30, and runs around 33 tokens per second because it always thinks. Mistral's coding model, Medium 3.5, scores 77.6% on SWE-Bench Verified by Mistral's own count and costs $1.50 in and $7.50 out, half of K3 on both sides. Large 3 and Small 4 go lower still, at $0.50 and $0.15 in. Moonshot's cheaper Kimi K2.6 costs $0.95 in and $4 out.

Self-hosting shows the gap in practice. K3 takes a 64+ accelerator cluster, and its custom license adds a commercial agreement above $20M in hosting revenue plus a branding clause at large scale. Medium 3.5 runs on as few as four GPUs with NVIDIA NIM containers, and Large 3 ships under Apache 2.0. Availability also differs: Moonshot paused new API subscriptions on July 19 after demand overran its GPUs, while Mistral sells through its API, Azure, Bedrock, Vertex AI, Snowflake Cortex and watsonx, with EU or US regions. K3 is the pick when repo-scale context and top coding quality justify the cost and latency.

## What each one does

### Mistral AI

Mistral AI is a Paris lab that sells its models through La Plateforme, its own API, and releases most of them as open weights. It consolidated the lineup in 2026. Mistral Medium 3.5, released April 28, is a dense 128B model that merges instruction following, reasoning and coding into one set of weights, and it replaced both Devstral 2 and the Magistral reasoning models. It costs $1.50 in and $7.50 out per million tokens and scores 77.6% on SWE-Bench Verified by Mistral's count. Mistral Small 4, a 119B mixture-of-experts model with 6.5B active, costs $0.15 in and $0.60 out. Mistral Large 3, a 675B MoE under Apache 2.0, runs $0.50 in and $1.50 out. All three carry a 256K context window.

### Moonshot AI

Moonshot AI is the Beijing lab behind the Kimi models. Its flagship Kimi K3 launched July 16, 2026 as a 2.8 trillion parameter mixture-of-experts model that activates 16 of 896 experts per token, with native vision and a 1M token context. It is the first open model in the 3T class, and full weights landed on Hugging Face on July 27. The hosted API costs $3 in and $15 out per million tokens, with cached input at $0.30, and it runs through an OpenAI-compatible endpoint, Kimi Code in the terminal, OpenRouter and Cloudflare Workers AI.

## Which is best, and when

### Choose Mistral AI for

- Self-hosting on four GPUs instead of a large cluster
- Cost-sensitive coding at half K3's rates
- Enterprise buying through major clouds

### Choose Moonshot AI for

- Top open-weight coding quality on large repos
- Document-heavy agents that need 1M context
- Visual agent work with native vision

## At a glance

| Attribute | Mistral AI | Moonshot AI |
|---|---|---|
| Model access | Open weights, plus closed Codestral | Open weights, custom license |
| Flagship models | Mistral Medium 3.5, Small 4, Large 3 | Kimi K3, Kimi K2.6 |
| Speed | - | ~33 tok/s on Kimi K3 |
| Price | $0.15–$1.50 in, $0.60–$7.50 out per 1M | $3 in, $15 out (Kimi K3) |
| Customization | Forge (enterprise); fine-tuning API deprecated | Open weights to fine-tune |
| Deployment | API, Azure, Bedrock, Vertex, self-host | API, Kimi Code, OpenRouter |
| Long context | 256K | 1M |

## FAQ

### What is the difference between Mistral AI and Moonshot AI?

Kimi K3 is the most capable open-weight model, at $3 in and $15 out. Mistral's lineup is far cheaper and far easier to self-host, but tops out at 256K context.

### When should I choose Mistral AI over Moonshot AI?

Self-hosting on four GPUs instead of a large cluster; Cost-sensitive coding at half K3's rates; Enterprise buying through major clouds.

### When should I choose Moonshot AI over Mistral AI?

Top open-weight coding quality on large repos; Document-heavy agents that need 1M context; Visual agent work with native vision.

### Is Mistral AI or Moonshot AI cheaper?

Mistral AI: $0.15–$1.50 in, $0.60–$7.50 out per 1M. Moonshot AI: $3 in, $15 out (Kimi K3). The cheaper choice depends on the model and workload.

### Which has more context, Mistral AI or Moonshot AI?

Mistral AI: 256K. Moonshot AI: 1M.

## Running long-horizon agents?

If your agents run past 200K tokens, compare both against Subconscious: [Subconscious vs Mistral AI](https://www.subconscious.dev/compare/subconscious-vs-mistral-ai.md), [Subconscious vs Moonshot AI](https://www.subconscious.dev/compare/subconscious-vs-moonshot-ai.md).

Full profiles: [Mistral AI](https://www.subconscious.dev/providers/mistral-ai.md), [Moonshot AI](https://www.subconscious.dev/providers/moonshot-ai.md).
