# Moonshot AI vs RunInfra

> A frontier open-weight lab against a young host of mid-size models with $10 monthly coding plans. K3 wins on capability; RunInfra wins on cost and deployment automation.

Canonical: https://www.subconscious.dev/compare/moonshot-ai-vs-runinfra · By The Subconscious Team · Updated September 30, 2026

## How they compare

RunInfra and Moonshot both court developers who live in agent CLIs, at very different price and quality points. RunInfra's coding plans start at $10 a month with limits that reset every five hours and weekly, and they plug into Claude Code, Codex, OpenCode, Cline, Aider and dozens more. Its hosted library is tiny and centered on mid-size models like Nemotron 3.5 Lightning 30B and Qwen 3.8 27B, far from frontier quality. Moonshot bills Kimi K3 per token at $3 in and $15 out, and K3 sits near the top of independent coding and intelligence rankings.

RunInfra's other product is an agent that builds deployments. It picks a model, benchmarks it across GPUs from L4 to B200, searches quantized variants, and ships an OpenAI-compatible endpoint that scales to zero with cold starts under two seconds. Paid plans accept custom uploads up to 50 GB, so the service is built for models far smaller than K3. Moonshot's weak points are speed, around 33 tokens per second, and self-hosting that takes a 64+ accelerator cluster. RunInfra's are youth and a thin track record.

## What each one does

### Moonshot AI

Moonshot AI is the Beijing lab behind the Kimi models. Its flagship Kimi K3 launched July 16, 2026 as a 2.8 trillion parameter mixture-of-experts model that activates 16 of 896 experts per token, with native vision and a 1M token context. It is the first open model in the 3T class, and full weights landed on Hugging Face on July 27. The hosted API costs $3 in and $15 out per million tokens, with cached input at $0.30, and it runs through an OpenAI-compatible endpoint, Kimi Code in the terminal, OpenRouter and Cloudflare Workers AI.

### RunInfra

RunInfra pitches open models built for agents, with two ways in. Its hosted Model APIs serve a small curated library, including Nemotron 3.5 Lightning 30B, Qwen 3.8 27B and Ornith 1.5 35B, behind one key that works with both the OpenAI and Anthropic SDKs. Cached context bills at a discount. Coding plans start at $10 a month with limits that reset every five hours and every week, and they plug into Claude Code, Codex, OpenCode, Cline, Aider and dozens of other agent CLIs.

## Which is best, and when

### Choose Moonshot AI for

- Hard coding tasks that mid-size models fail
- Repo-scale agents that need 1M context
- Teams that want top open-weight quality

### Choose RunInfra for

- Cheap flat-rate coding in popular agent CLIs
- Small teams deploying a tuned mid-size model
- Voice pipelines that chain speech, LLM and TTS

## At a glance

| Attribute | Moonshot AI | RunInfra |
|---|---|---|
| Model access | Open weights, custom license | Open weights |
| Flagship models | Kimi K3, Kimi K2.6 | Nemotron 3.5 Lightning 30B, Qwen 3.8 27B |
| Speed | ~33 tok/s on Kimi K3 | Cold starts under 2s |
| Price | $3 in, $15 out (Kimi K3) | Coding plans from $10 a month |
| Customization | Open weights to fine-tune | Uploads up to 50 GB; auto-quantization |
| Deployment | API, Kimi Code, OpenRouter | Model APIs, agent-built endpoints |
| Long context | 1M | Varies by model |

## FAQ

### What is the difference between Moonshot AI and RunInfra?

A frontier open-weight lab against a young host of mid-size models with $10 monthly coding plans. K3 wins on capability; RunInfra wins on cost and deployment automation.

### When should I choose Moonshot AI over RunInfra?

Hard coding tasks that mid-size models fail; Repo-scale agents that need 1M context; Teams that want top open-weight quality.

### When should I choose RunInfra over Moonshot AI?

Cheap flat-rate coding in popular agent CLIs; Small teams deploying a tuned mid-size model; Voice pipelines that chain speech, LLM and TTS.

### Is Moonshot AI or RunInfra cheaper?

Moonshot AI: $3 in, $15 out (Kimi K3). RunInfra: Coding plans from $10 a month. The cheaper choice depends on the model and workload.

### Which has more context, Moonshot AI or RunInfra?

Moonshot AI: 1M. RunInfra: Varies by model.

## Running long-horizon agents?

If your agents run past 200K tokens, compare both against Subconscious: [Subconscious vs Moonshot AI](https://www.subconscious.dev/compare/subconscious-vs-moonshot-ai.md), [Subconscious vs RunInfra](https://www.subconscious.dev/compare/subconscious-vs-runinfra.md).

Full profiles: [Moonshot AI](https://www.subconscious.dev/providers/moonshot-ai.md), [RunInfra](https://www.subconscious.dev/providers/runinfra.md).
