Novita AI vs StepFun
StepFun is a Shanghai lab with cheap Apache 2.0 multimodal models. Novita is a San Francisco host with 200+ models that adds new open releases quickly.
By The Subconscious Team · Updated
Novita AI vs StepFun: key differences
StepFun makes models; Novita hosts other people's. StepFun's Step 3.7 Flash is a 198B mixture-of-experts vision-language model with 11B active parameters, 256K context and an Apache 2.0 license, at $0.20 in and $1.15 out on StepFun's API. Its first-party inference is China-hosted, with thin Western distribution. Novita serves 200+ open models from San Francisco and is known for day-zero support of new releases, such as launching Gemma 4 as Google's partner. Step 3.7 Flash is not named in Novita's catalog, but Novita's dedicated endpoints run any Hugging Face model.
So the practical question is where to run a Step model. Calling StepFun directly is simplest and gets reasoning levels, tool use and structured outputs as designed. Running Apache 2.0 weights on a Novita dedicated endpoint keeps inference off China-hosted servers and adds LoRA hot swapping, at a 99.5% SLA. Novita also covers text, image, video and speech generation from other labs, which StepFun does not. StepFun trails frontier models on hard multimodal reasoning.
What Novita AI and StepFun do
Novita AI
Novita AI is a San Francisco inference cloud founded in late 2023 by Frank Lewis and Junyu Huang, and it competes on price and breadth. Its serverless API covers 200+ open models across LLMs, image, video, speech, voice cloning and embeddings, with LLM prices starting at $0.02 per million tokens. The API speaks both OpenAI and Anthropic formats. It became an official Hugging Face Inference Partner in April 2026 and was the day-zero launch partner for Google's Gemma 4.
Example models: DeepSeek V4 Pro, Gemma 4
Full Novita AI profileStepFun
StepFun is a Shanghai AI lab known for efficient multimodal models, with a mix of proprietary API models and open-weight releases. Its current workhorse, Step 3.7 Flash, came out in May 2026 as a 198B mixture-of-experts vision-language model with only 11B active parameters. It has 256K context, selectable reasoning levels, tool use and structured outputs, and it ships under Apache 2.0. StepFun's own API prices it at $0.20 in and $1.15 out per million tokens, and OpenRouter carries it too.
Example models: Step 3.7 Flash, Step3
Full StepFun profileShould you choose Novita AI or StepFun?
Novita AI
Choose Novita AI for
- Hosting open weights outside China
- Mixing many labs' models on one API
- Image and video generation, not just understanding
StepFun
Choose StepFun for
- Cheap vision and video understanding at 256K
- First-party Step models with reasoning levels
- Self-hosting a small-active-parameter model
Novita AI vs StepFun at a glance
| Attribute | ||
|---|---|---|
| Model access | Open weights | Open (Apache 2.0) and API models |
| Flagship models | DeepSeek V4 Pro, Gemma 4 | Step 3.7 Flash, Step3 |
| Speed | ~36 tok/s on DeepSeek V4 Pro | ~128 tok/s on Step 3.7 Flash |
| Price | From $0.02 per 1M; batch 50% off | $0.20 in, $1.15 out (Step 3.7 Flash) |
| Customization | Hot-swappable LoRA adapters | Open weights to fine-tune |
| Deployment | Serverless, GPU cloud, dedicated | First-party API, OpenRouter |
| Long context | Full 1M on DeepSeek V4 Pro | 256K |
Frequently asked questions
What is the difference between Novita AI and StepFun?
StepFun is a Shanghai lab with cheap Apache 2.0 multimodal models. Novita is a San Francisco host with 200+ models that adds new open releases quickly.
When should I choose Novita AI over StepFun?
Hosting open weights outside China; Mixing many labs' models on one API; Image and video generation, not just understanding.
When should I choose StepFun over Novita AI?
Cheap vision and video understanding at 256K; First-party Step models with reasoning levels; Self-hosting a small-active-parameter model.
Is Novita AI or StepFun cheaper?
Novita AI: From $0.02 per 1M; batch 50% off. StepFun: $0.20 in, $1.15 out (Step 3.7 Flash). The cheaper choice depends on the model and workload.
Which has more context, Novita AI or StepFun?
Novita AI: Full 1M on DeepSeek V4 Pro. StepFun: 256K.
Related comparisons
Running long-horizon agents?
If your agents run past 200K tokens, compare both against Subconscious. Our inference stack treats a long-horizon trace as the primary workload, so speed, cost, and accuracy hold up deep into the trace.