Step 3.5 Flash API pricing
Step 3.5 Flash from StepFun costs $0.10 per 1M input tokens and $0.30 per 1M output tokens.
That puts it 7th cheapest of the 60 models tracked here on a blended 3:1 input-to-output rate, or $0.150 per 1M blended tokens. Verified 2026-08-16.
- Input
- $0.10per 1M tokens
- Output
- $0.30per 1M tokens
- Cached input
- None not published
- Batch discount
- None not published
- Price rank
- #7of 60, cheapest first
- Tokens per $1
- 6,666,667at a 3:1 mix
What Step 3.5 Flash costs on a real workload
Per-million-token rates are hard to reason about, so here is the same model priced against four concrete monthly workloads. Each row uses this model's own input and output rates against a fixed token mix, with no caching or batch discount applied.
| Workload | Assumption | Monthly cost | Annual |
|---|---|---|---|
| Support chatbot | 50,000 conversations/month at 3K input and 500 output tokens each | $22.50 | $270.00 |
| RAG document search | 200,000 queries/month at 6K retrieved-context input and 400 output tokens | $144.00 | $1,728 |
| Coding agent | 5,000 runs/month at 60K input and 8K output tokens per run | $42.00 | $504.00 |
| Bulk classification | 5,000,000 items/month at 400 input and 20 output tokens each | $230.00 | $2,760 |
Your mix will differ. The API cost calculator takes your own token counts, and the cost per request calculator scales a single call up to per-1,000 and per-month totals.
Cheaper alternatives to Step 3.5 Flash
These are the closest cheaper options from other providers, ordered by how near they sit to Step 3.5 Flash on the blended rate. Closer is usually a more realistic swap: the further down this list you go, the more capability you are likely trading away.
| Model | Provider | Input / 1M | Output / 1M | Cheaper by |
|---|---|---|---|---|
| Llama 4 Scout | Meta | $0.08 | $0.30 | 1.1x |
| Nova Lite 1.0 | Amazon | $0.06 | $0.24 | 1.4x |
| Command R7B | Cohere | $0.0375 | $0.15 | 2.3x |
| Llama 3.1 8B Instant (via Groq) | Groq | $0.05 | $0.08 | 2.6x |
| Qwen3.7 Flash | Alibaba | $0.03 | $0.13 | 2.7x |
Before switching, price the move properly: the model switching savings calculator puts two models against the same workload and shows the annual difference.
Models priced near Step 3.5 Flash
If cost is roughly fixed and you are choosing on capability instead, these are the models sitting closest to Step 3.5 Flash on price.
| Model | Provider | Input / 1M | Output / 1M | Blended |
|---|---|---|---|---|
| Llama 4 Scout | Meta | $0.08 | $0.30 | $0.135 |
| GPT-4.1 nano | OpenAI | $0.10 | $0.40 | $0.175 |
| Gemini 2.5 Flash-Lite | $0.10 | $0.40 | $0.175 | |
| DeepSeek V4 Flash | DeepSeek | $0.14 | $0.28 | $0.175 |
About StepFun
StepFun is a Chinese lab shipping low-cost Flash-class models competitive with the cheapest tiers from larger providers.
See every StepFun model and how the lineup is tiered on the StepFun pricing page, or put Step 3.5 Flash against all 60 models from all 17 providers on the comparison table.
Step 3.5 Flash pricing FAQ
How much does Step 3.5 Flash cost per 1M tokens?
Step 3.5 Flash costs $0.10 per 1M input tokens and $0.30 per 1M output tokens. There is no published cached-input rate for this model.
Is Step 3.5 Flash expensive compared to other models?
It ranks 7 of 60 on a blended rate that weights input and output 3:1, so 6 tracked models are cheaper and 53 are more expensive. That works out to 2.7x the blended rate of Qwen3.7 Flash, the cheapest model tracked here.
What does a real workload cost on Step 3.5 Flash?
A support chatbot handling 50,000 conversations a month at 3K input and 500 output tokens each comes to about $22.50 a month. A coding agent doing 5,000 runs at 60K input and 8K output per run comes to about $42.00. Run your own numbers in the API cost calculator.
Why is Step 3.5 Flash output more expensive than input?
Output on this model is 3.0x the input rate. Every output token needs its own forward pass through the model, while input tokens are processed in parallel, so output costs more to serve across essentially every provider. It also means a workload's input:output mix, not just its total token count, drives the bill.
What is a cheaper alternative to Step 3.5 Flash?
Llama 4 Scout from Meta is the closest cheaper option at $0.08 / $0.30, roughly 1.1x cheaper on a blended basis. Whether it is a real substitute depends on whether your task actually needs the extra capability.