stepfun logo

StepFun: Step 3.7 Flash

stepfun/step-3.7-flash

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

Modalities

TextImageVideoText

In / out price

$0.2 / $1.15 per 1M

Context

262K

Released

May 28, 2026

Why use Step 3.7 Flash

  • Understands images
  • Step-by-step reasoning mode
  • Tool and function calling (agent-ready)
  • Structured (JSON) outputs

Pricing

Input$0.2 / 1M tokens
Output$1.15 / 1M tokens
Cached input (read)$0.040 / 1M tokens

List prices via OpenRouter, checked daily. Real cost depends on the provider and on reasoning tokens.

How it ranks

Coding index39.6

Source: Artificial Analysis (artificialanalysis.ai) via OpenRouter (openrouter.ai/rankings).

Cost calculator

Pick a task, set how often you run it, and compare what it costs on each model.

ModelPer runPer monthPer 1,000 runs
StepFun: Step 3.7 Flash$0.0025$2.46$2.46

Estimates from list prices via OpenRouter; real bills vary by provider and reasoning tokens.

Will it fit? Context window simulator

Each square is about one page (500 words). Colored squares are your content; grey squares are free space in the model's context window.

Your content

  • Novel (90,000 words)

Total: about 120,000 tokens (≈ 90,226 words). Token counts are estimates.

Models

StepFun: Step 3.7 Flash

262,144 tokens ≈ 395 pages

Fits: uses 46% of the window (room for 2× this much).

Write with the best model for the job

WordGPT picks and switches models for you: drafting, editing and files in one place. Free plan, no credit card.

Try it free

More from stepfun