Skip to content

Step 3.5 Flash

Step 3.5 Flash is a sparse Mixture-of-Experts model by StepFun with 196.81B total parameters (196B backbone + 0.81B MTP head) and ~11B active per token. Built on a 45-layer transformer with 288 routed experts (Top-8 selection), 3:1 SWA attention ratio, and 256K context. Achieves 100–300 tok/s throughput (peaking at 350 tok/s for coding) for frontier reasoning and agentic tasks.

Providers
Capabilities
Nous Research
nousresearch-byok
$0
$0
Unavailable
StepFun
stepfun-byok
$0
$0
Unavailable
StepFun
stepplan-byok
$0
$0
Unavailable
CommandCode
commandcode-byok
$0
$0
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
Step 3.5 Flash share card
Step 3.5 Flash
Nous Research upstream share card
Nous Research upstream
StepFun upstream share card
StepFun upstream
StepFun upstream share card
StepFun upstream
CommandCode upstream share card
CommandCode upstream
Credits
Use your own key

Run Step 3.5 Flash on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length256,000 tokens
Max output204,800 tokens
ArchitectureTransformer
Categorytext
ReleasedJun 1, 2026
Modalities
Capabilities