Skip to content

GPT-4o Mini

GPT-4o Mini is OpenAI's fast, affordable small model for focused tasks. It processes text and image inputs to generate text outputs (including structured outputs) and is particularly well suited for fine-tuning and distilling larger model results into cost-effective production solutions. Supports a 128,000-token input context with 16,384 output tokens, plus streaming, function calling, and predicted outputs. Priced at $0.15 / $0.60 per 1M input/output tokens ($0.075 cached input).

Providers
Capabilities
OpenAI
openai-byok
$0
$0
$0
Unavailable
OpenRouter
openrouter-byok
$0
$0
$0
Unavailable
Cline
cline-byok
$0
$0
—
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
GPT-4o Mini share card
GPT-4o Mini
OpenAI upstream share card
OpenAI upstream
OpenRouter upstream share card
OpenRouter upstream
Cline upstream share card
Cline upstream
Credits
Use your own key

Run GPT-4o Mini on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length128,000 tokens
Max output16,384 tokens
ArchitectureTransformer
Categorymultimodal
ReleasedJul 18, 2024
Modalities
→
Capabilities