Skip to content

GPT-5 Mini

GPT-5 Mini delivers near-frontier intelligence for cost-sensitive, low-latency, high-volume workloads. It supports a 400,000-token input context with 128,000 output tokens, text and image inputs, and exposes streaming, function calling, structured outputs, reasoning tokens, and tool integrations including web search, file search, code interpreter, and MCP. OpenAI recommends GPT-5.4 Mini for most new low-latency workloads.

Providers
Capabilities
OpenAI
openai-byok
$0
$0
$0
Unavailable
OpenRouter
openrouter-byok
$0
$0
—
Unavailable
Cline
cline-byok
$0
$0
—
Unavailable
Usage analytics

Loading usage…

API & code
Uptime & Health
No uptime data yet

These providers haven't been health-probed for this model yet. The router still routes around upstreams that fail live requests — uptime fills in once probe history accrues.

Share cards
GPT-5 Mini share card
GPT-5 Mini
OpenAI upstream share card
OpenAI upstream
OpenRouter upstream share card
OpenRouter upstream
Cline upstream share card
Cline upstream
Credits
Use your own key

Run GPT-5 Mini on your own key — your requests are billed by the provider. Pool callers pay AnyRouter credits.

No BYOK keys configured for this model yet.

Share a key with the pool to earn credits for every request it serves, covering your plan cost.

Text generation
Context length400,000 tokens
Max output128,000 tokens
ArchitectureTransformer
Categorymultimodal
ReleasedAug 7, 2025
Modalities
→
Capabilities