GPT-4o mini
GPT-4o mini is OpenAI's newest model after GPT-4 Omni, supporting both text and image inputs with text outputs.
As their most advanced small model, it is many multiples more affordable than other recent frontier models, and more than 60% cheaper than GPT-3.5 Turbo. It maintains SOTA intelligence, while being significantly more cost-effective.
GPT-4o mini achieves an 82% score on MMLU and presently ranks higher than GPT-4 on chat preferences common leaderboards.
Check out the launch announcement to learn more.
#multimodal
- Organization
- OpenAI
- Family
- gpt-mini
- Providers
- 20
- Context
- 128,000
- Output limit
- 16,384
- Knowledge
- 2023-09
- Release
- 2024-07-18
- Updated
- 2024-07-18
- Weights
- Closed
- Input
- text, image, pdf
- Output
- text
- Capabilities
- Tools, Structured, Temperature
Providers
Compare every cataloged offer for this model, including its provider, route, token limits, price, and capabilities.
| Provider | Model ID | Context | Output | Input / output · 1M | Reasoning | Tools | Structured | Details |
|---|---|---|---|---|---|---|---|---|
| Abacus | gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Azure | gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | — | View |
| Azure Cognitive Services | gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | — | View |
| Cloudflare AI Gateway | openai/gpt-4o-mini | 128,000 | 16,384 | $0.075 / $0.30 | No | Yes | Yes | View |
| Cortecs | gpt-4o-mini | 128,000 | 16,000 | $0.159 / $0.638 | No | Yes | No | View |
| CrossModel | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| DevPass (LLM Gateway) | gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Eden AI | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Impossibl | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Kilo Gateway | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| LLM Gateway | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Merge Gateway | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| NanoGPT | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | No | No | View |
| Ofox | openai/gpt-4o-mini | 128,000 | 16,384 | $0.12 / $0.48 | No | Yes | Yes | View |
| Azurevia OpenRouter | openai/gpt-4o-mini unknown | 128,000 | 16,384 | $0.165 / $0.66 | No | No | Yes | View |
| Azurevia OpenRouter | openai/gpt-4o-mini unknown | 128,000 | 16,384 | $0.15 / $0.60 | No | No | Yes | View |
| OpenAIvia OpenRouter | openai/gpt-4o-mini unknown | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| OrcaRouter | openai/gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Pioneer | gpt-4o-mini | 128,000 | 16,384 | $0.15 / $0.60 | No | Yes | Yes | View |
| Requesty | gpt-4o-mini@eu | 128,000 | 16,000 | $0.165 / $0.66 | No | Yes | Yes | View |
| Venice AI | openai-gpt-4o-mini-2024-07-18 | 128,000 | 16,384 | $0.1875 / $0.75 | No | Yes | Yes | View |
| Vercel AI Gateway | openai/gpt-4o-mini-fast | 128,000 | 16,384 | $0.25 / $1.00 | No | Yes | Yes | View |
Prices and limits apply to each offer. A dash means the value is not provided. OpenRouter offers show the hosting provider and routing channel separately.
Pricing
Compare token rates and cache charges across available routes and hosting configurations.
Lowest listed input and output prices across the offers above; they may belong to different providers. These are base rates, before conditional pricing, discounts or additional fees.
openai/gpt-4o-miniObserved - Input
- $0.15
- Output
- $0.60
- Cache read
- $0.075
Hosting prices through OpenRouter All endpoints and cache rates
| Endpoint | Input / 1M | Output / 1M | Cache read / 1M | Cache write / 1M |
|---|---|---|---|---|
| Azureazure/swedencentral | $0.165 | $0.66 | $0.0825 | — |
| Azureazure | $0.15 | $0.60 | $0.075 | — |
| OpenAIopenai | $0.15 | $0.60 | $0.075 | — |
Each row is a hosting configuration. Open an offer in Providers for its full pricing conditions and additional fees. Endpoint observations may differ from the routing catalog's base price.
Performance
Review reported latency and output-throughput percentiles from the latest observed rolling window.
Time to first token · milliseconds
All latency percentiles 3 endpoints
| Endpoint | P50 | P75 | P90 | P99 |
|---|---|---|---|---|
| Azureazure/swedencentral | 713.5 ms | 763.5 ms | 912.2 ms | 1,761.52 ms |
| Azureazure | 1,273 ms | 1,538 ms | 1,864 ms | 3,166.27 ms |
| OpenAIopenai | 565 ms | 692 ms | 900.9 ms | 3,310.72 ms |
Output throughput · tokens per second
All throughput percentiles 3 endpoints
| Endpoint | P50 | P75 | P90 | P99 |
|---|---|---|---|---|
| Azureazure/swedencentral | 119 t/s | 127 t/s | 134.4 t/s | 145.08 t/s |
| Azureazure | 58 t/s | 71 t/s | 82 t/s | 102 t/s |
| OpenAIopenai | 71 t/s | 85 t/s | 97 t/s | 112 t/s |
Percentiles describe the observed request distribution. These are snapshots of a rolling window, not a historical time series.
Uptime
Compare successful request rates across the latest 5-minute, 30-minute, and 24-hour observations.
Reported by OpenRouter at the time of collection; rate-limited requests are excluded. A dash means no measurement was reported. This is not a live availability check.
Benchmarks
Review attributed evaluation results on each source's original scale.
openai/gpt-4o-miniObserved Sources: Artificial Analysis and Design Arena, via OpenRouter. Scores retain their original scales and are not OpenAI Suite ratings.
Detailed evaluations
| Evaluation | Configuration | Source | Results | Details |
|---|---|---|---|---|
| Composite indices | GPT-4o miniopenai/gpt-4o-mini | Artificial Analysis | Intelligence — · Coding 11.4 · Agentic — | Details |
| gpqa diamond | OpenAI: GPT-4o-miniopenai/gpt-4o-mini | OpenRouter | 43.21% | Details |
| tau bench verified airline | OpenAI: GPT-4o-miniopenai/gpt-4o-mini | OpenRouter | 28.444% | Details |
Sources: Artificial Analysis, Design Arena and OpenRouter. Evaluation configurations and scales differ; results are not interchangeable. Updated Oct 4, 2026.
Apps & session costs
Compare observed median session costs for applications using this model, grouped by conversation length.
| Application | Model route | Turns | Median cost / session | Details |
|---|---|---|---|---|
| Hermes Agent | openai/gpt-4o-mini | 10–49 turns | $0.035612 | Details |
| Hermes Agent | openai/gpt-4o-mini | 1 turn | $0.001317 | Details |
| Hermes Agent | openai/gpt-4o-mini | 50+ turns | $0.207479 | Details |
| Hermes Agent | openai/gpt-4o-mini | 2–9 turns | $0.004249 | Details |
Source: OpenRouter, as of Oct 2, 2026. Published under CC BY 4.0. These are observed costs per session, not token prices or a forecast for your workload. Apps are compared separately.
Activity
Daily prompt and completion tokens reported for this model's routes in OpenRouter's top 50. Missing days have no published value.
48,416,326,303 tokens
47,652,451,213 tokens
34,848,609,408 tokens
34,939,138,687 tokens
25,099,684,638 tokens
23,745,470,796 tokens
30,171,682,097 tokens
43,447,624,664 tokens
40,563,725,410 tokens
34,953,711,764 tokens
35,232,317,368 tokens
36,040,347,515 tokens
61,943,325,124 tokens
47,531,784,070 tokens
47,278,277,541 tokens
35,330,722,716 tokens
35,849,612,246 tokens
37,054,189,041 tokens
22,865,443,467 tokens
25,066,564,961 tokens
35,198,347,246 tokens
35,203,795,174 tokens
34,766,158,474 tokens
32,638,002,324 tokens
29,948,314,529 tokens
23,053,255,963 tokens
24,597,929,556 tokens
32,510,518,687 tokens
31,956,892,855 tokens
31,890,108,970 tokens
30,608,719,620 tokens
31,511,674,112 tokens
25,247,486,635 tokens
25,987,452,163 tokens
32,144,692,015 tokens
32,232,021,495 tokens
30,092,412,627 tokens
31,700,734,002 tokens
28,855,005,029 tokens
24,826,537,850 tokens
23,622,764,587 tokens
32,881,995,719 tokens
33,664,202,059 tokens
34,697,717,276 tokens
37,779,475,582 tokens
34,444,105,097 tokens
30,700,118,517 tokens
26,319,378,523 tokens
35,594,703,703 tokens
36,397,681,040 tokens
36,548,432,744 tokens
35,590,098,729 tokens
33,227,906,195 tokens
25,862,652,268 tokens
32,410,566,376 tokens
35,185,682,502 tokens
32,711,218,391 tokens
33,666,012,348 tokens
41,385,452,908 tokens
48,101,380,955 tokens
Daily observations · 60 days
| Date (UTC) | Reported tokens | Reported routes |
|---|---|---|
| Sep 16, 2026 | 48,101,380,955 | 1 |
| Sep 8, 2026 | 41,385,452,908 | 1 |
| Sep 4, 2026 | 33,666,012,348 | 1 |
| Sep 2, 2026 | 32,711,218,391 | 1 |
| Sep 1, 2026 | 35,185,682,502 | 1 |
| Aug 31, 2026 | 32,410,566,376 | 1 |
| Aug 30, 2026 | 25,862,652,268 | 1 |
| Aug 28, 2026 | 33,227,906,195 | 1 |
| Aug 27, 2026 | 35,590,098,729 | 1 |
| Aug 26, 2026 | 36,548,432,744 | 1 |
| Aug 25, 2026 | 36,397,681,040 | 1 |
| Aug 24, 2026 | 35,594,703,703 | 1 |
| Aug 23, 2026 | 26,319,378,523 | 1 |
| Aug 22, 2026 | 30,700,118,517 | 1 |
| Aug 21, 2026 | 34,444,105,097 | 1 |
| Aug 20, 2026 | 37,779,475,582 | 1 |
| Aug 19, 2026 | 34,697,717,276 | 1 |
| Aug 18, 2026 | 33,664,202,059 | 1 |
| Aug 17, 2026 | 32,881,995,719 | 1 |
| Aug 16, 2026 | 23,622,764,587 | 1 |
| Aug 15, 2026 | 24,826,537,850 | 1 |
| Aug 14, 2026 | 28,855,005,029 | 1 |
| Aug 13, 2026 | 31,700,734,002 | 1 |
| Aug 12, 2026 | 30,092,412,627 | 1 |
| Aug 11, 2026 | 32,232,021,495 | 1 |
| Aug 10, 2026 | 32,144,692,015 | 1 |
| Aug 9, 2026 | 25,987,452,163 | 1 |
| Aug 8, 2026 | 25,247,486,635 | 1 |
| Aug 7, 2026 | 31,511,674,112 | 1 |
| Aug 6, 2026 | 30,608,719,620 | 1 |
| Aug 5, 2026 | 31,890,108,970 | 1 |
| Aug 4, 2026 | 31,956,892,855 | 1 |
| Aug 3, 2026 | 32,510,518,687 | 1 |
| Aug 2, 2026 | 24,597,929,556 | 1 |
| Aug 1, 2026 | 23,053,255,963 | 1 |
| Jul 31, 2026 | 29,948,314,529 | 1 |
| Jul 30, 2026 | 32,638,002,324 | 1 |
| Jul 29, 2026 | 34,766,158,474 | 1 |
| Jul 28, 2026 | 35,203,795,174 | 1 |
| Jul 27, 2026 | 35,198,347,246 | 1 |
| Jul 26, 2026 | 25,066,564,961 | 1 |
| Jul 25, 2026 | 22,865,443,467 | 1 |
| Jul 24, 2026 | 37,054,189,041 | 1 |
| Jul 23, 2026 | 35,849,612,246 | 1 |
| Jul 22, 2026 | 35,330,722,716 | 1 |
| Jul 21, 2026 | 47,278,277,541 | 1 |
| Jul 20, 2026 | 47,531,784,070 | 1 |
| Jul 19, 2026 | 61,943,325,124 | 1 |
| Jul 18, 2026 | 36,040,347,515 | 1 |
| Jul 17, 2026 | 35,232,317,368 | 1 |
| Jul 16, 2026 | 34,953,711,764 | 1 |
| Jul 15, 2026 | 40,563,725,410 | 1 |
| Jul 14, 2026 | 43,447,624,664 | 1 |
| Jul 13, 2026 | 30,171,682,097 | 1 |
| Jul 12, 2026 | 23,745,470,796 | 1 |
| Jul 11, 2026 | 25,099,684,638 | 1 |
| Jul 10, 2026 | 34,939,138,687 | 1 |
| Jul 9, 2026 | 34,848,609,408 | 1 |
| Jul 8, 2026 | 47,652,451,213 | 1 |
| Jul 7, 2026 | 48,416,326,303 | 1 |
Source: OpenRouter (openrouter.ai/rankings), as of Oct 4, 2026. CC BY 4.0. Totals include only individually published routes. Tokenizers differ by provider; the aggregated “other” category is never assigned to a model.
Technical details & API
Explore specifications, supported parameters, reasoning controls, and API identifiers for each OpenRouter route.
openai/gpt-4o-miniObserved - Context
- 128,000
- Maximum output
- 16,384
- Tokenizer
- GPT
- Input
- text, image, file
- Output
- text
- Moderation
- Yes
- Knowledge cutoff
- 2023-10-31
- Added to OpenRouter
- 2024-07-18
https://openrouter.ai/api/v1openai/gpt-4o-miniopenai/gpt-4o-miniSupported parameters
frequency_penaltylogit_biaslogprobsmax_completion_tokensmax_tokenspredictionpresence_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_logprobstop_pweb_search_options
Frequently asked questions
Answers based on the model metadata and serving offers currently in the catalog.
How much does GPT-4o mini cost?
Listed rates start at $0.075 per million input tokens and $0.30 per million output tokens. Prices vary by provider and configuration; see Pricing for details.
Which providers offer GPT-4o mini?
There are 22 cataloged offers across 20 providers. Compare model IDs, prices and limits in Providers.
What is the context window?
The cataloged model context is 128,000 tokens. Each serving endpoint may apply a different limit.
Does it support tools and structured output?
Tool calling: Yes. Structured output: Yes. Support can vary by endpoint.
Catalog history · 2 recorded revisions
Metadata observations since this model was first cataloged. These are not model release versions.
- ·
9d3a23eb6f8a - ·
87b0ce2e0337