OpenAI
GPT-4o Mini
gpt-4o-miniContext
128K
Input / 1M tokens
$0.1500
Output / 1M tokens
$0.6000
Cached input / 1M tokens
$0.0863
Cache write / 1M tokens
β (bills at input price)
About this model
GPT-4o Mini
GPT-4o Mini β the cheap, fast OpenAI tier. Strong for most tasks where you don't need Opus/Sonnet quality. The standard pairing for routers that escalate hard requests to a larger model.
Key facts
- Maker
- OpenAI
- Context window
- 128K tokens
- Max output
- 16K tokens
- Lane
- Guarded: the provider has confirmed it does not train on your data
- Modality
- Text and image input
- Input price
- $0.1500 per 1M tokens
- Output price
- $0.6000 per 1M tokens
- Processing region(s)
- us
- Route
- Via OpenAI's API
- Cheapest plan
- Available from the Pro plan
- Price comparison
- Priced below 196 of 236 pro-band guarded models on Anoman (blended price: 1 part input to 3 parts output).
Released
2024-07
Training cutoff
2023-10
Parameters
Not disclosed
Ways to call
Ways to call GPT-4o Mini
Each model ID below is callable through the Anoman AI API. Price and processing region depend on the route.
| Model ID | Route | Input / output per 1M tokens | Region(s) | Lane | Context |
|---|---|---|---|---|---|
| gpt-4o-mini | Via OpenAI's API | $0.1500 / $0.6000 | us | Guarded | 128K |
| gpt-4o-mini-openrouter | Via OpenRouter | $0.1500 / $0.6000 | global | Guarded | 128K |
Best for
Use cases
- High-volume chat where cost matters
- Classification + extraction
- Routing in front of larger models
- Cheap-tier customer support
Strengths
What it does well
- Sub-second TTFT
- Same tool calling reliability as GPT-4o
- 128K context window despite small price
- Strong default for OpenAI-shaped workflows
Limitations
Know the trade-offs
- Reasoning + code trails Claude Haiku 4.5 / Qwen 2.5 Coder
- More hallucination than larger models
- No vision β must use GPT-4o for multimodal
Benchmarks
Published scores
Scores from official model cards and public leaderboards. Higher is better unless noted.
| Benchmark | Score | Measures |
|---|---|---|
| MMLU | 82.0 | General knowledge across 57 subjects |
| HumanEval | 87.2 | Python code generation, pass@1 |
| MATH | 70.2 | Mathematics, mixed difficulty |
| MGSM | 87.0 | Multilingual grade school math |
Code example
Chat completion with GPT-4o Mini
from openai import OpenAI
client = OpenAI(
base_url="https://api.anoman.io/v1",
api_key="anm-sk-..."
)
response = client.chat.completions.create(
model="gpt-4o-mini",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)Alternatives
Similar models in our catalog
Same maker first, then the closest price to GPT-4o Mini in the Anoman catalog.