15% off all models 🎉 Every model at 85% of OpenRouter list price.Browse models →
Maker: OpenAI
GPT-4o mini
A compact GPT-4o model designed for fast, affordable general-purpose and multimodal tasks.
- Interface model ID
- gpt-4o-mini
- Context
- 128K context
- Max output
- Not published
- API provider
- 2 API providers
- Fast mode
- 0 fast routes
ApiFlux uptimeStatus unknown
Model transit availability
ApiFlux pricing by access route
USD per 1M tokens
- Input
- $0.1275
- Output
- $0.5100
OpenAI
Not published
Cache writes are free on the reviewed route; other routes may differ.
Input $0.1275Output $0.5100Cache read $0.0638Cache write $0.00001-hour cache write —128K Context · Not published max outputProvider statusOfficial statusAzure OpenAI
Not published
Cache writes are free on the reviewed route; other routes may differ.
Input $0.1275Output $0.5100Cache read $0.0638Cache write $0.00001-hour cache write —128K Context · Not published max outputProvider statusOfficial status
Pricing last updated: July 25, 2026
Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.
Suitable for
- Text and chat requests that fit the published context window.
- Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.
Unsuitable for
- Native image, audio, or video generation priced with non-token units.
- Workloads that require an unpublished cache or output-limit guarantee.