15% off all models 🎉 Every model at 85% of OpenRouter list price.Browse models →

Maker: OpenAI

GPT-4o mini

A compact GPT-4o model designed for fast, affordable general-purpose and multimodal tasks.

text
Interface model ID
gpt-4o-mini
Context
128K context
Max output
Not published
API provider
2 API providers
Fast mode
0 fast routes
ApiFlux uptimeStatus unknown
Past 24hToday

Model transit availability

ApiFlux pricing by access route

USD per 1M tokens

Input
$0.1275
Output
$0.5100
  • OpenAI

    ActiveFast: Not supported

    Not published

    Cache writes are free on the reviewed route; other routes may differ.

    Input $0.1275Output $0.5100Cache read $0.0638Cache write $0.00001-hour cache write
    128K Context · Not published max output
    Provider statusOfficial status

    Loading official status…

  • Azure OpenAI

    ActiveFast: Not supported

    Not published

    Cache writes are free on the reviewed route; other routes may differ.

    Input $0.1275Output $0.5100Cache read $0.0638Cache write $0.00001-hour cache write
    128K Context · Not published max output
    Provider statusOfficial status

    Loading official status…

Pricing last updated: July 25, 2026

Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.

Suitable for

  • Text and chat requests that fit the published context window.
  • Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.

Unsuitable for

  • Native image, audio, or video generation priced with non-token units.
  • Workloads that require an unpublished cache or output-limit guarantee.

Enterprise AI Gateway for routing, securing, and observing every model call.