15% off all models 🎉 Every model at 85% of OpenRouter list price.Browse models →

Maker: Google

Gemini 3.5 Flash

A high-efficiency Gemini model that targets strong coding and reasoning at Flash-tier speed and cost.

text
Interface model ID
gemini-3.5-flash
Context
1.049M context
Max output
Not published
API provider
2 API providers
Fast mode
0 fast routes
ApiFlux uptimeStatus unknown
Past 24hToday

Model transit availability

ApiFlux pricing by access route

USD per 1M tokens

Input
$1.2750
Output
$7.6500
  • Gemini API

    ActiveFast: Not supported

    Not published

    Cache reads are priced; a cache write price was not published.

    Input $1.2750Output $7.6500Cache read $0.1275Cache write 1-hour cache write
    1.049M Context · Not published max output
    Provider statusOfficial status

    Loading official status…

  • Vertex AI

    ActiveFast: Not supported

    Not published

    Cache reads are priced; a cache write price was not published.

    Input $1.2750Output $7.6500Cache read $0.1275Cache write 1-hour cache write
    1.049M Context · Not published max output
    Provider statusOfficial status

    Loading official status…

Pricing last updated: July 25, 2026

Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.

Suitable for

  • Text and chat requests that fit the published context window.
  • Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.

Unsuitable for

  • Native image, audio, or video generation priced with non-token units.
  • Workloads that require an unpublished cache or output-limit guarantee.

Enterprise AI Gateway for routing, securing, and observing every model call.