15% off all models 🎉 Every model at 85% of OpenRouter list price.Browse models →

Maker: DeepSeek

DeepSeek V4 Flash

An efficiency-focused DeepSeek mixture-of-experts model for fast inference and long-context work.

text
Interface model ID
deepseek-v4-flash
Context
1.049M context
Max output
Not published
API provider
1 API provider
Fast mode
0 fast routes
ApiFlux uptimeStatus unknown
Past 24hToday

Model transit availability

ApiFlux pricing by access route

USD per 1M tokens

Input
$0.0765
Output
$0.1530
  • DeepSeek

    ActiveFast: Not supported

    Not published

    Cache read and write prices reflect the reviewed route; other routes may differ.

    Input $0.0765Output $0.1530Cache read $0.0153Cache write $0.07651-hour cache write
    1.049M Context · Not published max output
    Provider statusOfficial status

    Loading official status…

Pricing last updated: July 25, 2026

Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.

Suitable for

  • Text and chat requests that fit the published context window.
  • Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.

Unsuitable for

  • Native image, audio, or video generation priced with non-token units.
  • Workloads that require an unpublished cache or output-limit guarantee.

Enterprise AI Gateway for routing, securing, and observing every model call.