15% off all models 🎉 Every model at 85% of OpenRouter list price.Browse models →

Maker: Anthropic

Claude Haiku 4.5

Anthropic's fast, efficient Haiku model for responsive applications and high-volume workloads.

text
Interface model ID
claude-haiku-4.5
Context
200K context
Max output
Not published
API provider
3 API providers
Fast mode
0 fast routes
ApiFlux uptimeStatus unknown
Past 24hToday

Model transit availability

ApiFlux pricing by access route

USD per 1M tokens

Input
$0.8500
Output
$4.2500
  • Anthropic

    ActiveFast: Not supported

    Not published

    Cache pricing includes read, standard write, and one-hour write rates.

    Input $0.8500Output $4.2500Cache read $0.0850Cache write $1.06251-hour cache write $1.7000
    200K Context · Not published max output
    Provider statusOfficial status

    Loading official status…

  • Amazon Bedrock

    ActiveFast: Not supported

    Not published

    Cache pricing includes read, standard write, and one-hour write rates.

    Input $0.8500Output $4.2500Cache read $0.0850Cache write $1.06251-hour cache write $1.7000
    200K Context · Not published max output
    Provider statusOfficial status

    Loading official status…

  • Vertex AI

    ActiveFast: Not supported

    Not published

    Cache pricing includes read, standard write, and one-hour write rates.

    Input $0.8500Output $4.2500Cache read $0.0850Cache write $1.06251-hour cache write $1.7000
    200K Context · Not published max output
    Provider statusOfficial status

    Loading official status…

Pricing last updated: July 25, 2026

Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.

Suitable for

  • Text and chat requests that fit the published context window.
  • Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.

Unsuitable for

  • Native image, audio, or video generation priced with non-token units.
  • Workloads that require an unpublished cache or output-limit guarantee.

Enterprise AI Gateway for routing, securing, and observing every model call.