15% off all models 🎉 Every model at 85% of the maker's official list price.Browse models →

Maker: Alibaba Qwen

Qwen3.5-Flash

An efficient Qwen Flash model combining linear attention with sparse mixture-of-experts routing.

textvision
Interface model ID
qwen3.5-flash
Context
1M context
Max output
Not published
API provider
1 API provider
Fast mode
0 fast routes
ApiFlux uptimeStatus unknown
Past 24hToday

Model transit availability

ApiFlux pricing by access route

USD per 1M tokens15% off

Input
$0.0553$0.0651
Output
$0.221$0.26
  • Alibaba Cloud Model Studio

    ActiveFast: Not supported

    Not published

    Explicit cache read and write prices were not published.

    Input $0.0553Output $0.221Cache read Cache write 1-hour cache write
    1M Context · Not published max output
    Provider statusOfficial status unavailable

    Official status history unavailable

Pricing last updated: July 25, 2026

Long-context pricing applies only after the threshold shown. Cache and token prices are listed per 1M tokens.

Suitable for

  • Text and image requests that fit the published context window.
  • Workloads that benefit from a unified OpenAI-compatible ApiFlux endpoint.

Unsuitable for

  • Native image, audio, or video generation priced with non-token units.
  • Workloads that require an unpublished cache or output-limit guarantee.

Enterprise AI Gateway for routing, securing, and observing every model call.