Qwen: Qwen3 Max

qwen
$3.9output / 1M tokens · input $0.78/M

Best price — Output ($/1M tokens)

Context pricing

Schedule for qwen

The headline price is the base bracket; longer input context is billed at the rates below.

ContextInput $/MOutput $/MCache read $/MSource
≤32Kbase$1.2$6$0.12qwen-cloud
32K–128K$2.4$12$0.24qwen-cloud
>128K$3$15$0.3qwen-cloud

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    openrouter → qwen
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.78
    Output $ / 1M tokens
    $3.9
    Cache read $ / 1M tokens
    $0.16
    -35.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    qwen
    Class
    Official
    $ / 1M tokens
    Input $ / 1M tokens
    $1.2
    Output $ / 1M tokens
    $6
    3 context tiers
    Cache read $ / 1M tokens
    $0.12
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    deepinfra
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $1.2
    Output $ / 1M tokens
    $6
    Cache read $ / 1M tokens
    $0.24
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    deepinfra → qwen
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $1.2
    Output $ / 1M tokens
    $6
    Cache read $ / 1M tokens
    $0.24
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    novita → qwen
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $2.11
    Output $ / 1M tokens
    $8.45
    Cache read $ / 1M tokens
    +40.8% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    alibaba-cloud
    Class
    Official
    Disabled by alibaba-cloud · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $1.2
    Output $ / 1M tokens
    $6
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    2.5s
    Throughput
    62.5 tps
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    novita
    Class
    Direct host
    Disabled by novita · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $2.11
    Output $ / 1M tokens
    $8.45
    Cache read $ / 1M tokens
    +40.8% vs consensus
    TTFT
    116.2s
    Throughput
    46.6 tps
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…