Qwen3-VL-32B-Thinking

qwen
$0.64output / 1M tokens · input $0.16/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    alibaba-cloud
    Class
    Official
    $ / 1M tokens
    Input $ / 1M tokens
    $0.16
    Output $ / 1M tokens
    $0.64
    Cache read $ / 1M tokens
    -40.2% vs consensus
    TTFT
    2.5s
    reasoning 2.7s
    Throughput
    89.3 tps
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    siliconflow
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.2
    Output $ / 1M tokens
    $1.5
    Cache read $ / 1M tokens
    +40.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…