Qwen: Qwen3 8B

qwen
$0.06output / 1M tokens · input $0.06/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    siliconflow
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.06
    Output $ / 1M tokens
    $0.06
    Cache read $ / 1M tokens
    -86.8% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    fireworks
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.2
    Output $ / 1M tokens
    $0.2
    Cache read $ / 1M tokens
    -56.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    openrouter → qwen
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.12
    Output $ / 1M tokens
    $0.46
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    alibaba-cloud
    Class
    Official
    $ / 1M tokens
    Input $ / 1M tokens
    $0.18
    Output $ / 1M tokens
    $0.7
    Cache read $ / 1M tokens
    +53.8% vs consensus
    TTFT
    3.7s
    reasoning 3.9s
    Throughput
    39.2 tps
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    qwen
    Class
    Official
    $ / 1M tokens
    Input $ / 1M tokens
    $0.18
    Output $ / 1M tokens
    $0.7
    Cache read $ / 1M tokens
    +53.8% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    eigenai
    Class
    Direct host
    Disabled by eigenai · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.04
    Output $ / 1M tokens
    $0.2
    Cache read $ / 1M tokens
    -56.0% vs consensus
    TTFT
    765ms
    reasoning 1.9s
    Throughput
    287.2 tps
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    openrouter → atlas-cloud
    Class
    Reseller
    Disabled by openrouter · disabled · fp8
    $ / 1M tokens
    Input $ / 1M tokens
    $0.05
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.05
    -12.1% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…