DeepSeek: DeepSeek V3

deepseek
$0.28output / 1M tokens · input $0.14/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    deepseek
    Class
    Official
    $ / 1M tokens
    Input $ / 1M tokens
    $0.14
    Output $ / 1M tokens
    $0.28
    Cache read $ / 1M tokens
    $0.03
    -66.9% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    gmi
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.21
    Output $ / 1M tokens
    $0.31
    Cache read $ / 1M tokens
    -63.4% vs consensus
    TTFT
    2.6s
    Throughput
    51.6 tps
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    deepinfra
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.26
    Output $ / 1M tokens
    $0.38
    Cache read $ / 1M tokens
    -55.0% vs consensus
    TTFT
    1.5s
    Throughput
    23.7 tps
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    novita
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.27
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    -52.7% vs consensus
    TTFT
    3.1s
    reasoning 2.0s
    Throughput
    39.5 tps
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    siliconflow
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.27
    Output $ / 1M tokens
    $0.42
    Cache read $ / 1M tokens
    -50.3% vs consensus
    TTFT
    2.7s
    reasoning 2.1s
    Throughput
    19.4 tps
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    zeroeval → deepseek
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.28
    Output $ / 1M tokens
    $0.42
    Cache read $ / 1M tokens
    -50.3% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    digitalocean
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.25
    Output $ / 1M tokens
    $0.8
    Cache read $ / 1M tokens
    -5.3% vs consensus
    TTFT
    1.0s
    reasoning 983ms
    Throughput
    41.0 tps
    Uptime 24h
    Requests 24h
  • #
    8
    Provider
    openrouter → deepinfra
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.32
    Output $ / 1M tokens
    $0.89
    Cache read $ / 1M tokens
    +5.3% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    9
    Provider
    openrouter → deepseek
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.25
    Output $ / 1M tokens
    $0.95
    Cache read $ / 1M tokens
    $0.13
    +12.4% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    10
    Provider
    openrouter → streamlake
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.26
    Output $ / 1M tokens
    $1.03
    Cache read $ / 1M tokens
    +21.7% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    11
    Provider
    friendli-ai
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.5
    Output $ / 1M tokens
    $1.5
    Cache read $ / 1M tokens
    +77.5% vs consensus
    TTFT
    1.1s
    Throughput
    62.2 tps
    Uptime 24h
    Requests 24h
  • #
    12
    Provider
    fireworks
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.56
    Output $ / 1M tokens
    $1.68
    Cache read $ / 1M tokens
    +98.8% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    13
    Provider
    azure → azure-ai
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.58
    Output $ / 1M tokens
    $1.68
    Cache read $ / 1M tokens
    +98.8% vs consensus
    TTFT
    1.9s
    Throughput
    146.7 tps
    Uptime 24h
    Requests 24h
  • #
    14
    Provider
    amazon-bedrock
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.62
    Output $ / 1M tokens
    $1.85
    Cache read $ / 1M tokens
    +118.9% vs consensus
    TTFT
    1.5s
    reasoning 1.5s
    Throughput
    59.4 tps
    Uptime 24h
    Requests 24h
  • #
    15
    Provider
    sambanova
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $3
    Output $ / 1M tokens
    $4.5
    Cache read $ / 1M tokens
    +432.5% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    16
    Provider
    parasail
    Class
    Direct host
    Disabled by parasail · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.28
    Output $ / 1M tokens
    $0.45
    Cache read $ / 1M tokens
    $0.13
    -46.7% vs consensus
    TTFT
    1.9s
    Throughput
    81.5 tps
    Uptime 24h
    Requests 24h
  • #
    17
    Provider
    nebius
    Class
    Direct host
    Disabled by nebius · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.3
    Output $ / 1M tokens
    $0.45
    Cache read $ / 1M tokens
    -46.7% vs consensus
    TTFT
    2.0s
    reasoning 2.4s
    Throughput
    104.4 tps
    Uptime 24h
    Requests 24h
  • #
    18
    Provider
    openrouter → novita
    Class
    Reseller
    Disabled by openrouter · disabled · fp8
    $ / 1M tokens
    Input $ / 1M tokens
    $0.4
    Output $ / 1M tokens
    $1.3
    Cache read $ / 1M tokens
    +53.8% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    19
    Provider
    google
    Class
    Direct host
    Disabled by google · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.56
    Output $ / 1M tokens
    $1.68
    Cache read $ / 1M tokens
    $0.06
    +98.8% vs consensus
    TTFT
    5.6s
    reasoning 2.3s
    Throughput
    17.1 tps
    Uptime 24h
    Requests 24h
  • #
    20
    Provider
    eigenai
    Class
    Direct host
    Disabled by eigenai · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $1.85
    Cache read $ / 1M tokens
    +118.9% vs consensus
    TTFT
    1.4s
    reasoning 1.5s
    Throughput
    136.8 tps
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…