DeepSeek: R1 Distill Llama 70B

deepseek
$0.67output / 1M tokens · input $0.25/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    ovhcloud
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.67
    Output $ / 1M tokens
    $0.67
    Cache read $ / 1M tokens
    -16.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    gmi
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.25
    Output $ / 1M tokens
    $0.75
    Cache read $ / 1M tokens
    -6.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    deepinfra
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.7
    Output $ / 1M tokens
    $0.8
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    novita → deepseek
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.8
    Output $ / 1M tokens
    $0.8
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    openrouter → deepseek
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.8
    Output $ / 1M tokens
    $0.8
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    openrouter → novita
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.8
    Output $ / 1M tokens
    $0.8
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    fireworks
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.9
    Output $ / 1M tokens
    $0.9
    Cache read $ / 1M tokens
    +12.5% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    8
    Provider
    vercel-ai-gateway → deepseek
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.75
    Output $ / 1M tokens
    $0.99
    Cache read $ / 1M tokens
    +23.8% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    9
    Provider
    sambanova
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.7
    Output $ / 1M tokens
    $1.4
    Cache read $ / 1M tokens
    +75.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    10
    Provider
    together
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $2
    Output $ / 1M tokens
    $2
    Cache read $ / 1M tokens
    +150.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    11
    Provider
    scaleway
    Class
    Direct host
    Disabled by scaleway · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $1.05
    Output $ / 1M tokens
    $1.05
    Cache read $ / 1M tokens
    +31.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…