Llama-4 Scout 17B 16E Instruct

meta
$0.1output / 1M tokens · input $0.05/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    lambda-ai
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.05
    Output $ / 1M tokens
    $0.1
    Cache read $ / 1M tokens
    -83.1% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    deepinfra → meta
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.08
    Output $ / 1M tokens
    $0.3
    Cache read $ / 1M tokens
    -49.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    deepinfra
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.1
    Output $ / 1M tokens
    $0.3
    Cache read $ / 1M tokens
    -49.2% vs consensus
    TTFT
    821ms
    Throughput
    40.6 tps
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    gmi
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.08
    Output $ / 1M tokens
    $0.5
    Cache read $ / 1M tokens
    -15.3% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    novita
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.18
    Output $ / 1M tokens
    $0.59
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    904ms
    Throughput
    43.8 tps
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    together
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.18
    Output $ / 1M tokens
    $0.59
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    novita → meta
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.18
    Output $ / 1M tokens
    $0.59
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    8
    Provider
    together → meta
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.18
    Output $ / 1M tokens
    $0.59
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    9
    Provider
    amazon-bedrock
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.17
    Output $ / 1M tokens
    $0.66
    Cache read $ / 1M tokens
    +11.9% vs consensus
    TTFT
    815ms
    Throughput
    174.7 tps
    Uptime 24h
    Requests 24h
  • #
    10
    Provider
    wandb → meta
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.17
    Output $ / 1M tokens
    $0.66
    Cache read $ / 1M tokens
    +11.9% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    11
    Provider
    google
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.25
    Output $ / 1M tokens
    $0.7
    Cache read $ / 1M tokens
    +18.6% vs consensus
    TTFT
    691ms
    Throughput
    150.2 tps
    Uptime 24h
    Requests 24h
  • #
    12
    Provider
    google-vertex
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.25
    Output $ / 1M tokens
    $0.7
    Cache read $ / 1M tokens
    +18.6% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    13
    Provider
    sambanova
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.4
    Output $ / 1M tokens
    $0.7
    Cache read $ / 1M tokens
    +18.6% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    14
    Provider
    oci
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.72
    Output $ / 1M tokens
    $0.72
    Cache read $ / 1M tokens
    +22.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    15
    Provider
    azure → azure-ai
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.2
    Output $ / 1M tokens
    $0.78
    Cache read $ / 1M tokens
    +32.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    16
    Provider
    cloudflare
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.27
    Output $ / 1M tokens
    $0.85
    Cache read $ / 1M tokens
    +44.1% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    17
    Provider
    compactifai
    Class
    Direct host
    Disabled by compactifai · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.1
    Output $ / 1M tokens
    $0.14
    Cache read $ / 1M tokens
    -76.3% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    18
    Provider
    groq
    Class
    Direct host
    Disabled by groq · stale
    $ / 1M tokens
    Input $ / 1M tokens
    $0.11
    Output $ / 1M tokens
    $0.34
    Cache read $ / 1M tokens
    -42.4% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…