meta/llama-3.2-90b-vision-instruct

meta
freeoutput / 1M tokens · input free/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    google-vertex
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    free
    Output $ / 1M tokens
    free
    Cache read $ / 1M tokens
    -100.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    oci
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $2
    Output $ / 1M tokens
    $2
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    azure → azure-ai
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $2.04
    Output $ / 1M tokens
    $2.04
    Cache read $ / 1M tokens
    +2.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    deepinfra
    Class
    Direct host
    Disabled by deepinfra · disabled · bfloat16
    $ / 1M tokens
    Input $ / 1M tokens
    $0.35
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    -80.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…