Z.ai: GLM 4.7 Flash

z-ai
$0.4output / 1M tokens · input $0.06/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    deepinfra
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.06
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.01
    +0.0% vs consensus
    TTFT
    604ms
    Throughput
    57.1 tps
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    openrouter → deepinfra
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.06
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.01
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    openrouter → venice
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.06
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.01
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    openrouter → z-ai
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.06
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.01
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    openrouter → cloudflare
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.06
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    amazon-bedrock
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.07
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    993ms
    reasoning 1.2s
    Throughput
    203.2 tps
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    bedrock
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.07
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    8
    Provider
    bedrock-converse
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.07
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    9
    Provider
    gmi
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.07
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    10
    Provider
    novita
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.07
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.01
    +0.0% vs consensus
    TTFT
    6.5s
    reasoning 1.9s
    Throughput
    69.4 tps
    Uptime 24h
    Requests 24h
  • #
    11
    Provider
    openrouter → novita
    Class
    Reseller
    Disabled by openrouter · disabled · bf16
    $ / 1M tokens
    Input $ / 1M tokens
    $0.07
    Output $ / 1M tokens
    $0.4
    Cache read $ / 1M tokens
    $0.01
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    12
    Provider
    openrouter → phala
    Class
    Reseller
    Disabled by openrouter · disabled · unknown
    $ / 1M tokens
    Input $ / 1M tokens
    $0.1
    Output $ / 1M tokens
    $0.43
    Cache read $ / 1M tokens
    +7.5% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…