MoonshotAI: Kimi K2 Thinking

moonshotai
$1.2output / 1M tokens · input $0.6/M

Best price — Output ($/1M tokens)

Markets

One row per access path (venue → underlying provider). Prices are each venue's published list price; performance is measured by Artificial Analysis. Your actual bill may vary by routing settings, fallbacks, region / ZDR, long-context tiers, and request features (JSON mode, tools, vision).

  • #
    1
    Provider
    gmi
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.8
    Output $ / 1M tokens
    $1.2
    Cache read $ / 1M tokens
    -52.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    2
    Provider
    gmi → moonshot
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.8
    Output $ / 1M tokens
    $1.2
    Cache read $ / 1M tokens
    -52.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    3
    Provider
    moonshot
    Class
    Official
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    $0.15
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    4
    Provider
    fireworks
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    5
    Provider
    google
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    751ms
    Throughput
    232.7 tps
    Uptime 24h
    Requests 24h
  • #
    6
    Provider
    novita
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    1.7s
    Throughput
    39.9 tps
    Uptime 24h
    Requests 24h
  • #
    7
    Provider
    azure → azure-ai
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    1.6s
    Throughput
    160.8 tps
    Uptime 24h
    Requests 24h
  • #
    8
    Provider
    baseten → moonshot
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    9
    Provider
    novita → moonshot
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    $0.15
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    10
    Provider
    openrouter → google-vertex
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    11
    Provider
    openrouter → moonshot
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    $0.15
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    12
    Provider
    openrouter → novita
    Class
    Reseller
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    $0.15
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    13
    Provider
    amazon-bedrock
    Class
    Direct host
    $ / 1M tokens
    Input $ / 1M tokens
    $0.73
    Output $ / 1M tokens
    $3.03
    Cache read $ / 1M tokens
    +21.2% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    14
    Provider
    deepinfra
    Class
    Direct host
    Disabled by deepinfra · disabled · fp4
    $ / 1M tokens
    Input $ / 1M tokens
    $0.47
    Output $ / 1M tokens
    $2
    Cache read $ / 1M tokens
    $0.14
    -20.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
  • #
    15
    Provider
    openrouter → atlas-cloud
    Class
    Reseller
    Disabled by openrouter · disabled · int4
    $ / 1M tokens
    Input $ / 1M tokens
    $0.6
    Output $ / 1M tokens
    $2.5
    Cache read $ / 1M tokens
    $0.6
    +0.0% vs consensus
    TTFT
    Throughput
    Uptime 24h
    Requests 24h
Source ecosystem
Loading…