Moonshot AI / Kimi Models & API

Compare the Moonshot AI / Kimi models available on TokenHub, including Kimi K3, Kimi K2.6, Kimi K2.5, and Kimi K2. Review pricing, context, capabilities, and model IDs, then call Moonshot AI, Moonshot, and Kimi through one OpenAI-compatible API workflow.

Moonshot AI / Kimi Models & API Pricing

Compare current Moonshot AI / Kimi API pricing, input and output token costs, context windows, endpoints, and model IDs. The catalog may include Kimi K3, Kimi K2.6, Kimi K2.5, and Kimi K2; availability follows the live TokenHub model list.

Compare input, output, and cache-read prices for the available Moonshot AI / Kimi models. Open a model page to confirm the current billing unit and model ID.

モデル名入力/ 百万トークン出力/ 百万トークンキャッシュ読み取り/ 百万トークン

Moonshot

Kimi K2.6kimi-k2.6
$0.9286$3.8571$0.1571

Moonshot

Kimi K2.7 Codekimi-k2.7-code
$0.9286$3.8571$0.1857

Moonshot

Kimi K3kimi-k3
$2.8571$14.2857$0.2857

Compare context windows, maximum output, reasoning, tool calling, endpoints, and release dates across Kimi K3, Kimi K2.6, Kimi K2.5, and Kimi K2.

モデル名モダリティコンテキストウィンドウMax outputReasoningTool callingEndpointリリース日

Moonshot

Kimi K2.6kimi-k2.6
262.1K
262.1K
Anthropic / OpenAI
2026年4月21日

Moonshot

Kimi K2.7 Codekimi-k2.7-code
262.1K
262.1K
OpenAI / Anthropic
2026年6月12日

Moonshot

Kimi K3kimi-k3
1M
131.1K
OpenAI / Anthropic
2026年7月16日
Catalog data

Pricing, model availability, context windows, and endpoint support reflect the current TokenHub catalog. Confirm the model detail page before a production release because specifications and availability can change.

Last updated: 2026-08-25

Get Started With the Moonshot AI / Kimi API

Create a TokenHub API key, copy a published Moonshot AI / Kimi model ID, and call it with OpenAI Chat Completions, Responses, or Claude Messages. Existing OpenAI SDK projects can usually switch by changing the API key, Base URL, and model ID.

  1. 01Create a TokenHub API key
  2. 02Choose an available Moonshot AI / Kimi model
  3. 03Choose an API protocol
  4. 04Copy the published Moonshot AI / Kimi model ID
  5. 05Send your first request

Choose an API protocol

import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.chat.completions.create({
  "model": "kimi-k2.6",
  "messages": [
    {
      "role": "user",
      "content": "Explain why low latency matters for an AI product in one sentence."
    }
  ]
})
console.log(result.choices[0]?.message?.content)
import json
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ.get("TOKENHUB_API_KEY"),
    base_url="https://us-api.tokenhub.com/v1",
)

request = json.loads("{\"model\":\"kimi-k2.6\",\"messages\":[{\"role\":\"user\",\"content\":\"Explain why low latency matters for an AI product in one sentence.\"}]}")
result = client.chat.completions.create(**request)
print(result.choices[0].message.content)
curl 'https://us-api.tokenhub.com/v1/chat/completions' \
  -X 'POST' \
  -H "Authorization: Bearer $TOKENHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "kimi-k2.6",
  "messages": [
    {
      "role": "user",
      "content": "Explain why low latency matters for an AI product in one sentence."
    }
  ]
}'
const response = await fetch("https://us-api.tokenhub.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.TOKENHUB_API_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "kimi-k2.6",
    "messages": [
      {
        "role": "user",
        "content": "Explain why low latency matters for an AI product in one sentence."
      }
    ]
  }),
})
const data = await response.json()
console.log(data)
import os
import requests

response = requests.request(method="POST", url="https://us-api.tokenhub.com/v1/chat/completions",
    headers={
        "Authorization": f"Bearer {os.environ['TOKENHUB_API_KEY']}",
        "Content-Type": "application/json",
    },
    json=__import__("json").loads("{\"model\":\"kimi-k2.6\",\"messages\":[{\"role\":\"user\",\"content\":\"Explain why low latency matters for an AI product in one sentence.\"}]}"),
)
response.raise_for_status()
print(response.json())
import OpenAI from "openai"

const client = new OpenAI({
  apiKey: process.env.TOKENHUB_API_KEY,
  baseURL: "https://us-api.tokenhub.com/v1",
})

const result = await client.responses.create({
  "model": "kimi-k2.6",
  "input": "Explain why low latency matters for an AI product in one sentence."
})
console.log(result.output_text)
import json
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ.get("TOKENHUB_API_KEY"),
    base_url="https://us-api.tokenhub.com/v1",
)

request = json.loads("{\"model\":\"kimi-k2.6\",\"input\":\"Explain why low latency matters for an AI product in one sentence.\"}")
result = client.responses.create(**request)
print(result.output_text)
curl 'https://us-api.tokenhub.com/v1/responses' \
  -X 'POST' \
  -H "Authorization: Bearer $TOKENHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "kimi-k2.6",
  "input": "Explain why low latency matters for an AI product in one sentence."
}'
const response = await fetch("https://us-api.tokenhub.com/v1/responses", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.TOKENHUB_API_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "kimi-k2.6",
    "input": "Explain why low latency matters for an AI product in one sentence."
  }),
})
const data = await response.json()
console.log(data)
import os
import requests

response = requests.request(method="POST", url="https://us-api.tokenhub.com/v1/responses",
    headers={
        "Authorization": f"Bearer {os.environ['TOKENHUB_API_KEY']}",
        "Content-Type": "application/json",
    },
    json=__import__("json").loads("{\"model\":\"kimi-k2.6\",\"input\":\"Explain why low latency matters for an AI product in one sentence.\"}"),
)
response.raise_for_status()
print(response.json())
curl 'https://us-api.tokenhub.com/v1/messages' \
  -X 'POST' \
  -H "Authorization: Bearer $TOKENHUB_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "kimi-k2.6",
  "max_tokens": 1024,
  "messages": [
    {
      "role": "user",
      "content": "Explain why low latency matters for an AI product in one sentence."
    }
  ]
}'
const response = await fetch("https://us-api.tokenhub.com/v1/messages", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.TOKENHUB_API_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "kimi-k2.6",
    "max_tokens": 1024,
    "messages": [
      {
        "role": "user",
        "content": "Explain why low latency matters for an AI product in one sentence."
      }
    ]
  }),
})
const data = await response.json()
console.log(data)
import os
import requests

response = requests.request(method="POST", url="https://us-api.tokenhub.com/v1/messages",
    headers={
        "Authorization": f"Bearer {os.environ['TOKENHUB_API_KEY']}",
        "Content-Type": "application/json",
    },
    json=__import__("json").loads("{\"model\":\"kimi-k2.6\",\"max_tokens\":1024,\"messages\":[{\"role\":\"user\",\"content\":\"Explain why low latency matters for an AI product in one sentence.\"}]}"),
)
response.raise_for_status()
print(response.json())

Which Moonshot AI / Kimi Model Should You Choose?

Use this Moonshot AI / Kimi model selection guide for long-context reasoning, coding, Kimi CLI workflows, agents, research, and general chat. Compare price, context, reasoning, tool calling, and endpoint support before shipping.

API workloadStart withWhen it fitsVerify before launch
High-volume Moonshot AI / Kimi API callsKimi K2.6Fast iteration, chat, extraction, summarization, and batch workloads using Moonshot AI / Kimi.Moonshot AI / Kimi API pricing, throughput, and output-token cost
Moonshot AI / Kimi reasoning and codingKimi K2.6Multi-step analysis, code generation, debugging, and agent tasks where answer quality matters.Reasoning support, tool calling, latency, and total request cost
General Moonshot AI / Kimi applicationsKimi K2.6Everyday assistants, content generation, information extraction, and product features.Task accuracy, endpoint support, and production price
Long-context Moonshot AI / Kimi workloadsKimi K3Documents, codebases, research, and conversations where input length is the constraint.Context window, input-token price, maximum output, and streaming
Quality-first Moonshot AI / Kimi workloadsKimi K2.6Production outputs where capability matters more than the lowest unit cost.Task quality, latency, and the price difference versus a faster model

Moonshot AI / Kimi API FAQ

What are Moonshot AI, Moonshot, and Kimi?

Moonshot AI, Moonshot, and Kimi are names and search terms associated with this model family. TokenHub uses the published provider and model IDs shown in the live catalog.

Which Moonshot AI / Kimi models are available on TokenHub?

The live list above shows current availability. Relevant model searches include Kimi K3, Kimi K2.6, Kimi K2.5, and Kimi K2, but only models displayed in the TokenHub catalog can be called through this page.

How do I get a Moonshot AI / Kimi API key?

Moonshot AI / Kimi API and Moonshot AI / Kimi API key searches lead to the same TokenHub workflow: create an API key, select a published Moonshot AI / Kimi model ID, and use the TokenHub Base URL in your application.

How is Moonshot AI / Kimi API pricing calculated?

Pricing depends on the selected model and billing type. Compare input, output, cache, or per-request prices in the table and confirm the model detail page before production use.

Can I use the OpenAI SDK with Moonshot AI / Kimi?

Yes. Use the TokenHub Base URL, API key, and published Moonshot AI / Kimi model ID with the OpenAI Python or Node.js SDK. Endpoint support is shown for each model.

Which Moonshot AI / Kimi model should I use?

Choose based on the workload: long-context reasoning, coding, Kimi CLI workflows, agents, research, and general chat. Compare actual price, context, reasoning, tool calling, and output limits in the live catalog.

Start Building With Moonshot AI / Kimi

Choose a Moonshot AI / Kimi model, copy its model ID, create a TokenHub API key, and send your first request through one compatible API workflow.