与 Kimi K2.7 Code 相同, 输出速度约 5-6 倍 (常规编程场景约 180 Token/s)
Model ID: kimi-k2.7-code-highspeed · Type: chat · Provider: Moonshot
Endpoints: /v1/chat/completions · /v1/messages · /v1/responses
| Input (per 1M tokens) | $1.045 USD |
| Output (per 1M tokens) | $4.4 USD |
| Cache read (per 1M tokens) | $0.209 USD |
from openai import OpenAI
client = OpenAI(api_key="sk-...", base_url="https://api.router.ai/v1")
resp = client.chat.completions.create(
model="kimi-k2.7-code-highspeed",
messages=[{"role": "user", "content": "Hello"}],
)
print(resp.choices[0].message.content)Kimi K2.7 Code HighSpeed (`kimi-k2.7-code-highspeed`) is billed per usage at $1.045/1M in · $4.4/1M out, in USD. Current pricing is always listed at https://yigekey.com/models/kimi-k2.7-code-highspeed.
Send a request to https://api.router.ai/v1/v1/chat/completions with the header `Authorization: Bearer <your API key>` and `"model": "kimi-k2.7-code-highspeed"`. The API is OpenAI-compatible, so any OpenAI SDK works by changing base_url to https://api.router.ai/v1 — no other code change.
Kimi K2.7 Code HighSpeed can be called on: /v1/chat/completions; /v1/messages; /v1/responses.
Kimi K2.7 Code HighSpeed accepts up to 262,144 input tokens and can return up to 32,768 output tokens. Requests exceeding the input limit are rejected before reaching the model.
Kimi K2.7 Code HighSpeed supports: function_calling, prompt_caching, reasoning.
Kimi K2.7 Code HighSpeed is a chat model from Moonshot, available through the 聚合 AI 中转站 gateway with the same API key as every other model.
Call it through the 聚合 AI 中转站 OpenAI-compatible endpoint (API base: https://api.router.ai/v1). AI agents can discover and call every model on this gateway through MCP (https://mcp.router.ai/mcp) with no manual integration.