All models
Kimi

Kimi K2.7 Code HighSpeed

moonshotai/kimi-k2.7-code-highspeed

Kimi K2.7 Code HighSpeed is the high-throughput tier of Moonshot AI's coding-focused Kimi K2.7 Code, rated by Moonshot at roughly 5-6x the output speed of the standard tier, with a 256K-token context window.

Pricing

Input$0.95per 1M tokens
Output$4.00per 1M tokens
Cached input$0.19per 1M tokens read from cache

Specifications

Context window262Ktokens
Max output66Ktokens
ProviderKimi
CapabilitiesVision · Reasoning · Function calling
Input typestext, image

Use Kimi K2.7 Code HighSpeed

Point your existing OpenAI SDK at xKiro and pass moonshotai/kimi-k2.7-code-highspeed as the model. Nothing else in your code changes.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.xkiro.com/v1",
    api_key="YOUR_XKIRO_KEY",
)

response = client.chat.completions.create(
    model="moonshotai/kimi-k2.7-code-highspeed",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Supported parameters

max_tokens · temperature · top_p · stop · frequency_penalty · presence_penalty · seed · stream · tools · tool_choice · response_format · structured_outputs · reasoning · include_reasoning

Related models

Call Kimi K2.7 Code HighSpeed through the same endpoints you already use.

Get an API key