MiniMax
MiniMax M2.7 Highspeed
minimax/minimax-m2.7-highspeed
MiniMax-M2.7 with identical performance and significantly faster, lower-latency inference.
Pricing
| Input | $0.6per 1M tokens |
|---|---|
| Output | $2.40per 1M tokens |
| Cached input | $0.06per 1M tokens read from cache |
Specifications
| Context window | 205Ktokens |
|---|---|
| Max output | 66Ktokens |
| Provider | MiniMax |
| Capabilities | Reasoning · Function calling |
| Input types | text |
Use MiniMax M2.7 Highspeed
Point your existing OpenAI SDK at xKiro and pass minimax/minimax-m2.7-highspeed as the model. Nothing else in your code changes.
from openai import OpenAI
client = OpenAI(
base_url="https://api.xkiro.com/v1",
api_key="YOUR_XKIRO_KEY",
)
response = client.chat.completions.create(
model="minimax/minimax-m2.7-highspeed",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Supported parameters
max_tokens · temperature · top_p · stream · tools · tool_choice
Related models
- MiniMax M2$0.3
- MiniMax M2.1$0.3
- MiniMax M2.1 Highspeed$0.6
- MiniMax M2.5$0.3
- MiniMax M2.5 Highspeed$0.6
- MiniMax M2.7$0.3
Call MiniMax M2.7 Highspeed through the same endpoints you already use.
Get an API key