Google
Gemini 3.8 Flash
google/gemini-3.8-flash
Google's most intelligent Flash model, with significant gains over 3.7 Flash on software engineering, agentic tasks, and multi-step reasoning.
Pricing
| Input | $0.75per 1M tokens |
|---|---|
| Output | $3.75per 1M tokens |
Specifications
| Context window | 1Mtokens |
|---|---|
| Max output | 66Ktokens |
| Provider | |
| Capabilities | Vision · Reasoning · Function calling |
| Input types | text, image |
Use Gemini 3.8 Flash
Point your existing OpenAI SDK at xKiro and pass google/gemini-3.8-flash as the model. Nothing else in your code changes.
from openai import OpenAI
client = OpenAI(
base_url="https://api.xkiro.com/v1",
api_key="YOUR_XKIRO_KEY",
)
response = client.chat.completions.create(
model="google/gemini-3.8-flash",
messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)Supported parameters
max_tokens · temperature · top_p · stop · frequency_penalty · presence_penalty · seed · stream · tools · tool_choice · reasoning · include_reasoning
Related models
- Gemini 2.5 Flash$0.195
- Gemini 2.5 Pro$0.8125
- Gemini 3 Flash$0.325
- Gemini 3.1 Pro$1.30
- Gemini 3.5 Flash$0.975
- Gemini 3.6 Flash$0.975
Call Gemini 3.8 Flash through the same endpoints you already use.
Get an API key