All models
QwenFree

Qwen3 Omni Flash

qwen/qwen3-omni-flash

Qwen3 Omni Flash is Alibaba’s natively omni-modal model, accepting text, image, video and audio input with strong results across audio and audio-visual benchmarks. Context 262K.

Pricing

InputFreeper 1M tokens
OutputFreeper 1M tokens

Specifications

Context window262Ktokens
Max output66Ktokens
ProviderQwen
CapabilitiesVision · Reasoning · Function calling
Input typestext, image, video

Use Qwen3 Omni Flash

Point your existing OpenAI SDK at xKiro and pass qwen/qwen3-omni-flash as the model. Nothing else in your code changes.

from openai import OpenAI

client = OpenAI(
    base_url="https://api.xkiro.com/v1",
    api_key="YOUR_XKIRO_KEY",
)

response = client.chat.completions.create(
    model="qwen/qwen3-omni-flash",
    messages=[{"role": "user", "content": "Hello"}],
)
print(response.choices[0].message.content)

Supported parameters

max_tokens · temperature · top_p · stop · frequency_penalty · presence_penalty · seed · stream · tools · tool_choice · reasoning · include_reasoning

Related models

Qwen3 Omni Flash costs $0 per token on xKiro. Create a key and call it today.

Get an API key