qwen3-vl-flash
NixAPI price, from
USD per 1M tokens · via group Alibaba-1 · Official list price $0.05 / $0.4Save 16%
About qwen3-vl-flash
A small-scale visual understanding model in the Qwen3 series that effectively combines thinking and non-thinking modes. Outperforms the open-source Qwen3-VL-30B-A3B with fast response. Comprehensively upgraded image/video understanding with support for ultra-long contexts including long videos and documents, spatial awareness, and universal object recognition; features 2D/3D visual grounding for complex real-world tasks.
Price by routing group
Pick a group in the console when you create a key. Status is live from the router. Live status by group →
| Group | Group status | Model status | Multiplier | Input | Output | Cache read | Cache write |
|---|---|---|---|---|---|---|---|
Alibaba-1 | … | ×0.845 | $0.0423 | $0.338 | — | — | |
Alibaba-2 | … | ×1.2675 | $0.0634 | $0.507 | — | — | |
Alibaba-3 | … | ×1.859 | $0.093 | $0.7436 | — | — |
USD per 1M tokens
Long-context pricing
- Prompts over 32,000 tokens: input ×2, output ×2
- Prompts over 128,000 tokens: input ×4, output ×4
Supported endpoints
Call qwen3-vl-flash in one request
Works with the OpenAI SDK and any OpenAI-compatible client. Replace the key with your own.
curl https://api.nixapi.com/v1/chat/completions \
-H "Authorization: Bearer $NIXAPI_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3-vl-flash",
"messages": [{"role": "user", "content": "Hello!"}]
}'from openai import OpenAI
client = OpenAI(base_url="https://api.nixapi.com/v1", api_key="sk-…")
resp = client.chat.completions.create(
model="qwen3-vl-flash",
messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)import OpenAI from 'openai';
const client = new OpenAI({ baseURL: 'https://api.nixapi.com/v1', apiKey: process.env.NIXAPI_KEY });
const resp = await client.chat.completions.create({
model: "qwen3-vl-flash",
messages: [{ role: 'user', content: 'Hello!' }],
});
console.log(resp.choices[0].message.content);Questions about qwen3-vl-flash
How much does qwen3-vl-flash cost on NixAPI?
From $0.0423 per 1M input tokens and $0.338 per 1M output tokens in the Alibaba-1 group, about 16% below the official list price. Prices are pay-as-you-go in USD with no monthly minimum.
How do I call qwen3-vl-flash?
Point any OpenAI-compatible client at https://api.nixapi.com/v1, use your NixAPI key and set the model to "qwen3-vl-flash".
Which routing groups serve qwen3-vl-flash?
qwen3-vl-flash is available in 3 routing groups: Alibaba-1, Alibaba-2, Alibaba-3. The lowest price is in Alibaba-1.