AlibabaPaidOperational

Qwen3 VL Flash

API model name: qwen3-vl-flash

Qwen3 VL Flash is Alibaba's chat model, served on the Api.Airforce unified API. Capabilities include Reasoning. It is priced at $0.04 per million input tokens and $0.32 per million output tokens. That is below the provider's $0.02 official input rate. Access it through the OpenAI-compatible API with one key, alongside 100+ other models on Api.Airforce.

Pricing

Input / 1M tokens
$0.04
Output / 1M tokens
$0.32
Cache read / 1M tokens
$0.01
Official input rate
$0.02
Official output rate
$0.22

Api.Airforce price vs. the provider's official rate.

Specifications

Provider
Alibaba
Type
chat model

Capabilities

Reasoning

What is Qwen3 VL Flash used for?

  • Chatbots & assistants — conversational AI, drafting, summarizing and Q&A.
  • Complex reasoning — math, coding and step-by-step problem solving.

Qwen3 VL Flash vs. similar models

ModelIntelligenceContextInput / 1MOutput / 1M
Qwen3 VL Flash$0.04$0.32
Qvq Max Latest$0.84$3.36
Qvq Plus$0.28$0.69
Qwen Coder Plus$0.46$0.91

Prices are Api.Airforce pay-as-you-go rates per 1M tokens. Context is the maximum input length.

Related models

Qwen3 VL Flash — frequently asked questions

How much does Qwen3 VL Flash cost?
Qwen3 VL Flash is billed pay-as-you-go at $0.04 per 1M input tokens and $0.32 per 1M output tokens. There is no subscription — you only pay for what you use.
What can Qwen3 VL Flash do?
Qwen3 VL Flash supports Reasoning.
Is Qwen3 VL Flash free to use?
Qwen3 VL Flash is a paid, pay-as-you-go model — no subscription, you are only charged for usage.
How do I use Qwen3 VL Flash via the API?
Qwen3 VL Flash is OpenAI-compatible. Point any OpenAI SDK at https://api.airforce/v1 and pass the model ID qwen3-vl-flash with your Api.Airforce API key.
Who makes Qwen3 VL Flash?
Qwen3 VL Flash is Alibaba's chat model, served through the unified Api.Airforce gateway alongside 100+ other models.