Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
Specifications
| Provider | Qwen (Alibaba) |
|---|---|
| Context window | 128,000 tokens (128K) |
| Input price | $0.8 / 1M tokens |
| Output price | $1 / 1M tokens |
| Input modalities | text, image |
| Output modalities | text |
| Model ID | qwen/qwen2.5-vl-72b-instruct |
Pricing is per 1M tokens, indicative pricing via OpenRouter (updated 2026-08-22); a provider's own list price may differ.
Frequently asked questions
How much does Qwen2.5 VL 72B Instruct cost?
Indicative pricing (via OpenRouter, updated 2026-08-22) is $0.8 per 1M input tokens and $1 per 1M output tokens. A provider's own list price may differ. On Vincony, Qwen2.5 VL 72B Instruct is available on credit-based pricing alongside 800+ other models.
What is Qwen2.5 VL 72B Instruct's context window?
Qwen2.5 VL 72B Instruct supports a context window of up to 128,000 tokens (128K), which is how much text it can consider at once.
Who makes Qwen2.5 VL 72B Instruct?
Qwen2.5 VL 72B Instruct is built by Qwen (Alibaba). You can access it — and compare it against other providers' models — through Vincony's multi-model platform.
How do I use Qwen2.5 VL 72B Instruct?
You can use Qwen2.5 VL 72B Instruct directly through Qwen (Alibaba), or via Vincony, which aggregates 800+ models (including Qwen2.5 VL 72B Instruct) into one interface with side-by-side comparison and credit-based pricing.