Browse all Qwen text models
Provider logo

Qwen25 VL 72b

qwen/qwen2.5-vl-72b-instruct
Back
Provider logo

Qwen25 VL 72b

qwen/qwen2.5-vl-72b-instruct
Back

Qwen25 VL 72b model with 32k context window

Added May 10, 2025

Context Window

32.0K

Max Output

32.8K

Avg output tokens (7d)

254 tokens

19%

Input Price (Auto)

$0.70/1M

Output Price (Auto)

$0.70/1M

Capabilities

Benchmarks

Performance metrics and benchmarks

No benchmark data is available yet for this model.

Providers

Provider information for this model’s automatic routing. These routes cannot be selected individually.

Loading provider options…

Compare Qwen25 VL 72b with similar models from the same provider or model family.

Qwen2 72B Dracarys

abacusai/Dracarys-72B-Instruct

A Qwen2 72B Instruct finetune by Abacus.AI focused on improving coding performance.

Qwen2.5 72B

qwen/qwen-2.5-72b-instruct

Great multilingual support, strong at mathematics and coding, supports roleplay and chatbots.

Qwen2.5 VL 72B TEE

TEE/qwen2.5-vl-72b-instruct

Qwen2.5 Vision-Language 72B model with multimodal capabilities. Running inside a TEE (Trusted Execution Environment), with provider attestation support.

Qwen3 VL 235B A22B Instruct Original

qwen3-vl-235b-a22b-instruct-original

Routed across multiple providers, which may include Alibaba, a Chinese entity; provider-specific privacy and logging guarantees may vary. Qwen3 Vision‑Language model (235B MoE, ≈22B active) tuned for instruction following and grounded visual QA. Excels at image understanding, dense OCR, charts and diagrams, and multi‑image context. Use this variant when you want concise, direct answers grounded in the visuals.

Qwen3 Next 80B A3B (Instruct)

qwen/qwen3-next-80b-a3b-instruct

Based on the new Qwen3‑Next architecture (hybrid attention, highly sparse MoE, training‑stability optimizations, and multi‑token prediction), the Qwen3‑Next‑80B‑A3B‑Instruct model delivers extreme efficiency with only 3B active parameters per pass. It performs comparably to Qwen3‑235B‑A22B‑Instruct‑2507 and shows clear advantages on ultra‑long context tasks (up to 256K tokens).

Qwen3 Coder 30B A3B Instruct

qwen/qwen3-coder-30b-a3b-instruct

Qwen3 Coder 30B with 3B active parameters, optimized for code generation and technical tasks