All Models

qwen2.5-vl-72b-instruct

Reasoning Tool Calling Attachments Open Weights Structured Output

Qwen2.5-VL is a powerful vision-language model with advanced capabilities in visual understanding, long video reasoning, and structured output generation.

Providers 9
Released Sep 1, 2024
Input Modalities text, image, video
Output Modalities text
Tarsk Use coding

Available Providers (9)

Provider Model ID Input Cost Output Cost Context Max Output Docs
Cortecs qwen2.5-vl-72b-instruct $0.25/MTok $0.75/MTok 32K 32K
Nebius Token Factory Qwen/Qwen2.5-VL-72B-Instruct $0.25/MTok $0.75/MTok 128K 8.2K
DevPass (LLM Gateway) qwen2-5-vl-72b-instruct $0.25/MTok $0.75/MTok 32K 8.2K
OpenRouter qwen/qwen2.5-vl-72b-instruct $0.80/MTok $1/MTok 128K 128K
NovitaAI qwen/qwen2.5-vl-72b-instruct $0.80/MTok $0.80/MTok 32.8K 32.8K
Kilo Gateway qwen/qwen2.5-vl-72b-instruct $0.80/MTok $1/MTok 128K 128K
OVHcloud AI Endpoints qwen2.5-vl-72b-instruct $1.01/MTok $1.01/MTok 32.8K 32.8K
Alibaba (China) qwen2-5-vl-72b-instruct $2.29/MTok $6.88/MTok 131.1K 8.2K
Alibaba qwen2-5-vl-72b-instruct $2.80/MTok $8.40/MTok 131.1K 8.2K

Capabilities

Reasoning
Tool Calling
Attachments
Open Weights
Structured Output