All Models

Qwen3 VL 8B Instruct

qwen Tool Calling Attachments Open Weights Structured Output

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Providers 2
Released Oct 14, 2025
Input Modalities image, text
Output Modalities text
Tarsk Use coding

Available Providers (2)

Provider Model ID Input Cost Output Cost Context Max Output Docs
OpenRouter qwen/qwen3-vl-8b-instruct $0.12/MTok $0.46/MTok 262.1K 32.8K
Kilo Gateway qwen/qwen3-vl-8b-instruct $0.12/MTok $0.46/MTok 131.1K 32.8K

Capabilities

Reasoning
Tool Calling
Attachments
Open Weights
Structured Output