All Models

voxtral-small-2507

Tool Calling Attachments Structured Output

Voxtral Small is a multimodal model with audio input, combining advanced speech capabilities with strong text performance for transcription, translation, and audio understanding.

Providers 1
Released Feb 2, 2026
Input Modalities text, audio
Output Modalities text
Tarsk Use coding

Available Providers (1)

Provider Model ID Input Cost Output Cost Context Max Output Docs
Cortecs voxtral-small-2507 $0.11/MTok $0.33/MTok 32K 32K

Capabilities

Reasoning
Tool Calling
Attachments
Open Weights
Structured Output