All Models
voxtral-small-2507
Voxtral Small is a multimodal model with audio input, combining advanced speech capabilities with strong text performance for transcription, translation, and audio understanding.
Available Providers (1)
| Provider | Model ID | Input Cost | Output Cost | Context | Max Output | Docs |
|---|---|---|---|---|---|---|
| | voxtral-small-2507 | $0.11/MTok | $0.33/MTok | 32K | 32K |
Capabilities
Reasoning
Tool Calling
Attachments
Open Weights
Structured Output