Voxtral-Small-2507
Voxtral-Small-2507 is a Multimodal model.
Voxtral Small is a multimodal model with audio input, combining advanced speech capabilities with strong text performance for transcription, translation, and audio understanding.
Vision
Audio
Tool Use
Reasoning
Citations
Key Specs
Context
32k
Max Output
32k
Best Price (Input / Output)
Chat with Model
$0.11 / $0.33 per 1M tokens
Pricing Comparison
| Provider | Input (1M) | Output (1M) | Image (1k) |
|---|---|---|---|
| cortecs (Ver 2507) | $0.11 | $0.33 | - |
Prices are per 1 million tokens unless otherwise noted. Image pricing is per 1000 images if applicable.
Price History
Prompt Price
Completion Price
Price trend per 1 million tokens over time.