Data verified: Aug 15, 2026, 11:13 AM · Prices sourced from official provider pages. Verify current rates before production use.
Data verified: Aug 15, 2026, 11:13 AM · Prices sourced from official provider pages. Verify current rates before production use.
Price, context and performance head to head. Data synced from the provider catalog.
Cheaper
Voxtral Small 24B 2507
Larger context
DeepSeek V3.1 Terminus
Faster
Voxtral Small 24B 2507
Higher quality
DeepSeek V3.1 Terminus
| Feature | DeepSeek V3.1 Terminus | Voxtral Small 24B 2507 |
|---|---|---|
| Provider | DeepSeek | Mistral |
| Tier | Mid-tier | Budget |
| Input per 1M tokens | $0.27 | $0.1 |
| Output per 1M tokens | $0.95 | $0.3 |
| Cached input per 1M | $0.13 | $0.01 |
| Context window | 163.8K | 32K |
| Intelligence index | — | — |
| Coding index | 44/100 | — |
| Agentic index | — | — |
| GPQA Diamond | 1 | — |
| Tau-bench (airline) | 1 | — |
| Speed | Standard | Fast |
| Vision (image input) | No | No |
| Function calling | Yes | Yes |
| Batch API | No | No |
Enter how many requests per day you send with an average prompt (1K input + 1K output) and compare the monthly cost of both models.
Voxtral Small 24B 2507 saves $2.46/mo vs DeepSeek V3.1 Terminus
Want us to build it for you?
We integrate DeepSeek V3.1 Terminus or Voxtral Small 24B 2507 into your product with caching, observability and continuous evaluation — typically 40-80% cheaper than the obvious first pick.
Other combinations developers frequently compare in 2026.
What people ask us when comparing GPT, Claude, Gemini and the rest.
A token is the unit an AI model processes: usually between half a word and a full word. Rule of thumb: 1,000 tokens ≈ 750 English words. A 20-word sentence is about 26 tokens; a 300-word email is around 400. Models charge for input tokens (your prompt) and output tokens (their answer) separately.