TechCompare LogoTechCompare

Mistral Small 4 vs Mistral Medium 3.5: standard utility vs balanced logic costs

Mistral Small 4 is the clear budget winner. Use Mistral Medium 3.5 for tasks requiring a step up in translation or conversational quality without paying premium frontier prices.

Lightweight utility versus balanced medium-scale logic.

Balancing speed and quality often leads developers to compare Mistral Small 4 and Mistral Medium 3.5. Let's compare their cost structures.

By TechCompare · Updated

Cost Comparison

Based on 100,000 input tokens (50% cached), 5,000 output tokens, and 100 requests.

Option A
Mistral Small 4
Wins 3 of 4 compared specs
Option B
Mistral Medium 3.5
Wins 0 of 4 compared specs

Side-by-side specs

SpecMistral Small 4Mistral Medium 3.5
Input Cost (per M)$0.15 (better on this spec)$0.40
Output Cost (per M)$0.60 (better on this spec)$2.00
Cached Input (per M)$0.015 (better on this spec)$0.04
Batch Discount50%50%

How they differ

Mistral Small 4 costs $0.15 per million input tokens and $0.60 per million output tokens. Mistral Medium 3.5 is priced at $0.40 per million input tokens and $2.00 per million output tokens. Both support a 90% caching discount and a 50% batch discount.

Verdict

The base gap is roughly 2.7x on input ($0.15/M against $0.40) and 3.3x on output ($0.60/M against $2.00). Cached input widens that to about 2.7x at $0.015/M versus $0.04/M, and batch ties at 50% on both so the ratio holds in every mode. Medium 3.5 is the balanced middle tier; reach for it only when Small 4's output is too shallow for the translation or conversation work.

Which should you pick?

Choose Mistral Small 4

Sentiment analysis, high-caching search routing, and standard text summaries.

Choose Mistral Medium 3.5

Intermediate multilingual text generation, conversational assistants, and detailed reports.

Related comparisons

GPT-5.4 vs Claude Sonnet 4.6
The workhorse model pricing showdown.
Read comparison ➜
Claude Opus 4.7 vs DeepSeek V4 Pro
Frontier reasoning versus optimized price-performance.
Read comparison ➜
DeepSeek V4 Pro vs Mistral Large 3
Serverless pricing versus flagship open weights.
Read comparison ➜
Gemini 3.5 Flash vs GPT-5.4 Mini
Fast, lightweight multimodal models comparison.
Read comparison ➜
Gemini 3.1 Pro (<=200k) vs Claude Sonnet 4.6
Coding workhorses and reasoning model showdown.
Read comparison ➜
Gemini 3.5 Flash vs Claude Sonnet 4.6
Speedy utility model versus premium reasoning flagship.
Read comparison ➜

Frequently asked questions

Is Mistral Small 4 or Mistral Medium 3.5 cheaper for chatbots?
Mistral Small 4. At $0.15/M input and $0.60/M output it is well below Medium 3.5's $0.40/M and $2.00/M. Both share a 90% caching discount, so cached Small 4 is $0.015/M versus Medium's $0.04/M, a 2.7x gap. For conversational assistants and basic chatbots, Small 4 wins on cost.
When does Mistral Medium 3.5 earn its price premium?
For intermediate multilingual text generation, conversational assistants requiring deeper reasoning, and detailed report writing where Small 4's quality falls short. Medium 3.5 is the balanced tier, and is appropriate when Small 4 produces overly shallow outputs but Large 3 is overkill.
What does the spec table's batch discount tell us?
Both offer 50% batch discounts. Small 4 batch drops to $0.075/M input and $0.30/M output. Medium 3.5 batch drops to $0.20/M input and $1.00/M output. Small 4 stays cheaper in batch mode by roughly the same 2.7x ratio as live mode.