
Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on WildBench and Arena Hard, reduces infinite generations, and delivers gains in tool use and structured output tasks.
It supports image and text inputs with structured outputs, function/tool calling, and strong performance across coding (HumanEval+, MBPP), STEM (MMLU, MATH, GPQA), and vision benchmarks (ChartQA, DocVQA).
| $0.075 | $0.20 | -- | 0.55s | 25 tps | ||
| $0.09 | $0.30 | $0.05 | 0.22s | 46 tps | ||
| $0.09375 | $0.25 | -- | 0.78s | 26 tps | ||
Not used in Standard routing:Why these endpoints are not used | ||||||
| $0.10 | $0.30 | $0.01 | 1.69s | 32 tps | ||
P50, best across providers
P50, best provider
When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.