Model overview
Multimodal and multilingual work starts here: 0.15 per million input tokens and 0.6 for output make mistral-small-latest the least expensive Mistral model that accepts images. Against mistral-large-latest, the gap is 3.3 times on input (0.5) and 2.5 times on output (1.5) for an identical declared pair of multimodal and multilingual capabilities, which is the argument for starting at this tier and moving up only when quality forces it. The comparison with ministral-8b-latest is sharper: the same 0.15 to send, but 0.6 against 0.15 to receive, four times more, so answer length rather than prompt size decides between the two.
Billing conditions: Standard API price. Mistral publishes a 90% discount for cached input; cache-write prices are not published separately.
- Context window
- 0 tokens
- Knowledge cutoff
- —
- Price last verified
- First observed
Descriptions, capabilities, and tariffs come from the linked source. First-party prices take priority; OpenRouter fallbacks are labelled.