Questions? hello@viberation.devGet supportBlogDocsChangelog
Get started

Mistral

France's frontier lab, with the small models everyone self-hosts.

Mistral is the main model family from the French lab of the same name, and it splits in a way worth knowing about: the Small models publish their weights so you can download and run them, while Medium and Large are sold through the API only. Small 4 folds several older models into one and is cheap enough to use by default; Medium 3.5 is the one to reach for on agentic and coding work. Mistral's other lines are separate products with their own names — Devstral and Codestral for code, Ministral for tiny on-device models, and Voxtral for speech, which has its own entry here.

At a glance — Mistral Small 4

Cost
$0.15 in · $0.60 out per 1M tokensBudget
Memory
262K tokens — about 400 pages
Reads
Text · Images
Can do
  • Use tools (agents)
  • Structured output
  • Thinking mode
Good at
Coding: better than 26% · Agentic work: better than 4% · Overall: better than 16%
Reliability
100.00% uptime, last 24h
Released
Mar 16, 2026
Advanced details
API model ID
mistralai/mistral-small-2603
Max output
210K tokens
Cached input
$0.015 per 1M
Tokenizer
Mistral
Coding index
26.6
Agentic work index
0.8
Overall index
11.3

Providers

The same model, hosted by different companies. OpenRouter picks one per request and falls back to the next if it fails.

ProviderInput /1MOutput /1MCached /1MMax outputUptime 24h
Mistral · zdr$0.15$0.60$0.015210K99.97%
Mistral$0.15$0.60$0.015210K99.96%
Mistral · eu$0.165$0.66$0.016210K100.00%
Mistral · us$0.165$0.66$0.016210K100.00%

Supported parameters

  • frequency_penalty
  • include_reasoning
  • max_tokens
  • presence_penalty
  • reasoning
  • reasoning_effort
  • response_format
  • seed
  • stop
  • structured_outputs
  • temperature
  • tool_choice
  • tools
  • top_p

Live from OpenRouter, refreshed hourly. Scores from Artificial Analysis.

Key info

Pricing
Freemium
Category
Models
Best for
Intermediate
Made by
Mistral AI, in France
Weights
Open for Small and Nemo, closed for Medium and Large
Good for
A cheap default model, and EU-hosted inference

More in Models

  • Meta's current line, built for long multi-agent runs.

    Models
  • The open-weight family everyone else gets benchmarked against.

    Models
  • Google's open models — Gemini's research, small enough to self-host.

    Models
  • Z.ai's open-weight family, behind the cheap coding subscriptions.

    Models