ModelRiskIndex

Rankings / Mistral AI

Mistral Medium 3.1

Tier 140/100draft — pending re-verification

mistral-medium-2508 via api.mistral.ai

Usage share 0.01% · OpenRouter rankings API (daily token share, 2026-08-04)

Tier assessment
Tier 1 requirements
  • Published model card. A model card or equivalent technical documentation is published for this model.Model documentation is thinner than frontier-lab system cards.
  • Published safety evals. Safety evaluations for this model are published.Limited: moderation benchmarks and policy documentation rather than a full safety eval suite.
  • Documented safety policy. A documented safety, usage, or acceptable-use policy governs the model.
  • Enterprise data controls. Customer data is not used for training by default, or a documented opt-out exists.
Tier 2 requirements
  • External pre-deployment testing. No disclosed external pre-deployment testing.
  • Third-party certification. The operating organization holds verifiable third-party certification (e.g. SOC 2, ISO/IEC 42001).
  • Versioning with changelogs. Model versions are explicitly identified and changes are changelogged.Dated version labels (e.g. -2508); changelog quality is inconsistent.
  • Stated deprecation policy. A deprecation policy with notice windows is published.

Missing for Tier 2: external pre-deployment testing. The tier is computed from this checklist — satisfying these requirements moves the badge, automatically.

Risk analysisfive vectors · click a wedge for its evidence

Risk vectors — the receipts

Data governancepartial

What happens to your data: training-on-customer-data defaults, retention windows, residency options, and the enterprise-versus-consumer terms gap. A legal property, not a capability — it survives every model generation.

EU residency and GDPR posture are genuine strengths; paid La Plateforme traffic is not used for training by default, but free-tier and some product-surface defaults are less clean, and documentation is thinner than the majors'.

Receipts (1)

Operational stabilitypartial

Whether it changes without warning: versioning discipline, changelog quality, deprecation policy, and observed silent changes. The signal no one else tracks.

Dated version labels (e.g. -2508) and a deprecation table in docs; changelog quality is inconsistent.

Receipts (1)

Adversarial resistanceweak

Whether an attacker can make it misbehave — direct jailbreaks against the model's own policies and indirect prompt injection in agentic tool use. Graded to the weaker of the two, because an attacker takes the easier path.

Jailbreak resistanceweak

Independent security indexes have consistently placed Mistral's hosted models in the lower tier for attack resistance; moderation is available as a separate API rather than built into model behavior.

Prompt injection (agentic)weak

No published agentic-injection hardening or agent red-teaming results at time of entry.

Receipts (2)

Transparencypartial

Whether you can see how it was built and tested: model cards, published safety evals, external pre-deployment testing, and disclosure of changes. The mechanism behind the tier ladder.

Model announcements and basic cards exist, but no external pre-deployment testing and sparse published safety evals. Mistral publicly disputes the framing of company-level safety indexes.

Receipts (2)
  • Mistral Medium 3.1 announcement
    Mistral AI provider artifacts · provider artifact · source tier B · mistral.ai · retrieved 2026-08-03
    Performance accuracy on all benchmarks were obtained through the same internal evaluation pipeline.
  • FLI AI Safety Index (company-level context)
    FLI AI Safety Index · institutional · source tier D · safetyindex.ai · retrieved 2026-08-03
    Company-level assessment shown as context only; never sets model-level grades.

Compliance posturepartial

Whether it is certified and compliant: SOC 2, ISO/IEC 42001, HIPAA eligibility, EU AI Act readiness, and audit availability.

SOC 2 attested; EU AI Act code-of-practice signatory with EU jurisdiction as a structural advantage; thinner certification stack than hyperscaler-hosted rivals.

Receipts (1)
  • Mistral AI trust and terms
    Mistral AI provider artifacts · provider artifact · source tier B · mistral.ai · retrieved 2026-08-03

Governance & evidence

Where your data goes

  • EU

No regional pinning — the provider chooses where data is processed.

EU-hosted La Plateforme (primary selling point); hyperscaler deployments via Azure and Bedrock follow those platforms' regions.

Enterprise vs consumer terms

Enterprise vs consumer gap: narrow. Paid La Plateforme traffic is not used for training by default, but free-tier and some product-surface defaults are less clean.CONSUMERENTERPRISE / APIWORSE TERMS →
Narrow gap

Paid La Plateforme traffic is not used for training by default, but free-tier and some product-surface defaults are less clean.

Change cadence

no tracked changesNo tracked changes for Mistral Medium 3.1. Absence of detection is not evidence of stability.none

No tracked changes for this model. That is absence of detection, not evidence of stability — it may mean the model is unwatched, not that it is unchanging.

Score volatility

No dated score readings recorded for this model yet. Readings are only entered where multiple real, dated third-party values exist — never interpolated.

Receipts — what backs this assessment

8 evidence refs4 distinct sources1 independent
  • CF5 Labs CASI/ARS leaderboardindependent eval×1 reference
  • BMistral AI provider artifactsprovider artifact×5 references
  • DFLI AI Safety Indexinstitutional×1 reference
  • EOpenRouter model rankingsusage data×1 reference

retrieved 2026-08-03 — 2026-08-05

Compliance & deployment

Trains on customer data by default
No
SOC 2
Yes
ISO/IEC 42001
not verified
HIPAA eligible
not verified
Retention window
Documented retention for abuse monitoring; EU-hosted
Data residency
EU (primary selling point)
EU AI Act
EU-headquartered; GPAI Code of Practice signatory.
Deprecation policy
Deprecation table maintained in API documentation
Available via
La Plateforme (api.mistral.ai) · Azure AI Foundry · AWS Bedrock

Change timeline

No tracked changes yet for this model.

Compare this model: Mistral Medium 3.1 + open compare view →