Mistral
Mistral AI: open-weight and commercial LLMs, Vibe agent (ex-Le Chat), and La Plateforme API. Free, Pro $14.99, Team $24.99/user, Enterprise; EU-hosted.
Pricing
Contact Sales
freemium
Category
AI Models
0 features tracked
Quick Links
Overview
Mistral AI is a Paris-based frontier AI lab that ships open-weight and commercial large language models, a consumer agent product, and enterprise tooling for custom training and private deployment. The company is best known for models such as Mistral Large 3, Mistral Medium 3.5, Mistral Small 4, Devstral, Codestral, Magistral, Voxtral (speech), and OCR 4, plus the hosted agent experience at chat.mistral.ai.
In May 2026 Mistral unified its consumer surface: Le Chat became Vibe—one agent and one licence for everyday work and coding. Free, Pro, Team, and Enterprise plans carry over conversation history and settings. Developers integrate through the La Plateforme API at console.mistral.ai, while enterprises can use Studio (agents and apps), Forge (custom training and alignment), Compute (GPU clusters), and applied AI services. Servers for Mistral Cloud are hosted in the EU; open-weight models can also be self-hosted or run via cloud partners (AWS, Azure, Google Cloud, NVIDIA NIM, and others).
Mistral’s practical pitch in 2026 is three-legged: competitive multimodal and coding models, aggressive open-weight releases (including Medium 3.5 under a modified MIT licence and Small 4 under Apache 2.0), and European data-residency / on-prem options that matter for regulated buyers who will not send all traffic to US-only labs.
Key features
- Vibe (ex-Le Chat) — Work Mode for multi-step tasks (inbox, research, docs, connectors) and Code Mode for remote coding sessions that open PRs; web, iOS/Android, VS Code extension, and CLI. Install CLI:
curl -LsSf https://mistral.ai/vibe/install.sh | bashoruv tool install mistral-vibe. - Remote coding agents — Cloud-sandboxed sessions spawned from CLI or chat; parallel runs;
/teleportmoves a local session to the cloud with history and approvals intact (Pro+). - Mistral Medium 3.5 — Default flagship in Vibe/Le Chat: dense 128B multimodal model, 256k context, configurable reasoning effort, open weights. Reported SWE-Bench Verified ~77.6%; API
mistral-medium-latestat $1.50 / $7.50 per million input/output tokens. - Mistral Large 3 — Open-weight general-purpose multimodal flagship (
mistral-large-latest), API $0.50 / $1.50 per M tokens; large checkpoints on Hugging Face (e.g. 675B-class instruct variants). - Mistral Small 4 — Efficient multimodal / agentic model under Apache 2.0; API $0.15 / $0.60 per M tokens; strong default for cost-sensitive production.
- Specialist models — Devstral 2 (agentic software engineering), Codestral (low-latency FIM/completion), Magistral Medium/Small (transparent reasoning), Ministral 3 (3B/8B/14B edge), Voxtral (TTS + transcription), OCR 4 (document extraction; page-priced), embeddings and moderation APIs.
- La Plateforme API — Chat completions-style API, batch (−50%), prompt caching (−90% on cached input), agent tools (code interpreter, web search, image gen, libraries/OCR indexing), classifier fine-tuning, regional inference endpoints.
- Studio, Forge, Compute — Agent orchestration and evals (Studio); custom training, RL, distillation (Forge); dedicated GPU infrastructure (Compute); enterprise SAML SSO, audit logs, white-label, and private deployments.
- Open weights + self-host — Official Hugging Face org releases; local deployment via vLLM, TensorRT-LLM, TGI, NVIDIA NIM; Vibe CLI can target OpenAI-compatible local endpoints for offline coding.
Tip: Treat chat Pro and API usage as separate products. Vibe Pro is a monthly seat with usage multipliers; API is prepaid/usage-based on console.mistral.ai and does not replace a Pro seat for all-day coding limits.
Pricing
Pricing below is taken from Mistral’s public pages (consumer plans at mistral.ai/pricing; API at mistral.ai/pricing/api). Taxes may apply by region. Confirm live figures before budgeting.
Vibe / chat plans (consumer & teams)
| Plan | Price (USD) | What you get (high level) |
|---|---|---|
| Free | $0 | Web + mobile; SOTA model access with limited messages, web searches, coding sessions, image gen; 100+ connectors; task scheduling up to 5; limited libraries/uploads. |
| Pro | $14.99 / month | Up to ~6× messages and ~5× web search vs free; ~40× image gen; all-day coding in CLI, IDE, and web; libraries up to 15 GB; unlimited task scheduling (fair use); chat + email support; PAYG credits to extend usage. |
| Team | $24.99 / user / month | Collaborative workspace; up to 30 GB storage per user; domain verification; data export; Pro-class usage multipliers; admin-oriented controls. |
| Education | $5.99 | Student plan (eligibility via student upgrade flow). |
| Enterprise | Contact sales | Custom models/agents/workflows; audit logs; SAML SSO; white label; private deployments and training options. |
API (selected models, per 1M tokens unless noted)
| Model / service | Input | Output / unit | Notes |
|---|---|---|---|
| Mistral Medium 3.5 | $1.50 | $7.50 | Default agentic/coding flagship; open weights |
| Mistral Large 3 | $0.50 | $1.50 | Open-weight multimodal generalist |
| Mistral Small 4 | $0.15 | $0.60 | Apache 2.0 open; efficient default |
| Devstral 2 | $0.40 | $2.00 | Agentic software engineering |
| Devstral Small 2 | $0.10 | $0.30 | Lightweight coding agents (Labs) |
| Codestral | $0.30 | $0.90 | Low-latency completion / FIM |
| Magistral Medium | $2.00 | $5.00 | Premier reasoning / “thinking” model |
| Magistral Small | $0.50 | $1.50 | Lighter reasoning tier |
| Ministral 3 (3B / 8B / 14B) | $0.10–$0.20 | $0.10–$0.20 | Edge-oriented open models |
| OCR 4 | — | $4 / 1k pages (OCR); $5 / 1k pages (Document AI) | Document extraction stack |
| Voxtral TTS | — | $0.016 / 1k characters | Speech synthesis / voice cloning |
| Mistral Embed / Codestral Embed | $0.10 / $0.15 | — | Text / code embeddings |
Batch processing is advertised at −50%; cached input tokens at −90% on input. Agent add-ons (examples): code execution ~$30 per 1k calls, web search ~$30 per 1k calls, image generation ~$100 per 1k images, premium news ~$50 per 1k calls, libraries OCR/indexing separate. Enterprise API adds regional controls, SLAs, higher rate limits, and premium support via sales.
Watch out: Model IDs rotate often (-latest aliases, dated snapshots, Labs vs Premier vs Open). Check deprecation tables in docs—many 2024–2025 endpoints were scheduled for retirement through mid-2026 with Medium 3.5 / Small 4 as replacements.
Limits & gotchas
- Free tier caps — Messages, web search, coding sessions, and image generation are limited; Free is for trying Vibe, not all-day agent work.
- Fair use on “unlimited” — Pro/Team scheduling and higher multipliers still sit under fair-use terms; heavy agent loops can still hit soft caps or need PAYG credits.
- Separate billing surfaces — Chat Pro ≠ API wallet. Teams building products need console credits; individuals coding in Vibe need Pro/Team seats for remote agents and higher quotas.
- Model lifecycle churn — Docs list long deprecation/retirement schedules (e.g. older Small, Medium, Devstral, Magistral, Pixtral, OCR 2). Pin versioned model IDs in production.
- Naming & product confusion — Le Chat, Vibe, Vibe CLI, Code Mode, Work Mode, Studio, and “La Plateforme” overlap in marketing; community threads often note hard-to-navigate docs and endpoint names (Devstral aliases, Labs prefixes).
- Not always the absolute frontier — HN and Reddit users often place Mistral behind OpenAI/Anthropic/Google on pure closed-model leaderboards, while praising price, open weights, EU hosting, and “good enough” quality.
- Self-hosting cost — Medium 3.5 (~128B) is marketable as multi-GPU self-host (Mistral cites as few as four GPUs), but Large-class and unquantized checkpoints still need serious hardware and ops.
- Connectors / MCP still maturing — Pricing table marks connectors and custom MCP connectors as beta; enterprise knowledge search quality depends on connector setup and permissions.
Community sentiment
On Hacker News, Mistral draws a consistent dual narrative. Positive: EU data residency and GDPR-friendly posture; strong cost-to-quality for API work; open-weight releases (7B lineage through Small 4 / Medium 3.5 / Large 3) that local and enterprise teams can actually run; Vibe CLI and Devstral for agentic coding; Le Chat Enterprise / on-prem announcements (2025) treated as a real differentiator for privacy-sensitive buyers. Negative: periodic claims that Mistral “fell behind” closed US labs after 2025; skepticism that EU branding alone is enough at frontier training cost; complaints about API reliability blips and confusing model catalogs.
Reddit (r/LocalLLaMA, r/MistralAI, r/MachineLearning) tends to treat Mistral as a core open-weights brand—Ministral/Small for local inference, Devstral and coding models for agents, OCR and Voxtral as useful specialists—while chat Pro users compare Vibe quotas and agent polish to ChatGPT Plus and Claude Pro rather than to pure API ranklists. Developers on GitHub/Hugging Face follow the official org for checkpoints (Medium 3.5 128B, Small 4 ~119B-class, Large 3 variants, Voxtral, Leanstral labs).
“I like Mistral—it hits the sweet spot between cost and my data staying in the EU, without a significant drop in quality—but their model lineup is easy to get lost in.”
— paraphrased developer sentiment common on HN threads around Forge, Devstral, and API naming
Who should use it
- EU / regulated teams that need EU-hosted SaaS, GDPR narrative, or on-prem / VPC deployment with open or commercial models.
- Builders who want open weights — Self-host Medium 3.5, Large 3, Small 4, Ministral, or Devstral with vLLM/NIM and still fall back to hosted API for spikes.
- Cost-sensitive API products — Small 4 and Large 3 price points undercut many frontier APIs for high-volume text; batch + cache deepen savings.
- Developers using agentic coding — Vibe CLI + VS Code + remote agents for refactors, tests, and PR-oriented workflows without fully committing to Claude Code or Cursor-only stacks.
- Document / speech pipelines — OCR 4 for extraction; Voxtral for TTS and realtime transcription; embeddings for RAG.
- Not ideal if you only want the single “best closed chat” experience with maximum brand ecosystem (plugins, voice, shopping) and never need open weights or EU residency—ChatGPT or Claude may fit better.
Alternatives
- ChatGPT (OpenAI) — Broader consumer ecosystem and often stronger closed multimodal demos; weaker open-weight story.
- Claude (Anthropic) — Preferred by many for long-context coding/writing and safety posture; chat Pro typically pricier than Mistral Pro.
- Gemini (Google) — Deep Google Workspace integration and long context; different privacy/geo tradeoffs.
- DeepSeek — Aggressive API pricing and open releases; different hosting and compliance profile.
- Llama (Meta) / Ollama — Pure local open-weight stacks without Mistral’s commercial agent suite.
- GitHub Copilot / Cursor / Claude Code — IDE-first coding agents if you only need code, not a full EU model lab.
Verdict
Mistral in mid-2026 is not “just another ChatGPT clone.” It is a full stack: Vibe (the Le Chat successor) at $14.99 Pro for agents and all-day coding, a broad usage-priced API from edge Ministrals to Medium 3.5 and Large 3, and a serious open-weight + EU/on-prem story that still differentiates against purely closed US labs. Medium 3.5 as a 128B open default with competitive SWE-Bench numbers, Small 4 under Apache 2.0, OCR/Voxtral specialists, and remote Vibe agents make it a credible primary stack for European enterprises and open-minded developers—or a strong secondary provider for multi-model products.
Choose Mistral when data residency, self-hosting, and $/token matter as much as peak closed-model scores. Pin model versions, budget chat seats and API wallets separately, and re-check pricing pages before large contracts—the catalog moves fast, and deprecations are real.
Alternatives
Best Alternatives to Mistral
Head-to-Head
Compare Mistral Side-by-Side
More in AI Models