Skip to content

Mistral API Pricing

Last updated: 47m ago

Complete pricing breakdown for all 0 Mistral API models. Compare input and output costs per million tokens for Mistral Large, Mistral Medium, Mistral Small, Codestral, and every other model. Includes a cost calculator and side-by-side comparison with OpenAI and Anthropic.

Cheapest ($/1M out)
Most Expensive ($/1M out)
0
Free Models
Free
Avg Output $/1M

Mistral Model Pricing -- All Models

0 models sorted by output price
ModelInput $/1MOutput $/1M

Mistral vs OpenAI vs Anthropic -- Pricing Comparison

See how Mistral API pricing stacks up against OpenAI (GPT) and Anthropic (Claude) models. All prices in USD per million tokens.

Mistral

0 models
ModelInOut

OpenAI

97 models
ModelInOut
gpt-oss-20b (free)FreeFree
SoraFreeFree
gpt-oss-20b$0.030$0.130
gpt-oss-120b$0.037$0.170
GPT-5 Nano (batch)$0.025$0.200
GPT-4.1 Nano (batch)$0.050$0.200
gpt-oss-safeguard-20b$0.075$0.300
GPT-4o-mini (batch)$0.075$0.300

Anthropic

28 models

Mistral API Cost Calculator

Mistral API cost projections at various request volumes. Based on ~1,000 input tokens and ~500 output tokens per request. Self-hosting open-weight models eliminates these per-token costs.

Note: Actual costs vary with prompt length, response length, and batch processing. Mistral offers competitive pricing and batch API discounts for high-volume usage. Try the interactive calculator for custom estimates.

Understanding Mistral API Pricing

Token-Based Billing

Mistral uses a SentencePiece tokenizer shared across their model family. Token counts are consistent whether you use Mistral Large or Small - only the per-token price changes. Prices are quoted per million tokens. As an open-weight provider, Mistral also lets you self-host models to eliminate per-token API costs entirely for high-volume workloads.

Mistral Large vs Small vs Codestral

Mistral Large is the most powerful model for complex reasoning, multilingual tasks, and code generation. Mistral Small offers an excellent balance of performance and cost for most production use cases. Codestral is specialized for code completion and generation, supporting 80+ programming languages with optimized performance.

Open-Weight Models

Mistral is known for releasing open-weight models that can be self-hosted. While the API provides managed access with pay-per-token pricing, you can also download and run models like Mistral 7B, Mixtral, and others on your own infrastructure, potentially reducing costs for high-volume workloads.

Saving on API Costs

Start with Mistral Small for simple tasks and only upgrade to Large when the task demands it. Use shorter prompts and set appropriate max_tokens limits to control output costs. For high-volume workloads, consider self-hosting open-weight models or using batch API endpoints for discounted pricing.

Mistral API Pricing FAQ

Mistral Large pricing is Free/1M input tokens and Free/1M output tokens. Mistral Large is Mistral AI's most capable model, designed for complex reasoning, multilingual tasks, and code generation.

Mistral offers 0 free models via their API. Mistral provides competitive pricing with some of the most affordable models in the industry. Mistral Small is one of the cheapest high-quality options for production workloads, while Mistral Large offers premium capabilities at a fraction of the cost of comparable models from competitors.

Pricing data is currently being updated. Check back soon for the latest Mistral model costs.

Mistral and OpenAI target different price points. Mistral's average output price is Free/1M tokens across 0 paid models. OpenAI offers 97 models with varying price points. Mistral models are generally more affordable than OpenAI equivalents, especially for multilingual and code-generation use cases. For the most accurate comparison, see our side-by-side pricing table above.

Mistral charges per token with pricing that reflects their model hierarchy - from the affordable Mistral Small to the premium Mistral Large. As an open-weight provider, their API pricing competes with self-hosting costs. Codestral is priced separately for code-specific workloads. All models share the same tokenizer, so token counts are consistent across the lineup.

Codestral is Mistral's specialized code generation model, optimized for programming tasks like code completion, refactoring, and generation. It offers competitive pricing compared to general-purpose models and excels at code-related tasks with support for 80+ programming languages. It's available via the Mistral API and through various IDE integrations.

Mistral API Pricing - All Model Costs (2026) | LM Market Cap