Mistral's flagship 123B model with strong multilingual and coding performance. Supports 128K context.
Context Window
128K
tokens
Max Output
8K
tokens
Input Price
$2
per 1M tokens
Output Price
$6
per 1M tokens
Speed
78
tokens/sec
Released
Nov 2024
2024-11-18
Blended Cost
$4.00
per 1M tokens
Value Score
19.8
quality per $
Capabilities
ChatVisionCode GenerationFunction Calling
Benchmarks
Quality Index
79
MMLU Pro
84
HumanEval (Coding)
84.6
MATH
71.2
Arena ELO
1250
Price vs Quality
Cheaper is left, more capable is up. 22 of 62 priced models are both cheaper and score at least as high. Self-host models are left off the price axis — their cost depends on the hardware you run them on, not on tokens.
Context Window in Context
Grok 3131K
▸ Mistral Large 2128K
GPT-4o128K
GPT-4o Mini128K
DeepSeek V3128K
DeepSeek R1128K
Command R+128K
Gemma 4 27B128K
Bar length is logarithmic — each full step is 10× the tokens. Shown against the models with the nearest context windows, not the directory extremes.
Ограничено 1024 токенами вывода и 20 запросами в час. Полный контроль — в playground.
About Mistral Large 2
Mistral Large 2 is a flagship AI model by Mistral AI, released on November 18, 2024. It supports a context window of 128K tokens and can generate up to 8K output tokens.
At $2 per million input tokens and $6 per million output tokens, its blended cost of $4.00/1M tokens puts it in the mid-range pricing tier. Its value score of 19.8 reflects the balance of quality and cost.
Using Mistral Large 2 with Swfte
Access Mistral Large 2 through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try Mistral Large 2 in our AI Playground before integrating.
Deploy a model with Swfte Connect
One gateway, every provider, per-token cost visibility. Swap models without touching your code.