Qwen 2.5 72B

Alibaba CloudbalancedOpen Source

Alibaba's flagship open-source model. Competitive with GPT-4o class models on benchmarks at a fraction of the cost.

Context Window

131K

tokens

Max Output

8K

tokens

Input Price

$0.3

per 1M tokens

Output Price

$0.9

per 1M tokens

Speed

85

tokens/sec

Released

Sep 2024

2024-09-19

Blended Cost

$0.60

per 1M tokens

Value Score

133.3

quality per $

Capabilities

ChatCode GenerationFunction Calling

Benchmarks

Quality Index
80
MMLU Pro
85.3
HumanEval (Coding)
86.4
MATH
78.9
Arena ELO
1255

Price vs Quality

6080100$1.00$10.00$100.00Blended cost per 1M tokens (log scale) →GPT-4o — $6.25/1M, quality 85GPT-4o Mini — $0.375/1M, quality 72o3 Mini — $2.75/1M, quality 88o3 — $5.00/1M, quality 94GPT-4.1 — $5.00/1M, quality 89Claude Opus 4 — $45.00/1M, quality 91Claude Sonnet 4 — $9.00/1M, quality 88Claude 3.5 Haiku — $2.40/1M, quality 75Gemini 2.5 Pro — $5.63/1M, quality 92Gemini 2.0 Flash — $0.250/1M, quality 74Llama 4 Maverick — $0.400/1M, quality 80Llama 4 Scout — $0.275/1M, quality 71Mistral Large 2 — $4.00/1M, quality 79Codestral — $0.600/1M, quality 76DeepSeek V3 — $0.685/1M, quality 86DeepSeek R1 — $1.37/1M, quality 91Grok 3 — $9.00/1M, quality 87Grok 3 Mini — $0.400/1M, quality 78Command R+ — $6.25/1M, quality 68Amazon Nova Pro — $2.00/1M, quality 70Qwen 2.5 Coder 32B — $0.300/1M, quality 74Sonar Pro — $9.00/1M, quality 78GPT-5.5 — $17.50/1M, quality 97GPT-5.5 Pro — $105.00/1M, quality 96Claude Opus 4.7 — $15.00/1M, quality 96Gemini 3.1 Pro — $7.00/1M, quality 96DeepSeek V4 Pro — $1.32/1M, quality 89DeepSeek V4 Flash — $0.210/1M, quality 84Qwen 3.6 Plus — $3.50/1M, quality 86Grok 4.3 — $1.88/1M, quality 93Kimi K2.6 — $2.11/1M, quality 92Kimi K3 — $9.00/1M, quality 97Claude Sonnet 4.6 — $9.00/1M, quality 90GLM-5.1 — $2.03/1M, quality 88Claude Opus 4.8 — $15.00/1M, quality 98Qwen 3.7 Max — $5.00/1M, quality 94Claude Fable 5 — $30.00/1M, quality 100MiniMax M3 — $1.50/1M, quality 89GPT-5.6 Sol — $17.50/1M, quality 98GPT-5.6 Terra — $7.00/1M, quality 93GPT-5.6 Luna — $0.700/1M, quality 83Gemini 3.6 Flash — $2.25/1M, quality 90Gemini 3.5 Flash-Lite — $1.40/1M, quality 79Claude Sonnet 5 — $9.00/1M, quality 93GLM-5.2 — $2.90/1M, quality 91Claude Opus 5 — $15.00/1M, quality 99Claude Mythos 5 — $30.00/1M, quality 100Grok 4.5 — $4.00/1M, quality 94Qwen3.8 Max — $4.00/1M, quality 96Qwen 2.5 72B — $0.600/1M, quality 80Qwen 2.5 72B
Cheaper is left, more capable is up. 2 of 49 priced models are both cheaper and score at least as high. Self-host models are left off the price axis — their cost depends on the hardware you run them on, not on tokens.

Context Window in Context

  • Qwen 2.5 72B131K
  • Grok 3131K
  • Grok 3 Mini131K
  • Qwen 2.5 Coder 32B131K
  • GPT-4o128K
  • GPT-4o Mini128K
  • Mistral Large 2128K
  • DeepSeek V3128K
Bar length is logarithmic — each full step is 10× the tokens. Shown against the models with the nearest context windows, not the directory extremes.

Open Source: Licensed under Apache 2.0

尚未接入我们的网关

Qwen 2.5 72B 是开放权重模型,你可以自行部署。Swfte Connect 会把你的部署变成一个托管端点,使用与本目录中其他模型相同的 API。

请求
curl https://api.swfte.com/agents/v2/gateway/chat/completions \
  -H "Authorization: Bearer $SWFTE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "alibaba:qwen-2.5-72b-instruct",
    "messages": [{"role": "user", "content": "Hello"}]
  }'
使用 Swfte Connect 部署此模型

About Qwen 2.5 72B

Qwen 2.5 72B is a balanced AI model by Alibaba Cloud, released on September 19, 2024. It supports a context window of 131K tokens and can generate up to 8K output tokens.

At $0.3 per million input tokens and $0.9 per million output tokens, its blended cost of $0.60/1M tokens makes it one of the most affordable models available. Its value score of 133.3 reflects the balance of quality and cost.

Qwen 2.5 72B is available as an open-source model under the Apache 2.0 license, meaning you can self-host it for predictable costs or use it through API providers like Swfte Connect.

Using Qwen 2.5 72B with Swfte

Access Qwen 2.5 72B through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try Qwen 2.5 72B in our AI Playground before integrating.

Deploy a model with Swfte Connect

One gateway, every provider, per-token cost visibility. Swap models without touching your code.