Key Specifications

Vendoralibaba
Version2.5-72b
Release Date2024-09-19
Context Window131072 tokens
Input Modalitiestext
Output Modalitiestext
LicenseQwen License
Documentationhttps://qwenlm.github.io/

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU86.1%2024-09-195-shotview
HUMANEVAL86.6pass@12024-09-19view
GSM8K88.4%2024-09-190-shot CoTview
MATH83.1%2024-09-190-shot CoTview
BBH82.4%2024-09-193-shot CoTview

Pricing

TierPriceCurrency
Input$0.5 / MtokUSD
Output$0.8 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://help.aliyun.com/zh/model-studio/getting-started/models · as of 2024-09-19

Compliance

  • Data Residency: CN
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

Qwen2.5 72B

Model Overview

Qwen2.5 72B is Alibaba’s flagship open-weights model, released on September 19, 2024 as part of the Qwen2.5 series. The model delivers leading Chinese-language understanding, strong multilingual coverage across 29+ languages, and a 131K context window. On academic benchmarks, Qwen2.5 72B matches or exceeds Llama 3.1 70B on MMLU (86.1) and significantly outperforms it on math (83.1 vs 71.8) and coding (86.6 HumanEval pass@1). Released under the Qwen License — which permits commercial use with minimal restrictions — Qwen2.5 72B has become a top choice for Chinese-market deployments, bilingual RAG systems, and code-generation pipelines where Chinese context matters. The model ships with native support for structured output, tool calling, and long-document summarization.

Key Specifications

AttributeValue
VendorAlibaba
Version2.5-72b
Release Date2024-09-19
Context Window131,072 tokens
Input Modalitiestext
Output Modalitiestext
LicenseQwen License
Documentationhttps://qwenlm.github.io/

Benchmark Performance

BenchmarkScoreUnitNotes
MMLU86.1%5-shot
HumanEval86.6pass@1
GSM8K88.4%0-shot CoT
MATH83.1%0-shot CoT
BBH82.4%3-shot CoT

Pricing

TierPrice (per 1M tokens)
Input (Alibaba Cloud)$0.50
Output (Alibaba Cloud)$0.80
Cache Read$0.00
Cache Write$0.00

Pricing source: https://help.aliyun.com/zh/model-studio/getting-started/models (as of 2024-09-19). Open-weights license also allows self-hosting.

Strengths

  • Best-in-class Chinese language understanding among open-weights models.
  • Strong coding and math performance, surpassing Llama 3.1 70B on multiple benchmarks.
  • 131K context window handles long Chinese documents and codebases.
  • Competitive API pricing at $0.50/$0.80 per million tokens.

Weaknesses

  • Qwen License has minor commercial restrictions vs Apache 2.0.
  • Vendor ecosystem is China-centric; community tooling outside Alibaba Cloud is less mature.
  • No native multimodal support in the 72B text variant (Qwen2.5-VL is a separate model).

Use Cases

  • Chinese-market customer service and content generation.
  • Bilingual RAG systems requiring balanced EN/ZH retrieval.
  • Code generation pipelines targeting Chinese developer audiences.
  • Academic research needing strong math and reasoning at lower cost.

References