Key Specifications

Vendormeta
Version3.2-3b
Release Date2024-09-25
Context Window128000 tokens
Input Modalitiestext
Output Modalitiestext
LicenseLlama 3.2 Community License
Documentationhttps://llama.meta.com/docs/

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU52.5%2024-09-255-shotview
HUMANEVAL42.5pass@12024-09-25view
GSM8K46.4%2024-09-250-shot CoTview
MATH10%2024-09-250-shot CoTview
BBH46%2024-09-253-shot CoTview
GPQA18.8%2024-09-250-shotview
IFEVAL48.6%2024-09-25prompt_strictview
ARC76.3%2024-09-25challengeview
MUSR29.3%2024-09-250-shotview
WINOGRANDE67.8%2024-09-250-shotview

Pricing

TierPriceCurrency
Input$0.15 / MtokUSD
Output$0.15 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://ai.meta.com/blog/ · as of 2024-09-25

Compliance

  • Data Residency: self-host
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

Llama 3.2 3B

Model Overview

Meta Llama 3.2 3B 轻量开源模型, 128K 上下文, 3B 参数, 适合移动设备与边缘部署。

Core Specifications

VendorVersionRelease DateContext WindowInput ModalitiesOutput ModalitiesLicense
Meta3.2-3b2024-09-25128KtexttextLlama 3.2 Community License

Benchmark Performance

BenchmarkScoreUnitNotes
MMLU (Massive Multitask Language Understanding)52.5%5-shot
HumanEval42.5pass@1
GSM8K (Grade School Math 8K)46.4%0-shot CoT
MATH10.0%0-shot CoT
BBH (BIG-Bench Hard)46.0%3-shot CoT
GPQA18.8%0-shot
IFEval48.6%prompt_strict
ARC76.3%challenge
MUSR29.3%0-shot
WinoGrande67.8%0-shot

Pricing

InputOutputCache ReadCache Write

per million tokens

Strengths

  • Reliable general-purpose model.

Weaknesses

  • MMLU 52.5, weak knowledge reasoning.
  • HumanEval 42.5, coding weak.
  • Proprietary, not self-hostable.

Use Cases

  • General chat and Q&A

References