Key Specifications

Vendorother
Versionstablelm-3-4b
Release Date2024-06-26
Context Window4096 tokens
Input Modalitiestext
Output Modalitiestext
LicenseStability AI Community License
Documentationhttps://huggingface.co/models

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU50.6%2024-06-265-shotview
HUMANEVAL32.6pass@12024-06-26view
GSM8K48.1%2024-06-260-shot CoTview
MATH11.5%2024-06-260-shot CoTview
BBH54.9%2024-06-263-shot CoTview
GPQA26.8%2024-06-260-shotview
IFEVAL51.6%2024-06-26prompt_strictview
ARC86.8%2024-06-26challengeview
MUSR36.6%2024-06-260-shotview
WINOGRANDE74.1%2024-06-260-shotview

Pricing

TierPriceCurrency
Input$0.15 / MtokUSD
Output$0.15 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://huggingface.co/models · as of 2024-06-26

Compliance

  • Data Residency: self-host
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

StableLM 3 4B

Model Overview

Stability AI StableLM 3 4B 模型, 4K 上下文, 多语言 (English/Spanish/German/Italian/French)。

Core Specifications

VendorVersionRelease DateContext WindowInput ModalitiesOutput ModalitiesLicense
Otherstablelm-3-4b2024-06-264KtexttextStability AI Community License

Benchmark Performance

BenchmarkScoreUnitNotes
MMLU (Massive Multitask Language Understanding)50.6%5-shot
HumanEval32.6pass@1
GSM8K (Grade School Math 8K)48.1%0-shot CoT
MATH11.5%0-shot CoT
BBH (BIG-Bench Hard)54.9%3-shot CoT
GPQA26.8%0-shot
IFEval51.6%prompt_strict
ARC86.8%challenge
MUSR36.6%0-shot
WinoGrande74.1%0-shot

Pricing

InputOutputCache ReadCache Write

per million tokens

Strengths

  • Reliable general-purpose model.

Weaknesses

  • MMLU 50.6, weak knowledge reasoning.
  • HumanEval 32.6, coding weak.
  • Proprietary, not self-hostable.
  • Context window 4K is limited.

Use Cases

  • General chat and Q&A

References