Key Specifications

Vendorother
Versionmpt-30b
Release Date2023-06-22
Context Window8192 tokens
Input Modalitiestext
Output Modalitiestext
LicenseApache 2.0
Documentationhttps://huggingface.co/models

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU55.2%2023-06-225-shotview
HUMANEVAL49.7pass@12023-06-22view
GSM8K59.6%2023-06-220-shot CoTview
MATH21.2%2023-06-220-shot CoTview
BBH49.2%2023-06-223-shot CoTview
GPQA24.2%2023-06-220-shotview
IFEVAL51.2%2023-06-22prompt_strictview
ARC80.7%2023-06-22challengeview
MUSR28.9%2023-06-220-shotview
WINOGRANDE77.9%2023-06-220-shotview

Pricing

TierPriceCurrency
Input$0.4 / MtokUSD
Output$0.4 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://huggingface.co/models · as of 2023-06-22

Compliance

  • Data Residency: self-host
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

MPT 30B

Model Overview

MosaicML MPT 30B 开源模型, 8K 上下文, 300 亿参数, 改进长上下文处理, Apache 2.0 可商用。

Core Specifications

VendorVersionRelease DateContext WindowInput ModalitiesOutput ModalitiesLicense
Othermpt-30b2023-06-228KtexttextApache 2.0

Benchmark Performance

BenchmarkScoreUnitNotes
MMLU (Massive Multitask Language Understanding)55.2%5-shot
HumanEval49.7pass@1
GSM8K (Grade School Math 8K)59.6%0-shot CoT
MATH21.2%0-shot CoT
BBH (BIG-Bench Hard)49.2%3-shot CoT
GPQA24.2%0-shot
IFEval51.2%prompt_strict
ARC80.7%challenge
MUSR28.9%0-shot
WinoGrande77.9%0-shot

Pricing

InputOutputCache ReadCache Write

per million tokens

Strengths

  • Reliable general-purpose model.

Weaknesses

  • MMLU 55.2, weak knowledge reasoning.
  • HumanEval 49.7, coding weak.
  • Proprietary, not self-hostable.
  • Context window 8K is limited.

Use Cases

  • General chat and Q&A

References