Key Specifications

Vendoropenai
Version3.5
Release Date2022-11-30
Context Window4096 tokens
Input Modalitiestext
Output Modalitiestext
LicenseProprietary
Documentationhttps://platform.openai.com/docs/models

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU50.3%2022-11-305-shotview
HUMANEVAL28.2pass@12022-11-30view
GSM8K56.6%2022-11-300-shot CoTview
MATH29.4%2022-11-300-shot CoTview
BBH63.9%2022-11-303-shot CoTview
GPQA26.9%2022-11-300-shotview
IFEVAL52.8%2022-11-30prompt_strictview
ARC81.8%2022-11-30challengeview
MUSR40.4%2022-11-300-shotview
WINOGRANDE72.7%2022-11-300-shotview

Pricing

TierPriceCurrency
Input$2 / MtokUSD
Output$2 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://openai.com/api/pricing/ · as of 2022-11-30

Compliance

  • Data Residency: US
  • SOC2: ✓
  • HIPAA: ✓
  • GDPR: ✓
  • ISO 27001: ✓

GPT-3.5

Model Overview

OpenAI GPT-3.5 基础模型, 4K 上下文, ChatGPT 初始版本所用模型, 推理能力优于 GPT-3。

Core Specifications

VendorVersionRelease DateContext WindowInput ModalitiesOutput ModalitiesLicense
Openai3.52022-11-304KtexttextProprietary

Benchmark Performance

BenchmarkScoreUnitNotes
MMLU (Massive Multitask Language Understanding)50.3%5-shot
HumanEval28.2pass@1
GSM8K (Grade School Math 8K)56.6%0-shot CoT
MATH29.4%0-shot CoT
BBH (BIG-Bench Hard)63.9%3-shot CoT
GPQA26.9%0-shot
IFEval52.8%prompt_strict
ARC81.8%challenge
MUSR40.4%0-shot
WinoGrande72.7%0-shot

Pricing

InputOutputCache ReadCache Write

per million tokens

Strengths

  • Reliable general-purpose model.

Weaknesses

  • MMLU 50.3, weak knowledge reasoning.
  • HumanEval 28.2, coding weak.
  • Proprietary, not self-hostable.
  • Context window 4K is limited.

Use Cases

  • General chat and Q&A

References