Key Specifications

Vendoropenai
Version3.5
Release Date2022-11-30
Context Window4096 tokens
Input Modalitiestext
Output Modalitiestext
LicenseProprietary
Documentationhttps://platform.openai.com/docs/models

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU50.3%2022-11-305-shotview
HUMANEVAL28.2pass@12022-11-30view
GSM8K56.6%2022-11-300-shot CoTview
MATH29.4%2022-11-300-shot CoTview
BBH63.9%2022-11-303-shot CoTview
GPQA26.9%2022-11-300-shotview
IFEVAL52.8%2022-11-30prompt_strictview
ARC81.8%2022-11-30challengeview
MUSR40.4%2022-11-300-shotview
WINOGRANDE72.7%2022-11-300-shotview

Pricing

TierPriceCurrency
Input$2 / MtokUSD
Output$2 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://openai.com/api/pricing/ · as of 2022-11-30

Compliance

  • Data Residency: US
  • SOC2: ✓
  • HIPAA: ✓
  • GDPR: ✓
  • ISO 27001: ✓

GPT-3.5

Modeloverzicht

OpenAI GPT-3.5 基础模型, 4K 上下文, ChatGPT 初始版本所用模型, 推理能力优于 GPT-3。

Kernspecificaties

LeverancierVersieReleasedatumContextvensterInvoermodaliteitenUitvoermodaliteitenLicentie
Openai3.52022-11-304KtexttextProprietary

Benchmarkprestaties

BenchmarkScoreEenheidNotities
MMLU (Massive Multitask Language Understanding)50.3%5-shot
HumanEval28.2pass@1
GSM8K (Grade School Math 8K)56.6%0-shot CoT
MATH29.4%0-shot CoT
BBH (BIG-Bench Hard)63.9%3-shot CoT
GPQA26.9%0-shot
IFEval52.8%prompt_strict
ARC81.8%challenge
MUSR40.4%0-shot
WinoGrande72.7%0-shot

Prijzen

InvoerUitvoerCache-lezenCache-schrijven

per miljoen tokens

Sterktes

  • 可靠的通用模型。

Zwaktes

  • MMLU 仅 50.3,知识推理偏弱。
  • HumanEval 28.2,代码能力较弱。
  • 闭源专有模型,不支持自托管。
  • 上下文窗口 4K 偏小。

Gebruiksscenario’s

  • 通用对话与问答

Referenties