Key Specifications

Vendorother
Versionyi-large
Release Date2024-05-13
Context Window32768 tokens
Input Modalitiestext
Output Modalitiestext
LicenseProprietary
Documentationhttps://huggingface.co/models

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU82.8%2024-05-135-shotview
HUMANEVAL72.2pass@12024-05-13view
GSM8K81.3%2024-05-130-shot CoTview
MATH55.9%2024-05-130-shot CoTview
BBH77.1%2024-05-133-shot CoTview
GPQA35.2%2024-05-130-shotview
IFEVAL73.4%2024-05-13prompt_strictview
ARC94.8%2024-05-13challengeview
MUSR59.4%2024-05-130-shotview
WINOGRANDE84.1%2024-05-130-shotview

Pricing

TierPriceCurrency
Input$3 / MtokUSD
Output$3 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://huggingface.co/models · as of 2024-05-13

Compliance

  • Data Residency: self-host
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

Yi Large

Modellübersicht

01.AI Yi Large 旗舰闭源模型, 32K 上下文, 中英文能力突出, 在 LMSYS 排行榜上接近 GPT-4。

Kernspezifikationen

AnbieterVersionVeröffentlichungsdatumKontextfensterEingabemodalitätenAusgabemodalitätenLizenz
Otheryi-large2024-05-1332KtexttextProprietary

Benchmark-Leistung

BenchmarkErgebnisEinheitNotizen
MMLU (Massive Multitask Language Understanding)82.8%5-shot
HumanEval72.2pass@1
GSM8K (Grade School Math 8K)81.3%0-shot CoT
MATH55.9%0-shot CoT
BBH (BIG-Bench Hard)77.1%3-shot CoT
GPQA35.2%0-shot
IFEval73.4%prompt_strict
ARC94.8%challenge
MUSR59.4%0-shot
WinoGrande84.1%0-shot

Preise

EingabeAusgabeCache-LesenCache-Schreiben

pro Million Token

Stärken

  • MMLU score 82.8, strong knowledge reasoning.

Schwächen

  • 闭源专有模型,不支持自托管。

Anwendungsfälle

  • 代码生成与调试
  • Agent 工作流与工具调用

Referenzen