Key Specifications

Vendoralibaba
Version2.5-3b
Release Date2024-09-19
Context Window32768 tokens
Input Modalitiestext
Output Modalitiestext
LicenseQwen License
Documentationhttps://qwenlm.github.io/

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU56.8%2024-09-195-shotview
HUMANEVAL49pass@12024-09-19view
GSM8K31.3%2024-09-190-shot CoTview
MATH17%2024-09-190-shot CoTview
BBH48.8%2024-09-193-shot CoTview
GPQA18%2024-09-190-shotview
IFEVAL56%2024-09-19prompt_strictview
ARC76.5%2024-09-19challengeview
MUSR23.3%2024-09-190-shotview
WINOGRANDE61.4%2024-09-190-shotview

Pricing

TierPriceCurrency
Input$0.1 / MtokUSD
Output$0.1 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://help.aliyun.com/zh/model-studio/getting-started/models · as of 2024-09-19

Compliance

  • Data Residency: CN
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

Qwen2.5 3B

Aperçu du modèle

阿里巴巴 Qwen2.5 3B 轻量开源模型, 32K 上下文, 3B 参数, 适合移动设备与边缘部署。

Spécifications principales

FournisseurVersionDate de sortieFenêtre de contexteModalités d’entréeModalités de sortieLicence
Alibaba2.5-3b2024-09-1932KtexttextQwen License

Performance aux benchmarks

BenchmarkScoreUnitéNotes
MMLU (Massive Multitask Language Understanding)56.8%5-shot
HumanEval49.0pass@1
GSM8K (Grade School Math 8K)31.3%0-shot CoT
MATH17.0%0-shot CoT
BBH (BIG-Bench Hard)48.8%3-shot CoT
GPQA18.0%0-shot
IFEval56.0%prompt_strict
ARC76.5%challenge
MUSR23.3%0-shot
WinoGrande61.4%0-shot

Tarification

EntréeSortieLecture cacheÉcriture cache

par million de jetons

Forces

  • 可靠的通用模型。

Faiblesses

  • MMLU 仅 56.8,知识推理偏弱。
  • HumanEval 49.0,代码能力较弱。
  • 闭源专有模型,不支持自托管。

Cas d’usage

  • 通用对话与问答

Références