RAG Pipeline Recommendation: Claude 3.5 Sonnet (2024-10-22)
RAG Pipeline → Claude 3.5 Sonnet (2024-10-22)
Anwendungsfallübersicht
Claude 3.5 Sonnet (2024-10-22) by Anthropic is a recommended option for the RAG Pipeline use case.
Empfohlene Modelle
| Rang | Anbieter | Kontextfenster | Score |
|---|
| #3 | Anthropic | 200K | 37/100 |
Leistungsanalyse
Benchmark-Leistung
| Benchmark | Ergebnis |
|---|
| MMLU | 87.2 |
| HUMANEVAL | 86.1 |
| GSM8K | 94.0 |
| MATH | 75.1 |
| BBH | 84.2 |
| GPQA | 50.9 |
Stärken
- MMLU score 87.2, strong knowledge reasoning.
- HumanEval 86.1, excellent code generation.
- GSM8K 94.0, robust math reasoning.
- 上下文窗口 200K,支持长文本。
Anforderungen
- 需要 anthropic 的 API 密钥
- 输入长度须在 200K 上下文窗口内