Key Specifications

Vendorother
Versionhermes-3-llama-3-1-405b
Release Date2024-08-12
Context Window131072 tokens
Input Modalitiestext
Output Modalitiestext
LicenseLlama 3.1 Community License
Documentationhttps://huggingface.co/models

Benchmark Performance

BenchmarkScoreUnitEvaluated AtNotesSource
MMLU83.9%2024-08-125-shotview
HUMANEVAL89pass@12024-08-12view
GSM8K88.9%2024-08-120-shot CoTview
MATH65.7%2024-08-120-shot CoTview
BBH82.1%2024-08-123-shot CoTview
GPQA40.5%2024-08-120-shotview
IFEVAL78.2%2024-08-12prompt_strictview
ARC94.5%2024-08-12challengeview
MUSR57.7%2024-08-120-shotview
WINOGRANDE86.5%2024-08-120-shotview

Pricing

TierPriceCurrency
Input$3 / MtokUSD
Output$3 / MtokUSD
Cache Read$0 / MtokUSD
Cache Write$0 / MtokUSD

Source: https://huggingface.co/models · as of 2024-08-12

Compliance

  • Data Residency: self-host
  • SOC2: ✗
  • HIPAA: ✗
  • GDPR: ✗
  • ISO 27001: ✗

Hermes 3 Llama 3.1 405B

Tổng quan mô hình

NousResearch Hermes 3 Llama 3.1 405B 微调模型, 131K 上下文, 改进中立性与推理, 性能超越基座模型。

Thông số kỹ thuật cốt lõi

Nhà cung cấpPhiên bảnNgày phát hànhCửa sổ ngữ cảnhPhương thức đầu vàoPhương thức đầu raGiấy phép
Otherhermes-3-llama-3-1-405b2024-08-12131KtexttextLlama 3.1 Community License

Hiệu suất benchmark

BenchmarkĐiểmĐơn vịGhi chú
MMLU (Massive Multitask Language Understanding)83.9%5-shot
HumanEval89.0pass@1
GSM8K (Grade School Math 8K)88.9%0-shot CoT
MATH65.7%0-shot CoT
BBH (BIG-Bench Hard)82.1%3-shot CoT
GPQA40.5%0-shot
IFEval78.2%prompt_strict
ARC94.5%challenge
MUSR57.7%0-shot
WinoGrande86.5%0-shot

Giá

Đầu vàoĐầu raĐọc bộ nhớ đệmGhi bộ nhớ đệm

mỗi triệu token

Điểm mạnh

  • MMLU score 83.9, strong knowledge reasoning.
  • HumanEval 89.0, excellent code generation.
  • GSM8K 88.9, robust math reasoning.

Điểm yếu

  • 闭源专有模型,不支持自托管。

Trường hợp sử dụng

  • 代码生成与调试

Tham chiếu