This model is deprecated. We only continue performance benchmarking for the default 10k input token workload. Results for other workloads are historical and no longer updated.
Nous Research has launched a newer model, Hermes 4 70B. We suggest considering it instead.
For more information, see comparison of Hermes 4 70B to other models and API provider benchmarks for Hermes 4 70B.
Hermes 3 - Llama-3.1 70B 智能、性能与价格分析
模型摘要
智能
速度
输入价格
输出价格
冗长度
Hermes 3 - Llama-3.1 70B 的智能水平低于平均水平,价格也尤其昂贵;比较对象为其他规模相近的开放权重非推理模型。 该模型支持文本输入,可输出文本,上下文窗口为 128k 个 token,知识截止至 2023年12月。
Hermes 3 - Llama-3.1 70B 在 Artificial Analysis Intelligence Index 上的得分为 5,在同类模型中低于平均水平(中位数:7)。
Hermes 3 - Llama-3.1 70B 每 100 万输入 token 的价格为 $0.70(昂贵,中位数:$0.18),每 100 万输出 token 的价格为 $0.70(有些昂贵,中位数:$0.40)。
Hermes 3 - Llama-3.1 70B 的速度为每秒 33 个 token,明显较慢(96)。
| 推理 | 否 此页面展示该模型的非推理版本。 可能还存在推理版本。 |
|---|---|
| 输入模态 | 支持:文本 |
| 输出模态 | 支持:文本 |
| 知识截止日期 | 2023年12月1日 |
| 上下文窗口 | 128k 约 192 页 A4 纸(12 号 Arial 字体) |
| 总参数量 | 70.6B |
| 许可证 | LLAMA 3.1 COMMUNITY LICENSE AGREEMENT |
| 模型权重 | Hugging Face |
指标与同类别模型进行比较:
- 非推理模型 → 仅与其他非推理模型比较
- 推理模型 → 同时与推理和非推理模型比较
- 开放权重模型 → 仅与规模类别相同的其他开放权重模型比较:
- 微型:≤4B 个参数
- 小型:4B–40B 个参数
- 中型:40B–150B 个参数
- 大型:>150B 个参数
- 专有模型 → 与价格区间相同的专有模型和开放权重模型比较,采用输入/输出价格 3:1 的混合比例:
- 每 100 万 token <$0.15
- 每 100 万 token $0.15–$1
- 每 100 万 token >$1
亮点
智能
Artificial Analysis Intelligence Index
Artificial Analysis Intelligence Index by Open Weights / Proprietary
Intelligence Evaluations
Agentic real-world work tasks, (Elo-500)/2000
Agentic tool use
Agentic coding & terminal use
Coding
Reasoning & knowledge
Scientific reasoning
Physics reasoning
Knowledge
1 - hallucination rate
Long context reasoning
Agentic knowledge work, Elo
Agentic SaaS workflows
Legal agentic work, criterion pass rate
Agentic business operations
Instruction following
Long-horizon agentic tasks
Kubernetes incident root-cause analysis
Visual reasoning
开放性指数
Artificial Analysis Openness Index: Score
Intelligence Index 比较
Intelligence Index vs. Cost per Intelligence Index Task
成本
Cost per Intelligence Index Task
Cost to Run Artificial Analysis Intelligence Index
Pricing: Cache Hit, Input, and Output
上下文窗口
Context Window
速度
按输出速度(每秒 token 数)衡量
Output Speed
Time per Intelligence Index Task
延迟
按首 Token 延迟(秒)衡量
Latency: Time To First Answer Token
端到端响应时间
Seconds to output 500 tokens, calculated based on time to first token, 'thinking' time for reasoning models, and output speed
End-to-End Response Time
模型规模(仅开放权重模型)
Model Size: Total and Active Parameters
常见问题
关于 Hermes 3 - Llama-3.1 70B 的常见问题
Hermes 3 - Llama-3.1 70B 发布于 2024年8月15日。
Hermes 3 - Llama-3.1 70B 由 Nous Research 开发。
Hermes 3 - Llama-3.1 70B 在 Artificial Analysis Intelligence Index 上的估算得分为 5,在其他规模相近的开放权重非推理模型中低于平均水平(中位数:7)。
Hermes 3 - Llama-3.1 70B 以每秒 32.8 个 token 的速度生成输出(基于提供该模型的各服务商中位数),与其他规模相近的开放权重非推理模型相比处于末段(中位数:95.7 t/s)。
Hermes 3 - Llama-3.1 70B 的首 Token 延迟(TTFT)为 1.97 秒(基于提供该模型的各服务商中位数),与其他规模相近的开放权重非推理模型相比略高于平均水平(中位数:1.77 秒)。
Hermes 3 - Llama-3.1 70B 每 100 万输入 token 的价格为 $0.70(略高于平均水平,中位数:$0.46),每 100 万输出 token 的价格为 $0.70(优于平均水平,中位数:$0.71);基于提供该模型的各服务商中位数。
Hermes 3 - Llama-3.1 70B 每 100 万输入 token 的价格为 $0.70,每 100 万输出 token 的价格为 $0.70(基于提供该模型的各服务商中位数)。按缓存命中/输入/输出为 7:2:1 的比例计算,混合价格为每 100 万 token $0.70。价格可能因服务商而异。 比较服务商价格
否,Hermes 3 - Llama-3.1 70B 不是推理模型。它不进行扩展的思维链推理,而是直接作答。
Hermes 3 - Llama-3.1 70B 支持文本输入。
Hermes 3 - Llama-3.1 70B 支持文本输出。
否,Hermes 3 - Llama-3.1 70B 不支持图像输入,只能处理文本。
否,Hermes 3 - Llama-3.1 70B 不是多模态模型,仅支持文本输入。
Hermes 3 - Llama-3.1 70B 的上下文窗口为 130k 个 token。这决定了模型在单次请求中可以处理多少文本和对话历史。
是,Hermes 3 - Llama-3.1 70B 是开放权重模型,其模型权重已公开,可供下载并自行托管。
Hermes 3 - Llama-3.1 70B 有 70.6B 参数。
Hermes 3 - Llama-3.1 70B 基于 LLAMA 3.1 COMMUNITY LICENSE AGREEMENT 许可证发布,该许可证允许商业使用。 查看许可证
Hermes 3 - Llama-3.1 70B 在 Artificial Analysis Intelligence Index 上的得分为 5。这项综合基准测试评估模型的推理、知识、数学和编程能力。
Hermes 3 - Llama-3.1 70B 的知识截止日期为 2023年12月,模型训练数据包含截至该日期的信息。
是,可通过 1 家服务商的 API 使用 Hermes 3 - Llama-3.1 70B。 比较 API 服务商
可通过 1 家 API 服务商使用 Hermes 3 - Llama-3.1 70B。 比较服务商
