A 7.3B parameter model that outperforms Llama 2 13B on all benchmarks, with optimizations for speed and context length.
Model details →Mistral: Mistral 7B Instruct v0.1 vs NousResearch: Hermes 2 Pro - Llama-3 8B
NousResearch: Hermes 2 Pro - Llama-3 8B wins on 2 of 6 axes — pricing and capability skew in its favour for most workloads.
Hermes 2 Pro is an upgraded, retrained version of Nous Hermes 2, consisting of an updated and cleaned version of the OpenHermes 2.5 Dataset, as well as a newly introduced Function Calling and JSON Mode dataset developed in-house.
Model details →Side-by-side comparison
| Capability | Mistral: Mistral 7B Instruct v0.1 | NousResearch: Hermes 2 Pro - Llama-3 8B | Winner |
|---|---|---|---|
| Context window Maximum number of input tokens the model can attend to in a single request. | 3K | 8K | 🏆 NousResearch: Hermes 2 Pro - Llama-3 8B |
| Input price (per 1M) Cost per million input tokens billed by the provider. | $0.11 | $0.14 | 🏆 Mistral: Mistral 7B Instruct v0.1 |
| Output price (per 1M) Cost per million output tokens billed by the provider. | $0.19 | $0.14 | 🏆 NousResearch: Hermes 2 Pro - Llama-3 8B |
| Tool / function calling First-class support for emitting structured tool calls. | — | — | Tie |
| Vision input Accepts image inputs alongside text. | — | — | Tie |
| Reasoning mode Internal chain-of-thought / extended-thinking support. | — | — | Tie |
Frequently asked questions
Is Mistral: Mistral 7B Instruct v0.1 better than NousResearch: Hermes 2 Pro - Llama-3 8B?
NousResearch: Hermes 2 Pro - Llama-3 8B wins on 2 of 6 axes — pricing and capability skew in its favour for most workloads.
What's the price difference between Mistral: Mistral 7B Instruct v0.1 and NousResearch: Hermes 2 Pro - Llama-3 8B?
Input: $0.11 vs $0.14 per 1M tokens. Output: $0.19 vs $0.14 per 1M tokens.
What context windows do Mistral: Mistral 7B Instruct v0.1 and NousResearch: Hermes 2 Pro - Llama-3 8B support?
Mistral: Mistral 7B Instruct v0.1 supports up to 3K tokens. NousResearch: Hermes 2 Pro - Llama-3 8B supports up to 8K tokens.
Do both Mistral: Mistral 7B Instruct v0.1 and NousResearch: Hermes 2 Pro - Llama-3 8B support tool calling?
Mistral: Mistral 7B Instruct v0.1: not advertised. NousResearch: Hermes 2 Pro - Llama-3 8B: not advertised.