A large LLM created by combining two fine-tuned Llama 70B models into one 120B model. Combines Xwin and Euryale. Credits to - [@chargoddard](https://huggingface.co/chargoddard) for developing the framework used to merge the model - [mergek…
Model details →Goliath 120B vs SorcererLM 8x22B
SorcererLM 8x22B wins on 2 of 6 axes — pricing and capability skew in its favour for most workloads.
SorcererLM is an advanced RP and storytelling model, built as a Low-rank 16-bit LoRA fine-tuned on [WizardLM-2 8x22B](/microsoft/wizardlm-2-8x22b). - Advanced reasoning and emotional intelligence for engaging and immersive interactions - V…
Model details →Side-by-side comparison
| Capability | Goliath 120B | SorcererLM 8x22B | Winner |
|---|---|---|---|
| Context window Maximum number of input tokens the model can attend to in a single request. | 6K | 16K | 🏆 SorcererLM 8x22B |
| Input price (per 1M) Cost per million input tokens billed by the provider. | $3.75 | $4.50 | 🏆 Goliath 120B |
| Output price (per 1M) Cost per million output tokens billed by the provider. | $7.50 | $4.50 | 🏆 SorcererLM 8x22B |
| Tool / function calling First-class support for emitting structured tool calls. | — | — | Tie |
| Vision input Accepts image inputs alongside text. | — | — | Tie |
| Reasoning mode Internal chain-of-thought / extended-thinking support. | — | — | Tie |
Frequently asked questions
Is Goliath 120B better than SorcererLM 8x22B?
SorcererLM 8x22B wins on 2 of 6 axes — pricing and capability skew in its favour for most workloads.
What's the price difference between Goliath 120B and SorcererLM 8x22B?
Input: $3.75 vs $4.50 per 1M tokens. Output: $7.50 vs $4.50 per 1M tokens.
What context windows do Goliath 120B and SorcererLM 8x22B support?
Goliath 120B supports up to 6K tokens. SorcererLM 8x22B supports up to 16K tokens.
Do both Goliath 120B and SorcererLM 8x22B support tool calling?
Goliath 120B: not advertised. SorcererLM 8x22B: not advertised.