AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Showing 61 - 75 of 548 matching models from 548 total

DeepSeek
Variable tiertext -> text

DeepSeek V4 Flash

Ultra context

DeepSeek V4 Flash — 1M context, thinking and non-thinking modes; see DeepSeek pricing/docs for current capabilities.

1M+ context
Context1M
InputCustom / 1M
OutputCustom / 1M
DeepSeek
Variable tiertext -> text

DeepSeek V4 Pro

Ultra context

DeepSeek V4 Pro — flagship API tier with discounts published on DeepSeek pricing page; verify live rates before billing.

1M+ context
Context1M
InputCustom / 1M
OutputCustom / 1M
DeepSeek
Budget tiertext -> text

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team.

Long contextReasoningTool calling
Context164K
Input$0.20 / 1M
Output$0.77 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek: DeepSeek V3.1

Standard context

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates.

ReasoningTool callingBudget input
Context33K
Input$0.15 / 1M
Output$0.75 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and...

Long contextReasoningTool calling
Context164K
Input$0.21 / 1M
Output$0.79 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and...

Long contextReasoningTool calling
Context164K
Input$0.21 / 1M
Output$0.79 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance.

Long contextReasoningTool calling
Context164K
Input$0.26 / 1M
Output$0.38 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures.

Long contextReasoningTool calling
Context164K
Input$0.27 / 1M
Output$0.41 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3.2-Speciale is a high-compute variant of DeepSeek-V3.2 optimized for maximum reasoning and agentic performance.

Long contextReasoningBudget input
Context164K
Input$0.40 / 1M
Output$1.20 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek: R1

Standard context

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens.

ReasoningTool callingBudget input
Context64K
Input$0.70 / 1M
Output$2.50 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek: R1 0528

Long context

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens.

Long contextReasoningTool calling
Context164K
Input$0.45 / 1M
Output$2.15 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1.

Long contextReasoningBudget input
Context131K
Input$0.70 / 1M
Output$0.80 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1.

ReasoningBudget inputHugging Face
Context33K
Input$0.29 / 1M
Output$0.29 / 1M
Eleutherai
Budget tiertext -> text

EleutherAI: Llemma 7b

Standard context

Llemma 7B is a language model for mathematics. It was initialized with Code Llama 7B weights, and trained on the Proof-Pile-2 for 200B tokens.

Budget inputHugging FaceConfigurable API
Context4K
Input$0.80 / 1M
Output$1.20 / 1M
Essentialai
Budget tiertext -> text

Rnj-1 is an 8B-parameter, dense, open-weight model family developed by Essential AI and trained from scratch with a focus on programming, math, and scientific reasoning.

Tool callingBudget inputHugging Face
Context33K
Input$0.15 / 1M
Output$0.15 / 1M