AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Clear Filters

Showing 46 - 60 of 235 matching models from 548 total for query "reasoning"

DeepSeek
Budget tiertext -> text

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures.

Long contextReasoningTool calling
Context164K
Input$0.27 / 1M
Output$0.41 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3.2-Speciale is a high-compute variant of DeepSeek-V3.2 optimized for maximum reasoning and agentic performance.

Long contextReasoningBudget input
Context164K
Input$0.40 / 1M
Output$1.20 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek: R1

Standard context

DeepSeek R1 is here: Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens.

ReasoningTool callingBudget input
Context64K
Input$0.70 / 1M
Output$2.50 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek: R1 0528

Long context

May 28th update to the original DeepSeek R1 Performance on par with OpenAI o1, but open-sourced and with fully open reasoning tokens.

Long contextReasoningTool calling
Context164K
Input$0.45 / 1M
Output$2.15 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1.

Long contextReasoningBudget input
Context131K
Input$0.70 / 1M
Output$0.80 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1.

ReasoningBudget inputHugging Face
Context33K
Input$0.29 / 1M
Output$0.29 / 1M
Eleutherai
Budget tiertext -> text

EleutherAI: Llemma 7b

Standard context

Llemma 7B is a language model for mathematics. It was initialized with Code Llama 7B weights, and trained on the Proof-Pile-2 for 200B tokens.

Budget inputHugging FaceConfigurable API
Context4K
Input$0.80 / 1M
Output$1.20 / 1M
Essentialai
Budget tiertext -> text

Rnj-1 is an 8B-parameter, dense, open-weight model family developed by Essential AI and trained from scratch with a focus on programming, math, and scientific reasoning.

Tool callingBudget inputHugging Face
Context33K
Input$0.15 / 1M
Output$0.15 / 1M
Google
Budget tiertext + image + file + audio + video -> text

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks.

1M+ contextReasoningTool calling
Context1.0M
Input$0.30 / 1M
Output$2.50 / 1M
Google
Budget tiertext + image + file + audio + video -> text

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.

1M+ contextReasoningTool calling
Context1.0M
Input$0.10 / 1M
Output$0.40 / 1M
Google
Budget tiertext + image + file + audio + video -> text

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency.

1M+ contextReasoningTool calling
Context1.0M
Input$0.10 / 1M
Output$0.40 / 1M
Google
Standard tiertext + image + file + audio + video -> text

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

1M+ contextReasoningTool calling
Context1.0M
Input$1.25 / 1M
Output$10.00 / 1M
Google
Standard tiertext + image + file + audio + video -> text

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

1M+ contextReasoningTool calling
Context1.0M
Input$1.25 / 1M
Output$10.00 / 1M
Google
Standard tiertext + image + file + audio -> text

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks.

1M+ contextReasoningTool calling
Context1.0M
Input$1.25 / 1M
Output$10.00 / 1M
Google
Budget tiertext + image + file + audio + video -> text

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance.

1M+ contextReasoningTool calling
Context1.0M
Input$0.50 / 1M
Output$3.00 / 1M