AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Clear Filters

Showing 106 - 120 of 235 matching models from 548 total for query "reasoning"

Mistral AI
Budget tiertext + image -> text

Mistral Small 3.1 24B Instruct is an upgraded variant of Mistral Small 3 (2501), featuring 24 billion parameters with advanced multimodal capabilities.

Long contextTool callingVision
Context128K
Input$0.00 / 1M
Output$0.00 / 1M
Mistral AI
Budget tiertext + image -> text

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system.

Long contextReasoningTool calling
Context262K
Input$0.15 / 1M
Output$0.60 / 1M
Mistral AI
Standard tiertext -> text

Mistral's official instruct fine-tuned version of Mixtral 8x22B.

Tool callingHugging FaceConfigurable API
Context66K
Input$2.00 / 1M
Output$6.00 / 1M
Moonshotai
Budget tiertext -> text

Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass.

Long contextTool callingBudget input
Context131K
Input$0.57 / 1M
Output$2.30 / 1M
Moonshotai
Budget tiertext -> text

Kimi K2 0905 is the September update of Kimi K2 0711. It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters...

Long contextTool callingBudget input
Context131K
Input$0.40 / 1M
Output$2.00 / 1M
Moonshotai
Budget tiertext -> text

Kimi K2 0905 is the September update of Kimi K2 0711. It is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters...

Long contextTool callingBudget input
Context262K
Input$0.60 / 1M
Output$2.50 / 1M
Moonshotai
Budget tiertext -> text

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning.

Long contextReasoningTool calling
Context131K
Input$0.47 / 1M
Output$2.00 / 1M
Nousresearch
Standard tiertext -> text

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation,...

Long contextHugging FaceConfigurable API
Context131K
Input$1.00 / 1M
Output$1.00 / 1M
Nousresearch
Budget tiertext -> text

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation,...

Long contextBudget inputHugging Face
Context131K
Input$0.00 / 1M
Output$0.00 / 1M
Nousresearch
Budget tiertext -> text

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation,...

Long contextBudget inputHugging Face
Context131K
Input$0.30 / 1M
Output$0.30 / 1M
Nousresearch
Standard tiertext -> text

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research.

Long contextReasoningHugging Face
Context131K
Input$1.00 / 1M
Output$3.00 / 1M
Nousresearch
Budget tiertext -> text

Nous: Hermes 4 70B

Long context

Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B.

Long contextReasoningBudget input
Context131K
Input$0.13 / 1M
Output$0.40 / 1M
Nvidia
Budget tiertext -> text

Llama-3.1-Nemotron-Ultra-253B-v1 is a large language model (LLM) optimized for advanced reasoning, human-interactive chat, retrieval-augmented generation (RAG), and tool-calling tasks.

Long contextReasoningBudget input
Context131K
Input$0.60 / 1M
Output$1.80 / 1M
Nvidia
Budget tiertext -> text

Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context.

Long contextReasoningTool calling
Context131K
Input$0.10 / 1M
Output$0.40 / 1M
Nvidia
Budget tiertext -> text

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems.

Long contextReasoningTool calling
Context262K
Input$0.05 / 1M
Output$0.20 / 1M