AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Clear Filters

Showing 76 - 90 of 235 matching models from 548 total for query "reasoning"

Moonshotai
Budget tiertext + image -> text

Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm.

Long contextReasoningTool calling
Context262K
Input$0.42 / 1M
Output$2.20 / 1M
OpenAI
Budget tiertext -> text

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases.

Long contextReasoningTool calling
Context131K
Input$0.04 / 1M
Output$0.19 / 1M
OpenAI
Budget tiertext -> text

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license.

Long contextReasoningTool calling
Context131K
Input$0.03 / 1M
Output$0.11 / 1M
Qwen
Budget tiertext -> text

Qwen2.5-Coder-7B-Instruct is a 7B parameter instruction-tuned language model optimized for code-related tasks such as code generation, reasoning, and bug fixing.

Budget inputHugging FaceConfigurable API
Context33K
Input$0.03 / 1M
Output$0.09 / 1M
Qwen
Budget tiertext + image -> text

Qwen2.5 VL 7B is a multimodal LLM from the Qwen Team with the following key enhancements: - SoTA understanding of images of various resolution & ratio: Qwen2.5-VL achieves...

VisionBudget inputHugging Face
Context33K
Input$0.20 / 1M
Output$0.20 / 1M
Qwen
Budget tiertext -> text

Qwen: Qwen3 14B

Standard context

Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue.

ReasoningTool callingBudget input
Context41K
Input$0.06 / 1M
Output$0.24 / 1M
Qwen
Budget tiertext -> text

Qwen3-30B-A3B-Instruct-2507 is a 30.5B-parameter mixture-of-experts language model from Qwen, with 3.3B active parameters per inference.

Long contextTool callingBudget input
Context262K
Input$0.09 / 1M
Output$0.30 / 1M
Qwen
Budget tiertext -> text

Qwen: Qwen3 32B

Standard context

Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue.

ReasoningTool callingBudget input
Context41K
Input$0.08 / 1M
Output$0.24 / 1M
Qwen
Budget tiertext -> text

Qwen: Qwen3 8B

Standard context

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue.

ReasoningTool callingBudget input
Context41K
Input$0.05 / 1M
Output$0.40 / 1M
Qwen
Budget tiertext + image -> text

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video.

Long contextTool callingVision
Context131K
Input$0.08 / 1M
Output$0.50 / 1M
Inception
Budget tiertext -> text

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).

Long contextReasoningTool calling
Context128K
Input$0.25 / 1M
Output$0.75 / 1M
Liquid
Budget tiertext -> text

LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices.

ReasoningBudget inputHugging Face
Context33K
Input$0.00 / 1M
Output$0.00 / 1M
Meituan
Budget tiertext -> text

LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input.

Long contextTool callingBudget input
Context131K
Input$0.20 / 1M
Output$0.80 / 1M
Meta
Budget tiertext + image -> text

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data.

Long contextVisionBudget input
Context131K
Input$0.05 / 1M
Output$0.05 / 1M
Meta
Budget tiertext -> text

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization.

Long contextBudget inputHugging Face
Context131K
Input$0.00 / 1M
Output$0.00 / 1M