AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Clear Filters

Showing 196 - 210 of 235 matching models from 548 total for query "reasoning"

Qwen
Budget tiertext + image + video -> text

The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving...

Long contextReasoningTool calling
Context262K
Input$0.26 / 1M
Output$2.08 / 1M
Qwen
Budget tiertext + image + video -> text

Qwen: Qwen3.5-27B

Long context

The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance.

Long contextReasoningTool calling
Context262K
Input$0.20 / 1M
Output$1.56 / 1M
Qwen
Budget tiertext + image + video -> text

The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model,...

Long contextReasoningTool calling
Context262K
Input$0.16 / 1M
Output$1.30 / 1M
Qwen
Budget tiertext + image + video -> text

Qwen: Qwen3.5-9B

Long context

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture.

Long contextReasoningTool calling
Context256K
Input$0.05 / 1M
Output$0.15 / 1M
Qwen
Budget tiertext + image + video -> text

Qwen: Qwen3.5-Flash

Ultra context

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving...

1M+ contextReasoningTool calling
Context1M
Input$0.07 / 1M
Output$0.26 / 1M
Qwen
Budget tiertext -> text

Qwen 3.6 Plus Preview is the next-generation evolution of the Qwen Plus series, featuring an advanced hybrid architecture that improves efficiency and scalability.

1M+ contextReasoningTool calling
Context1M
Input$0.00 / 1M
Output$0.00 / 1M
Qwen
Budget tiertext -> text

Qwen: QwQ 32B

Long context

QwQ is the reasoning model of the Qwen series. Compared with conventional instruction-tuned models, QwQ, which is capable of thinking and reasoning, can achieve significantly...

Long contextReasoningTool calling
Context131K
Input$0.15 / 1M
Output$0.58 / 1M
Qwen
Budget tiertext -> text

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen).

Budget inputHugging FaceConfigurable API
Context33K
Input$0.66 / 1M
Output$1.00 / 1M
Raifle
Standard tiertext -> text

SorcererLM 8x22B

Standard context

SorcererLM is an advanced RP and storytelling model, built as a Low-rank 16-bit LoRA fine-tuned on WizardLM-2 8x22B.

Hugging FaceConfigurable API
Context16K
Input$4.50 / 1M
Output$4.50 / 1M
Rekaai
Budget tiertext -> text

Reka: Flash 3

Standard context

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka.

ReasoningBudget inputHugging Face
Context66K
Input$0.10 / 1M
Output$0.20 / 1M
Relace
Standard tiertext -> text

The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request.

Long contextTool callingConfigurable API
Context256K
Input$1.00 / 1M
Output$3.00 / 1M
Sao10k
Budget tiertext -> text

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3.

Budget inputHugging FaceConfigurable API
Context8K
Input$0.04 / 1M
Output$0.05 / 1M
Stepfun
Budget tiertext -> text

Step 3.5 Flash is StepFun's most capable open-source foundation model.

Long contextReasoningTool calling
Context262K
Input$0.10 / 1M
Output$0.30 / 1M
Stepfun
Budget tiertext -> text

Step 3.5 Flash is StepFun's most capable open-source foundation model.

Long contextReasoningTool calling
Context256K
Input$0.00 / 1M
Output$0.00 / 1M
Switchpoint
Budget tiertext -> text

Switchpoint Router

Long context

Switchpoint AI's router instantly analyzes your request and directs it to the optimal AI from an ever-evolving library.

Long contextReasoningBudget input
Context131K
Input$0.85 / 1M
Output$3.40 / 1M