AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Showing 496 - 510 of 548 matching models from 548 total

Qwen
Budget tiertext + image + video -> text

Qwen: Qwen3.5-9B

Long context

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture.

Long contextReasoningTool calling
Context256K
Input$0.05 / 1M
Output$0.15 / 1M
Qwen
Budget tiertext + image + video -> text

Qwen: Qwen3.5-Flash

Ultra context

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving...

1M+ contextReasoningTool calling
Context1M
Input$0.07 / 1M
Output$0.26 / 1M
Qwen
Budget tiertext -> text

Qwen 3.6 Plus Preview is the next-generation evolution of the Qwen Plus series, featuring an advanced hybrid architecture that improves efficiency and scalability.

1M+ contextReasoningTool calling
Context1M
Input$0.00 / 1M
Output$0.00 / 1M
Qwen
Budget tiertext -> text

Qwen: QwQ 32B

Long context

QwQ is the reasoning model of the Qwen series. Compared with conventional instruction-tuned models, QwQ, which is capable of thinking and reasoning, can achieve significantly...

Long contextReasoningTool calling
Context131K
Input$0.15 / 1M
Output$0.58 / 1M
Qwen
Budget tiertext -> text

Qwen2.5 72B Instruct

Standard context

Qwen2.5 72B is the latest series of Qwen large language models.

Tool callingBudget inputHugging Face
Context33K
Input$0.12 / 1M
Output$0.39 / 1M
Qwen
Budget tiertext -> text

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen).

Budget inputHugging FaceConfigurable API
Context33K
Input$0.66 / 1M
Output$1.00 / 1M
Raifle
Standard tiertext -> text

SorcererLM 8x22B

Standard context

SorcererLM is an advanced RP and storytelling model, built as a Low-rank 16-bit LoRA fine-tuned on WizardLM-2 8x22B.

Hugging FaceConfigurable API
Context16K
Input$4.50 / 1M
Output$4.50 / 1M
Reka
Budget tiertext + image + video -> text

Reka Edge

Standard context

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs.

Tool callingVisionVideo
Context16K
Input$0.10 / 1M
Output$0.10 / 1M
Rekaai
Budget tiertext -> text

Reka: Flash 3

Standard context

Reka Flash 3 is a general-purpose, instruction-tuned large language model with 21 billion parameters, developed by Reka.

ReasoningBudget inputHugging Face
Context66K
Input$0.10 / 1M
Output$0.20 / 1M
Relace
Budget tiertext -> text

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files.

Long contextBudget inputConfigurable API
Context256K
Input$0.85 / 1M
Output$1.25 / 1M
Relace
Standard tiertext -> text

The relace-search model uses 4-12 `view_file` and `grep` tools in parallel to explore a codebase and return relevant files to the user request.

Long contextTool callingConfigurable API
Context256K
Input$1.00 / 1M
Output$3.00 / 1M
Sao10k
Budget tiertext -> text

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3.

Budget inputHugging FaceConfigurable API
Context8K
Input$0.04 / 1M
Output$0.05 / 1M
Sao10k
Standard tiertext -> text

Euryale 70B v2.1 is a model focused on creative roleplay from Sao10k.

Tool callingHugging FaceConfigurable API
Context8K
Input$1.48 / 1M
Output$1.48 / 1M
Sao10k
Standard tiertext -> text

This is Sao10K's experiment over Euryale v2.2.

Hugging FaceConfigurable API
Context16K
Input$3.00 / 1M
Output$3.00 / 1M
Sao10k
Budget tiertext -> text

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from Sao10k.

Long contextTool callingBudget input
Context131K
Input$0.85 / 1M
Output$0.85 / 1M