AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Showing 46 - 60 of 548 matching models from 548 total

Baidu
Budget tiertext + image -> text

A powerful multimodal Mixture-of-Experts chat model featuring 28B total parameters with 3B activated per token, delivering exceptional text and vision understanding through its...

ReasoningTool callingVision
Context30K
Input$0.14 / 1M
Output$0.56 / 1M
Baidu
Budget tiertext + image -> text

ERNIE-4.5-VL-424B-A47B is a multimodal Mixture-of-Experts (MoE) model from Baidu’s ERNIE 4.5 series, featuring 424B total parameters with 47B active per token.

ReasoningVisionBudget input
Context123K
Input$0.42 / 1M
Output$1.25 / 1M
ByteDance
Budget tiertext + image -> text

UI-TARS-1.5 is a multimodal vision-language agent optimized for GUI-based environments, including desktop interfaces, web browsers, mobile systems, and games.

Long contextVisionBudget input
Context128K
Input$0.10 / 1M
Output$0.20 / 1M
Bytedance Seed
Budget tiertext + image + video -> text

Seed 1.6 is a general-purpose model released by the ByteDance Seed team.

Long contextReasoningTool calling
Context262K
Input$0.25 / 1M
Output$2.00 / 1M
Bytedance Seed
Budget tiertext + image + video -> text

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding.

Long contextReasoningTool calling
Context262K
Input$0.07 / 1M
Output$0.30 / 1M
Bytedance Seed
Budget tiertext + image + video -> text

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a...

Long contextReasoningTool calling
Context262K
Input$0.25 / 1M
Output$2.00 / 1M
Bytedance Seed
Budget tiertext + image + video -> text

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment.

Long contextReasoningTool calling
Context262K
Input$0.10 / 1M
Output$0.40 / 1M
Cognitivecomputations
Budget tiertext -> text

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai.

Budget inputHugging FaceConfigurable API
Context33K
Input$0.00 / 1M
Output$0.00 / 1M
Cohere
Standard tiertext -> text

Cohere: Command A

Long context

Command A is an open-weights 111B parameter model with a 256k context window focused on delivering great performance across agentic, multilingual, and coding use cases.

Long contextModeratedHugging Face
Context256K
Input$2.50 / 1M
Output$10.00 / 1M
Cohere
Budget tiertext -> text

command-r-08-2024 is an update of the Command R with improved performance for multilingual retrieval-augmented generation (RAG) and tool use.

Long contextTool callingBudget input
Context128K
Input$0.15 / 1M
Output$0.60 / 1M
Cohere
Standard tiertext -> text

command-r-plus-08-2024 is an update of the Command R+ with roughly 50% higher throughput and 25% lower latencies as compared to the previous Command R+ version, while keeping the...

Long contextTool callingModerated
Context128K
Input$2.50 / 1M
Output$10.00 / 1M
Cohere
Budget tiertext -> text

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024.

Long contextBudget inputModerated
Context128K
Input$0.04 / 1M
Output$0.15 / 1M
Deepcogito
Standard tiertext -> text

Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models.

Long contextReasoningConfigurable API
Context128K
Input$1.25 / 1M
Output$1.25 / 1M
DeepSeek
Budget tiertext -> text

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions.

Long contextTool callingBudget input
Context164K
Input$0.32 / 1M
Output$0.89 / 1M
DeepSeek
Variable tiertext -> text

Legacy compatibility id mapping to DeepSeek V4 Flash thinking mode (per DeepSeek pricing notes).

1M+ context
Context1M
InputCustom / 1M
OutputCustom / 1M