AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Showing 301 - 315 of 548 matching models from 548 total

Ibm Granite
Budget tiertext -> text

Granite-4.0-H-Micro is a 3B parameter from the Granite 4 family of models.

Long contextBudget inputHugging Face
Context131K
Input$0.02 / 1M
Output$0.11 / 1M
Inception
Budget tiertext -> text

Inception: Mercury

Long context

Mercury is the first diffusion large language model (dLLM).

Long contextTool callingBudget input
Context128K
Input$0.25 / 1M
Output$0.75 / 1M
Inception
Budget tiertext -> text

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM).

Long contextReasoningTool calling
Context128K
Input$0.25 / 1M
Output$0.75 / 1M
Inception
Budget tiertext -> text

Mercury Coder is the first diffusion large language model (dLLM).

Long contextTool callingBudget input
Context128K
Input$0.25 / 1M
Output$0.75 / 1M
Inflection
Standard tiertext -> text

Inflection 3 Pi powers Inflection's Pi chatbot, including backstory, emotional intelligence, productivity, and safety.

Configurable API
Context8K
Input$2.50 / 1M
Output$10.00 / 1M
Kwaipilot
Budget tiertext -> text

KAT-Coder-Pro V1 is KwaiKAT's most advanced agentic coding model in the KAT-Coder series.

Long contextTool callingBudget input
Context256K
Input$0.21 / 1M
Output$0.83 / 1M
Kwaipilot
Budget tiertext -> text

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration.

Long contextTool callingBudget input
Context256K
Input$0.30 / 1M
Output$1.20 / 1M
Liquid
Budget tiertext -> text

LiquidAI: LFM2-2.6B

Standard context

LFM2 is a new generation of hybrid models developed by Liquid AI, specifically designed for edge AI and on-device deployment.

Budget inputHugging FaceConfigurable API
Context33K
Input$0.01 / 1M
Output$0.02 / 1M
Liquid
Budget tiertext -> text

LiquidAI: LFM2-24B-A2B

Standard context

LFM2-24B-A2B is the largest model in the LFM2 family of hybrid architectures designed for efficient on-device deployment.

Budget inputHugging FaceConfigurable API
Context33K
Input$0.03 / 1M
Output$0.12 / 1M
Liquid
Budget tiertext -> text

LiquidAI: LFM2-8B-A1B

Standard context

LFM2-8B-A1B is an efficient on-device Mixture-of-Experts (MoE) model from Liquid AI’s LFM2 family, built for fast, high-quality inference on edge hardware.

Budget inputHugging FaceConfigurable API
Context33K
Input$0.01 / 1M
Output$0.02 / 1M
Liquid
Budget tiertext -> text

LFM2.5-1.2B-Instruct is a compact, high-performance instruction-tuned model built for fast on-device AI.

Budget inputHugging FaceConfigurable API
Context33K
Input$0.00 / 1M
Output$0.00 / 1M
Liquid
Budget tiertext -> text

LFM2.5-1.2B-Thinking is a lightweight reasoning-focused model optimized for agentic tasks, data extraction, and RAG—while still running comfortably on edge devices.

ReasoningBudget inputHugging Face
Context33K
Input$0.00 / 1M
Output$0.00 / 1M
Mancer
Budget tiertext -> text

Mancer: Weaver (alpha)

Standard context

An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory.

Budget inputConfigurable API
Context8K
Input$0.75 / 1M
Output$1.00 / 1M
Meituan
Budget tiertext -> text

LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input.

Long contextTool callingBudget input
Context131K
Input$0.20 / 1M
Output$0.80 / 1M