AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Clear Filters

Showing 61 - 75 of 235 matching models from 548 total for query "reasoning"

Google
Standard tiertext + image + file + audio + video -> text

Gemini 3 Pro is Google’s flagship frontier model for high-precision multimodal reasoning, combining strong performance across text, image, video, audio, and code with a 1M-token...

1M+ contextReasoningTool calling
Context1.0M
Input$2.00 / 1M
Output$12.00 / 1M
Google
Budget tiertext + image + file + audio + video -> text

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases.

1M+ contextReasoningTool calling
Context1.0M
Input$0.25 / 1M
Output$1.50 / 1M
Google
Standard tiertext + image + file + audio + video -> text

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage...

1M+ contextReasoningTool calling
Context1.0M
Input$2.00 / 1M
Output$12.00 / 1M
Google
Standard tiertext + image + file + audio + video -> text

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient...

1M+ contextReasoningTool calling
Context1.0M
Input$2.00 / 1M
Output$12.00 / 1M
Google
Budget tiertext -> text

Google: Gemma 2 27B

Standard context

Gemma 2 27B by Google is an open model built from the same research and technology used to create the Gemini models.

Budget inputHugging FaceConfigurable API
Context8K
Input$0.65 / 1M
Output$0.65 / 1M
Google
Budget tiertext + image -> text

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Long contextVisionBudget input
Context131K
Input$0.04 / 1M
Output$0.13 / 1M
Google
Budget tiertext + image -> text

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

VisionBudget inputHugging Face
Context33K
Input$0.00 / 1M
Output$0.00 / 1M
Google
Budget tiertext + image -> text

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Long contextVisionBudget input
Context131K
Input$0.08 / 1M
Output$0.16 / 1M
Google
Budget tiertext + image -> text

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Long contextVisionBudget input
Context131K
Input$0.00 / 1M
Output$0.00 / 1M
Google
Budget tiertext + image -> text

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

Long contextVisionBudget input
Context131K
Input$0.04 / 1M
Output$0.08 / 1M
Google
Budget tiertext + image -> text

Gemma 3 introduces multimodality, supporting vision-language input and text outputs.

VisionBudget inputHugging Face
Context33K
Input$0.00 / 1M
Output$0.00 / 1M
Google
Budget tiertext -> text

Gemma 3n E2B IT is a multimodal, instruction-tuned model developed by Google DeepMind, designed to operate efficiently at an effective parameter size of 2B while leveraging a 6B...

Budget inputHugging FaceConfigurable API
Context8K
Input$0.00 / 1M
Output$0.00 / 1M
Google
Budget tiertext + image -> text + image

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed.

ReasoningVisionBudget input
Context66K
Input$0.50 / 1M
Output$3.00 / 1M
Google
Standard tiertext + image -> text + image

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro.

ReasoningVisionConfigurable API
Context66K
Input$2.00 / 1M
Output$12.00 / 1M
Meta
Budget tiertext -> text

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization.

Budget inputHugging FaceConfigurable API
Context80K
Input$0.05 / 1M
Output$0.34 / 1M