AI Models Directory

OpenRouter-style model index with deep metadata: pricing, context windows, architecture, and parameters.

Showing 376 - 390 of 548 matching models from 548 total

Nousresearch
Budget tiertext -> text

Nous: Hermes 4 70B

Long context

Hermes 4 70B is a hybrid reasoning model from Nous Research, built on Meta-Llama-3.1-70B.

Long contextReasoningBudget input
Context131K
Input$0.13 / 1M
Output$0.40 / 1M
Nousresearch
Budget tiertext -> text

Hermes 2 Pro is an upgraded, retrained version of Nous Hermes 2, consisting of an updated and cleaned version of the OpenHermes 2.5 Dataset, as well as a newly introduced Function...

Budget inputHugging FaceConfigurable API
Context8K
Input$0.14 / 1M
Output$0.14 / 1M
Nvidia
Standard tiertext -> text

NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses.

Long contextTool callingHugging Face
Context131K
Input$1.20 / 1M
Output$1.20 / 1M
Nvidia
Budget tiertext -> text

Llama-3.1-Nemotron-Ultra-253B-v1 is a large language model (LLM) optimized for advanced reasoning, human-interactive chat, retrieval-augmented generation (RAG), and tool-calling tasks.

Long contextReasoningBudget input
Context131K
Input$0.60 / 1M
Output$1.80 / 1M
Nvidia
Budget tiertext -> text

Llama-3.3-Nemotron-Super-49B-v1.5 is a 49B-parameter, English-centric reasoning/chat model derived from Meta’s Llama-3.3-70B-Instruct with a 128K context.

Long contextReasoningTool calling
Context131K
Input$0.10 / 1M
Output$0.40 / 1M
Nvidia
Budget tiertext -> text

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems.

Long contextReasoningTool calling
Context262K
Input$0.05 / 1M
Output$0.20 / 1M
Nvidia
Budget tiertext -> text

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems.

Long contextReasoningTool calling
Context256K
Input$0.00 / 1M
Output$0.00 / 1M
Nvidia
Budget tiertext -> text

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications.

Long contextReasoningTool calling
Context262K
Input$0.10 / 1M
Output$0.50 / 1M
Nvidia
Budget tiertext -> text

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications.

Long contextReasoningTool calling
Context262K
Input$0.00 / 1M
Output$0.00 / 1M
Nvidia
Budget tiertext + image + video -> text

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence.

Long contextReasoningVision
Context131K
Input$0.20 / 1M
Output$0.60 / 1M
Nvidia
Budget tiertext + image + video -> text

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence.

Long contextReasoningTool calling
Context128K
Input$0.00 / 1M
Output$0.00 / 1M
Nvidia
Budget tiertext -> text

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks.

Long contextReasoningTool calling
Context131K
Input$0.04 / 1M
Output$0.16 / 1M
Nvidia
Budget tiertext -> text

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks.

Long contextReasoningTool calling
Context128K
Input$0.00 / 1M
Output$0.00 / 1M
OpenAI
Standard tiertext + audio -> text + audio

OpenAI: GPT Audio

Long context

The gpt-audio model is OpenAI's first generally available audio model.

Long contextAudioModerated
Context128K
Input$2.50 / 1M
Output$10.00 / 1M
OpenAI
Budget tiertext + audio -> text + audio

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency.

Long contextAudioBudget input
Context128K
Input$0.60 / 1M
Output$2.40 / 1M