Gemini Flash 2.0 offers a significantly faster time to first token (TTFT) compared to [Gemini Flash 1.5](/google/gemini-flash-1.5), while maintaining quality on par with larger models like [Gemini Pro 1.5](/google/gemini…
Model details →Best AI voice / audio models
Audio models handle either transcription, audio reasoning, or speech synthesis. Some support real-time streaming.
77Fit score
77Fit score
Gemini 2.0 Flash Lite offers a significantly faster time to first token (TTFT) compared to [Gemini Flash 1.5](/google/gemini-flash-1.5), while maintaining quality on par with larger models like [Gemini Pro 1.5](/google/g…
Model details →77Fit score
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide…
Model details →Top 25 models for speech-to-text
Frequently asked questions
Which model is best for transcription?
GPT-4o audio variants, Gemini audio, and Whisper-class models lead. Whisper is open-weight and self-hostable.
Related use cases
AI models for codingAI reasoning modelsAI vision modelslong-context AI models (128K+)AI models for function callingAI models for structured / JSON outputsAI models for agentsfree AI modelsCheapest AI models per tokenmultilingual AI modelsembedding modelssmall / on-device AI modelsAI image generation modelsAI models for mathAI models for writingopen-source AI modelsAI models for RAGAI models for summarizationAI models for translationAI models for enterprise