Models
Explore models and compare pricing across providers. Select a creator to browse their models.
Chirp 2
GoogleGoogle's latest speech recognition model with improved accuracy across 100+ languages.
Cloud TTS
GoogleGoogle's text-to-speech service supporting 75+ languages with WaveNet and Neural2 voices.
Gemini 2.0 Flash
GoogleFast and efficient Gemini model for high-throughput workloads.
Gemini 2.5 Flash
GoogleSpeed-optimized Gemini with strong reasoning and multimodal capabilities.
Gemini 2.5 Flash Image
GoogleGemini 2.5 Flash with native image generation and editing capabilities.
Gemini 2.5 Flash Lite
GoogleGoogle's Gemini 2.5 Flash Lite — lowest-cost Gemini 2.5 tier with multimodal input and a 1M-token context window.
Gemini 2.5 Pro
GoogleHigh-capability Gemini model for complex reasoning and coding tasks.
Gemini 3 Flash
GoogleFast and efficient Gemini 3 model for high-throughput workloads.
Gemini 3 Pro
GoogleHigh-capability Gemini 3 model. Deprecated in favor of 3.1 Pro.
Gemini 3 Pro Image
GoogleGoogle's most advanced image generation model built on Gemini 3 Pro.
Gemini 3.1 Flash Image
GoogleGemini 3.1 Flash Image model with Pro-level quality and accurate text rendering.
Gemini 3.1 Flash Lite
GoogleGemini 3.1 Flash Lite, a proprietary Google model available via OpenRouter.
Gemini 3.1 Flash Lite Image
GoogleFastest, most cost-efficient Nano Banana variant — ~4-second generation with a 1K output cap, optimized for high-volume workflows, with legible in-image text and character consistency.
Gemini 3.1 Pro
GoogleGoogle's current flagship model with top benchmark scores and 1M context.
Gemini 3.5 Flash
GoogleGemini 3.5 Flash, a proprietary Google model available via OpenRouter.
Gemini 3.5 Flash Lite
GoogleGoogle's Gemini 3.5 Flash Lite — lowest-cost Gemini 3.5 tier, multimodal input (text, image, audio, video, file) with a 1M-token context window.
Gemini 3.6 Flash
GoogleGoogle's Gemini 3.6 Flash — fast multimodal model (text, image, audio, video, file input) with a 1M-token context window.
Gemini 3.7 Flash
GoogleGoogle's Gemini 3.7 Flash — fast multimodal model (text, image, audio, video, file input) with a 1M-token context window.
Gemini 3.8 Flash
GoogleGoogle's Gemini 3.8 Flash — fast multimodal model (text, image, audio, video, file input) with a 1M-token context window.
Gemini Embedding 2
GoogleFirst natively multimodal embedding model supporting text, images, video, audio, and documents. 3072 dimensions, 100+ languages.
Gemma 3 12B
GoogleMid-size open-weight Gemma model with vision support.
Gemma 3 27B
GoogleLargest Gemma 3 model with strong reasoning and instruction following.
Gemma 3 4B
GoogleCompact open-weight model for edge and mobile deployment.
Gemma 4 12B
GoogleLatest Gemma generation optimized for reasoning and agentic workflows.
Gemma 4 26B A4B
GoogleGemma 4 26B A4B, an open-weight Google model available via OpenRouter.
Gemma 4 31B
GoogleGoogle's Gemma 4 31B — largest dense open-weight Gemma 4 model, with image input and a 262K-token context window.
Imagen 4
GoogleGoogle's latest image generation model with photorealistic output and strong text rendering.
Imagen 4 Fast
GoogleSpeed-optimized variant of Imagen 4 for faster generation.
Imagen 4 Ultra
GoogleHighest quality Imagen 4 variant with maximum detail and resolution.
Lyria 3
GoogleGoogle DeepMind music model with deep musical awareness and structural coherence. SynthID watermarked.
Lyria 3 Pro
GoogleAdvanced Lyria generating full 3-minute structured songs with verses, choruses, and vocals.
Veo 2
GoogleGoogle's second generation video model with cinematic quality output.
Veo 3
GooglePrevious generation Google video model with audio generation. Predecessor to Veo 3.1.
Veo 3.1
GoogleGoogle's best video generation model with native 4K, audio generation, and superior lip sync.