Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
27 of 351 models
- Language
Gemini 3.7 Flash
google/gemini-3.7-flash
Gemini 3.7 Flash is the high-efficiency, cost-effective powerhouse of the Gemini 3 family. It delivers Pro-level agentic capabilities, major leaps in code generation and terminal execution. 3.7 Flash serves as the primary agentic workhorse in the Gemini 3 family, bridging the gap between deep-reasoning Pro models and high-throughput Flash-Lite models while delivering high token efficiency and multi-step multimodal processing.
- Context
- 1M
- In / 1M
- $0.750
- Out / 1M
- $3.75
- Reasoning
- Tool use
- Vision
- +4
- Language
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-lite
Gemini 3.5 Flash Lite features upgraded agentic capabilities, making the model ideal for subagents in complex workflows.
- Context
- 1M
- In / 1M
- $0.300
- Out / 1M
- $2.50
- Reasoning
- Tool use
- Vision
- +4
- Language
Gemini 3.6 Flash
google/gemini-3.6-flash
Gemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.
- Context
- 1M
- In / 1M
- $0.750
- Out / 1M
- $3.75
- Reasoning
- Tool use
- Vision
- +4
- Language
Gemini 3.5 Flash
google/gemini-3.5-flash
Google's latest model, highly optimized for coding proficiency and parallel agentic execution loops. Defaults to medium thinking effort for faster and more cost-efficient responses.
- Context
- 1M
- In / 1M
- $1.50
- Out / 1M
- $9.00
- Reasoning
- Tool use
- Vision
- +4
- Language
Gemini 3.1 Flash Lite
google/gemini-3.1-flash-lite
Gemini 3.1 Flash Lite outperforms 2.5 Flash Lite on overall quality and lands close to 2.5 Flash performance across key capability areas. It is a workhorse model for high-volume use cases, with improvements across audio input/ASR, RAG snippet ranking, translation, data extraction, and code completion.
- Context
- 1M
- In / 1M
- $0.250
- Out / 1M
- $1.50
- Reasoning
- Tool use
- Vision
- +3
- Language
Gemini 3.1 Pro Preview
google/gemini-3.1-pro-preview
This model improves upon Gemini 2.5 Pro and is catered towards challenging tasks, especially those involving complex reasoning or agentic workflows. Improvements highlighted include use cases for coding, multi-step function calling, planning, reasoning, deep knowledge tasks, and instruction following.
- Context
- 1M
- In / 1M
- $2.00
- Out / 1M
- $12.00
- Reasoning
- Tool use
- Vision
- +3
- Language
Gemini 3 Flash
google/gemini-3-flash
Google's most intelligent model built for speed, combining frontier intelligence with superior search and grounding.
- Context
- 1M
- In / 1M
- $0.500
- Out / 1M
- $3.00
- Reasoning
- Tool use
- Vision
- +3
- Language
Gemini 2.5 Flash Lite
google/gemini-2.5-flash-lite
Gemini 2.5 Flash-Lite is a balanced, low-latency model with configurable thinking budgets and tool connectivity (e.g., Google Search grounding and code execution). It supports multimodal input and offers a 1M-token context window.
- Context
- 1.0M
- In / 1M
- $0.100
- Out / 1M
- $0.400
- Reasoning
- Tool use
- Vision
- +3
- Language
Gemini 2.5 Flash
google/gemini-2.5-flash
Gemini 2.5 Flash is a thinking model that offers great, well-rounded capabilities. It is designed to offer a balance between price and performance with multimodal support and a 1M token context window.
- Context
- 1M
- In / 1M
- $0.300
- Out / 1M
- $2.50
- Reasoning
- Tool use
- Vision
- +3
- Language
Gemini 2.5 Pro
google/gemini-2.5-pro
Gemini 2.5 Pro is our most advanced reasoning Gemini model, capable of solving complex problems. Gemini 2.5 Pro can comprehend vast datasets and challenging problems from different information sources, including text, audio, images, video, and even entire code repositories.
- Context
- 1.0M
- In / 1M
- $1.25
- Out / 1M
- $10.00
- Reasoning
- Tool use
- Vision
- +3
- Language
Gemini 3.1 Flash Lite Image (Nano Banana 2 Lite)
google/gemini-3.1-flash-lite-image
Gemini 3.1 Flash-Lite Image (Nano Banana 2 Lite) is Google's fastest image generation model enabling rapid creation and iteration.
- Context
- 66K
- In / 1M
- $0.250
- Out / 1M
- $1.50
- Reasoning
- Vision
- Web search
- +2
- Language
Gemini 3.1 Flash Image (Nano Banana 2)
google/gemini-3.1-flash-image
Gemini 3.1 Flash Image is optimized for image understanding and generation and offers a balance of price and performance.
- Context
- 131K
- In / 1M
- $0.500
- Out / 1M
- $3.00
- Reasoning
- Vision
- Web search
- +2
- Language
Google Gemma 4 26B A4B
google/gemma-4-26b-a4b-it
Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on small models) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.
- Context
- 262K
- In / 1M
- $0.150
- Out / 1M
- $0.600
- Reasoning
- Tool use
- Vision
- +2
- Language
Gemini 3.1 Flash Image Preview (Nano Banana 2)
google/gemini-3.1-flash-image-preview
Gemini 3.1 Flash Image (Nano Banana 2) is optimized for image understanding and generation and offers a balance of price and performance.
- Context
- 131K
- In / 1M
- $0.500
- Out / 1M
- $3.00
- Reasoning
- Vision
- Web search
- +2
- Language
Gemini Omni Flash Preview
google/gemini-omni-flash-preview
Gemini Omni Flash (Preview) is a multimodal model designed for video, image, and text tasks. It is optimized for video generation, offering video output alongside text responses in a single model.
- Context
- 1M
- In / 1M
- $1.50
- Out / 1M
- $9.00
- Reasoning
- Vision
- File input
- +1
- Language
Gemma 4 31B IT
google/gemma-4-31b-it
Gemma 4 31B is engineered to tackle the most demanding enterprise workloads and complex reasoning tasks. With an expansive 256K-token context window, the 31B model can effortlessly ingest entire codebases, and massive sets of images in a single prompt.
- Context
- 262K
- In / 1M
- $0.140
- Out / 1M
- $0.400
- Reasoning
- Tool use
- Vision
- +1
- Language
Nano Banana Pro (Gemini 3 Pro Image)
google/gemini-3-pro-image
Nano Banana Pro (Gemini 3 Pro Image) builds on Nano Banana's generation capabilities into a new era of studio-quality, functional design to help you create and edit high-fidelity, production-ready visuals with unparalleled precision and control. Improvements include enhanced world knowledge and reasoning, dynamic text and translation, and studio level controls.
- Context
- 66K
- In / 1M
- $2.00
- Out / 1M
- $12.00
- Vision
- Web search
- Image generation
- +1
- Language
Nano Banana (Gemini 2.5 Flash Image)
google/gemini-2.5-flash-image
Nano Banana (Gemini 2.5 Flash Image) is Google's first fully hybrid reasoning model, letting developers turn thinking on or off and set thinking budgets to balance quality, cost, and latency. Upgraded for rapid creative workflows, it can generate interleaved text and images and supports conversational, multi‑turn image editing in natural language. It’s also locale‑aware, enabling culturally and linguistically appropriate image generation for audiences worldwide.
- Context
- 33K
- In / 1M
- $0.300
- Out / 1M
- $2.50
- Vision
- Web search
- Image generation
- Video
Veo 3.1 Lite Generate
google/veo-3.1-lite-generate-001
Veo 3.1 Lite Preview is a high-efficiency, developer-first video model providing high-fidelity video generation, editing, and cinematic control. It leverages the state-of-the-art Veo 3.1 model to democratize professional-grade video AI by offering a scalable, programmable interface for creators and enterprises.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Embedding
Gemini Embedding 2
google/gemini-embedding-2
Google’s first fully multimodal Embedding model that is capable of mapping text, image, video, audio, and PDFs and their interleaved combinations thereof into a single, unified vector space. Built on the Gemini architecture, it supports 100+ languages.
- Context
- —
- In / 1M
- $0.200
- Out / 1M
- —
- Video
Veo 3.1
google/veo-3.1-generate-001
Veo 3.1 is Google's state-of-the-art model for generating high-fidelity, 8-second 720p, 1080p or 4k videos featuring stunning realism and natively generated audio.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Veo 3.1 Fast Generate
google/veo-3.1-fast-generate-001
Veo 3.1 Fast is a specialized, high-speed variant of Google DeepMind’s Veo 3.1 text-to-video model, optimized for rapid generation of 8-second, high-fidelity videos. It is designed to create cinematic, 1080p, or 720p content with improved prompt adherence and native audio, making it ideal for creating quick, high-quality video clips, social media content, and ad creatives.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Video
Veo 3.0 Fast Generate
google/veo-3.0-fast-generate-001
Veo 3 Fast is a quicker and more cost effective version of Veo 3, allowing developers to create videos with sound while maintaining high quality and optimizing for speed and business use cases. Veo 3 Fast offers both text-to-video and image-to-video modalities.
- Context
- —
- In / 1M
- —
- Out / 1M
- —
- Embedding
Gemini Embedding 001
google/gemini-embedding-001
State-of-the-art embedding model with excellent performance across English, multilingual and code tasks.
- Context
- —
- In / 1M
- $0.150
- Out / 1M
- —
Showing 24 of 27
