Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
351 of 351 models
- Language
GPT 5.6 Luna
openai/gpt-5.6-luna
GPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.
- Context
- 1.1M
- In / 1M
- $0.200
- Out / 1M
- $1.20
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.6 Luna (Fast)
openai/gpt-5.6-luna-fast
GPT-5.6 Luna is a fast, affordable GPT-5.6 model that brings strong capability at the lowest cost in the series.
- Context
- 1.1M
- In / 1M
- $0.400
- Out / 1M
- $2.40
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.6 Sol
openai/gpt-5.6-sol
GPT-5.6 Sol is the flagship of OpenAI's GPT-5.6 series, its most capable model for long-horizon agentic work across coding, biology, and cybersecurity.
- Context
- 1.1M
- In / 1M
- $2.00
- Out / 1M
- $10.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.6 Sol (Fast)
openai/gpt-5.6-sol-fast
GPT-5.6 Sol is the flagship of OpenAI's GPT-5.6 series, its most capable model for long-horizon agentic work across coding, biology, and cybersecurity.
- Context
- 1.1M
- In / 1M
- $4.00
- Out / 1M
- $20.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.6 Terra
openai/gpt-5.6-terra
GPT-5.6 Terra is a balanced GPT-5.6 model for everyday work, with performance comparable to the previous generation at half the cost.
- Context
- 1.1M
- In / 1M
- $2.00
- Out / 1M
- $12.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.6 Terra (Fast)
openai/gpt-5.6-terra-fast
GPT-5.6 Terra is a balanced GPT-5.6 model for everyday work, with performance comparable to the previous generation at half the cost.
- Context
- 1.1M
- In / 1M
- $4.00
- Out / 1M
- $24.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.5
openai/gpt-5.5
GPT‑5.5 understands what you’re trying to do faster and can carry more of the work itself. It excels at writing and debugging code, researching online, analyzing data, creating documents and spreadsheets, operating software, and moving across tools until a task is finished. Instead of carefully managing every step, you can give GPT‑5.5 a messy, multi-part task and trust it to plan, use tools, check its work, navigate through ambiguity, and keep going.
- Context
- 1M
- In / 1M
- $5.00
- Out / 1M
- $30.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.5 (Fast)
openai/gpt-5.5-fast
GPT‑5.5 understands what you’re trying to do faster and can carry more of the work itself. It excels at writing and debugging code, researching online, analyzing data, creating documents and spreadsheets, operating software, and moving across tools until a task is finished. Instead of carefully managing every step, you can give GPT‑5.5 a messy, multi-part task and trust it to plan, use tools, check its work, navigate through ambiguity, and keep going.
- Context
- 1M
- In / 1M
- $12.50
- Out / 1M
- $75.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.4 Mini
openai/gpt-5.4-mini
GPT-5.4 Mini brings the strengths of GPT-5.4 to a faster, more efficient model designed for high-volume workloads.
- Context
- 400K
- In / 1M
- $0.750
- Out / 1M
- $4.50
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.4 Mini (Fast)
openai/gpt-5.4-mini-fast
GPT-5.4 Mini brings the strengths of GPT-5.4 to a faster, more efficient model designed for high-volume workloads.
- Context
- 400K
- In / 1M
- $1.50
- Out / 1M
- $9.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.4
openai/gpt-5.4
GPT-5.4 is OpenAI's best general-purpose model, part of the GPT-5 flagship model family. It's their most intelligent model yet for both general and agentic tasks.
- Context
- 1.1M
- In / 1M
- $2.50
- Out / 1M
- $15.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.4 (Fast)
openai/gpt-5.4-fast
GPT-5.4 is OpenAI's best general-purpose model, part of the GPT-5 flagship model family. It's their most intelligent model yet for both general and agentic tasks.
- Context
- 1.1M
- In / 1M
- $5.00
- Out / 1M
- $30.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.3 Codex
openai/gpt-5.3-codex
GPT-5.3-Codex advances both the frontier coding performance of GPT‑5.2-Codex and the reasoning and professional knowledge capabilities of GPT‑5.2, together in one model, which is also 25% faster. This enables it to take on long-running tasks that involve research, tool use, and complex execution.
- Context
- 400K
- In / 1M
- $1.75
- Out / 1M
- $14.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.3 Codex (Fast)
openai/gpt-5.3-codex-fast
GPT-5.3-Codex advances both the frontier coding performance of GPT‑5.2-Codex and the reasoning and professional knowledge capabilities of GPT‑5.2, together in one model, which is also 25% faster. This enables it to take on long-running tasks that involve research, tool use, and complex execution.
- Context
- 400K
- In / 1M
- $3.50
- Out / 1M
- $28.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.2
openai/gpt-5.2
GPT-5.2 is OpenAI's best general-purpose model, part of the GPT-5 flagship model family. It's their most intelligent model yet for both general and agentic tasks.
- Context
- 400K
- In / 1M
- $1.75
- Out / 1M
- $14.00
- Reasoning
- Tool use
- Vision
- +5
- Language
GPT 5.2 (Fast)
openai/gpt-5.2-fast
GPT-5.2 is OpenAI's best general-purpose model, part of the GPT-5 flagship model family. It's their most intelligent model yet for both general and agentic tasks.
- Context
- 400K
- In / 1M
- $3.50
- Out / 1M
- $28.00
- Reasoning
- Tool use
- Vision
- +5
- Language
Gemini 3.7 Flash
google/gemini-3.7-flash
Gemini 3.7 Flash is the high-efficiency, cost-effective powerhouse of the Gemini 3 family. It delivers Pro-level agentic capabilities, major leaps in code generation and terminal execution. 3.7 Flash serves as the primary agentic workhorse in the Gemini 3 family, bridging the gap between deep-reasoning Pro models and high-throughput Flash-Lite models while delivering high token efficiency and multi-step multimodal processing.
- Context
- 1M
- In / 1M
- $0.750
- Out / 1M
- $3.75
- Reasoning
- Tool use
- Vision
- +4
- Language
Sakana Namazu
sakana/namazu
Sakana Namazu is a Japanese-specialized LLM that combines a deep understanding of Japanese culture and business customs with high-performance language capabilities. Built on the open model Kimi K2.6 and refined with Sakana AI's in-house data for Japanese language and business workflows, it handles complex tasks using web search and code execution. Unlike Fugu, which orchestrates multiple frontier models, Sakana Namazu provides a single in-house model as an API.
- Context
- 256K
- In / 1M
- $0.950
- Out / 1M
- $4.00
- Reasoning
- Tool use
- Vision
- +4
- Language
Claude Opus 5
anthropic/claude-opus-5
Claude Opus 5 is the latest model in Anthropic's Opus family and a step-change improvement over Opus 4.8. It delivers major gains over Opus 4.8 in agentic coding, professional knowledge work, and long-horizon reasoning, and it is stronger per token across effort levels.
- Context
- 1M
- In / 1M
- $5.00
- Out / 1M
- $25.00
- Reasoning
- Tool use
- Vision
- +4
- Language
Claude Opus 5 (Fast)
anthropic/claude-opus-5-fast
Claude Opus 5 is the latest model in Anthropic's Opus family and a step-change improvement over Opus 4.8. It delivers major gains over Opus 4.8 in agentic coding, professional knowledge work, and long-horizon reasoning, and it is stronger per token across effort levels.
- Context
- 1M
- In / 1M
- $10.00
- Out / 1M
- $50.00
- Reasoning
- Tool use
- Vision
- +4
- Language
Gemini 3.5 Flash Lite
google/gemini-3.5-flash-lite
Gemini 3.5 Flash Lite features upgraded agentic capabilities, making the model ideal for subagents in complex workflows.
- Context
- 1M
- In / 1M
- $0.300
- Out / 1M
- $2.50
- Reasoning
- Tool use
- Vision
- +4
- Language
Gemini 3.6 Flash
google/gemini-3.6-flash
Gemini 3.6 Flash delivers higher quality across coding, agentic workflows, and web development with reduced token consumption and fewer model calls compared to previous model iterations.
- Context
- 1M
- In / 1M
- $0.750
- Out / 1M
- $3.75
- Reasoning
- Tool use
- Vision
- +4
- Language
Claude Opus 4.8
anthropic/claude-opus-4.8
Opus 4.8 is a focused upgrade to Opus 4.7 and is Anthropic's best generally available model for coding, agentic tasks, and enterprise workflows. It builds on the strengths of previous Opus models with stronger performance on complex, multi-step coding tasks. Anthropic recommends using it on long-horizon coding and agentic tasks. It is also stronger on professional work, including document drafting, data analysis, and presentations.
- Context
- 1M
- In / 1M
- $5.00
- Out / 1M
- $25.00
- Reasoning
- Tool use
- Vision
- +4
- Language
Claude Opus 4.8 (Fast)
anthropic/claude-opus-4.8-fast
Opus 4.8 is a focused upgrade to Opus 4.7 and is Anthropic's best generally available model for coding, agentic tasks, and enterprise workflows. It builds on the strengths of previous Opus models with stronger performance on complex, multi-step coding tasks. Anthropic recommends using it on long-horizon coding and agentic tasks. It is also stronger on professional work, including document drafting, data analysis, and presentations.
- Context
- 1M
- In / 1M
- $10.00
- Out / 1M
- $50.00
- Reasoning
- Tool use
- Vision
- +4
Showing 24 of 351
