Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
9 of 351 models
- Language
Muse Spark 1.2
meta/muse-spark-1.2
A coding-optimized model purpose-built for agentic workflows. Improvements to code generation, debugging, and codebase understanding — with a 1M context window that handles your entire project in one session.
- Context
- 1.0M
- In / 1M
- $1.25
- Out / 1M
- $4.25
- Reasoning
- Tool use
- Vision
- +2
- Language
Muse Spark 1.2 Contributor
meta/muse-spark-1.2-contributor
A coding-optimized model with pricing designed for builders. Same model, same capabilities — up to 95% less than Standard. Your inputs and outputs are used to train and improve Meta's AI models.
- Context
- 1.0M
- In / 1M
- $0.100
- Out / 1M
- $0.200
- Reasoning
- Tool use
- Vision
- +2
- Language
Muse Spark 1.1
meta/muse-spark-1.1
Muse Spark 1.1 is strongest at agentic performance, tool use, and computer use. It does well on long-running tasks with 1M token context window, can delegate execution to sub-agents running in parallel, and is trained to use computer interfaces on desktop, mobile, or browser.
- Context
- 1.0M
- In / 1M
- $1.25
- Out / 1M
- $4.25
- Reasoning
- Tool use
- Vision
- +2
- Language
Muse Glimmer 30B
meta/muse-glimmer-30b
- Context
- 131K
- In / 1M
- $0.350
- Out / 1M
- $1.50
- Reasoning
- Tool use
- Vision
- +1
- Language
Llama 4 Maverick 17B Instruct
meta/llama-4-maverick
As a general purpose LLM, Llama 4 Maverick contains 17 billion active parameters, 128 experts, and 400 billion total parameters, offering high quality at a lower price compared to Llama 3.3 70B.
- Context
- 128K
- In / 1M
- $0.240
- Out / 1M
- $0.970
- Tool use
- Vision
- Language
Llama 4 Scout 17B Instruct
meta/llama-4-scout
Llama 4 Scout is the best multimodal model in the world in its class and is more powerful than our Llama 3 models, while fitting in a single H100 GPU. Additionally, Llama 4 Scout supports an industry-leading context window of up to 10M tokens.
- Context
- 128K
- In / 1M
- $0.170
- Out / 1M
- $0.660
- Tool use
- Vision
- Language
Llama 3.3 70B Instruct
meta/llama-3.3-70b
Where performance meets efficiency. This model supports high-performance conversational AI designed for content creation, enterprise applications, and research, offering advanced language understanding capabilities, including text summarization, classification, sentiment analysis, and code generation.
- Context
- 128K
- In / 1M
- $0.720
- Out / 1M
- $0.720
- Tool use
- Language
Llama 3.1 70B Instruct
meta/llama-3.1-70b
An update to Meta Llama 3 70B Instruct that includes an expanded 128K context length, multilinguality and improved reasoning capabilities.
- Context
- 128K
- In / 1M
- $0.720
- Out / 1M
- $0.720
- Tool use
- Language
Llama 3.1 8B Instruct
meta/llama-3.1-8b
An update to Meta Llama 3 8B Instruct that includes an expanded 128K context length, multilinguality and improved reasoning capabilities.
- Context
- 128K
- In / 1M
- $0.220
- Out / 1M
- $0.220
- Tool use
