Model catalog
One endpoint, every model. Compare context windows, capabilities and real per-token pricing side by side, then swap the model string in your request — nothing else changes.
- 351
- Models
- 35
- Providers
- 11
- Free to try
- $0.0001
- Cheapest input / 1M
5 of 351 models
- Language
Command A
cohere/command-a
Command A is Cohere's most performant model to date, excelling at tool use, agents, retrieval augmented generation (RAG), and multilingual use cases. Command A has a context length of 256K, only requires two GPUs to run, and has 150% higher throughput compared to Command R+ 08-2024.
- Context
- 256K
- In / 1M
- $2.50
- Out / 1M
- $10.00
- Tool use
- Reranking
Cohere Rerank 4 Fast
cohere/rerank-v4-fast
A light version of Rerank 4 Pro, this is a multilingual model that allows for re-ranking English and non-english documents and semi-structured data (JSON). This model is better suited for low latency and high throughput use-cases than its pro variant.
- Context
- 32K
- In / 1M
- —
- Out / 1M
- —
- Reranking
Cohere Rerank 4 Pro
cohere/rerank-v4-pro
A multilingual model that allows for re-ranking English and non-english documents and semi-structured data (JSON). This model is better suited for state-of-the-art quality and complex use-cases than its fast variant.
- Context
- 32K
- In / 1M
- —
- Out / 1M
- —
- Embedding
Embed v4.0
cohere/embed-v4.0
A model that allows for text, images, or mixed content to be classified or turned into embeddings.
- Context
- 128K
- In / 1M
- $0.120
- Out / 1M
- —
- Reranking
Cohere Rerank 3.5
cohere/rerank-v3.5
A model that allows for re-ranking English Language documents and semi-structured data (JSON).
- Context
- 4K
- In / 1M
- —
- Out / 1M
- —
