Mistral Nemo 12B
Mistral
A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese, Korean, Arabic, and Hindi. It supports function calling and is released under the Apache 2.0 license.
- Language
- Tool use
- Vision
- Context
- 128K
- Max output
- 128K
- Input / 1M
- $0.150
- Output / 1M
- $0.150
Quickstart
Point any OpenAI-compatible client at Maxynetic Systems and pass mistral/mistral-nemo as the model.
curl https://api.maxynetic.com/v1/chat/completions \
-H "Authorization: Bearer $MAXYNETIC_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "mistral/mistral-nemo",
"messages": [{ "role": "user", "content": "Explain routing in one line." }]
}'Pricing
Billed per token with no markup on the underlying provider rate.
Input
Per 1M prompt tokens
$0.150
Output
Per 1M completion tokens
$0.150
A 1M-token prompt with 250K tokens of output costs about $0.188.
Rates for this model vary by upstream provider; Maxynetic Systems routes to the cheapest healthy one by default.
