Maxynetic Systems
To be released soon

Google Gemma 4 26B A4B

Google

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on small models) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

  • Language
  • File input
  • Reasoning
  • Tool use
  • Vision
  • Implicit caching
Context
262K
Max output
131K
Input / 1M
$0.150
Output / 1M
$0.600

Quickstart

Point any OpenAI-compatible client at Maxynetic Systems and pass google/gemma-4-26b-a4b-it as the model.

curl https://api.maxynetic.com/v1/chat/completions \
  -H "Authorization: Bearer $MAXYNETIC_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "google/gemma-4-26b-a4b-it",
    "messages": [{ "role": "user", "content": "Explain routing in one line." }]
  }'

Pricing

Billed per token with no markup on the underlying provider rate.

  • Input

    Per 1M prompt tokens

    $0.150

  • Output

    Per 1M completion tokens

    $0.600

  • Cache read

    Per 1M cached prompt tokens

    $0.015

A 1M-token prompt with 250K tokens of output costs about $0.300.

Rates for this model vary by upstream provider; Maxynetic Systems routes to the cheapest healthy one by default.

More from Google