Sovereign inference — Gemma 4 26B A4B

Your sovereign, private Gemma 4 26B A4B

A reasoning model in the small tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.

AU · EU · US resident0.17 UoI/1M input0.51 UoI/1M output262K context

How Gemma 4 26B A4B compares

Independent intelligence index for Gemma 4 26B A4B against every peer in the Small tier (and the tier above where needed) — same source, same tests.

Gemma 4 26B A4B (this model)16.7Qwen3.8 27B33.7Qwen 3.6 27B21.4Llama 3.1 8b6.9NVIDIA Nemotron Nano 12B v25.8Gemma 3 12B3.8

Latency — end-to-end seconds for a 500-token answer

Qwen3.8 27B62.3sQwen 3.6 27B111.0sLlama 3.1 8b4.5sNVIDIA Nemotron Nano 12B v23.6s

Run Gemma 4 26B A4B privately

One key, one endpoint, three jurisdictions.

  • Size tierSmall
  • Input rate0.17 UoI/1M
  • Output rate0.51 UoI/1M
  • Cached input rate—
  • Context window262K tokens
  • Max output262K tokens
  • ReasoningYes — emits thinking traces
  • Vision & video inputSupports image input
  • Tool callingSupported
  • Data residencyAU, EU and US endpoints — processing stays inside the region you call