Sovereign inference — Gemma 4 31B
Your sovereign, private Gemma 4 31B
A reasoning model in the medium tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.
AU · EU · US resident0.43 UoI/1M input1.51 UoI/1M output131K context
How Gemma 4 31B compares
Independent intelligence index for Gemma 4 31B against every peer in the Medium tier (and the tier above where needed) — same source, same tests.
Latency — end-to-end seconds for a 500-token answer
Capability indexes
Domain-weighted agentic performance for Gemma 4 31B.
Capability index vs medium peers
EngineeringEconomics
Run Gemma 4 31B privately
One key, one endpoint, three jurisdictions.
- Size tierMedium
- Input rate0.43 UoI/1M
- Output rate1.51 UoI/1M
- Cached input rate—
- Context window131K tokens
- Max output66K tokens
- ReasoningYes — emits thinking traces
- Vision & video inputSupports image input
- Tool callingSupported
- Data residencyAU, EU and US endpoints — processing stays inside the region you call


