Sovereign inference — GLM 4.7
Your sovereign, private GLM 4.7
A reasoning model in the extra-large tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.
AU · EU · US resident1.56 UoI/1M input4.68 UoI/1M output203K context
How GLM 4.7 compares
Independent intelligence index for GLM 4.7 against every peer in the Extra-large tier (and the tier above where needed) — same source, same tests.
Latency — end-to-end seconds for a 500-token answer
Run GLM 4.7 privately
One key, one endpoint, three jurisdictions.
- Size tierExtra-large
- Input rate1.56 UoI/1M
- Output rate4.68 UoI/1M
- Cached input rate—
- Context window203K tokens
- Max output203K tokens
- ReasoningYes — emits thinking traces
- Vision & video inputText only
- Tool callingSupported
- Data residencyAU, EU and US endpoints — processing stays inside the region you call


