Sovereign inference — Kimi K2.5

Your sovereign, private Kimi K2.5

A reasoning model in the large tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.

AU · EU · US resident0.84 UoI/1M input2.10 UoI/1M output262K context

How Kimi K2.5 compares

Independent intelligence index for Kimi K2.5 against every peer in the Large tier (and the tier above where needed) — same source, same tests.

Kimi K2.5 (this model)23.5MiniMax 2.522.8MiniMax 2.722.8GPT OSS 120B11.6Qwen3 Coder Next9.2

Latency — end-to-end seconds for a 500-token answer

MiniMax 2.531.0sMiniMax 2.750.5sGPT OSS 120B14.3sQwen3 Coder Next6.8s

Run Kimi K2.5 privately

One key, one endpoint, three jurisdictions.

  • Size tierLarge
  • Input rate0.84 UoI/1M
  • Output rate2.10 UoI/1M
  • Cached input rate—
  • Context window262K tokens
  • Max output33K tokens
  • ReasoningYes — emits thinking traces
  • Vision & video inputSupports image input
  • Tool callingSupported
  • Data residencyAU, EU and US endpoints — processing stays inside the region you call