Sovereign inference — Qwen3 Coder 30B A3B

Your sovereign, private Qwen3 Coder 30B A3B

A reasoning model in the medium tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.

AU · EU · US resident0.43 UoI/1M input1.51 UoI/1M output262K context

How Qwen3 Coder 30B A3B compares

Independent intelligence index for Qwen3 Coder 30B A3B against every peer in the Medium tier (and the tier above where needed) — same source, same tests.

Qwen3 Coder 30B A3B (this model)9.6GLM 5.3 Flash41.8DeepSeek V4.1 Flash39.5DeepSeek V4 Flash Official34.3Minimax 329.2Gemma 4 31B19.0GPT OSS 20B9.0

Latency — end-to-end seconds for a 500-token answer

Qwen3 Coder 30B A3B9.4sGLM 5.3 Flash58.6sDeepSeek V4.1 Flash11.7sDeepSeek V4 Flash Official12.5sMinimax 322.5sGemma 4 31B64.1sGPT OSS 20B14.7s

Run Qwen3 Coder 30B A3B privately

One key, one endpoint, three jurisdictions.

  • Size tierMedium
  • Input rate0.43 UoI/1M
  • Output rate1.51 UoI/1M
  • Cached input rate—
  • Context window262K tokens
  • Max output33K tokens
  • ReasoningYes — emits thinking traces
  • Vision & video inputText only
  • Tool callingSupported
  • Data residencyAU, EU and US endpoints — processing stays inside the region you call