Sovereign inference — Qwen3.8 27B

Your sovereign, private Qwen3.8 27B

A reasoning model in the small tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.

AU · EU · US resident0.17 UoI/1M input0.51 UoI/1M output262K context

How Qwen3.8 27B compares

Independent intelligence index for Qwen3.8 27B against every peer in the Small tier (and the tier above where needed) — same source, same tests.

Qwen3.8 27B (this model)33.7Qwen 3.6 27B21.4Gemma 4 26B A4B16.7Llama 3.1 8b6.9NVIDIA Nemotron Nano 12B v25.8Gemma 3 12B3.8

Latency — end-to-end seconds for a 500-token answer

Qwen3.8 27B62.3sQwen 3.6 27B111.0sGemma 4 26B A4B4.5sLlama 3.1 8b3.6s

Capability indexes

Domain-weighted agentic performance for Qwen3.8 27B.

Finance & AccountingStrategy & OpsLegalHealthcare & MedicalEngineeringEconomics

Capability index vs small peers

10.721.432.142.834.437.134.634.433.342.8Qwen3.8 27B21.316.923.323.331.4Qwen 3.6 27B3.82.93.84.85.3Gemma 3 12B
Finance & AccountingStrategy & OpsLegalHealthcare & MedicalEngineeringEconomics

The evaluations behind the index

Independent benchmark results for Qwen3.8 27B — Elo scores as published; pass rates shown as percentages.

AutomationBench-AA48.2% %Terminal-Bench 4.05.6% %SciCode46.6% %Humanity's Last Exam33.9% %GDP.pdf16.6% %CritPt5.4% %AA-Omniscience-998.3% %AA-LCR (long context)82.0% %

Per-evaluation economics

EvaluationScoreTime / taskOutput tokens / task
AA-Briefcase1403.1346.9 min142232
GDPval-AA1408.9433.2 min100519
AutomationBench-AA48.2%10.6 min32226
Terminal-Bench 4.05.6%48.9 min148327
SciCode46.6%6.9 min20839
Humanity's Last Exam33.9%14.6 min44126
GDP.pdf16.6%3.7 min11365
CritPt5.4%35.1 min106387
AA-Omniscience-998.3%1.2 min3697
AA-LCR (long context)82.0%0.9 min2809

Run Qwen3.8 27B privately

One key, one endpoint, three jurisdictions.

  • Size tierSmall
  • Input rate0.17 UoI/1M
  • Output rate0.51 UoI/1M
  • Cached input rate—
  • Context window262K tokens
  • Max output16K tokens
  • ReasoningYes — emits thinking traces
  • Vision & video inputSupports image input
  • Tool callingSupported
  • Data residencyAU, EU and US endpoints — processing stays inside the region you call