Sovereign inference — DeepSeek V4.1 Flash

Your sovereign, private DeepSeek V4.1 Flash

A reasoning model in the medium tier, served on ToothFairyAI's AU, EU and US endpoints — private by default, billed in Units of Intelligence at the published rates below. No training on your data, no vendor lock-in.

AU · EU · US resident0.43 UoI/1M input1.51 UoI/1M output0.01 UoI/1M cached input1049K context

How DeepSeek V4.1 Flash compares

Independent intelligence index for DeepSeek V4.1 Flash against every peer in the Medium tier (and the tier above where needed) — same source, same tests.

DeepSeek V4.1 Flash (this model)39.5GLM 5.3 Flash41.8DeepSeek V4 Flash Official34.3Minimax 329.2Gemma 4 31B19.0Qwen3 Coder 30B A3B9.6GPT OSS 20B9.0

Latency — end-to-end seconds for a 500-token answer

DeepSeek V4.1 Flash11.7sGLM 5.3 Flash58.6sDeepSeek V4 Flash Official12.5sMinimax 322.5sGemma 4 31B64.1sQwen3 Coder 30B A3B9.4sGPT OSS 20B14.7s

Capability indexes

Domain-weighted agentic performance for DeepSeek V4.1 Flash.

Finance & AccountingStrategy & OpsLegalHealthcare & MedicalEngineeringEconomics

Capability index vs medium peers

12.925.938.851.745.051.741.340.639.344.4DeepSeek V4.1 Flash42.045.139.944.843.848.9GLM 5.3 Flash37.543.137.735.434.940.3DeepSeek V4 Flash Official30.125.031.529.731.243.7Minimax 316.720.6Gemma 4 31B7.34.87.36.611.711.3GPT OSS 20B
Finance & AccountingStrategy & OpsLegalHealthcare & MedicalEngineeringEconomics

The evaluations behind the index

Independent benchmark results for DeepSeek V4.1 Flash — Elo scores as published; pass rates shown as percentages.

AutomationBench-AA68.9% %Terminal-Bench 4.026.8% %SciCode51.9% %Humanity's Last Exam39.2% %GDP.pdf12.8% %CritPt14.3% %AA-Omniscience-530.0% %AA-LCR (long context)84.0% %

Per-evaluation economics

EvaluationScoreTime / taskOutput tokens / task
AA-Briefcase1431.799.1 min170690
GDPval-AA1600.006.2 min116108
AutomationBench-AA68.9%4.6 min87009
Terminal-Bench 4.026.8%15.2 min285252
SciCode51.9%0.5 min9753
Humanity's Last Exam39.2%1.8 min34685
GDP.pdf12.8%0.8 min15040
CritPt14.3%5.9 min111624
AA-Omniscience-530.0%0.4 min8127
AA-LCR (long context)84.0%0.2 min3097

Run DeepSeek V4.1 Flash privately

One key, one endpoint, three jurisdictions.

  • Size tierMedium
  • Input rate0.43 UoI/1M
  • Output rate1.51 UoI/1M
  • Cached input rate0.01 UoI/1M
  • Context window1049K tokens
  • Max output16K tokens
  • ReasoningYes — emits thinking traces
  • Vision & video inputSupports image input
  • Tool callingSupported
  • Data residencyAU, EU and US endpoints — processing stays inside the region you call