TRULYSOVEREIGN AI · RADAR

DeepSeek V4.1 Flash pricing up 50% prompt, 20% completion; output limit cut 58%

OpenRouter model catalogue · published 2026-09-23 21:12 UTC
ACTION NEEDED BUDGET PRICINGLIMITSMODEL
WHAT CHANGEDOpenRouter's DeepSeek V4.1 Flash model pricing increased from $0.0000001 to $0.00000015 per prompt token (50% increase) and $0.0000005 to $0.0000006 per completion token (20% increase). Maximum completion tokens dropped from 943,718 to 393,216 tokens (58% reduction).
WHY IT MATTERSApplications generating long outputs will hit the new 393K token ceiling where they previously could produce up to 943K tokens. Monthly costs will increase proportionally for all usage: a workload consuming 1B prompt tokens and 200M completion tokens now costs $270/month instead of $200/month.
WHAT TO DOCalculate your current monthly token usage for this model and reforecast costs with the new rates. Check whether any workflows rely on outputs exceeding 393,216 tokens and either chunk the requests or switch models.
CONFIDENCE 95% Primary source ↗ written by anthropic/claude-sonnet-4.5
Show the reasoning — raw diff, triage verdict, confidence
RAW DIFF
 id: deepseek/deepseek-v4.1-flash
 name: DeepSeek: DeepSeek V4.1 Flash
-pricing.prompt: 0.0000001
-pricing.completion: 0.0000005
+pricing.prompt: 0.00000015
+pricing.completion: 0.0000006
 context_length: 1048576
-top_provider.max_completion_tokens: 943718
+top_provider.max_completion_tokens: 393216
HOW IT WAS JUDGED
triage verdictMATERIAL at 0.98, needed 0.5
triage reasoningPricing increased and output capacity reduced. Teams budgeting or selecting models need to know.
signalspricing.prompt increased 50% · pricing.completion increased 20% · max_completion_tokens decreased 58% · three numeric changes affecting cost and capability
suppression
triage modelanthropic/claude-haiku-4.5 ($0.00000)
writer modelanthropic/claude-sonnet-4.5 ($0.00000)