Updated 8 Oct 2026: Distinguish confirmed old-Flash alias migration from contradictory V4 Pro reroute announcement; live changelog and pricing table retain distinct V4 Pro.

Key details

  1. DeepSeek-V4.1-Flash was announced on September 10, 2026.
  2. It is a 552B MoE model with 8B active input parameters and 16B active output parameters according to DeepSeek.
  3. Native multimodal visual understanding is built into V4.1 Flash.
  4. Legacy deepseek-v4-flash and deepseek-v4-flash-vision-exp compatibility names temporarily route to V4.1 Flash.
  5. DeepSeek says the new model uses one quarter of the HBM and one eighth of the SSD KV-cache storage of the previous generation.
  6. API pricing was reduced on September 10 while peak/off-peak pricing remains.
  7. The original V4 Pro reroute announcement conflicts with the current API changelog, which says V4 Pro continues after September 14 at unchanged billing.
  8. The live rate card separately lists deepseek-v4-pro as V4-Pro-0813 and deepseek-flash as V4.1 Flash.

What builders should take away

  1. Re-run production evaluations even if your existing model string still works; the backend model may have changed materially.
  2. Measure total task cost and token usage rather than assuming the lower list price automatically means lower end-to-end cost.
  3. If you use the old Vision-Exp path, verify image behavior and tool workflows against V4.1 Flash rather than assuming exact compatibility.
  4. Treat provider compatibility aliases as mutable routing layers and record observed model versions in production telemetry where possible.
  5. If you depend on V4 Pro, verify the current model identity and actual billing; do not assume the original September 14 reroute occurred.

What changed

On September 10, 2026, DeepSeek released V4.1 Flash, a 552B MoE with 8B active input and 16B active output parameters, native vision, new architecture and lower prices. Legacy deepseek-v4-flash and deepseek-v4-flash-vision-exp aliases route to V4.1 Flash. The original release announcement also said V4 Pro would reroute to Flash from September 14, but the current official API Change Log says V4 Pro service continues with unchanged billing, and the live pricing page lists V4-Pro-0813 separately. The two primary documents contradict one another on Pro; do not assume the Pro reroute happened.

Why it matters

The Flash migration is real for users of the old Flash aliases: their backend model and pricing have changed. Builders should re-benchmark behavior, tool use, vision, latency and cost. But teams using V4 Pro should not conflate the confirmed Flash alias migration with an unverified Pro migration: current provider documentation shows Pro continuing separately.

V4.1 Flash replaces the old split between text Flash and Vision-Exp

DeepSeek V4.1 Flash now combines native multimodal understanding with the Flash family’s text and agent use cases. DeepSeek says the older V4 Flash and V4 Flash Vision experimental compatibility names temporarily route to the new model, so existing integrations may receive different model behavior without a model-string change.

The new architecture is designed around cheaper serving

V4.1 Flash is a 552B MoE model, but DeepSeek says only 8B parameters are active for input and 16B for output. The company also says its KV cache needs one quarter of the HBM and one eighth of the SSD storage required by the prior generation. Those are vendor claims, but they explain the lower pricing and higher-throughput positioning.

Pricing drops while time-of-day pricing remains

DeepSeek says V4.1 Flash launched with lower API rates on September 10 and retains peak/off-peak pricing, with off-peak rates at half peak rates. Builders with deferrable workloads can still combine model choice, caching and scheduling as cost-control levers.

V4 Pro remains listed separately despite the older reroute announcement

The September 10 release announcement says Pro requests would move to Flash from September 14. DeepSeek's current API Change Log instead says it decided to continue Pro service with unchanged billing; the live Models & Pricing page lists deepseek-v4-pro as V4-Pro-0813 with distinct rates. Without authenticated request traces, the safe statement is that the official documents conflict and current operational docs retain Pro.

Existing users should treat this as a backend migration

Because compatibility model names can point to a new model family, regression testing should cover instruction following, tool use, multimodal behavior, latency, token consumption and task success. The safe assumption is that a stable model string does not guarantee a stable underlying model.

What to watch next

  • DeepSeek reconciling the older V4 Pro reroute announcement with its current API changelog.
  • V4.1 Pro release and pricing, and whether it replaces V4-Pro-0813.
  • The duration of retired Flash compatibility aliases.
  • Independent V4.1 Flash evaluations and deployment support.
  • Changes to peak/off-peak rates.

Still unclear

  • Architecture and benchmark claims are vendor-reported.
  • Retired Flash aliases definitely route to V4.1 Flash, but DeepSeek's official pages contradict one another on the V4 Pro endpoint.
  • No authenticated Pro request or invoice was independently tested.
  • V4.1 Pro launch timing remains unannounced.

Sources

Direct reading behind this dossier.

3 sources
Models & Pricing
DeepSeek primary

Live model and pricing table lists V4-Pro-0813 as a separate Pro endpoint and confirms old Flash aliases route to V4.1 Flash.

DeepSeek API Change Log
DeepSeek primary

Current September 10 entry says V4 Pro service continues after September 14 with unchanged billing; conflicts with older release announcement.

Discussion

Discussion is reader-contributed. Comments are not part of the BTN dossier or its editorial evidence.

0 visible comments

Join the discussion

Keep comments useful and relevant. Reader contributions may be moderated and are not BTN editorial evidence.

Sign in to comment