DeepSeek V4.1 Flash Launches With Lower Prices and Native Vision

DeepSeek has officially introduced V4.1 Flash, replacing its previous V4 Flash and V4 Flash Vision Experimental models while cutting API costs and adding native image understanding.

TechnologyNews Info Wire4 min read
DeepSeek V4.1 Flash Launches With Lower Prices and Native Vision

DeepSeek has officially introduced V4.1 Flash, replacing its previous V4 Flash and V4 Flash Vision Experimental models while cutting API costs and adding native image understanding.

Article outline

  1. What happened
  2. What comes next
  3. The key numbers
  4. Background
  5. The details
  6. The bottom line

Key points

  • While Terminal Bench 2.1 rose from 82.7 to 90.6, its DeepSWE score rose from 54.4 on V4 Flash to 74.2.
  • V4.1 Flash uses a redesigned architecture with around 748 billion total parameters, including a 552B backbone and 196B Engram parameters.
  • Starting September 14 at 04: 00 UTC, requests sent to deepseek-v4-pro will automatically be redirected to V4.1 Flash and charged at the new Flash rates.
  • Anthropic's Secret Model is Even Better Than Mythos 5 But You Won't Obtain It.
  • V4.1 Flash additionally scored 88.1 on CyberGym and 54.8 on AutomationBench.

Notably, the new model became available on September 10, 2026, under the API name deepseek-flash. Older Flash model names will continue to work as aliases but now route requests to V4.1 Flash. Faster Architecture and Native Vision.

Despite its size, only around 8B parameters per token are active during prefill and 16B during decoding, helping keep inference costs lower.

In practice, the model additionally introduces native multimodal capabilities. Its DeepSeek-ViT encoder can process images alongside text, with backing for resolutions of up to roughly 1344 × 1344 pixels.

V4.1 Flash supports a context window of up to 1 million tokens and output lengths of up to 384, 000 tokens.

Developers can additionally adjust reasoning effort from 1 to 100, allowing them to balance cost and accuracy.

Anthropic's Secret Model is Even Better Than Mythos 5 But You Won't Obtain It. Specification DeepSeek V4.1 Flash. Release Date September 10, 2026. API Model Name deepseek-flash. Total Parameters Approx. 748B. Active Parameters ~8B during prefill, ~16B during decode. Architecture Mixture-of-Experts, Causal Encoder-Decoder. Context Window Up to 1 million tokens. Maximum Output Up to 384K tokens. Vision Backing Native multimodal image understanding. Attention System Compressed Sparse Attention 2. Reasoning Control Adjustable from 1-100. Training Data 45 trillion multimodal tokens. License MIT, open weights. Models Replaced V4 Flash, V4 Flash Vision Exp.

Next Model V4.1 Pro planned. Stronger Agent and Coding Performance.

DeepSeek's own benchmarks show major improvements over V4 Flash and, in a number of areas, V4 Pro.

Independent testing has additionally been positive. While Vals.ai provided the model a 57.86% Vals Index, placing it at the top among open-weight models in that comparison, artificial Analysis documented a 68.9% score on AutomationBench-AA.

Nevertheless, V4.1 Flash does not lead every benchmark. It trails some frontier models on newer Terminal Bench tests and performs poorly on certain specialized evaluations such as SRE Bench and Harvey's Legal Agent benchmark. DeepSeek has additionally reduced Flash pricing significantly.

During off-peak hours, cached input now costs $0.003 per million tokens, uncached input costs $0.15, and output costs $0.60.

Peak pricing doubles those figures to $0.006 for cached input, $0.30 for uncached input, and $1.20 for output per million tokens.

In practice, the biggest reduction applies to cached input. It should particularly benefit agent workloads that repeatedly reuse sizeable context windows.

New DeepSeek V4 Pro is Almost as Good as Claude Fable 5 and Super Cheap. V4 Pro Is Additionally Being Retired. DeepSeek is additionally preparing to phase out V4 Pro.

That arrangement will remain in place until V4.1 Pro becomes available.

Although the new architecture may behave differently in areas such as prompting, tool calling, and response style, the migration could substantially reduce costs for existing V4 Pro users. Stay Connected with ProPakistani.

Obtain the latest tech news, telecom insights, and product launches wherever you prefer. Follow on Google Discover.

Add as a preferredSource on Google Follow on Google News Join WhatsApp.

Add ProPakistani to Preferred Sources and see more of our stories in Google Search and Top Stories.

Technology and Automotive Specialist covering the latest cars, smartphones, AI breakthroughs, and.

Taken together, the developments around deepSeek V4.1 Flash Launches With Lower Prices and Native Vision point to a situation that is still moving, and the coming days should bring more clarity.

Leave a Reply

Your email address will not be published. Required fields are marked *