Frontier Models
DeepSeek-V4.1-Flash Secures 14th Place in Code Arena WebDev at 1620 Points
The September 10 release delivers a 38-point gain over prior variants, multimodal support, and pricing that positions it competitively among open models while routing legacy requests to the new architecture.
DeepSeek-V4.1-Flash is the smallest model in DeepSeek's new architecture family, featuring native visual understanding and released on September 10, 2026.
DeepSeek introduced the V4.1-Flash model on September 10, 2026.
The announcement described the model as smarter, faster, and more efficient than prior versions.
The release targets greater capability along with faster inference and higher throughput.
It forms the base for scaling to larger models in the same family.
What background information explains the DeepSeek-V4.1-Flash launch?
The V4 series previously included separate Flash and Pro variants.
The Pro served as the flagship model before the architecture update.
DeepSeek now retires the older Pro in favor of the new family.
The transition allows the company to consolidate around a single efficient line.
What performance metrics did DeepSeek-V4.1-Flash achieve in Code Arena?
The model scored 1620 points in the WebDev category.
This score places it at approximately 14th overall.
It ranks fourth among open models.
The result sits within 11 points of the third open model, Qwen3.8-Flash-Next.
The prior V4-Flash variant scored 1582 points at rank 20.
The prior V4-Pro variant scored 1580 points at rank 21.
What are the technical specifications of the new model?
The model uses a Mixture-of-Experts design with 552 billion total parameters.
Active parameters stand at 8 billion for input and 16 billion for output.
The architecture follows a Causal Encoder-Decoder structure.
A 1 million token context window is supported.
Multimodal inputs include both text and images for native visual understanding.
| Model | Code Arena Score | Overall Rank | Open Rank | Input Price per M Tokens | Output Price per M Tokens |
|---|---|---|---|---|---|
| DeepSeek-V4.1-Flash | 1620 | 14 | 4 | $0.30 | $1.20 |
| DeepSeek-V4-Flash | 1582 | 20 | N/A | N/A | N/A |
| DeepSeek-V4-Pro | 1580 | 21 | N/A | N/A | N/A |
| Qwen3.8-Flash-Next | ~1631 | N/A | 3 | N/A | N/A |
What rollout timeline applies to the model and legacy variants?
- DeepSeek-V4.1-Flash launched on September 10, 2026 and is accessible via the deepseek-flash model name.
- Legacy V4-Flash variants began routing to the new model immediately upon release.
- V4-Pro requests will route to the Flash model starting on September 14, 2026 at the Flash pricing rates.
- The V4.1-Pro model is scheduled to launch at a later date to complete the family.
How does the pricing of DeepSeek-V4.1-Flash impact the market?
The API pricing stands at 0.30 dollars per million input tokens.
Output tokens cost 1.20 dollars per million.
These rates apply at peak usage.
The structure encourages high volume usage due to the low input cost.
This pricing undercuts many rivals in the open model category.
Developers gain access to advanced capabilities at reduced expense.
The combination with 1M context supports complex applications.
Stakeholders in software development may shift workloads to this option.
Official partners such as WorkBuddy and CodeBuddy have added support.
OpenCode also integrates the model fully.
This ecosystem support accelerates adoption.
The efficiency surge benefits the broader open source community.
What reactions have emerged from experts and platforms regarding the release?
Arena.ai noted the significant improvement over prior variants.
The platform highlighted the 38 point gain and the competitive standing.
The release signals ongoing competition in model efficiency.
DeepSeek-V4.1-Flash by @deepseek_ai just landed ~#14 overall in Code Arena: WebDev with 1620 pts (AutoEval)! Among open models, DeepSeek-V4.1-Flash landed at ~#4 within 11 pts of Qwen3.8-Flash-Next. This release is a significant improvement compared to DeepSeek-V4 variants: +38 pts vs. V4-Flash (High) at #20 (1582 pts) +40 pts vs. V4-Pro (High) #21 (1580 pts)Arena.ai, AI evaluation platform
What developments are anticipated following the V4.1-Flash launch?
The company plans to launch the V4.1-Pro variant in the near term.
Routing changes will standardize access to the new architecture.
Users can expect continued improvements in the series.
The focus remains on efficiency and capability scaling.
The native multimodal support opens new use cases.
Integration with existing tools will expand over time.
The 1M context window enables longer document processing.
Overall the release strengthens the position of open models in competitive benchmarks.
Frequently asked
What score did DeepSeek-V4.1-Flash receive in Code Arena WebDev?
The model received 1620 points. This corresponds to the 14th position overall. It ranks fourth among open models in the benchmark.
Sources
- Arena.ai — DeepSeek-V4.1-Flash scored 1620 pts and ranked 14th overall in Code Arena WebDev.
- DeepSeek — The pricing for DeepSeek-V4.1-Flash is $0.30 per million input tokens and $1.20 per million output tokens.
- DeepSeek — DeepSeek-V4.1-Flash is now live on the DeepSeek API with native multimodal support. Set your model to `deepseek-flash`.
- Hugging Face — DeepSeek-V4.1-Flash is a multimodal Mixture-of-Experts model with 552B backbone parameters and support for contexts of up to one million tokens.
- DeepSeek — 🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient. 🔹 Introducing the smallest model in our new architecture family, with native visual understanding. 🔹 Designed for greater capability, faster…