Sunday, July 26, 2026

Today’s Edition

AI Intel Report

MARKETS

Frontier Models

DeepSeek V4 Enters General Availability with Peak API Pricing

The official launch introduces optimizations for V4-Pro and V4-Flash while retiring legacy endpoints and applying new pricing structures to the open-weight models.

4 MIN READ
A vast expansive interior of a modern high security data center facility features multiple parallel rows of tall black server racks extending into the distance each rack densely packed with visible networking cables power supplies and cooling vents humming with operational activity several anonymous technicians wearing neutral colored lab coats and safety gloves stand at various positions along the aisles performing routine hardware inspections one technician kneels to examine connections at the base of a rack while another stands on a small platform accessing upper modules with tools in hand the floor consists of raised perforated tiles allowing airflow beneath the entire setup overhead industrial lighting illuminates the space evenly with visible dust particles in the air and environmental monitoring sensors mounted on walls the scene captures the infrastructure supporting large scale AI model deployments including optimized versions for high performance inference and rapid response tasks alongside the transition away from older systems through the presence of upgraded hardware configurations and efficient cable management representing new pricing models applied to accessible open weight AI systems the technicians move methodically among the equipment symbolizing the rollout of general availability for advanced AI APIs with focus on V4 series enhancements for professional and accelerated variants the environment includes backup power units visible through open rack doors bundles of fiber optic cables color coded for organization cooling fans spinning steadily and distant views of additional server banks indicating massive computational capacity dedicated to handling peak API demands in a controlled corporate technology environment with no visible markings or identifiers on any surfaces emphasizing the physical reality of scaling AI services from development to widespread commercial access through optimized hardware deployments and legacy system replacements the detailed view shows precise arrangements of components like GPU modules memory banks and interconnects all contributing to reliable service delivery for users accessing these models globally.
Illustration: AI Intel Report

DeepSeek V4 is a series of frontier AI models from DeepSeek that utilize mixture-of-experts architecture to deliver high performance with efficient parameter activation and one million token context support.

DeepSeek moved its V4 series from preview to official general availability in mid-July 2026. The release incorporates feature optimizations and performance improvements across the models. Peak-hour API pricing entered effect at the same time as the official launch.

What background led to the DeepSeek V4 official release?

DeepSeek launched the preview version of its V4 series on April 24, 2026. The preview made V4-Pro and V4-Flash available to users through the API and as open weights. Both models received immediate support for one million token context length by default. The preview announcement highlighted the cost-effective nature of the one million token context capability.

The preview established the mixture-of-experts design for the series. DeepSeek positioned the models as strong performers in the frontier category. Users accessed the models through the DeepSeek API during the preview period. The company indicated plans for an official version during the preview phase.

What technical specifications define the V4 models?

DeepSeek-V4-Pro contains 1.6 trillion total parameters with 49 billion active parameters per token. DeepSeek-V4-Flash contains 284 billion total parameters with 13 billion active parameters per token. The design allows efficient computation while maintaining high capacity through the mixture-of-experts structure.

Key specifications of DeepSeek V4 model variants released in 2026
ModelTotal ParametersActive ParametersContext LengthLicense
DeepSeek-V4-Pro1.6T49B1M tokensMIT
DeepSeek-V4-Flash284B13B1M tokensMIT

The models operate under the MIT license. Open weights for both variants became available on the Hugging Face platform. This release method enables direct download and local deployment by developers and researchers.

What timeline applies to the API changes and legacy retirement?

  1. April 24, 2026: Preview launch of V4-Pro and V4-Flash models with one million token context.
  2. Mid-July 2026: Official general availability release with feature optimizations and peak-hour API pricing introduction.
  3. July 24, 2026 at 15:59 UTC: Discontinuation of legacy API aliases deepseek-chat and deepseek-reasoner.

The legacy aliases previously routed traffic to V4-Flash modes. Applications using the old names required updates before the July 24, 2026 cutoff. The change log from DeepSeek documented the three-month notice period starting from the April preview date.

What performance results does DeepSeek-V4-Pro show on benchmarks?

The 80.6 percent result applies specifically to the Max reasoning mode configuration. This benchmark measures software engineering task resolution. The score reflects the model's capability in coding-related evaluations.

The official version of DeepSeek V4 is planned to launch in mid-July.DeepSeek, Company

How does the release affect market economics for open-weight models?

The MIT license and open weights place the models in direct competition with other open frontier releases. Peak and off-peak API pricing creates differentiated cost structures for high-demand periods. The combination of large context length and efficient active parameters supports broader adoption in cost-sensitive applications.

Stakeholders in the AI industry gain additional options for high-context workloads. The retirement of legacy endpoints standardizes access through the current model names. Developers must adjust integration points to maintain continuity after the July 24, 2026 date.

What implications arise for users and the industry?

Users benefit from the one million token context in both Pro and Flash variants without additional configuration. The open release supports fine-tuning and modification by third parties. Performance improvements in the official version build on the preview foundation.

The discontinuation enforces migration to supported endpoints. Peak-hour pricing introduces planning requirements for API usage patterns. The overall release strengthens the position of open-weight models in the frontier segment.

What developments may follow the DeepSeek V4 launch?

Further iterations could incorporate additional optimizations based on user feedback from the official release. The community can contribute to the models through the open weights on Hugging Face. Continued competition in the open-weight space may influence pricing strategies across providers.

Frequently asked

When did the DeepSeek V4 preview launch?

The preview launched on April 24, 2026, introducing V4-Pro and V4-Flash with one million token context support.

When were the legacy API aliases discontinued?

The legacy API aliases deepseek-chat and deepseek-reasoner were discontinued on July 24, 2026, at 15:59 UTC.

What license governs the DeepSeek V4 models?

The models are released under the MIT license with open weights available on Hugging Face.

Sources

  1. DeepSeek — The two legacy API model names, deepseek-chat and deepseek-reasoner, will be discontinued in three months (2026-07-24).
  2. DeepSeek-AI — We present a preview version of DeepSeek-V4 series, including two strong Mixture-of-Experts (MoE) language models — DeepSeek-V4-Pro with 1.6T parameters (49B activated) and DeepSeek-V4-Flash with 284B parameters (13B activated) — both supporting a context length of one million tokens. DeepSeek-V4-Pro resolves 80.6% of tasks on SWE-bench Verified in Max reasoning mode.
  3. DeepSeek — DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length.
  4. ExplainX — The official version of DeepSeek V4 is planned to launch in mid-July.